You can also connect your own LLM — see Bring Your Own LLM. Use the
/v1/providers endpoint to always fetch the latest available model list.Groq
Ultra-low latency inference. Best for real-time conversational agents where response speed is the top priority.Azure OpenAI
Enterprise-grade OpenAI models hosted on Microsoft Azure. Best for teams with compliance, data residency, or enterprise SLA requirements.Open Router
Access to open-source models via a unified routing layer. Best for flexibility and experimenting with open-weight models.Livekit Inference
Edge-optimised inference via the Livekit infrastructure. Best for low-latency deployments close to the media layer.Choosing a Model
What’s Next?
Bring Your Own LLM
Connect a self-hosted, fine-tuned, or third-party LLM to your agents.
Prompting Strategies
Write system prompts that get the best out of any model.
Providers
See all STT, LLM, and TTS providers available on TruGen.