> For clean Markdown of any page, append .md to the page URL. > For a complete documentation index, see https://docs.jambonz.org/guides/features/bring-your-own-llm/llms.txt. > For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.jambonz.org/_mcp/server. # Bring Your Own LLM ## Docs - [Bring Your Own LLM](https://docs.jambonz.org/guides/features/bring-your-own-llm/overview.md): Configure jambonz to use your own credentials with any of 14 supported LLM providers. - [Anthropic](https://docs.jambonz.org/guides/features/bring-your-own-llm/anthropic.md): Configure jambonz to use Anthropic's Claude models. - [AWS Bedrock](https://docs.jambonz.org/guides/features/bring-your-own-llm/aws-bedrock.md): Configure jambonz to use Anthropic Claude, Meta Llama, Mistral, and Amazon Nova models on AWS Bedrock. - [Azure OpenAI](https://docs.jambonz.org/guides/features/bring-your-own-llm/azure-open-ai.md): Configure jambonz to use OpenAI models hosted in your Azure subscription. - [Baseten](https://docs.jambonz.org/guides/features/bring-your-own-llm/baseten.md): Configure jambonz to use Baseten's hosted open-weight model catalog or a dedicated Baseten deployment. - [DeepSeek](https://docs.jambonz.org/guides/features/bring-your-own-llm/deep-seek.md): Configure jambonz to use DeepSeek's V4 family models. - [Google (AI Studio)](https://docs.jambonz.org/guides/features/bring-your-own-llm/google-ai-studio.md): Configure jambonz to use Google's Gemini models via AI Studio. - [Groq](https://docs.jambonz.org/guides/features/bring-your-own-llm/groq.md): Configure jambonz to use Llama and Gemma models on Groq's LPU hardware for sub-100ms TTFT. - [HuggingFace](https://docs.jambonz.org/guides/features/bring-your-own-llm/hugging-face.md): One HuggingFace token, many backends — Groq, Together, Fireworks, Cerebras, Nebius all behind a single broker endpoint. - [Moonshot (Kimi)](https://docs.jambonz.org/guides/features/bring-your-own-llm/moonshot-kimi.md): Configure jambonz to use Moonshot AI's Kimi models over its OpenAI-compatible API. - [OpenAI](https://docs.jambonz.org/guides/features/bring-your-own-llm/open-ai.md): Configure jambonz to use OpenAI's chat completions API. - [Vertex AI — Gemini](https://docs.jambonz.org/guides/features/bring-your-own-llm/vertex-ai-gemini.md): Configure jambonz to use Gemini models via Google Cloud's Vertex AI. - [Vertex AI — Partner Models](https://docs.jambonz.org/guides/features/bring-your-own-llm/vertex-ai-partner-models.md): Configure jambonz to use Llama, Mistral, and other partner models hosted on Vertex AI's OpenAI-compatible endpoint. - [Z.ai (GLM)](https://docs.jambonz.org/guides/features/bring-your-own-llm/z-ai-glm.md): Configure jambonz to use Z.ai's GLM models over its OpenAI-compatible API. - [xAI (Grok)](https://docs.jambonz.org/guides/features/bring-your-own-llm/x-ai-grok.md): Configure jambonz to use xAI's Grok models — reasoning + vision over an OpenAI-compatible API.