> ## Documentation Index
> Fetch the complete documentation index at: https://phidatainc.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# Models

> Examples for all supported LLM providers in Agno.

| Example                                                      | Description                                                                                                                                     |
| ------------------------------------------------------------ | ----------------------------------------------------------------------------------------------------------------------------------------------- |
| [Aimlapi](/examples/models/aimlapi/overview)                 | AIML API examples for basic runs, multimodal input, memory, retries, structured output, and tool use.                                           |
| [Anthropic](/examples/models/anthropic/overview)             | Claude examples for multimodal input, context management, caching, knowledge, memory, thinking, structured output, server tools, and skills.    |
| [AWS](/examples/models/aws/overview)                         | Run Claude and Amazon Nova models on AWS Bedrock.                                                                                               |
| [Azure](/examples/models/azure/overview)                     | Run Claude and open-source models on Azure AI Foundry and OpenAI endpoints.                                                                     |
| [Cerebras](/examples/models/cerebras/overview)               | Cerebras examples: basic runs, storage, knowledge, structured output, retries, and tool use.                                                    |
| [Cerebras OpenAI](/examples/models/cerebras-openai/overview) | Run Cerebras models through the OpenAI-compatible endpoint: streaming, tools, structured output, storage, and knowledge.                        |
| [Clients](/examples/models/clients/overview)                 | Configure Agno's default sync httpx.Client with headers, logging, request IDs, timeouts, and error tracking.                                    |
| [Cohere](/examples/models/cohere/overview)                   | Run Cohere Command A and Aya Vision models with tools, knowledge, memory, retries, and structured output.                                       |
| [Cometapi](/examples/models/cometapi/overview)               | Run GPT, Claude, Gemini, DeepSeek, and Qwen models through CometAPI's OpenAI-compatible gateway.                                                |
| [Dashscope](/examples/models/dashscope/overview)             | Browse DashScope model examples with Qwen models, image analysis, knowledge tools, and retry patterns.                                          |
| [DeepInfra](/examples/models/deepinfra/overview)             | DeepInfra examples: basic agent runs, JSON output, tool use, and retries.                                                                       |
| [DeepSeek](/examples/models/deepseek/overview)               | Run DeepSeek models with reasoning, thinking mode, structured output, retries, and tool use.                                                    |
| [Fireworks](/examples/models/fireworks/overview)             | Run Fireworks models with streaming, structured output, web search, and retry configuration.                                                    |
| [Google](/examples/models/google/overview)                   | Use Gemini for audio, video, image, PDF, grounding, file search, and thinking-budget examples.                                                  |
| [Groq](/examples/models/groq/overview)                       | Groq examples for agents and teams, multimodal input, knowledge, reasoning, research, transcription, translation, structured output, and tools. |
| [Hugging Face](/examples/models/huggingface/overview)        | Hugging Face examples for basic and streaming runs, essay generation, retries, and web-search tool use.                                         |
| [IBM](/examples/models/ibm/overview)                         | IBM watsonx examples for model retries, storage, knowledge, structured output, and tools.                                                       |
| [Internlm](/examples/models/internlm/overview)               | Run InternLM models with basic responses, tools, knowledge, storage, retries, and structured output.                                            |
| [LangDB](/examples/models/langdb/overview)                   | Run LangDB models with basic responses, tools, retries, and structured output.                                                                  |
| [LiteLLM](/examples/models/litellm/overview)                 | Run agents through the LiteLLM gateway with tools, knowledge, structured output, and audio, image, and PDF input.                               |
| [LiteLLM OpenAI](/examples/models/litellm-openai/overview)   | Examples for LiteLLM with OpenAI-compatible models.                                                                                             |
| [Llama Cpp](/examples/models/llama-cpp/overview)             | Run agents against a local llama.cpp server serving `ggml-org/gpt-oss-20b-GGUF` at `http://127.0.0.1:8080/v1`.                                  |
| [Lmstudio](/examples/models/lmstudio/overview)               | LM Studio examples for local models, images, knowledge, memory, storage, retries, structured output, and tools.                                 |
| [Meta](/examples/models/meta/overview)                       | Llama and Llama OpenAI examples covering tool use, knowledge, memory, metrics, storage, and retries.                                            |
| [Mistral](/examples/models/mistral/overview)                 | Run Mistral models with image input, memory, structured output, retries, and tool use.                                                          |
| [Moonshot](/examples/models/moonshot/overview)               | Moonshot Kimi K2 agent examples: basic sync/streaming responses and web-search tool use.                                                        |
| [N1N](/examples/models/n1n/overview)                         | N1N gateway examples: running OpenAI models via N1N with basic streaming and web-search tool calls.                                             |
| [Nebius](/examples/models/nebius/overview)                   | Nebius model examples: basic runs, Postgres sessions, PgVector knowledge, retries, structured output, and tool use.                             |
| [Neosantara](/examples/models/neosantara/overview)           | Neosantara examples covering basic runs, structured output, and web-search tool use.                                                            |
| [Nexus](/examples/models/nexus/overview)                     | Nexus examples covering basic runs, retry configuration, and tool use.                                                                          |
| [NVIDIA](/examples/models/nvidia/overview)                   | NVIDIA API examples: basic runs, retry configuration, and tool use.                                                                             |
| [Ollama](/examples/models/ollama/overview)                   | Ollama Chat and Responses API examples for local and cloud models, knowledge, memory, reasoning, structured output, and tools.                  |
| [OpenAI](/examples/models/openai/overview)                   | OpenAI Chat and Responses API examples for multimodal input, tools, reasoning, structured output, storage, and streaming.                       |
| [OpenRouter](/examples/models/openrouter/overview)           | OpenRouter Chat and Responses API examples for model routing, retries, structured output, and tools.                                            |
| [Perplexity](/examples/models/perplexity/overview)           | Index of Perplexity sonar-pro agent examples: basic runs, knowledge, memory, retries, structured output, and web search.                        |
| [Portkey](/examples/models/portkey/overview)                 | Index of Agno examples routing agents through the Portkey AI gateway: basic runs, retries, structured output, and tool use.                     |
| [Requesty](/examples/models/requesty/overview)               | Requesty AI is an LLM gateway with AI governance. See their [website](https://www.requesty.ai) for more information.                            |
| [Sambanova](/examples/models/sambanova/overview)             | SambaNova model examples: basic sync/stream/async runs and retry configuration.                                                                 |
| [Siliconflow](/examples/models/siliconflow/overview)         | Examples for SiliconFlow model integration.                                                                                                     |
| [Together](/examples/models/together/overview)               | Run Together models with streaming, image input, reasoning, structured output, web search, and retry configuration.                             |
| [Vercel](/examples/models/vercel/overview)                   | Index of Agno examples running on Vercel's v0 model: basic runs, images, knowledge, retries, and web search.                                    |
| [Vertex AI](/examples/models/vertexai/overview)              | Vertex AI examples for Claude models, retries, multimodal input, knowledge, memory, caching, structured output, and tools.                      |
| [vLLM](/examples/models/vllm/overview)                       | vLLM is a fast and easy-to-use library for running LLM models locally.                                                                          |
| [xAI](/examples/models/xai/overview)                         | xAI model examples for building agents with Grok, including vision, web search, and financial analysis.                                         |
| [Cloudflare](/examples/models/cloudflare/overview)           | Cloudflare AI Gateway model examples.                                                                                                           |
| [Inception](/examples/models/inception/overview)             | Inception Labs Mercury model examples.                                                                                                          |
| [MiniMax](/examples/models/minimax/overview)                 | MiniMax M3 agent examples: basic runs, web search tool use, and JSON-mode structured output.                                                    |
| [Xiaomi MiMo](/examples/models/xiaomi/overview)              | Xiaomi MiMo model examples.                                                                                                                     |
