> Discover all available pages from the documentation index: https://mastra.zisheng.pro/en/llms.txt # ![Ollama Cloud logo](https://models.dev/logos/ollama-cloud.svg)Ollama Cloud Access 20 Ollama Cloud models through Mastra's model router. Authentication is handled automatically using the `OLLAMA_API_KEY` environment variable. Learn more in the [Ollama Cloud documentation](https://docs.ollama.com/cloud). ```bash OLLAMA_API_KEY=your-api-key ``` ```typescript import { Agent } from "@mastra/core/agent"; const agent = new Agent({ id: "my-agent", name: "My Agent", instructions: "You are a helpful assistant", model: "ollama-cloud/deepseek-v4-flash" }); // Generate a response const response = await agent.generate("Hello!"); // Stream a response const stream = await agent.stream("Tell me a story"); for await (const chunk of stream) { console.log(chunk); } ``` > **Info:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [Ollama Cloud documentation](https://docs.ollama.com/cloud) for details. ## Models | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M | | ------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- | | `ollama-cloud/deepseek-v4-flash` | 1.0M | | | | | | — | — | | `ollama-cloud/deepseek-v4-flash:0731` | 1.0M | | | | | | — | — | | `ollama-cloud/deepseek-v4-pro` | 1.0M | | | | | | — | — | | `ollama-cloud/gemma4:31b` | 262K | | | | | | — | — | | `ollama-cloud/glm-5.1` | 203K | | | | | | — | — | | `ollama-cloud/glm-5.2` | 976K | | | | | | — | — | | `ollama-cloud/gpt-oss:120b` | 131K | | | | | | — | — | | `ollama-cloud/gpt-oss:20b` | 131K | | | | | | — | — | | `ollama-cloud/kimi-k2.5` | 262K | | | | | | — | — | | `ollama-cloud/kimi-k2.6` | 262K | | | | | | — | — | | `ollama-cloud/kimi-k2.7-code` | 262K | | | | | | — | — | | `ollama-cloud/kimi-k3` | 1.0M | | | | | | — | — | | `ollama-cloud/minimax-m2.5` | 205K | | | | | | — | — | | `ollama-cloud/minimax-m2.7` | 197K | | | | | | — | — | | `ollama-cloud/minimax-m3` | 512K | | | | | | — | — | | `ollama-cloud/mistral-large-3:675b` | 262K | | | | | | — | — | | `ollama-cloud/nemotron-3-nano:30b` | 1.0M | | | | | | — | — | | `ollama-cloud/nemotron-3-super` | 262K | | | | | | — | — | | `ollama-cloud/nemotron-3-ultra` | 262K | | | | | | — | — | | `ollama-cloud/qwen3.5:397b` | 262K | | | | | | — | — | ## Advanced configuration ### Custom headers ```typescript const agent = new Agent({ id: "custom-agent", name: "custom-agent", model: { url: "https://ollama.com/v1", id: "ollama-cloud/deepseek-v4-flash", apiKey: process.env.OLLAMA_API_KEY, headers: { "X-Custom-Header": "value" } } }); ``` ### Dynamic model selection ```typescript const agent = new Agent({ id: "dynamic-agent", name: "Dynamic Agent", model: ({ requestContext }) => { const useAdvanced = requestContext.task === "complex"; return useAdvanced ? "ollama-cloud/qwen3.5:397b" : "ollama-cloud/deepseek-v4-flash"; } }); ```