InferX
Mastra のモデルルーターを通じて 6 個の InferX モデルを利用できます。認証には INFERX_API_KEY 環境変数が自動的に使用されます。
詳しくは、InferX のドキュメントを参照してください。
.env
INFERX_API_KEY=your-api-key
src/mastra/agents/my-agent.ts
import { Agent } from "@mastra/core/agent";
const agent = new Agent({
id: "my-agent",
name: "My Agent",
instructions: "You are a helpful assistant",
model: "inferx/google/gemma-4-31b-it-fp8"
});
// Generate a response
const response = await agent.generate("Hello!");
// Stream a response
const stream = await agent.stream("Tell me a story");
for await (const chunk of stream) {
console.log(chunk);
}
情報
Mastra は OpenAI 互換の /chat/completions endpoint を使用します。一部の Provider 固有機能は利用できない場合があります。詳しくは、InferX のドキュメントを確認してください。
モデルモデルへの直接リンク
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|---|---|---|---|---|---|---|---|---|
inferx/google/gemma-4-31b-it-fp8 | 262K | — | — | |||||
inferx/qwen/qwen3.5-122b-a10b-nvfp4 | 256K | — | — | |||||
inferx/qwen/qwen3.6-27b-fp8 | 262K | — | — | |||||
inferx/qwen/qwen3.6-35b-a3b-fp8 | 262K | — | — | |||||
inferx/qwen3-coder-next-fp8 | 256K | — | — | |||||
inferx/qwen3-coder-next-fp8-1m | 1.0M | — | — |
高度な設定高度な設定への直接リンク
カスタムヘッダーカスタムヘッダーへの直接リンク
src/mastra/agents/my-agent.ts
const agent = new Agent({
id: "custom-agent",
name: "custom-agent",
model: {
url: "https://model.inferx.net/endpoints/v1",
id: "inferx/google/gemma-4-31b-it-fp8",
apiKey: process.env.INFERX_API_KEY,
headers: {
"X-Custom-Header": "value"
}
}
});
動的なモデル選択動的なモデル選択への直接リンク
src/mastra/agents/my-agent.ts
const agent = new Agent({
id: "dynamic-agent",
name: "Dynamic Agent",
model: ({ requestContext }) => {
const useAdvanced = requestContext.task === "complex";
return useAdvanced
? "inferx/qwen3-coder-next-fp8-1m"
: "inferx/google/gemma-4-31b-it-fp8";
}
});