Fugu
Decision note
Fugu is an orchestration surface, not a standalone foundation model. Its learned coordinator assigns Thinker, Worker, and Verifier roles across a changing model pool, so evaluate the system on task quality, latency, billing predictability, and vendor control rather than treating its system-level results as one model's benchmark scores.
Fit
- Best fit
- Teams evaluating Agents or Coding workloads that want a managed, OpenAI-compatible multi-agent endpoint without building their own router, role prompts, or verification loop.
- Not for
- Teams that require a named fixed model, full visibility into per-request routing, or a coding agent that edits files and runs tools locally. Fugu is the model-orchestration API behind a client or agent surface, not that client or agent surface itself.
Overview
Sakana AI's hosted multi-agent orchestration API. A learned coordinator routes each request across Thinker, Worker, and Verifier roles drawn from a model pool that standard Fugu customers can constrain by provider or model; Fugu balances latency and quality, while Fugu Ultra coordinates a deeper expert-agent pool to prioritize quality. For pay-as-you-go usage, Sakana bills one active Fugu agent at the selected underlying model's standard rate; multi-agent runs use a single non-stacked rate based on the top-tier participating model.
Best LLMs for this agent
Pricing and Provider Comparison
Agent access tiers
| Tier | Price | Includes |
|---|---|---|
| Standard | $20/mo | Both Fugu and Fugu Ultra; baseline allowance for lightweight daily usage, occasional API calls, and small experiments. |
| Pro | $100/mo | Both Fugu and Fugu Ultra; 10x Standard usage for regular Coding and Agents workloads. |
| Max | $200/mo | Both Fugu and Fugu Ultra; intended for heavy, long-running workloads. Sakana's current pricing surfaces publish conflicting Max allowances, so confirm the allowance before subscribing. |
| Fugu via Sakana AI | Variable by model | Variable pay as you go with a 1M-token context: one active agent is billed at the selected underlying model's standard rate; multiple agents are billed one non-stacked rate based on the top-tier participating model. |
| Fugu Ultra via Sakana AI | $5 in / $30 out per 1M | $0.50/M cached input through 272K context; above 272K, $10/M input, $45/M output, and $1/M cached input. |
| Fugu Ultra via Vercel AI Gateway | $5 in / $30 out per 1M | $0.50/M cached input, 1M-token context, and Vercel pass-through pricing with no gateway markup. |
Featured Model Specs
Model Compatibility Matrix
Underlying Models
Capabilities
| Feature | Supported |
|---|---|
| Chat | Yes |
| Agent mode | Yes |
| Multi-file edit | No |
| Autocomplete | No |
| Terminal | No |
| Code review | No |
| Git integration | No |
| Deploy integration | No |
| Local models | No |
| Bring your own key | No |
| MCP support | No |
| Offline mode | No |
| IP indemnity | No |
| SSO | No |
| On-prem | No |