Dolphin 2.9.2 Qwen2-72B
Dolphin 2.9.2 Qwen2-72B is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.
Use it for
- Teams evaluating rag, agents, and long context
- Workloads that can use a 128k context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Family
- Dolphin
- Released
- 2024-05-27
- Context
- 128k
- Parameters
- 72B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Available
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 1 route · Fireworks AI
About
Dolphin 2.9.2 Qwen2-72B is Cognitive Computations' uncensored full-weight fine-tune of Qwen2-72B. The model card documents ChatML formatting, initial agentic abilities, and function calling support; it also notes the Qwen2 base has a 128K context window while the fine-tune used 8K training sequences. Because Dolphin removes much of the usual alignment filtering, deployments should add their own moderation and policy layer before exposing it as a public service.
Top use-case fit: coding, agents, and build tasks
RAG
Included by capability and metadata signals in the decision map.
Agents
Included by capability and metadata signals in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Fireworks AI | $0.900 | $0.900 | Provisioned |
Available via routers & gateways(1)
Capabilities
Benchmark peer barsfor RAG
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Cheapest of 1 route · Fireworks AI