Cogito Models by Deep Cogito

Deep CogitoLlama 3 CommunityOpen weights
Deep Cogito releases · 1 in the last 12 monthsChangelog →
10 models2025Up to 128k ctxFrom $0.100/1M input

Last refreshed 2026-06-29. Next refresh: weekly.

Details

ResearcherDeep Cogito
Commercial useCommercial use: conditional
Models10
Released2025
Max context128k

Capabilities

ReasoningAll models
JSON / Tool useAll models
Structured Outputs6 of 10 models
Code Execution1 of 10 models

About

The Cogito family is a series of hybrid open-weight reasoning models from Deep Cogito, trained with Iterated Distillation and Amplification (IDA). Models span 3B to 671B parameters, support both direct and extended-thinking (reasoning) modes, and are fine-tuned from Llama and Qwen base checkpoints (v1 Preview) and DeepSeek V3 Base (v2.1). Available via Fireworks AI, Together AI, and other inference providers.

Decision facts

Best fit
reasoningJSON / Tool usestructured outputs
Capability starting point
Cogito v2.1 671B with 128k context and reasoning, JSON / Tool use, and structured outputs
Lowest tracked input
Cogito v1 Preview Llama 3B · $0.100/1M · Fireworks AI
Closest related family
360Zhinao

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

10 in view

Use when the workload needs 128k context, 671B parameters, and reasoning.

2025-11128k context671B parametersreasoning

Use when the workload needs 671B parameters, reasoning, and JSON / Tool use.

2025-07671B parametersreasoningJSON / Tool use

Use when the workload needs 109B parameters, reasoning, and JSON / Tool use.

2025-07109B parametersreasoningJSON / Tool use

Use when the workload needs 405B parameters, reasoning, and JSON / Tool use.

2025-07405B parametersreasoningJSON / Tool use

Use when the workload needs 70B parameters, reasoning, and JSON / Tool use.

2025-0770B parametersreasoningJSON / Tool use

Use when the workload needs 128k context, 3B parameters, and reasoning.

2025-04128k context3B parametersreasoning

Use when the workload needs 128k context, 70B parameters, and reasoning.

2025-04128k context70B parametersreasoning

Use when the workload needs 128k context, 8B parameters, and reasoning.

2025-04128k context8B parametersreasoning

Use when the workload needs 128k context, 14B parameters, and reasoning.

2025-04128k context14B parametersreasoning

Use when the workload needs 128k context, 32B parameters, and reasoning.

2025-04128k context32B parametersreasoning

Release Timeline

3 release groups
2025-11
1 current
Cogito v2.1 671B
128k context671B parametersreasoning
Current
2025-07
4 current
Cogito v2 Preview DeepSeek 671B MoE
671B parametersreasoningJSON / Tool use
Current
Cogito v2 Preview Llama 109B MoE
109B parametersreasoningJSON / Tool use
Current
Cogito v2 Preview Llama 405B
405B parametersreasoningJSON / Tool use
Current
Cogito v2 Preview Llama 70B
70B parametersreasoningJSON / Tool use
Current
2025-04
5 current
Cogito v1 Preview Llama 3B
128k context3B parametersreasoning
Current
Cogito v1 Preview Llama 70B
128k context70B parametersreasoning
Current
Cogito v1 Preview Llama 8B
128k context8B parametersreasoning
Current
Cogito v1 Preview Qwen-14B
128k context14B parametersreasoning
Current
Cogito v1 Preview Qwen-32B
128k context32B parametersreasoning
Current

Specifications(10 models)

Cogito model specifications comparison
ModelReleasedContextParametersReasoningJSON / Tool useStructured OutputsCode Exec
Cogito v2.1 671B2025-11128k671BYesYesYesYes
Cogito v2 Preview DeepSeek 671B MoE2025-07—671BYesYesNoNo
Cogito v2 Preview Llama 109B MoE2025-07—109BYesYesNoNo
Cogito v2 Preview Llama 405B2025-07—405BYesYesNoNo
Cogito v2 Preview Llama 70B2025-07—70BYesYesNoNo
Cogito v1 Preview Llama 3B2025-04128k3BYesYesYesNo
Cogito v1 Preview Llama 70B2025-04128k70BYesYesYesNo
Cogito v1 Preview Llama 8B2025-04128k8BYesYesYesNo
Cogito v1 Preview Qwen-14B2025-04128k14BYesYesYesNo
Cogito v1 Preview Qwen-32B2025-04128k32BYesYesYesNo

Available From(1 provider)

Pricing

Cogito model pricing by provider
ModelProviderInput / 1MOutput / 1MType
Cogito v1 Preview Llama 3BFireworks AI$0.1$0.1Serverless
Cogito v1 Preview Llama 8BFireworks AI$0.2$0.2Serverless
Cogito v1 Preview Qwen-14BFireworks AI$0.2$0.2Serverless
Cogito v1 Preview Llama 70BFireworks AI$0.9$0.9Serverless
Cogito v1 Preview Qwen-32BFireworks AI$0.9$0.9Serverless