LLaDA2.2 Models by InclusionAI
Last refreshed 2026-09-07. Next refresh: weekly.
Details
Capabilities
About
LLaDA2.2 is InclusionAI's agentic MoE diffusion language-model series with Levenshtein Editing (DELETE/INSERT control tokens), Block Routing for long-context MoE diffusion, and L-EBPO agentic RL. First-party Hugging Face ships LLaDA2.2-flash (100B non-embedding) and LLaDA2.2-mini (16B / ~1.4B active), both with 128K context under Apache 2.0. Tech report and LICENSE live in github.com/inclusionAI/LLaDA2.X. Compare it for Agents, Coding, and Long context.
Decision facts
- Best fit
- agenticagentcoding
- Capability starting point
- LLaDA2.2 Mini with 131k context and JSON / Tool use
- Lowest tracked input
- Not tracked
- Closest related family
- K2 Horizon
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs agentic, 131k context, and JSON / Tool use.
Use when the workload needs agentic, 131k context, and 100B parameters.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| LLaDA2.2 Mini | Use when the workload needs agentic, 131k context, and JSON / Tool use. | 2026-09 | agentic131k contextJSON / Tool use | Current |
| LLaDA2.2 Flash | Use when the workload needs agentic, 131k context, and 100B parameters. | 2026-07 | agentic131k context100B parameters | Current |
Release Timeline
2 release groupsSpecifications(2 models)
| Model | Released | Context | Parameters | JSON / Tool use |
|---|---|---|---|---|
| LLaDA2.2 Mini | 2026-09 | 131k | 16B (1.4B active) | Yes |
| LLaDA2.2 Flash | 2026-07 | 131k | 100B | Yes |