Baichuan Models by Baichuan Intelligent Technology
Last refreshed 2026-05-19. Next refresh: weekly.
Details
About
Baichuan 2, developed by Baichuan Intelligent Technology, is an advanced series of open-source large language models. It has been trained on a substantial corpus of 2.6 trillion tokens, showing remarkable performance across various Chinese and English benchmarks. The model family comprises 7B and 13B parameter versions, offered in both base and chat configurations, and includes a 4-bit quantized chat model for efficient deployment. These models are available for both academic research and commercial use under an official license. Leveraging PyTorch 2.0's F.scaled_dot_product_attention, they ensure faster inference. Additionally, intermediate checkpoints from its training are provided, supporting research into model training dynamics 23.
Decision facts
- Best fit
- coding
- Capability starting point
- Baichuan 13B with 4k context
- Lowest tracked input
- Not tracked
- Closest related family
- Baichuan 4
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs 4k context and 13B parameters.
Use when the workload needs 4k context and 13B parameters.
Use when the workload needs 4k context and 7B parameters.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Baichuan 13B Chat | Use when the workload needs 4k context and 13B parameters. | 2023-06 | 4k context13B parameters | Current |
| Baichuan 13B | Use when the workload needs 4k context and 13B parameters. | 2023-06 | 4k context13B parameters | Current |
| Baichuan 7B | Use when the workload needs 4k context and 7B parameters. | 2023-06 | 4k context7B parameters | Current |
Release Timeline
1 release groupSpecifications(3 models)
| Model | Released | Context | Parameters |
|---|---|---|---|
| Baichuan 13B Chat | 2023-06 | 4k | 13B |
| Baichuan 13B | 2023-06 | 4k | 13B |
| Baichuan 7B | 2023-06 | 4k | 7B |



