LLM Reference

Aquila 2 Models by Beijing Academy of Artificial Intelligence (BAAI)

8 models2023Up to 16k ctx

Last refreshed 2026-05-19. Next refresh: weekly.

Details

LicenseProprietary
Commercial useCommercial use: conditional
Models8
Released2023
Max context16k

About

The Aquila 2 family is a series of bilingual large language models designed to proficiently handle Chinese and English languages. These models vary significantly in size, from 7 billion to 70 billion parameters, offering a broad range for different computational needs. They are developed using the advanced HeuriMentor (HM) framework, which enhances the training process by dynamically adjusting data distributions, thereby boosting efficiency and model performance. This framework includes components such as the Adaptive Training Engine (ATE) and Training State Monitor (TSM). The Aquila 2 models consistently perform well on key benchmarks and have been open-sourced to encourage further innovation. Notably, the Aquila2-34B model retains performance standards even when quantized to Int4 format, highlighting its efficiency and accuracy. The AquilaChat2 variants are specially fine-tuned for conversational applications, demonstrating the versatility of the Aquila 2 family 14.

Decision facts

Best fit
codingchatbot and role-playing use caseslong-context generation
Capability starting point
Aquila Chat 2 34B-16K with 16k context
Lowest tracked input
Not tracked
Closest related family
BGE

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

8 in view

Use when the workload needs 2k context and 70B parameters.

2023-112k context70B parameters

Use when the workload needs 2k context and 34B parameters.

2023-112k context34B parameters

Use when the workload needs 2k context and 7B parameters.

2023-112k context7B parameters

Use when the workload needs 16k context and 34B parameters.

2023-1116k context34B parameters

Use when the workload needs 16k context and 7B parameters.

2023-1116k context7B parameters

Use when the workload needs 2k context and 70B parameters.

2023-112k context70B parameters

Use when the workload needs 2k context and 34B parameters.

2023-112k context34B parameters

Use when the workload needs 2k context and 7B parameters.

2023-112k context7B parameters

Release Timeline

1 release group
2023-11
8 current
Aquila 2 34B
2k context34B parameters
Current
Aquila 2 70B Expressive
2k context70B parameters
Current
Aquila 2 7B
2k context7B parameters
Current
Aquila Chat 2 34B
2k context34B parameters
Current
Aquila Chat 2 34B-16K
16k context34B parameters
Current
Aquila Chat 2 70B Expressive
2k context70B parameters
Current
Aquila Chat 2 7B
2k context7B parameters
Current
Aquila Chat 2 7B-16K
16k context7B parameters
Current

Specifications(8 models)

Aquila 2 model specifications comparison
ModelReleasedContextParameters
Aquila Chat 2 70B Expressive2023-112k70B
Aquila Chat 2 34B2023-112k34B
Aquila Chat 2 7B2023-112k7B
Aquila Chat 2 34B-16K2023-1116k34B
Aquila Chat 2 7B-16K2023-1116k7B
Aquila 2 70B Expressive2023-112k70B
Aquila 2 34B2023-112k34B
Aquila 2 7B2023-112k7B

Popular comparisons in this family