SubQ 1M-Preview
SubQ 1M-Preview is worth evaluating for rag, agents, and long context when its provider route and context window match the workload.
Use it for
- Teams evaluating rag, agents, and long context
- Workloads that can use a 1m context window
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Family
- SubQ
- Released
- 2026-05-05
- Context
- 1m
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
- Training
- Pretrained
Cheapest of 1 route · SubQ API
About
SubQ 1M-Preview is Subquadratic's sub-quadratic sparse-attention LLM: linear O(n) attention versus O(n²) standard transformers, with 1M production context tested to 12M. Subquadratic claims roughly 50× faster/cheaper inference at long context—vendor-claimed, not independently verified. The model posts vendor-claimed SWE-bench Verified 81.8%, RULER 95.0%, and MRCR 65.9% in lab prose; LLMReference has no independent benchmark rows yet. OpenAI-compatible API access is available via SubQ AI.
Top use-case fit: coding, agents, and build tasks
RAG
Included by capability and metadata signals in the decision map.
Agents
Included by capability and metadata signals in the decision map.
Long context
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| SubQ API | - | - | ServerlessPartial |
Capabilities
Benchmark peer barsfor RAG
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Cheapest of 1 route · SubQ API