LLM Reference
SubQ API

Using SubQ 1M-Preview on SubQ API

Implementation guide · SubQ · Subquadratic

Serverless

SubQ API exposes SubQ 1M-Preview through model ID subq-1m-preview. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-30. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at SubQ API and generate an API key.
  2. 2
    Use the SubQ API SDK or REST API to call subq-1m-preview — see the documentation for request format.

Code Examples

See SubQ API documentation for integration details.

Pricing on SubQ API

Capabilities

ReasoningJSON / Tool use

About SubQ 1M-Preview

SubQ 1M-Preview is Subquadratic's first large language model, built on a fully sub-quadratic sparse-attention architecture that scales compute linearly with context length (O(n) vs. traditional O(n²)). Supports a production context window of 1M tokens (architecture tested to 12M). Achieves 81.8% on SWE-Bench Verified, 95.0% on RULER @128K, and 65.9% on MRCR v2 (8-needle, 1M). Claims 50x faster and 50x cheaper than leading frontier models at 1M context length. Available via OpenAI-compatible API with streaming and tool use support. Model is proprietary and not open-source; fine-tuning for customer-specific use cases is mentioned as a future capability.

Model Specs

Released2026-05-05
Context1m
ArchitectureDecoder Only

Provider

SubQ API
SubQ API

Subquadratic