LLM Reference

QuiverAI

Researched todayAI LabTier 3
RAGAgentsLong contextVisionJSON / Tool use

QuiverAI offers 2 tracked models (2 with output token pricing). This catalog covers rag, agents, and long context; open any model detail page for benchmarks, batch tiers, and migration prompts.

Covers 5 workload areas across 2 tracked models; last verified 2026-09-17.

Use it for

  • Teams comparing token and batch pricing across this provider's models
  • Operators routing rag, agents, and long context workloads through this API

Do not use it for

  • Final benchmark picks without opening the relevant model detail page

Tracked models

2

Models available through this provider

Priced output routes

2

Models with output token pricing tracked

Cheapest output

$20.00

Arrow 2 on this route

Batch-ready models

0

No batch pricing tracked

Latest model release

2026-09-07

10d since newest release

Freshness

2026-09-17

Researched today

fresh

Information

TypeAI Lab
TierTier 3
Models2
San Francisco, California, United States

First-party QuiverAI hosted inference for Arrow models via App and API Platform. API base https://api.quiver.ai/v1 (Bearer API key). Docs: OpenResponses POST /v1/responses plus native SVG endpoints (generations, vectorizations, edits, animations). Arrow 2 / Arrow 2 Telos generally available per docs; token_usage billing. Prepaid API balance separate from App subscription credits.

Read more ->

Where this host wins

  • RAG: 2 tracked models with ruler / needle retrieval benchmarks.
  • Agentic: 2 tracked models with BFCL, tau-bench, and SWE-bench tool-use coverage.
  • Long-context: 2 tracked models with context-token or InfiniteBench-class signal.
  • Vision: 2 tracked models with multimodal benchmark coverage.

Getting started

Verify: quotas and regions in the linked vendor documentation.

Platform Overview

First-party QuiverAI hosted inference for Arrow models via App and API Platform. API base https://api.quiver.ai/v1 (Bearer API key). Docs: OpenResponses POST /v1/responses plus native SVG endpoints (generations, vectorizations, edits, animations). Arrow 2 / Arrow 2 Telos generally available per docs; token_usage billing. Prepaid API balance separate from App subscription credits.

Available Models(2)

View all →

All models available as Serverless

ModelInput (per 1M)Output (per 1M)
Arrow 2$4.00$20.00
Arrow 2 Telos$6.00$30.00

Where else to run this