Last refreshed 2026-07-01. Next refresh: weekly.
Why use Cohere Rerank v4.0 Fast on Cohere API?
Cohere API offers Cohere Rerank v4.0 Fast with competitive pricing. Cohere is a leading enterprise AI company that specializes in developing large language models (LLMs) and Retrieval-Augmented Generation (RAG) capabilities.
Compare Cohere Rerank v4.0 Fast across 2 providers to find the best fit for your use caseInput / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced
Setup recipe
Docs fallbackInstall
Use the provider REST API or SDKAuth
Create a provider API keyCall
model: rerank-v4.0-fastModel ID
rerank-v4.0-fastRequest example
Curated snippets for this provider are not sourced yet. Use Cohere API documentation with model ID
rerank-v4.0-fast.Gotchas
- Use provider model ID "rerank-v4.0-fast", not the LLMReference slug "cohere-rerank-v4-0-fast".
Compare Cohere Rerank v4.0 Fast Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Microsoft Foundry | — | — |
| Cohere API | — | — |
Capabilities
No model capability flags are currently sourced.
About Cohere Rerank v4.0 Fast
Fast variant reranking model optimized for low latency and high throughput. Multilingual support for reranking English and non-English documents and semi-structured data (JSON). Provides good quality at faster inference speeds than the pro variant.
Get Started
Model Specs
Released2025-04-01
Context32k
ArchitectureEncoder Only