Cohere Rerank v4.0 Fast on Cohere API

Cohere Rerank · Cohere

Serverless

Last refreshed 2026-07-01. Next refresh: weekly.

Why use Cohere Rerank v4.0 Fast on Cohere API?

Cohere API offers Cohere Rerank v4.0 Fast with competitive pricing. Cohere is a leading enterprise AI company that specializes in developing large language models (LLMs) and Retrieval-Augmented Generation (RAG) capabilities.

Compare Cohere Rerank v4.0 Fast across 2 providers to find the best fit for your use case
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: rerank-v4.0-fast
Model ID
rerank-v4.0-fast

Request example

Curated snippets for this provider are not sourced yet. Use Cohere API documentation with model ID rerank-v4.0-fast.

Gotchas

  • Use provider model ID "rerank-v4.0-fast", not the LLMReference slug "cohere-rerank-v4-0-fast".

Compare Cohere Rerank v4.0 Fast Across Providers

ProviderInput (per 1M)Output (per 1M)
Microsoft Foundry——
Cohere API——

Capabilities

No model capability flags are currently sourced.

About Cohere Rerank v4.0 Fast

Fast variant reranking model optimized for low latency and high throughput. Multilingual support for reranking English and non-English documents and semi-structured data (JSON). Provides good quality at faster inference speeds than the pro variant.

Get Started