LLM Reference
Microsoft Foundry

Cohere Rerank v4.0 Fast on Microsoft Foundry

Cohere Rerank · Cohere

Serverless

Last refreshed 2026-07-01. Next refresh: weekly.

Why use Cohere Rerank v4.0 Fast on Microsoft Foundry?

Microsoft Foundry offers Cohere Rerank v4.0 Fast with competitive pricing. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.

Compare Cohere Rerank v4.0 Fast across 2 providers to find the best fit for your use case
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: cohere-rerank-v4-fast
Model ID
cohere-rerank-v4-fast

Request example

Curated snippets for this provider are not sourced yet. Use Microsoft Foundry documentation with model ID cohere-rerank-v4-fast.

Gotchas

  • Use provider model ID "cohere-rerank-v4-fast", not the LLMReference slug "cohere-rerank-v4-0-fast".

Compare Cohere Rerank v4.0 Fast Across Providers

ProviderInput (per 1M)Output (per 1M)
Microsoft Foundry——
Cohere API——

Pricing

TypePrice (per 1M)
Query$2.00

Capabilities

No model capability flags are currently sourced.

About Cohere Rerank v4.0 Fast

Fast variant reranking model optimized for low latency and high throughput. Multilingual support for reranking English and non-English documents and semi-structured data (JSON). Provides good quality at faster inference speeds than the pro variant.

Get Started

Model Specs

Released2025-04-01
Context32k
ArchitectureEncoder Only

Related Models on Microsoft Foundry