Last refreshed 2026-06-29. Next refresh: weekly.
Why use BGE M3 on Novita AI?
Novita AI offers BGE M3 with pay-as-you-go pricing at $0.01/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.
Compare BGE M3 across 2 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: bge-m3bge-m3Request example
Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare BGE M3 Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Cloudflare Workers AI | — | — |
| Novita AI | $0.01 | — |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.01 |
Capabilities
No model capability flags are currently sourced.
About BGE M3
BGE-M3 is BAAI's flagship multilingual embedding model that simultaneously performs dense retrieval, sparse (lexical) retrieval, and multi-vector (ColBERT-style) retrieval. It covers 100+ languages with an 8,192-token context window — far longer than most embedding models — making it effective for both short queries and long documents. Built on an extended XLM-RoBERTa architecture, it achieves state-of-the-art results on the MKQA and MLDR multilingual retrieval benchmarks and is available via NVIDIA NIM.