1024-dimension embeddings across 100+ languages, built for retrieval that has to work in more than English. The 8K window takes a whole page without chunking. Reach for it when the corpus is multilingual and the per-million rate matters.
Modalities
Text → Text
Usage
How much this model is actually called here.
Rank
#73
of 101 active models
Tokens served
72.8K
all-time
Platform share
0.0%
of all tokens
Pricing
Pay per token. Cached input bills at this model’s own cached rate, listed below.
+ Estimate your workload− Estimate your workload
Estimated monthly cost
$0.8100
$0.0270 / day on BGE-M3
Same workload on:
Estimates use list pricing. Actual bills depend on real token counts, and every response includes its exact cost.
How it compares
Against the peers people actually weigh it against.
| Spec | BGE-M3 | BGE Large EN v1.5 | BGE Base EN v1.5 |
|---|---|---|---|
| Input /1M | $0.0135 | $0.0135 | $0.00675 |
| Output /1M | $0.00 | $0.00 | $0.00 |
| Context | 8K | 1K | 1K |
| Tools | No | No | No |
| Reasoning | No | No | No |
| Throughput | 1311 tok/s | 1067 tok/s | 1393 tok/s |
Quick start
Up and running in under two minutes.
- 1
Create an API key
Sign up and grab a key from the dashboard — $0.50 free credit on the free tier, which covers BGE-M3. No card required.
Get API key → - 2
Make your first request
Submit a generation job and poll until it succeeds.
curl https://kymaapi.com/v1/embeddings \ -H "Authorization: Bearer YOUR_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "bge-m3", "input": ["first document", "second document"] }'
FAQ
Common questions about this model.
What is the context window of BGE-M3?
How much does the BGE-M3 API cost?
Does BGE-M3 support function calling?
How do I use BGE-M3?
More models by BAAI
| Model | Context | Input | Output |
|---|---|---|---|
| 1K | $0.00675 | $0.00 | |
| 1K | $0.0135 | $0.00 |