Qwen3 Embedding
qwen/qwen3-embeddingMultilingual embedding model priced for indexing everything, the budget default for RAG at scale.
Model overview
Context window
32K
tokens
Input price
—
quoted on request
Output price
—
quoted on request
Weekly volume
—
tokens / week
Open weights
No
API only
Capabilities
1
of 11 tags
Description
Qwen3 Embedding is an embeddings model from Alibaba Qwen. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Alibaba Qwen’s stated focus is the broadest open model family in the world. The weights are not published, so it is API-only.
Capabilities
- Batch
- Accepts offline batch jobs, which are processed asynchronously at a reduced rate.
Strengths & weaknesses
Strengths
- Lowest embedding price in the catalogue
- Strong multilingual recall
- Batch-friendly
Weaknesses
- Smaller context than Cohere Embed v4
- Text-only input
Pricing
Free| Rate | Price | Unit |
|---|---|---|
| Input | $0.00 | per 1M tokens |
| Output | — not charged | — |
Batch jobs are supported and settle at a reduced rate against the same unit. Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.
Context window
32Ktokens
covers prompt and response together, so a long input leaves less room for the answer.
Supported features
| Feature | Support |
|---|---|
| Tool calling | Not supported |
| JSON mode | Not supported |
| Streaming | Not supported |
| Vision input | Not supported |
| Audio input | Not supported |
| Long context | Not supported |
| Open weights | Not supported |
| Reasoning | Not supported |
| Fine-tunable | Not supported |
| Batch | Supported |
| Caching | Not supported |
Example use cases
Semantic search
Lowest embedding price in the catalogue
RAG indexing
Strong multilingual recall
Deduplication
Batch-friendly
Comparison
| Attribute | Qwen3 Embedding | Embed v4 | Gemini Embedding |
|---|---|---|---|
| Context window | 32K | 128K | 8K |
| Input price | $0.00 | $0.00 | $0.00 |
| Output price | — not charged | — not charged | — not charged |
| Tool calling | Not supported | Not supported | Not supported |
| JSON mode | Not supported | Not supported | Not supported |
| Vision input | Not supported | Supported | Not supported |
| Open weights | Not supported | Not supported | Not supported |
| Prompt caching | Not supported | Not supported | Not supported |
Best value in each row is highlighted. Illustrative placeholder data.
Provider information
View Alibaba Qwen →Alibaba Qwen · 通义千问 · Hangzhou, China · est. 2023 (Model lab; parent Alibaba founded 1999)
Alibaba's Qwen family spans every size class from edge models to frontier MoE systems, most released with open weights. Its breadth (chat, coding, vision, embeddings) and permissive licensing made it the default base model for much of the global open-source ecosystem.
Recent releases
Documentation
Frequently asked questions
Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Batch jobs settle at a reduced rate against the same unit. Every figure on this page is an illustrative placeholder.