Skip to content

Qwen3 Embedding

qwen/qwen3-embedding
Embeddings

Multilingual embedding model priced for indexing everything, the budget default for RAG at scale.

Model overview

Context window

32K

tokens

Input price

quoted on request

Output price

quoted on request

Weekly volume

tokens / week

Open weights

No

API only

Capabilities

1

of 11 tags

Description

Qwen3 Embedding is an embeddings model from Alibaba Qwen. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. Alibaba Qwen’s stated focus is the broadest open model family in the world. The weights are not published, so it is API-only.

Capabilities

Batch
Accepts offline batch jobs, which are processed asynchronously at a reduced rate.

Strengths & weaknesses

Strengths

  • Lowest embedding price in the catalogue
  • Strong multilingual recall
  • Batch-friendly

Weaknesses

  • Smaller context than Cohere Embed v4
  • Text-only input

Pricing

Free
Pricing for Qwen3 Embedding, per 1M tokens
RatePriceUnit
Input$0.00per 1M tokens
Output not charged

Batch jobs are supported and settle at a reduced rate against the same unit. Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.

Context window

32Ktokens

covers prompt and response together, so a long input leaves less room for the answer.

32K tokens against a peer range of 8K to 128K, median 32K. Compared across 3 embeddings models.

Supported features

Feature support for Qwen3 Embedding
FeatureSupport
Tool callingNot supported
JSON modeNot supported
StreamingNot supported
Vision inputNot supported
Audio inputNot supported
Long contextNot supported
Open weightsNot supported
ReasoningNot supported
Fine-tunableNot supported
BatchSupported
CachingNot supported

Example use cases

  • Semantic search

    Lowest embedding price in the catalogue

  • RAG indexing

    Strong multilingual recall

  • Deduplication

    Batch-friendly

Comparison

Qwen3 Embedding compared with Embed v4 and Gemini Embedding
AttributeQwen3 EmbeddingEmbed v4Gemini Embedding
Context window32K128K8K
Input price$0.00$0.00$0.00
Output price not charged not charged not charged
Tool callingNot supportedNot supportedNot supported
JSON modeNot supportedNot supportedNot supported
Vision inputNot supportedSupportedNot supported
Open weightsNot supportedNot supportedNot supported
Prompt cachingNot supportedNot supportedNot supported

Best value in each row is highlighted. Illustrative placeholder data.

Provider information

View Alibaba Qwen
Alibaba Qwen5 models

Alibaba Qwen · 通义千问 · Hangzhou, China · est. 2023 (Model lab; parent Alibaba founded 1999)

Alibaba's Qwen family spans every size class from edge models to frontier MoE systems, most released with open weights. Its breadth (chat, coding, vision, embeddings) and permissive licensing made it the default base model for much of the global open-source ecosystem.

Recent releases

    Documentation

    Frequently asked questions

    Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Batch jobs settle at a reduced rate against the same unit. Every figure on this page is an illustrative placeholder.