QwenQwen3 Embedding 8B

External Link

Qwen3 Embedding 8B is Qwen's open-weight multilingual text-embedding model with a 32K-token context window and configurable output dimensions up to 4,096.

Token context
33K tokens
Inputs
Text
Outputs
Embedding
Released
Jun 3, 2025

Providers & Pricing

Estimate workload cost →
ProviderInput / 1MOutput / 1MCache read / 1MCache write / 1MObservedEvidence
$0.010Sep 22, 2026Provider-reported

At a Glance

Model Facts

Identity

Status
active
Developer ID
Qwen/Qwen3-Embedding-8B
Released
Jun 3, 2025
Version
Qwen3 Embedding 8B
License
apache-2.0

Capacity

Token context
33K tokens
Maximum output
Unknown
Total parameters
8B
Active parameters
Unknown
Knowledge cutoff
Not reported

Interface and Access

Inputs
Text
Outputs
Embedding
API available
Yes
Open weights
Yes
Self-hostable
Yes
Reasoning
No
Vision
No
Tool calling
No
Capabilities
embeddings, multilingual, retrieval

This page represents one developer model product. Dated API snapshots, serving endpoints, pricing, regions, quantizations, and service tiers attach as versioned aliases or provider details and never create duplicate public model records.

Version Lineage

PredecessorUnknown
SuccessorsUnknown
Aliasesqwen3-embedding-8b, Qwen3 Embedding 8B, Qwen/Qwen3-Embedding-8B

Explicit Unknowns

  • Maximum output
  • Knowledge cutoff

Primary Evidence

Sources and Observation Date

Comparable Peers

Related Model Comparisons

All comparisons →
APairBContext
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vsfamily variants
vscross-developer peersembedding
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers
vscross-developer peers

Questions

Qwen3 Embedding 8B FAQs

What is Qwen3 Embedding 8B?+

Qwen3 Embedding 8B is Qwen's open-weight multilingual text-embedding model with a 32K-token context window and configurable output dimensions up to 4,096. It is developed by Qwen and its current sourced lifecycle status is active.

What inputs and outputs does Qwen3 Embedding 8B support?+

Qwen3 Embedding 8B's current record lists Text as input and Embedding as output.

How much context does Qwen3 Embedding 8B support?+

The current record lists a 33K-token context window and does not report a maximum output length.

Is Qwen3 Embedding 8B available through an API?+

Qwen3 Embedding 8B is marked as API-available. The directly sourced provider records currently include Deepinfra and Fireworks Ai.

How large is Qwen3 Embedding 8B?+

The current record lists 8B total parameters and an unknown active parameter count. Its architecture is Qwen3ForCausalLM.

What capabilities and tasks does Qwen3 Embedding 8B support?+

Qwen3 Embedding 8B's recorded capabilities are embeddings, multilingual, and retrieval. Its supported tasks are code-retrieval, embedding, retrieval, and text.

How much does Qwen3 Embedding 8B cost through an API?+

The lowest directly sourced prices currently attached to Qwen3 Embedding 8B are $0.010 per million input tokens through Deepinfra and an unknown output-token price. Prices are provider-specific and should be checked against each cited observation date.

Can Qwen3 Embedding 8B be self-hosted?+

Qwen3 Embedding 8B is an open-weight model. Self-hosting is supported, and the recorded license is apache-2.0.

Does Qwen3 Embedding 8B have related versions or aliases?+

No predecessor is recorded. No successor is recorded. Recorded aliases are qwen3-embedding-8b, Qwen3 Embedding 8B, and Qwen/Qwen3-Embedding-8B.

Which models can Qwen3 Embedding 8B be compared with?+

The current compatible peer set includes GPT-6 Astra, Claude Fable 5.1, Gemini 3.1 Pro, DeepSeek V4.1 Flash, and Grok 4.6. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons

Send Feedback