Qwen3.8 27B vs Llama 4 Maverick 17B 128E Instruct

Model Markets Rankings

Intelligence, Cost, and Efficiency

All rankings →
RankingQwen3.8-27BLlama-4-Maverick-17B-128E-Instruct
IntelligenceHigher is better · MM Intelligence v1.4#29 of 4653.3 score · 2/3-source partialUnrankedNot in the 46-model eligible cohort
CostLower is better · Published-token output estimate#18 of 45$0.086 per LiveBench caseUnrankedNot in the 45-model eligible cohort
EfficiencyHigher is better · MM Efficiency v1.1#23 of 3850.7 score · 2/3-source intelligenceUnrankedNot in the 38-model eligible cohort

Ranks come from current available eligible cohorts. Intelligence may use a labeled two-of-three partial core. Green highlights appear only when both models are ranked in the same metric, and the three dimensions are not collapsed into an overall winner.

Benchmark Performance

Available Benchmarks

BenchmarkQwen3.8-27BLlama-4-Maverick-17B-128E-Instruct
LMArena Text Arenatext-2026-09-01-011508720696 · arena_rating · leader1,437.35100% of row best · rating · qwen3.8-27b; 95% CI [1429.71618499, 1444.98971204]; votes 6626; rank 691,287.1390% of row best · rating · llama-4-maverick-17b-128e-instruct; 95% CI [1282.87677682, 1291.39199009]; votes 39349; rank 243
LMArena Vision Arenavision-2026-08-27-011508720696 · arena_rating · leader1,279.32100% of row best · rating · qwen3.8-27b; 95% CI [1267.46655532, 1291.16933290]; votes 2954; rank 281,141.4789% of row best · rating · llama-4-maverick-17b-128e-instruct; 95% CI [1132.42029588, 1150.51619722]; votes 6936; rank 102
Overall ResultCounted from the protocol-matched rows above2 benchmark winsOverall lead0 benchmark wins

Third-party benchmark Only like-for-like primary-publisher results are shown; raw scores, relative scores, configuration, and token spend remain visible.

FieldAt a Glance
Qwen · activeQwen3.8 27BVerified Aug 28, 2026
Meta · activeLlama 4 Maverick 17B 128E InstructVerified Aug 28, 2026

Technical Differences

Side-by-Side Facts

Indexable
FieldQwen3.8-27BLlama-4-Maverick-17B-128E-Instruct
DeveloperQwenMeta
FamilyQwen3 8 27bLlama 4 Maverick 17b 128e Instruct
ModelQwen3.8-27BLlama-4-Maverick-17B-128E-Instruct
VersionQwen3.8-27BLlama-4-Maverick-17B-128E-Instruct
Lifecycleactiveactive
ReleasedUnknown2025-04-05
Knowledge cutoffUnknownUnknown
Input modalitiesText, ImageText, Image
Output modalitiesTextText
Context window262K1,000K
Total parameters27.8B401.6B
Active parametersUnknown17B
Licenseapache-2.0other
Open weightsYesYes
API availableYesYes
Self-hostableYesYes
Provider accessDeepinfra (Standard), Hugging Face (Standard), Openrouter (Standard)Openrouter (Standard)
Capabilitieschat, generation, reasoning, toolschat, generation, tools

15 comparable fields · 9 material differences · Pair passes the primary-source comparison gate

Qwen3.8 27B Capabilities

chatgenerationreasoningtools
Input price$0.40
Output price$3.00
Serving providers3
Canonical IDQwen/Qwen3.8-27B

Llama 4 Maverick 17B 128E Instruct Capabilities

chatgenerationtools
Input price$0.20
Output price$0.696
Serving providers1
Canonical IDmeta-llama/Llama-4-Maverick-17B-128E-Instruct

Internal Comparison Graph

Related Comparisons

All image comparisons →
APairBContext
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peersimage, text
vsfamily variantsimage, text
vscross-developer peersimage, text
vscross-developer peersimage, text
vscross-developer peerstext
vscross-developer peerstext
vscross-developer peersimage, text
vscross-developer peersimage, text

Primary Evidence

Sources and Freshness

Questions

Qwen3.8 27B vs Llama 4 Maverick 17B 128E Instruct FAQs

Is Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct better for coding?+

This comparison does not currently contain a protocol-matched coding benchmark for both Qwen3.8 27B and Llama 4 Maverick 17B 128E Instruct, so Model Markets cannot name a coding leader from pricing, context size, or capability labels alone.

Which is cheaper, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

Qwen3.8 27B is $0.40 and Llama 4 Maverick 17B 128E Instruct is $0.20 per million tokens, so Llama 4 Maverick 17B 128E Instruct is cheaper on this metric. Qwen3.8 27B is $3.00 and Llama 4 Maverick 17B 128E Instruct is $0.696 per million tokens, so Llama 4 Maverick 17B 128E Instruct is cheaper on this metric.

Which has a larger context window, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

Llama 4 Maverick 17B 128E Instruct has the larger sourced context window. Qwen3.8 27B supports 262K and Llama 4 Maverick 17B 128E Instruct supports 1,000K.

Which performs better in benchmarks, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

There is no overall benchmark winner: An overall winner requires at least two decisive benchmarks from at least two original publishers.

Can Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct be self-hosted?+

Both models have the same recorded self-hosting status: supported. Qwen3.8 27B is open weight; Llama 4 Maverick 17B 128E Instruct is open weight.

Can Qwen3.8 27B and Llama 4 Maverick 17B 128E Instruct understand images?+

Qwen3.8 27B is documented with image input; Llama 4 Maverick 17B 128E Instruct is documented with image input. This reflects supported input modalities, not vision quality.

Which can generate longer answers, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

Neither has a larger sourced maximum output. Qwen3.8 27B is 131K and Llama 4 Maverick 17B 128E Instruct is —.

Do Qwen3.8 27B and Llama 4 Maverick 17B 128E Instruct support reasoning and tool use?+

Qwen3.8 27B: reasoning, tool calling, and image input. Llama 4 Maverick 17B 128E Instruct: tool calling and image input. Feature support does not establish relative quality.

Which is available from more inference providers, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

Qwen3.8 27B has 3 sourced provider routes; Llama 4 Maverick 17B 128E Instruct has 1, so Qwen3.8 27B has broader tracked availability.

Which offers better value, Qwen3.8 27B or Llama 4 Maverick 17B 128E Instruct?+

There is no universal value winner. Compare the input and output prices above with the matched benchmark result for your workload: cheaper tokens can be offset by different quality, token usage, latency, or provider availability.

Send Feedback