Aider Polyglot

Aider · polyglot-225-5dc9490bb35f · A 225-exercise coding benchmark across six programming languages, published with per-run configuration and pass counts by Aider.

Model Ranking

Source
#1Kimi-K2-InstructMoonshot AI59.11percent1,617.1225diff; aider 0.85.3.devtotalRecomputedwinner eligibleThird-party benchmark
#2GPT-4.1OpenAI52.44percent225diff; aider 0.81.4.devunknownRecomputedwinner eligibleThird-party benchmark
#3gpt-oss-120bOpenAI41.78percent3,806.6225diff; aider 0.85.3.devtotalRecomputedwinner eligibleThird-party benchmark
#4GPT-4.1 MiniOpenAI32.44percent225diff; aider 0.81.4.devunknownRecomputedwinner eligibleThird-party benchmark
#5GPT-4oOpenAI23.11percent225diff; aider 0.70.1.devunknownRecomputedwinner eligibleThird-party benchmark
#6Llama-4-Maverick-17B-128E-InstructMeta15.56percent225whole; aider 0.81.2.devunknownRecomputedwinner eligibleThird-party benchmark
#7Command ACohereSelected model12.00percent225whole; aider 0.77.1.devunknownRecomputedwinner eligibleThird-party benchmark
#8GPT-4.1 NanoOpenAI8.89percent225whole; aider 0.81.4.devunknownRecomputedwinner eligibleThird-party benchmark
#9GPT-4o MiniOpenAI3.56percent225whole; aider 0.69.2.devunknownRecomputedwinner eligibleThird-party benchmark

9 models ranked by the best winner-eligible current-version score when available, otherwise the best publisher-artifact score; highest first. Observed Sep 2, 2026. The model you came from is highlighted.

Methodology and Coverage

Ranking protocolpass rate 2 · percent. Each model appears once; verified winner-eligible runs take precedence, and missing models are not imputed.
Acquisition coverage9 published models from 69 source rows; 59 source identities remain quarantined.

Primary Evidence

Publisher Artifacts

Questions

Aider Polyglot FAQs

What does Aider Polyglot measure?+

A 225-exercise coding benchmark across six programming languages, published with per-run configuration and pass counts by Aider. Model Markets classifies it as a coding benchmark and preserves the publisher's polyglot-225-5dc9490bb35f release as a distinct comparison cohort.

How are models ranked on Aider Polyglot?+

Models are ordered by pass rate 2 in percent, with higher scores ranked first. Each model appears once; a verified winner-eligible current-version result takes precedence over a publisher-artifact-only result, and missing scores are not estimated.

Which model currently leads Aider Polyglot?+

Kimi-K2-Instruct leads the current verified table with 59.11 percent on polyglot-225-5dc9490bb35f. This is a benchmark-specific result, not a universal model-quality claim.

How many models have a published Aider Polyglot score?+

9 catalog models are published from 69 source rows. 59 source identities remain quarantined rather than guessed.

Can Aider Polyglot scores be compared with other benchmarks?+

Raw scores should be compared only within the same benchmark version, metric, and protocol. Model Markets normalizes eligible scores only for aggregate rankings, and groups models by an identical benchmark set before ranking them. View aggregate rankings

Why might a model be missing from Aider Polyglot?+

A model remains absent when the publisher has no current result, the source model identity is unresolved, the evaluation protocol is incompatible, or the evidence cannot be verified. Model Markets does not substitute a provider claim or infer a score from a related model.

How current is the Aider Polyglot leaderboard?+

The current Model Markets snapshot was observed Sep 2, 2026 from Aider artifacts. The exact publisher source and each retained result artifact are linked on this page.

Send Feedback