30% off open-source modelsends in 18:13:53

The 2026 DevPass Model Census

Which coding models are actually worth the money?

The 2026 DevPass Model Census rates coding models on value, quality, and speed — using only developers who shipped with them. Every rating is backed by at least 50 real requests through LLM Gateway in the past 30 days. No benchmarks, no vibes.

  • ✓ 50+ requests per rating
  • ✓ Paid members only
  • ✓ Aggregates, never individuals

DevPass · Model Census

Registry of coding models

Open · Q3Filing now
Entries filed
1,131
Developers reporting
915
Models on registry
30
Wave
Q3 · filing now
Listing rule
5+ verified ratings
Scale
1–5 averages

Doc. CS-2026 · issued by LLM Gateway

Now boarding · category leaders

Top of the registry

Highest average in each score across every model with 5+ verified ratings. Ties go to the model with more ratings.

Boarding pass · Best value

MiMo V2.5 Pro

mimo-v2.5-pro · Xiaomi

Gate
Agentic coding
Ratings
5
Recommend
100%
Value · Quality · Speed
5.0 · 4.6 · 4.4

Value

5.0

out of 5 for value

CS-2026-01

Boarding pass · Best quality

MiMo V2.5 Pro

mimo-v2.5-pro · Xiaomi

Gate
Agentic coding
Ratings
5
Recommend
100%
Value · Quality · Speed
5.0 · 4.6 · 4.4

Quality

4.6

out of 5 for quality

CS-2026-01

Boarding pass · Fastest

MiMo V2.5 Pro

mimo-v2.5-pro · Xiaomi

Gate
Agentic coding
Ratings
5
Recommend
100%
Value · Quality · Speed
5.0 · 4.6 · 4.4

Speed

4.4

out of 5 for speed

CS-2026-01

As of September 9, 2026, MiMo V2.5 Pro leads the 2026 DevPass Model Census on value for money (5.0/5 across 5 verified ratings), MiMo V2.5 Pro leads on output quality (4.6/5), and MiMo V2.5 Pro leads on speed (4.4/5). 915 DevPass developers have filed 1,131 verified ratings across 30 models. The most-rated model is GPT-5.6 Sol with 209 ratings and a 96% recommend rate.

Departures · full registry

Every model on the registry

Click a column to sort, or narrow the board by vendor, use case, and rating count. Filters live in the URL, so a view can be shared.

Showing 30 of 30 models · sorted by value descendingDOC. CS-2026

Departures · Coding models

Rank = registry position by value · refreshed every 5 min

  • 1

    MiMo V2.5 Pro

    Xiaomi · 5 ratings

    100%

    recommend

    Value
    5.0
    Quality
    4.6
    Speed
    4.4
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 2

    GLM-5.3 Flash

    Z.ai · 22 ratings

    100%

    recommend

    Value
    4.8
    Quality
    4.3
    Speed
    3.5
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 3

    DeepSeek V4 Pro

    DeepSeek · 64 ratings

    98%

    recommend

    Value
    4.6
    Quality
    3.8
    Speed
    3.9
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 4

    DeepSeek V4 Flash

    DeepSeek · 67 ratings

    99%

    recommend

    Value
    4.5
    Quality
    3.7
    Speed
    4.1
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 5

    GLM-5.3

    Z.ai · 7 ratings

    100%

    recommend

    Value
    4.4
    Quality
    4.1
    Speed
    3.7
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 6

    Qwen3.8 Max

    Alibaba · 8 ratings

    100%

    recommend

    Value
    4.4
    Quality
    4.3
    Speed
    3.4
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 7

    Gemini 3.8 Flash

    Google · 22 ratings

    95%

    recommend

    Value
    4.4
    Quality
    4.2
    Speed
    3.9
    Mostly writing testsCleared: 90% or more of raters would recommend this model
  • 8

    Gemini 3 Flash (Preview)

    Google · 36 ratings

    100%

    recommend

    Value
    4.3
    Quality
    3.9
    Speed
    3.9
    Mostly writing testsCleared: 90% or more of raters would recommend this model
  • 9

    Gemini 3.7 Flash

    Google · 107 ratings

    95%

    recommend

    Value
    4.3
    Quality
    4.0
    Speed
    4.0
    Mostly writing testsCleared: 90% or more of raters would recommend this model
  • 10

    Qwen3.8 Flash

    Alibaba · 5 ratings

    100%

    recommend

    Value
    4.2
    Quality
    3.6
    Speed
    4.0
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 11

    GPT-5.1

    OpenAI · 15 ratings

    87%

    recommend

    Value
    4.1
    Quality
    3.9
    Speed
    3.6
    Mostly writing testsBoarding: 75–89% of raters would recommend this model
  • 12

    Gemini 3.6 Flash

    Google · 91 ratings

    90%

    recommend

    Value
    4.0
    Quality
    3.8
    Speed
    4.3
    Mostly writing testsCleared: 90% or more of raters would recommend this model
  • 13

    GPT-5.6 Sol

    OpenAI · 209 ratings

    96%

    recommend

    Value
    4.0
    Quality
    4.3
    Speed
    3.7
    Mostly writing testsCleared: 90% or more of raters would recommend this model
  • 14

    GLM-5.2

    Z.ai · 64 ratings

    97%

    recommend

    Value
    4.0
    Quality
    4.1
    Speed
    3.6
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 15

    GPT-5.6 Luna

    OpenAI · 34 ratings

    91%

    recommend

    Value
    4.0
    Quality
    3.8
    Speed
    4.1
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 16

    Gemini 3.5 Flash

    Google · 21 ratings

    86%

    recommend

    Value
    4.0
    Quality
    3.8
    Speed
    4.0
    Mostly writing testsBoarding: 75–89% of raters would recommend this model
  • 17

    Claude Sonnet 5

    Anthropic · 19 ratings

    89%

    recommend

    Value
    3.9
    Quality
    4.1
    Speed
    3.6
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 18

    Gemini 3.1 Pro (Preview)

    Google · 128 ratings

    86%

    recommend

    Value
    3.8
    Quality
    3.7
    Speed
    3.3
    Mostly writing testsBoarding: 75–89% of raters would recommend this model
  • 19

    GPT-5.5

    OpenAI · 9 ratings

    89%

    recommend

    Value
    3.8
    Quality
    4.2
    Speed
    3.8
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 20

    GPT-5.6 Terra

    OpenAI · 11 ratings

    82%

    recommend

    Value
    3.7
    Quality
    3.6
    Speed
    3.9
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 21

    Claude Opus 4.6

    Anthropic · 61 ratings

    95%

    recommend

    Value
    3.7
    Quality
    4.2
    Speed
    3.5
    Mostly otherCleared: 90% or more of raters would recommend this model
  • 22

    Gemini 3.5 Flash Lite

    Google · 10 ratings

    80%

    recommend

    Value
    3.7
    Quality
    3.5
    Speed
    4.0
    Mostly writing testsBoarding: 75–89% of raters would recommend this model
  • 23

    MiniMax M3

    MiniMax · 9 ratings

    89%

    recommend

    Value
    3.7
    Quality
    3.9
    Speed
    3.7
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 24

    Kimi K3

    Moonshot · 39 ratings

    97%

    recommend

    Value
    3.6
    Quality
    4.4
    Speed
    4.0
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 25

    Claude Opus 5

    Anthropic · 18 ratings

    89%

    recommend

    Value
    3.3
    Quality
    4.1
    Speed
    3.1
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 26

    Grok 4.5

    xAI · 11 ratings

    91%

    recommend

    Value
    3.3
    Quality
    3.5
    Speed
    3.9
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 27

    Claude Fable 5

    Anthropic · 18 ratings

    89%

    recommend

    Value
    3.2
    Quality
    4.4
    Speed
    3.2
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 28

    Claude Opus 4.8

    Anthropic · 11 ratings

    91%

    recommend

    Value
    3.0
    Quality
    4.2
    Speed
    3.8
    Mostly agentic codingCleared: 90% or more of raters would recommend this model
  • 29

    Claude Sonnet 4.6

    Anthropic · 5 ratings

    80%

    recommend

    Value
    3.0
    Quality
    4.2
    Speed
    3.4
    Mostly agentic codingBoarding: 75–89% of raters would recommend this model
  • 30

    Gemini 3.1 Flash Lite

    Google · 5 ratings

    80%

    recommend

    Value
    2.8
    Quality
    3.2
    Speed
    4.4
    Mostly writing testsBoarding: 75–89% of raters would recommend this model

Scores are 1–5 averages. Models need 5+ verified ratings to be listed; totals refresh every few minutes. Last updated September 9, 2026 (UTC).

Conditions of entry

The rules of the registry

  1. Usage-verified, or it doesn't count

    You can only rate a model your DevPass workspace has hit with 50+ requests in the past 30 days. Nobody rates a model they read a thread about.

  2. Members only

    Every respondent has an active, paid DevPass plan. These are verdicts from people spending their own credits.

  3. No small-sample noise

    A model is published only after 5 or more developers rate it, and only aggregates ever leave the building.

  4. One reward per member per quarter

    The census runs in quarterly waves. Your first entry of each wave earns a free Reset Pass — rate as many models as you use, but nobody can farm passes.

Reading the census

Model Census FAQ

How are models scored in the DevPass Model Census?
Each entry rates one model from 1 to 5 on value for money, output quality, and speed, plus a yes-or-no “would you recommend it”. The registry shows the per-model averages and the share of raters who would recommend the model.
Who can rate a model?
Only paid DevPass members, and only for models their workspace has sent at least 50 requests to in the past 30 days. Ratings are tied to verified usage, not opinions from a thread.
How is the registry ranked?
Registry rank is the model's position by average value score, with ties broken by number of ratings. The sort and filter controls change what you see on the board but never the underlying rank.
Why is a model missing from the registry?
A model appears once 5 or more developers have filed a verified rating on it in 2026. Per-use-case breakdowns are additionally hidden below two entries so no single respondent can be identified.
How often does the census update?
Members file entries in quarterly waves and the yearly registry aggregates every wave. The published totals refresh every five minutes.
Your entryWave Q3 · 2026

Rate the models you ship with, stamp a free Reset Pass

Takes two minutes. Your first entry of the wave earns a Reset Pass on your tier, and every entry sharpens the registry for the next developer at the desk.