Census open · Q3 wave filing now

Which coding models are actually worth the money?

The 2026 DevPass Model Census: value, quality, and speed — rated only by developers who shipped with these models. Every rating is backed by at least 50 real requests through LLM Gateway in the past 30 days. No benchmarks, no vibes.

305
Entries filed
256
Developers reporting
15
Models on the registry

The registry · ranked by value score

Doc. CS-2026
1
deepseek-v4-pro
22 verified ratings · mostly Agentic coding
100%
recommend
Value
4.5
Quality
3.7
Speed
3.9
2
deepseek-v4-flash
14 verified ratings · mostly Other
100%
recommend
Value
4.4
Quality
3.6
Speed
4.1
3
claude-sonnet-5
6 verified ratings · mostly Agentic coding
83%
recommend
Value
4.3
Quality
4.3
Speed
3.5
4
gemini-3-flash-preview
18 verified ratings · mostly Writing tests
100%
recommend
Value
4.3
Quality
3.7
Speed
3.8
5
gemini-3.6-flash
22 verified ratings · mostly Writing tests
95%
recommend
Value
4.2
Quality
3.9
Speed
4.5
6
gpt-5.1
6 verified ratings · mostly Writing tests
83%
recommend
Value
4.2
Quality
4.0
Speed
3.8
7
gpt-5.6-sol
74 verified ratings · mostly Writing tests
93%
recommend
Value
4.2
Quality
4.3
Speed
4.0
8
gpt-5.5
6 verified ratings · mostly Agentic coding
100%
recommend
Value
4.0
Quality
4.5
Speed
3.3
9
gemini-3.1-pro-preview
55 verified ratings · mostly Writing tests
84%
recommend
Value
4.0
Quality
3.8
Speed
3.5
10
claude-opus-4-6
28 verified ratings · mostly Other
100%
recommend
Value
3.9
Quality
4.3
Speed
3.7
11
glm-5.2
19 verified ratings · mostly Agentic coding
95%
recommend
Value
3.8
Quality
3.8
Speed
3.1
12
gpt-5.6-luna
6 verified ratings · mostly Agentic coding
67%
recommend
Value
3.8
Quality
3.7
Speed
4.0
13
gemini-3.5-flash
12 verified ratings · mostly Writing tests
92%
recommend
Value
3.8
Quality
3.8
Speed
3.8
14
kimi-k3
8 verified ratings · mostly Agentic coding
100%
recommend
Value
3.5
Quality
4.1
Speed
3.8
15
claude-fable-5
9 verified ratings · mostly Agentic coding
89%
recommend
Value
3.0
Quality
4.1
Speed
3.2

Scores are 1–5 averages. Models need 5+ verified ratings to be listed; totals refresh every few minutes.

The rules of the registry

01
Usage-verified, or it doesn't count

You can only rate a model your DevPass workspace has hit with 50+ requests in the past 30 days. Nobody rates a model they read a thread about.

02
Members only

Every respondent has an active, paid DevPass plan. These are verdicts from people spending their own credits.

03
No small-sample noise

A model is published only after 5 or more developers rate it, and only aggregates ever leave the building.

04
One reward per member per quarter

The census runs in quarterly waves. Your first entry of each wave earns a free Reset Pass — rate as many models as you use, but nobody can farm passes.