Who ranks #1 on the Codegolf V2 2 leaderboard?
As of August 31, 2026, Gemma 3n E4B Instructed by Google ranks #1 on Codegolf V2 2 at 16.8%. API pricing is $20.00/M input and $40.00/M output.
As of August 31, 2026, Gemma 3n E4B Instructed is #1 for Codegolf V2 2 at 16.8%. Ranked by the Codegolf V2 2 score 2 models in this index have a published Codegolf V2 2 score. Codegolf V2 2 leaderboard: rank models by Codegolf V2 2 next to live API token prices.
Next step
Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.
The index
345 models across 36 providers. Search or jump to a lab — every model page stays linked here.









As of August 31, 2026, Gemma 3n E4B Instructed by Google is #1 for Codegolf V2 2 at 16.8%. Ranked by the Codegolf V2 2 score This board also tracks Codegolf V2 2. Next on the same board: Gemma 3n E2B Instructed. This codegolf v2 2 leaderboard ranks models by Codegolf V2 2. Scores come from public evals. Prices are the live API rates in the table above.
Sources: OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)
| Rank | Model | Codegolf V2 2 | Input /M | Output /M |
|---|---|---|---|---|
| 1 | Gemma 3n E4B Instructed | 16.8% | $20.00 | $40.00 |
| 2 | Gemma 3n E2B Instructed | 11% | $0.02 | $0.04 |
Rank one eval at a time. All LLM benchmarks.
As of August 31, 2026, Gemma 3n E4B Instructed by Google ranks #1 on Codegolf V2 2 at 16.8%. API pricing is $20.00/M input and $40.00/M output.
The current Codegolf V2 2 ranking as of August 31, 2026 is 1. Gemma 3n E4B Instructed at 16.8%; 2. Gemma 3n E2B Instructed at 11%.
Gemma 3n E2B Instructed is the cheapest scored model on this codegolf v2 2 leaderboard at $0.02/M input and $0.04/M output ($0.06 blended). Gemma 3n E4B Instructed still leads Codegolf V2 2 at 16.8%.
Not automatically. Gemma 3n E4B Instructed leads Codegolf V2 2, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh Codegolf V2 2 against input/output price, context window, and related evals.
Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 31, 2026. Treat it as a current index, not a one-off blog post.
Codegolf V2 2 is a public LLM eval (the Codegolf V2 2 score). This page ranks models that have published a score, next to live API prices.
This page is the Codegolf V2 2 leaderboard. Models are sorted by Codegolf V2 2, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.
Official eval pages own the methodology. This page keeps the published Codegolf V2 2 score next to live API $/M so you can pick a production SKU, not only a trophy number.