LiveCodeBench

    ๐Ÿ† Leaderboard

    As of August 31, 2026, DeepSeek-V4-Pro-Max is #1 for LiveCodeBench at 93.5%. Ranked by LiveCodeBench: contest programming on problems released after training. LiveCodeBench leaderboard for contamination-resistant coding scores, plus API pricing. Rank LLMs on fresh programming problems.

    Updated August 31, 2026345 models36 providers

    LiveCodeBench vs price

    137 models. Left is cheaper. Up is a higher score. The line is the best score you can buy at each price โ€” a dot under it is a worse deal than something on the line.

    Best score at each price#1 on this board
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek ยท Open Source
    93.5%โ€”โ€”โ€”$5.22
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek ยท Open Source
    91.6%โ€”โ€”โ€”$0.42
    Claude Fable 5
    Anthropic ยท Proprietary
    89.8%โ€”โ€”โ€”$60.00
    Claude Opus 5
    Anthropic ยท Proprietary
    89.0%โ€”โ€”โ€”$30.00
    Gemini 3.7 Flash
    Google ยท Proprietary
    88.7%โ€”โ€”โ€”$4.50
    Gemini 3.1 Pro
    Google ยท Proprietary
    88.5%โ€”โ€”โ€”$17.50
    DeepSeek-V4-Flash-0423OSS
    DeepSeek ยท Open Source
    88.4%โ€”โ€”โ€”$0.30
    Grok 4.6
    xAI ยท Proprietary
    88.2%โ€”โ€”โ€”$8.00
    Gemini 3.6 Flash
    Google ยท Proprietary
    88.1%โ€”โ€”โ€”$4.50
    GPT-5.2 Codex
    OpenAI ยท Proprietary
    88.0%โ€”โ€”โ€”$15.75
    Qwen3.8 MaxOSS
    Qwen ยท Open Source
    87.8%โ€”โ€”โ€”$8.00
    Claude Opus 4.8
    Anthropic ยท Proprietary
    87.8%โ€”โ€”โ€”$30.00
    Solar Pro 4
    Upstage ยท Proprietary
    87.8%โ€”โ€”โ€”$1.50
    Gemini 3.5 Flash
    Google ยท Proprietary
    87.6%โ€”โ€”โ€”$10.50
    DeepSeek-V4-Pro-0813OSS
    DeepSeek ยท Open Source
    87.5%โ€”โ€”โ€”$5.28
    Grok 4.5
    xAI ยท Proprietary
    87.3%โ€”โ€”โ€”$8.00
    GPT-5.3 Codex
    OpenAI ยท Proprietary
    87.3%โ€”โ€”โ€”$15.75
    DeepSeek-V4-Flash-0731OSS
    DeepSeek ยท Open Source
    87.3%โ€”โ€”โ€”$1.76
    Kimi K3OSS
    Moonshot AI ยท Open Source
    87.2%โ€”โ€”โ€”$18.00
    Qwen3.7 Max
    Qwen ยท Proprietary
    87.1%โ€”91.6%โ€”$5.00
    Kimi K2.6OSS
    Moonshot AI ยท Open Source
    86.8%โ€”89.6%โ€”$4.93
    GPT-5 mini
    OpenAI ยท Proprietary
    86.6%โ€”โ€”โ€”$2.25
    GPT-5.1
    OpenAI ยท Proprietary
    86.5%โ€”โ€”โ€”$11.25
    Gemini 3 Pro
    Google ยท Proprietary
    86.4%โ€”โ€”โ€”$14.00
    Nemotron 3 Ultra (550B A55B)OSS
    NVIDIA ยท Open Source ยท via OpenRouter
    86.0%โ€”89%โ€”$2.70
    Qwen3.6 Plus
    Qwen ยท Proprietary
    86.0%โ€”87.1%โ€”$3.50
    Inkling-SmallOSS
    Thinking Machines ยท Open Source ยท via Thinking Machines Lab
    85.9%โ€”โ€”โ€”$1.50
    GPT-5.6 Terra
    OpenAI ยท Proprietary
    85.9%โ€”โ€”โ€”$14.00
    Muse Spark 1.1
    Meta ยท Proprietary
    85.9%โ€”โ€”โ€”$5.50
    GPT-5
    OpenAI ยท Proprietary
    85.9%93.4%โ€”88%$11.25
    Gemini 3 Flash
    Google ยท Proprietary
    85.6%โ€”โ€”โ€”$3.50
    GPT-5.1 Codex
    OpenAI ยท Proprietary
    85.5%โ€”โ€”โ€”$11.25
    GPT-5.2
    OpenAI ยท Proprietary
    85.4%โ€”โ€”โ€”$15.75
    GPT-5.5
    OpenAI ยท Proprietary
    85.3%โ€”โ€”โ€”$35.00
    Claude Opus 4.7
    Anthropic ยท Proprietary
    85.1%โ€”โ€”โ€”$30.00
    GPT-5 Codex
    OpenAI ยท Proprietary
    84.7%โ€”โ€”โ€”$11.25
    Grok 4.3
    xAI ยท Proprietary
    84.5%โ€”โ€”โ€”$3.75
    GPT-5.4
    OpenAI ยท Proprietary
    84.1%โ€”โ€”โ€”$17.50
    GPT-5.4 nano
    OpenAI ยท Proprietary
    84.0%โ€”โ€”โ€”$1.45
    o3
    OpenAI ยท Proprietary
    83.9%โ€”โ€”81.3%$10.00
    DeepSeek-V3.2OSS
    DeepSeek ยท Open Source ยท via OpenRouter
    83.3%โ€”โ€”โ€”$0.57
    GPT OSS 120BOSS
    OpenAI ยท Open Source
    83.2%โ€”โ€”โ€”$0.54
    MiniMax M2OSS
    MiniMax ยท Open Source
    83%โ€”โ€”โ€”$1.50
    LongCat-Flash-Thinking-2601OSS
    Meituan ยท Open Source
    82.8%โ€”โ€”โ€”$1.50
    GPT-5.6 Sol
    OpenAI ยท Proprietary
    82.6%โ€”โ€”โ€”$35.00
    Claude Sonnet 5
    Anthropic ยท Proprietary
    82.4%โ€”โ€”โ€”$12.00
    GLM-4.7OSS
    Z AI ยท Open Source
    82.2%84.9%84.9%โ€”$2.80
    o4-mini
    OpenAI ยท Proprietary
    82.2%โ€”โ€”68.9%$5.50
    MiniMax M3OSS
    MiniMax ยท Open Source
    82.2%โ€”โ€”โ€”$1.50
    Claude Sonnet 4.6
    Anthropic ยท Proprietary
    82.1%โ€”โ€”โ€”$18.00
    Showing 1โ€“50 of 345 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    345 models across 36 providers. Search or jump to a lab โ€” every model page stays linked here.

    Baidu

    2 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Liquid AI

    2 models

    Nous Research

    1 models

    OpenBMB

    1 models

    Sakana AI

    1 models

    Sarvam AI

    2 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads LiveCodeBench right now?

    As of August 31, 2026, DeepSeek-V4-Pro-Max by DeepSeek is #1 for LiveCodeBench at 93.5%. Ranked by LiveCodeBench: contest programming on problems released after training. This board also tracks LiveCodeBench, HumanEval, LiveCodeBench v6, Alder Polyglot. Next on the same board: DeepSeek-V4-Flash-Max and Claude Fable 5. Related leaders: MiniCPM-SALA on HumanEval at 95.1%; Qwen3.8 Flash on LiveCodeBench v6 at 91.9%. This livecodebench leaderboard ranks models by LiveCodeBench. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: LiveCodeBench (https://livecodebench.github.io/); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 8 for livecodebench. Ranked by LiveCodeBench: contest programming on problems released after training. Input and output are dollars per million tokens.
    RankModelLiveCodeBenchInput /MOutput /M
    1DeepSeek-V4-Pro-Max93.5%$1.74$3.48
    2DeepSeek-V4-Flash-Max91.6%$0.14$0.28
    3Claude Fable 589.8%$10.00$50.00
    4Claude Opus 589.0%$5.00$25.00
    5Gemini 3.7 Flash88.7%$0.75$3.75
    6Gemini 3.1 Pro88.5%$2.50$15.00
    7DeepSeek-V4-Flash-042388.4%$0.10$0.20
    8Grok 4.688.2%$2.00$6.00

    LiveCodeBench FAQ

    Who ranks #1 on the LiveCodeBench leaderboard?

    As of August 31, 2026, DeepSeek-V4-Pro-Max by DeepSeek ranks #1 on LiveCodeBench at 93.5%. API pricing is $1.74/M input and $3.48/M output.

    What are the top models on LiveCodeBench?

    The current LiveCodeBench ranking as of August 31, 2026 is 1. DeepSeek-V4-Pro-Max at 93.5%; 2. DeepSeek-V4-Flash-Max at 91.6%; 3. Claude Fable 5 at 89.8%.

    Which livecodebench model is the cheapest?

    Gemma 3n E2B Instructed is the cheapest scored model on this livecodebench leaderboard at $0.02/M input and $0.04/M output ($0.06 blended). DeepSeek-V4-Pro-Max still leads LiveCodeBench at 93.5%.

    Should I always pick the #1 LiveCodeBench model?

    Not automatically. DeepSeek-V4-Pro-Max leads LiveCodeBench, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh LiveCodeBench against input/output price, context window, and related evals.

    How often is the LiveCodeBench leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 31, 2026. Treat it as a current index, not a one-off blog post.

    What is LiveCodeBench?

    LiveCodeBench scores models on new competitive-programming problems released after training cutoffs, which reduces memorization compared with HumanEval.

    How is LiveCodeBench different from SWE-bench?

    LiveCodeBench is contest-style code. SWE-bench is multi-file GitHub repair. Use LiveCodeBench for algorithms, SWE-bench for product engineering.

    Why show HumanEval next to LiveCodeBench?

    HumanEval is older and more saturated, but still widely cited. Seeing both columns makes it obvious when a model only looks strong on the easy set.