Qwen SWE-bench

    🏆 Leaderboard

    As of August 31, 2026, Qwen3.8 Max is #1 for Qwen SWE-bench at 80.7%. Ranked by the Qwen SWE-bench score 2 models in this index have a published Qwen SWE-bench score. Qwen SWE-bench leaderboard with live API prices. Qwen SWE-bench — Qwen's published software-engineering score on repository issue fixing.

    Updated August 31, 2026345 models36 providers
    Qwen3.8 MaxOSS
    Qwen · Open Source
    80.7%$8.00
    Qwen3.8-27BOSS
    Qwen · Open Source
    79%$3.65
    ChatGPT-4o Latest
    OpenAI · Proprietary
    $12.50
    Claude 3 Haiku
    Anthropic · Proprietary
    $1.50
    Claude 3 Opus
    Anthropic · Proprietary
    $90.00
    Claude 3 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.5 Haiku
    Anthropic · Proprietary
    $4.80
    Claude 3.5 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.5 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude 3.7 Sonnet
    Anthropic · Proprietary
    $18.00
    Claude Fable 5
    Anthropic · Proprietary
    $60.00
    Claude Haiku 4.5
    Anthropic · Proprietary
    $6.00
    Claude Mythos 5
    Anthropic · Proprietary
    $60.00
    Claude Mythos Preview
    Anthropic · Proprietary
    $60.00
    Claude Opus 4
    Anthropic · Proprietary
    $90.00
    Claude Opus 4.1
    Anthropic · Proprietary
    $90.00
    Claude Opus 4.5
    Anthropic · Proprietary
    $30.00
    Claude Opus 4.6
    Anthropic · Proprietary
    $30.00
    Claude Opus 4.7
    Anthropic · Proprietary
    $30.00
    Claude Opus 4.8
    Anthropic · Proprietary
    $30.00
    Claude Opus 5
    Anthropic · Proprietary
    $30.00
    Claude Sonnet 4
    Anthropic · Proprietary
    $18.00
    Claude Sonnet 4.5
    Anthropic · Proprietary
    $18.00
    Claude Sonnet 4.6
    Anthropic · Proprietary
    $18.00
    Claude Sonnet 5
    Anthropic · Proprietary
    $12.00
    Codestral-22BOSS
    Mistral · Open Source · via Mistral AI
    $1.20
    Command A+OSS
    Cohere · Open Source
    $12.50
    Command R+OSS
    Cohere · Open Source
    $1.25
    Composer 2
    Cursor · Proprietary
    $3.00
    Composer 2 Fast
    Cursor · Proprietary
    $9.00
    DeepSeek R1 Distill Llama 70BOSS
    DeepSeek · Open Source
    $0.50
    DeepSeek R1 Distill Qwen 32BOSS
    DeepSeek · Open Source
    $0.30
    DeepSeek VL2OSS
    DeepSeek · Open Source · via Replicate
    $2.80
    DeepSeek VL2 SmallOSS
    DeepSeek · Open Source · via Replicate
    $1.70
    DeepSeek VL2 TinyOSS
    DeepSeek · Open Source · via Replicate
    $0.90
    DeepSeek-R1OSS
    DeepSeek · Open Source
    $2.74
    DeepSeek-R1-0528OSS
    DeepSeek · Open Source
    $2.74
    DeepSeek-V2.5OSS
    DeepSeek · Open Source
    $0.42
    DeepSeek-V3OSS
    DeepSeek · Open Source
    $1.37
    DeepSeek-V3 0324OSS
    DeepSeek · Open Source
    $1.42
    DeepSeek-V3.1OSS
    DeepSeek · Open Source
    $1.27
    DeepSeek-V3.2OSS
    DeepSeek · Open Source · via OpenRouter
    $0.57
    DeepSeek-V3.2 (Non-thinking)OSS
    DeepSeek · Open Source
    $0.70
    DeepSeek-V3.2-ExpOSS
    DeepSeek · Open Source
    $0.68
    DeepSeek-V4-Flash-0423OSS
    DeepSeek · Open Source
    $0.30
    DeepSeek-V4-Flash-0731OSS
    DeepSeek · Open Source
    $1.76
    DeepSeek-V4-Flash-MaxOSS
    DeepSeek · Open Source
    $0.42
    DeepSeek-V4-Flash-Vision-Exp
    DeepSeek · Proprietary
    $0.88
    DeepSeek-V4-Pro-0813OSS
    DeepSeek · Open Source
    $5.28
    DeepSeek-V4-Pro-MaxOSS
    DeepSeek · Open Source
    $5.22
    Showing 150 of 345 models

    Next step

    You found the model. Now ship the product.

    Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.

    The index

    All Large Language Models

    345 models across 36 providers. Search or jump to a lab — every model page stays linked here.

    Baidu

    2 models

    Inception

    1 models

    inclusionAI

    1 models

    LG AI Research

    1 models

    Liquid AI

    2 models

    Nous Research

    1 models

    OpenBMB

    1 models

    Sakana AI

    1 models

    Sarvam AI

    2 models

    Tencent

    1 models

    Thinking Machines

    1 models

    Unisound

    1 models

    Upstage

    1 models

    Which model leads Qwen SWE-bench right now?

    As of August 31, 2026, Qwen3.8 Max by Qwen is #1 for Qwen SWE-bench at 80.7%. Ranked by the Qwen SWE-bench score This board also tracks Qwen SWE-bench. Next on the same board: Qwen3.8-27B. This qwen swe-bench leaderboard ranks models by Qwen SWE-bench. Scores come from public evals. Prices are the live API rates in the table above.

    Sources: OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)

    Top 2 for qwen swe-bench. Ranked by the Qwen SWE-bench score Input and output are dollars per million tokens.
    RankModelQwen SWE-benchInput /MOutput /M
    1Qwen3.8 Max80.7%$2.00$6.00
    2Qwen3.8-27B79%$0.45$3.20

    Qwen SWE-bench FAQ

    Who ranks #1 on the Qwen SWE-bench leaderboard?

    As of August 31, 2026, Qwen3.8 Max by Qwen ranks #1 on Qwen SWE-bench at 80.7%. API pricing is $2.00/M input and $6.00/M output.

    What are the top models on Qwen SWE-bench?

    The current Qwen SWE-bench ranking as of August 31, 2026 is 1. Qwen3.8 Max at 80.7%; 2. Qwen3.8-27B at 79%.

    Which qwen swe-bench model is the cheapest?

    Qwen3.8-27B is the cheapest scored model on this qwen swe-bench leaderboard at $0.45/M input and $3.20/M output ($3.65 blended). Qwen3.8 Max still leads Qwen SWE-bench at 80.7%.

    Should I always pick the #1 Qwen SWE-bench model?

    Not automatically. Qwen3.8 Max leads Qwen SWE-bench, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh Qwen SWE-bench against input/output price, context window, and related evals.

    How often is the Qwen SWE-bench leaderboard updated?

    Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled August 31, 2026. Treat it as a current index, not a one-off blog post.

    What is Qwen SWE-bench?

    Qwen SWE-bench — Qwen's published software-engineering score on repository issue fixing. This page ranks models that have published a Qwen SWE-bench score, with live API token prices on the same row.

    Where is the Qwen SWE-bench leaderboard?

    This page is the Qwen SWE-bench leaderboard. Models are sorted by Qwen SWE-bench, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.

    How is this Qwen SWE-bench ranking different from the official board?

    Official eval pages own the methodology. This page keeps the published Qwen SWE-bench score next to live API $/M so you can pick a production SKU, not only a trophy number.