Qwen Models

148 modelsGeneral models from $0.0206/M inputUp to 1.05M context

Usage

Last 30 days · 2026-08-26 to 2026-09-24

Tokens

265B

Requests

7.3M

Models in use

139 of 148

Tokens per day, stacked by model

010.6B21.2B08-2609-0209-0909-1609-232026-08-26 — 21,219,605,385 tokens qwen3.7-flash: 18,819,287,620 90 more models: 1,551,508,050 qwen3.7-plus: 459,769,510 qwen3.8-max: 326,822,250 qwen3.5-35b-a3b: 51,405,340 qwen3.5-plus: 10,812,440 qwen3-coder-30b-a3b-instruct: 1752026-08-27 — 16,929,402,775 tokens qwen3.7-flash: 14,636,196,755 90 more models: 1,306,557,575 qwen3.8-max: 448,246,100 qwen3.7-plus: 375,852,805 qwen3.8-flash: 101,836,125 qwen3.5-35b-a3b: 42,340,340 qwen3.5-plus: 18,373,0752026-08-28 — 9,992,357,565 tokens qwen3.7-flash: 4,571,307,665 qwen3.8-max: 2,705,194,310 qwen3.8-flash: 1,611,022,680 90 more models: 848,381,500 qwen3.5-plus: 143,532,075 qwen3.7-plus: 85,054,645 qwen3.5-35b-a3b: 27,864,6902026-08-29 — 3,865,700,735 tokens qwen3.8-flash: 1,304,691,330 qwen3.7-plus: 967,758,565 qwen3.7-flash: 582,730,885 90 more models: 521,146,480 qwen3.8-max: 351,256,495 qwen3.5-plus: 81,935,780 qwen3.5-35b-a3b: 56,181,2002026-08-30 — 4,613,177,950 tokens qwen3.8-flash: 2,874,334,035 90 more models: 729,600,575 qwen3.8-max: 362,585,665 qwen3.7-flash: 274,445,050 qwen3.7-plus: 229,640,335 qwen3.5-plus: 103,258,460 qwen3.5-35b-a3b: 39,306,915 qwen3-coder-30b-a3b-instruct: 6,9152026-08-31 — 5,175,736,935 tokens qwen3.7-flash: 2,658,587,475 qwen3.8-flash: 1,290,923,125 90 more models: 448,882,915 qwen3.8-max: 446,626,545 qwen3.7-plus: 173,928,785 qwen3.5-plus: 114,089,445 qwen3.5-35b-a3b: 42,698,110 qwen3-coder-30b-a3b-instruct: 5352026-09-01 — 5,909,231,430 tokens qwen3.7-flash: 3,064,866,360 90 more models: 1,276,599,830 qwen3.8-flash: 868,273,890 qwen3.8-max: 517,189,160 qwen3.7-plus: 95,539,205 qwen3.5-plus: 46,555,880 qwen3.5-35b-a3b: 40,206,810 qwen3-coder-30b-a3b-instruct: 2952026-09-02 — 2,624,590,710 tokens qwen3.8-max: 682,135,670 90 more models: 532,116,610 qwen3.7-flash: 474,126,635 qwen3.7-plus: 385,064,385 qwen3.8-flash: 321,819,470 qwen3.8-max-2026-09-02: 137,077,990 qwen3.5-plus: 50,666,670 qwen3.5-35b-a3b: 41,582,900 qwen3-coder-30b-a3b-instruct: 3802026-09-03 — 5,537,612,300 tokens qwen3.8-flash: 2,739,642,180 qwen3.7-flash: 998,547,835 qwen3.8-max-2026-09-02: 906,928,385 90 more models: 507,068,530 qwen3.8-max: 279,175,850 qwen3.5-35b-a3b: 43,131,965 qwen3.7-plus: 42,829,865 qwen3.5-plus: 20,287,6902026-09-04 — 12,488,881,260 tokens qwen3.8-flash: 9,993,403,840 90 more models: 1,034,015,530 qwen3.7-flash: 524,029,925 qwen3.8-max: 412,862,265 qwen3.8-max-2026-09-02: 368,870,925 qwen3.7-plus: 76,676,335 qwen3.5-35b-a3b: 42,483,000 qwen3.5-plus: 36,539,350 qwen3-coder-30b-a3b-instruct: 902026-09-05 — 4,022,671,485 tokens qwen3.8-flash: 2,243,267,610 qwen3.8-max: 625,610,780 qwen3.7-flash: 456,534,395 90 more models: 382,875,145 qwen3.8-max-2026-09-02: 121,533,475 qwen3.7-plus: 113,934,270 qwen3.5-35b-a3b: 42,708,250 qwen3.5-plus: 36,207,370 qwen3-coder-30b-a3b-instruct: 1902026-09-06 — 2,626,546,915 tokens qwen3.8-flash: 783,405,980 qwen3.7-flash: 525,941,750 90 more models: 397,854,340 qwen3.8-max: 389,019,335 qwen3.8-max-2026-09-02: 213,127,295 qwen3.5-plus: 155,086,305 qwen3.7-plus: 108,501,455 qwen3.5-35b-a3b: 53,610,335 qwen3-coder-30b-a3b-instruct: 1202026-09-07 — 9,239,671,715 tokens qwen3.8-flash: 6,190,494,710 qwen3.7-flash: 1,257,714,380 90 more models: 639,543,945 qwen3.8-max: 454,783,380 qwen3.7-plus: 332,139,620 qwen3.8-max-2026-09-02: 206,455,505 qwen3.5-plus: 119,908,500 qwen3.5-35b-a3b: 38,631,505 qwen3-coder-30b-a3b-instruct: 1702026-09-08 — 5,829,326,175 tokens qwen3.7-flash: 2,068,261,435 qwen3.8-flash: 1,577,516,590 90 more models: 1,028,654,850 qwen3.8-max: 622,859,425 qwen3.7-plus: 328,252,005 qwen3.5-plus: 95,913,220 qwen3.8-max-2026-09-02: 67,832,015 qwen3.5-35b-a3b: 39,886,465 qwen3-coder-30b-a3b-instruct: 150,1702026-09-09 — 5,185,373,400 tokens qwen3.8-flash: 1,970,807,585 90 more models: 1,062,492,455 qwen3.7-flash: 799,413,475 qwen3.8-max: 423,545,040 qwen3.7-plus: 351,726,740 qwen3.5-plus: 323,675,495 qwen3.8-max-2026-09-02: 206,650,860 qwen3.5-35b-a3b: 47,043,485 qwen3-coder-30b-a3b-instruct: 18,2652026-09-10 — 5,191,505,810 tokens qwen3.7-flash: 1,975,783,065 90 more models: 903,638,380 qwen3.7-plus: 630,277,955 qwen3.8-flash: 626,499,440 qwen3.8-max-2026-09-02: 473,250,630 qwen3.8-max: 446,450,955 qwen3.5-plus: 97,651,545 qwen3.5-35b-a3b: 37,953,590 qwen3-coder-30b-a3b-instruct: 2502026-09-11 — 8,244,349,505 tokens qwen3.7-flash: 4,781,081,620 90 more models: 1,377,062,805 qwen3.8-flash: 550,154,925 qwen3.8-max: 538,164,185 qwen3.8-max-2026-09-02: 325,398,580 qwen3.7-plus: 298,764,370 qwen3.5-35b-a3b: 214,642,580 qwen3.5-plus: 159,080,4402026-09-12 — 3,545,208,675 tokens qwen3.8-max: 826,539,465 qwen3.7-flash: 766,822,830 90 more models: 607,438,035 qwen3.8-max-2026-09-02: 423,241,700 qwen3.5-plus: 413,881,555 qwen3.8-flash: 270,429,405 qwen3.7-plus: 186,075,810 qwen3.5-35b-a3b: 50,779,705 qwen3-coder-30b-a3b-instruct: 1702026-09-13 — 5,054,764,075 tokens qwen3.8-max: 1,247,804,290 qwen3.7-flash: 1,210,605,070 qwen3.8-flash: 821,153,755 90 more models: 697,372,775 qwen3.8-max-2026-09-02: 609,629,460 qwen3.7-plus: 243,396,330 qwen3.5-plus: 197,609,160 qwen3.5-35b-a3b: 27,193,145 qwen3-coder-30b-a3b-instruct: 902026-09-14 — 7,398,658,815 tokens qwen3.8-flash: 2,362,250,595 qwen3.7-flash: 1,747,887,370 90 more models: 1,142,206,725 qwen3.8-max: 1,127,928,580 qwen3.8-max-2026-09-02: 607,695,115 qwen3.5-plus: 286,588,165 qwen3.5-35b-a3b: 100,327,290 qwen3.7-plus: 23,774,135 qwen3-coder-30b-a3b-instruct: 8402026-09-15 — 8,960,980,510 tokens qwen3.8-max: 4,256,658,605 qwen3.7-flash: 1,843,092,255 90 more models: 1,362,462,635 qwen3.8-flash: 684,446,835 qwen3.5-35b-a3b: 294,812,400 qwen3.5-plus: 270,195,265 qwen3.8-max-2026-09-02: 163,879,425 qwen3.7-plus: 85,433,0902026-09-16 — 7,173,950,185 tokens 90 more models: 2,243,275,475 qwen3.7-flash: 1,763,559,090 qwen3.8-max: 1,221,507,235 qwen3.8-flash: 996,883,320 qwen3.5-35b-a3b: 834,655,585 qwen3.8-max-2026-09-02: 57,504,095 qwen3.5-plus: 42,347,025 qwen3.7-plus: 14,217,910 qwen3-coder-30b-a3b-instruct: 4502026-09-17 — 8,577,738,830 tokens qwen3.8-flash: 3,696,625,615 qwen3.7-flash: 2,619,348,795 90 more models: 956,967,230 qwen3.8-max: 638,809,955 qwen3.7-plus: 475,708,395 qwen3.5-35b-a3b: 142,311,500 qwen3.8-max-2026-09-02: 27,516,610 qwen3.5-plus: 20,450,650 qwen3-coder-30b-a3b-instruct: 802026-09-18 — 17,347,481,805 tokens qwen3.8-flash: 12,710,425,175 qwen3.8-max: 1,610,283,420 qwen3.7-flash: 1,132,189,755 90 more models: 939,905,235 qwen3.7-plus: 559,943,425 qwen3.8-max-2026-09-02: 305,132,615 qwen3.5-35b-a3b: 46,458,720 qwen3.5-plus: 43,138,935 qwen3-coder-30b-a3b-instruct: 4,5252026-09-19 — 8,982,063,990 tokens qwen3.8-flash: 4,082,188,335 qwen3.7-flash: 1,665,555,715 90 more models: 1,437,269,905 qwen3.7-plus: 611,623,125 qwen3.8-max-2026-09-02: 478,414,715 qwen3.8-max: 440,095,410 qwen3.5-35b-a3b: 207,011,505 qwen3.5-plus: 59,903,050 qwen3-coder-30b-a3b-instruct: 2,2302026-09-20 — 11,191,614,860 tokens qwen3.8-flash: 4,649,415,195 qwen3.7-flash: 1,861,462,445 qwen3.5-35b-a3b: 1,505,951,595 qwen3.8-max-2026-09-02: 1,149,355,110 qwen3.8-max: 867,399,815 qwen3.7-plus: 675,817,580 90 more models: 443,690,525 qwen3.5-plus: 38,522,5952026-09-21 — 14,041,612,085 tokens qwen3.8-max: 7,060,875,885 qwen3.8-flash: 2,386,480,075 90 more models: 1,615,042,380 qwen3.5-35b-a3b: 1,127,355,685 qwen3.7-flash: 665,805,600 qwen3.7-plus: 576,448,760 qwen3.5-plus: 359,281,865 qwen3.8-max-2026-09-02: 250,321,8352026-09-22 — 11,646,872,895 tokens qwen3.8-flash: 4,606,537,145 qwen3.5-35b-a3b: 2,074,607,095 qwen3.8-max: 2,010,537,490 qwen3.7-flash: 1,346,858,120 qwen3.7-plus: 721,761,690 90 more models: 719,682,910 qwen3.8-max-2026-09-02: 97,415,690 qwen3.5-plus: 69,470,355 qwen3-coder-30b-a3b-instruct: 2,4002026-09-23 — 12,287,628,245 tokens qwen3.8-flash: 5,739,027,535 qwen3.8-max: 2,385,397,500 90 more models: 1,181,949,805 qwen3.7-flash: 844,419,445 qwen3.5-35b-a3b: 791,707,345 qwen3.8-max-2026-09-02: 503,298,145 qwen3-coder-30b-a3b-instruct: 490,868,810 qwen3.7-plus: 281,089,645 qwen3.5-plus: 69,870,0152026-09-24 — 20,195,964,720 tokens qwen3.8-flash: 11,504,040,330 qwen3-coder-30b-a3b-instruct: 3,006,674,535 qwen3.8-max: 2,745,371,645 90 more models: 2,100,895,005 qwen3.5-plus: 341,480,180 qwen3.8-max-2026-09-02: 301,598,275 qwen3.7-flash: 125,287,835 qwen3.5-35b-a3b: 44,402,025 qwen3.7-plus: 26,214,890
  • qwen3.8-flash
  • qwen3.7-flash
  • qwen3.8-max
  • qwen3.7-plus
  • qwen3.5-35b-a3b
  • qwen3.8-max-2026-09-02
  • qwen3.5-plus
  • qwen3-coder-30b-a3b-instruct
  • 90 more models

Which models that traffic went to

  1. Qwen3.8 Flash33.8%89.6B
  2. Qwen3.7 Flash28.7%76.1B
  3. Qwen3.8 Max13.8%36.5B
  4. Qwen3.7 Plus3.6%9.5B
  5. Qwen3.5 35B A3B3.1%8.1B
  6. Qwen3.8 Max 2026 09-023.0%8B
  7. Qwen3.5 Plus1.4%3.8B
  8. Qwen3 Coder 30B A3B Instruct1.3%3.5B
  9. 90 more models11.3%30B

Share of 265B tokens. 41 models with traffic report no token counts and cannot be ranked here, including qwen-audio-3.0-tts-flash and qwen-image-3.0 — they are in the request view.

The two views disagree on purpose: a model can take a large share of the calls and a small share of the tokens — many short requests — or the reverse. Which one matters depends on whether your cost is driven by call volume or by prompt length. Measured on AIHubMix over the last 30 days, counting the 148 model IDs listed on this page; traffic routed through upstream-specific IDs that are not in the public catalog is not included.

All 148 Qwen Models

Open in model list
Qwen models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
qwen3-coder-plusTakes text, returns text.1.05M66K$0.54$2.16/M$0.108/M—83 tok/s1.63 s
qwen3.6-plus-preview-freeTakes text, returns text.1M66KFreeFree/M————
qwen3.5-flashTakes text, vision, video, returns text.1M66K$0.0282$0.282/M$0.0028/M$0.0352/M80 tok/s1.29 s
qwen3.7-flashTakes text, vision, video, returns text.1M131K$0.0282$0.1128/M$0.0056/M$0.0352/M58 tok/s3.17 s
qwen3.5-plusTakes text, vision, video, returns text.1M66K$0.1096$0.6576/M$0.011/M$0.137/M31 tok/s4.69 s
qwen3.8-flashTakes text, vision, video, returns text.1M131K$0.1126$0.38/M$0.0141/M$0.1759/M39 tok/s4.11 s
qwen3.8-omni-flashTakes text, vision, audio, video, returns text.1M131K$0.1126$0.38/M$0.0141/M$0.1759/M44 tok/s14.77 s
qwen3-coder-flashTakes text, returns text.1M66K$0.136$0.544/M——110 tok/s1.85 s
qwen3.6-flashTakes text, vision, video, returns text.1M66K$0.169$1.014/M$0.0169/M$0.2112/M62 tok/s9.61 s
qwen3.8-27bTakes text, vision, video, returns text.1M131K$0.2$2.5/M$0.05/M—46 tok/s0.62 s
qwen3.6-plusTakes text, vision, video, returns text.1M66K$0.282$1.692/M$0.0282/M$0.3525/M55 tok/s1.11 s
qwen3.7-plusTakes text, vision, video, returns text.1M131K$0.282$1.128/M$0.0564/M$0.3525/M54 tok/s6.90 s
qwen3.8-max-previewTakes text, vision, video, returns text.1M131K$0.338$1.014/M$0.0676/M$0.4225/M48 tok/s2.40 s
qwen3.7-maxTakes text, returns text.1M131K$1.69$5.07/M$0.169/M$2.1125/M50 tok/s1.18 s
qwen3.8-maxTakes text, vision, video, returns text.1M131K$1.69$5.07/M$0.169/M$2.1125/M31 tok/s3.80 s
qwen3.8-max-2026-09-02Takes text, vision, video, returns text.1M131K$1.69$5.07/M$0.169/M$2.1125/M29 tok/s5.03 s
qwen3.8-2.4t-a95bTakes text, returns text.1M131K$2$6/M$0.5/M—48 tok/s2.40 s
qwen3-vl-flashTakes text, vision, video, returns text.262K33K$0.0206$0.206/M$0.0041/M—2 tok/s1.95 s
qwen3-vl-flash-2026-01-22Takes text, vision, video, returns text.262K33K$0.0206$0.206/M——64 tok/s0.34 s
qwen3.5-35b-a3bTakes text, vision, video, returns text.262K66K$0.0564$0.4512/M——71 tok/s1.27 s
qwen3.5-27bTakes text, vision, video, returns text.262K66K$0.0846$0.6768/M——53 tok/s3.98 s
qwen3.5-122b-a10bTakes text, vision, video, returns text.262K66K$0.1126$0.9008/M——71 tok/s1.27 s
qwen3-coder-nextTakes text, returns text.262K66K$0.137$0.548/M——16 tok/s0.51 s
qwen3-vl-plusTakes text, vision, video, returns text.262K33K$0.137$1.37/M$0.0274/M—67 tok/s2.32 s
qwen3.5-397b-a17bTakes text, vision, video, returns text.262K66K$0.1644$0.9864/M——3 tok/s1.42 s
qwen3-coder-30b-a3b-instructTakes text, returns text.262K262K$0.2$0.8/M——115 tok/s0.95 s
qwen3.6-35b-a3bTakes text, vision, video, returns text.262K66K$0.254$1.524/M——33 tok/s2.62 s
qwen3-235b-a22b-instruct-2507Takes text, vision, returns text.262K16K$0.28$1.12/M——96 tok/s1.03 s
qwen3-235b-a22b-thinking-2507Takes text, vision, returns text.262K131K$0.28$2.8/M——87 tok/s0.27 s
qwen3-235b-a22b-2507Takes text, returns text.262K—$0.35$1.4/M$0.07/M———
qwen3.6-27bTakes text, vision, video, returns text.262K66K$0.422$2.532/M——81 tok/s1.84 s
qwen3-maxTakes text, returns text.262K66K$0.4508$1.8032/M$0.0902/M$0.5635/M13 tok/s0.64 s
qwen3-max-2026-01-23Takes text, returns text.262K66K$0.4508$1.8032/M$0.0902/M$0.5635/M——
qwen3.6-max-previewTakes text, returns text.262K66K$1.268$7.608/M$0.1268/M$1.585/M185 tok/s0.65 s
qwen3-coder-480b-a35b-instructTakes text, returns text.262K66K$0.82$3.28/M——1655 tok/s0.92 s
qwen3-next-80b-a3b-instructTakes text, vision, returns text.256K33K$0.138$0.552/M——150 tok/s0.10 s
qwen3-next-80b-a3b-thinkingTakes text, vision, returns text.256K33K$0.142$1.42/M——227 tok/s0.52 s
qwen3-235b-a22bTakes text, returns text.131K128K$0.28$1.12/M——81 tok/s0.63 s
qwen3-vl-30b-a3b-instructTakes text, vision, video, returns text.131K33K$0.1028$0.4112/M——42 tok/s1.07 s
qwen3-vl-30b-a3b-thinkingTakes text, vision, video, returns text.131K33K$0.1028$1.028/M——42 tok/s1.49 s
qwen3-vl-235b-a22b-instructTakes text, vision, video, returns text.131K33K$0.274$1.096/M——57 tok/s1.13 s
qwen3-vl-235b-a22b-thinkingTakes text, vision, video, returns text.131K33K$0.274$2.74/M——58 tok/s2.22 s
qwen-3.8-27bTakes text, vision, video, returns text.131K—$1.1$1.65/M——831 tok/s0.19 s
bai-qwen3-vl-235b-a22b-instructTakes , returns text.131K—$0.274$1.096/M——3 tok/s1.77 s
kat-devTakes text, returns text.128K—$0.137$0.548/M————
Qwen/Qwen2.5-VL-72B-InstructTakes text, vision, video, returns text.128K—$0.5$0.5/M————
qwen3-coder-plus-2025-07-22Takes text, returns text.128K66K$0.54$2.16/M$0.108/M—83 tok/s1.63 s
qwen3-reranker-0.6bTakes text, vision. Output modality not published.16K8K$0.11$0.11/M————
qwen-mt-turboTakes text, returns text.16K8K$0.192$0.5349/M——13 tok/s0.41 s
qwen-mt-plusTakes text, returns text.16K8K$0.492$1.476/M——82 tok/s2.54 s
qwen-audio-3.0-tts-flashTakes text, returns audio.——Free$0.141/M————
qwen-audio-3.0-tts-plusTakes text, returns audio.——Free$0.197/M————
qwen-image-3.0Takes text, vision, returns vision.——FreeFree/M————
qwen-image-3.0-proTakes text, vision, returns vision.——FreeFree/M————
qwen-flash——$0.02$0.2/M——31 tok/s0.38 s
qwen-flash-2025-07-28——$0.02$0.2/M————
qwen-turboTakes text, returns text.——$0.046$0.092/M$0.0092/M—13 tok/s0.40 s
qwen-turbo-2024-11-01Takes text, returns text.——$0.046$0.092/M——94 tok/s0.66 s
qwen-turbo-2025-04-28Takes , returns text.——$0.046$0.092/M————
qwen-turbo-latestTakes , returns text.——$0.046$0.092/M$0.0092/M———
qwen3-0.6bTakes , returns text.——$0.046$0.46/M————
qwen3-1.7bTakes , returns text.——$0.046$0.46/M————
qwen3-4bTakes , returns text.——$0.046$0.46/M————
bce-reranker-baseTakes text, vision. Output modality not published.——$0.068$0.068/M————
qwen3-embedding-0.6bTakes text. Output modality not published.——$0.068$0.068/M————
qwen3-embedding-4bTakes text. Output modality not published.——$0.068$0.068/M————
qwen3-embedding-8bTakes text. Output modality not published.——$0.068$0.068/M————
Qwen/Qwen2-7B-Instruct——$0.08$0.08/M————
qwen3-8bTakes , returns text.——$0.08$0.8/M——58 tok/s7.91 s
text-embedding-v4Takes text. Output modality not published.——$0.08$0.08/M————
qwen-long——$0.1$0.4/M——55 tok/s0.67 s
qwen3-30b-a3b-instruct-2507——$0.1028$0.4112/M————
gte-rerank-v2Takes text, vision. Output modality not published.——$0.11$0.11/M————
qwen3-reranker-4bTakes text, vision. Output modality not published.——$0.11$0.11/M————
qwen3-reranker-8bTakes text, vision. Output modality not published.——$0.11$0.11/M————
qwen-plus——$0.1126$1.126/M$0.0225/M$0.1407/M62 tok/s0.89 s
qwen-plus-2025-04-28Takes , returns text.——$0.1126$1.126/M$0.0225/M$0.1407/M10 tok/s0.45 s
qwen-plus-2025-07-28——$0.1126$1.126/M$0.0225/M$0.1407/M——
qwen-plus-latestTakes , returns text.——$0.1126$1.126/M$0.0225/M$0.1407/M39 tok/s0.93 s
qwen3-30b-a3b——$0.12$1.2/M————
qwen3-30b-a3b-thinking-2507——$0.12$1.2/M————
gme-qwen2-vl-2b-instructTakes text, vision, video. Output modality not published.——$0.138$0.138/M————
Qwen/QwQ-32BTakes , returns text.——$0.14$0.56/M————
Qwen/Qwen2.5-Coder-32B-Instruct——$0.16$0.16/M————
Qwen/QwQ-32B-Preview——$0.16$0.16/M————
qwen3-14bTakes , returns text.——$0.16$1.6/M——45 tok/s0.41 s
qwen3-32b——$0.16$0.64/M————
Qwen/Qwen2-1.5B-Instruct——$0.2$0.2/M————
Qwen/Qwen3-8B——$0.2$0.2/M——11 tok/s3.31 s
qwen2.5-coder-1.5b-instruct——$0.2$0.4/M————
qwen2.5-coder-7b-instruct——$0.2$0.4/M————
qwen2.5-math-1.5b-instruct——$0.2$0.2/M————
qwen2.5-math-7b-instruct——$0.2$0.4/M————
Qwen/Qwen2-57B-A14B-Instruct——$0.24$0.24/M————
Qwen/Qwen2.5-VL-32B-InstructTakes text, vision, video, returns text.——$0.24$0.24/M————
qwen-3-235b-a22b-instruct-2507——$0.28$1.4/M————
qwen-3-235b-a22b-thinking-2507——$0.28$2.8/M————
Qwen2-VL-7B-InstructTakes text, vision, video. Output modality not published.——$0.28$0.7/M————
Qwen3-235B-A22B-Thinking-2507——$0.28$2.8/M——87 tok/s0.27 s
qwen-max——$0.38$1.52/M——28 tok/s0.41 s
qwen-max-0125——$0.38$1.52/M————
qwen-3-32b——$0.4$1.6/M————
qwen-qwq-32b——$0.4$0.8/M————
Qwen/Qwen2.5-7B-Instruct——$0.4$0.4/M————
Qwen/Qwen3-32B——$0.4$0.8/M————
qwen2.5-14b-instruct——$0.4$1.2/M————
qwen2.5-3b-instruct——$0.4$0.8/M————
qwen2.5-7b-instruct——$0.4$0.8/M————
Qwen/Qwen3-14B——$0.5$0.5/M————
Qwen/Qwen2.5-32B-Instruct——$0.6$0.6/M——24 tok/s1.62 s
qwen2.5-32b-instruct——$0.6$1.2/M————
Qwen/Qwen2-72B-Instruct——$0.8$0.8/M————
Qwen/Qwen2.5-72B-Instruct——$0.8$0.8/M————
Qwen/Qwen2.5-72B-Instruct-128K——$0.8$0.8/M————
qwen2.5-72b-instruct——$0.8$2.4/M————
qwen2.5-math-72b-instruct——$0.8$2.4/M————
qwen3-max-previewTakes text, vision, returns text.——$0.846$3.384/M$0.1692/M—10 tok/s0.53 s
Qwen/Qwen3-30B-A3B——$1$1/M————
Qwen/QVQ-72B-Preview——$1.2$1.2/M————
qwen-imageTakes text, vision, returns vision.——$2$2/M————
qwen-image-2.0Takes text, vision, returns vision.——$2$2/M————
qwen-image-2.0-proTakes text, vision, returns vision.——$2$2/M————
qwen-image-editTakes text, vision, returns vision.——$2$2/M————
qwen-image-maxTakes text, vision, returns vision.——$2$2/M————
wan2.7-imageTakes text, vision, returns vision.——$2$2/M————
wan2.7-image-proTakes text, vision, returns vision.——$2$2/M————
Qwen2-VL-72B-InstructTakes text, vision, video. Output modality not published.——$2.18$6.54/M————
qwen2.5-vl-72b-instructTakes text, vision, returns text.——$2.4$7.2/M————
qwen-max-longcontext——$7$21/M————
wan2.2-i2v-plusTakes text, vision, returns video.——$480$480/M————
wan2.5-i2v-previewTakes text, vision, returns video.——$480$480/M————
wan2.5-t2v-previewTakes text, returns video.——$480$480/M————
wan3.0-videoTakes text, vision, audio, video, returns video.——$480$480/M————
wan3.0-video-primeTakes text, vision, audio, video, returns video.——$480$480/M————
happyhorse-1.0-i2vTakes text, vision, returns video.——$720$720/M————
happyhorse-1.0-r2vTakes text, vision, returns video.——$720$720/M————
happyhorse-1.0-t2vTakes text, returns video.——$720$720/M————
happyhorse-1.0-video-editTakes text, vision, video, returns video.——$720$720/M————
happyhorse-1.1-i2vTakes text, vision, returns video.——$720$720/M————
happyhorse-1.1-r2vTakes text, vision, returns video.——$720$720/M————
happyhorse-1.1-t2vTakes text, returns video.——$720$720/M————
wan2.6-i2vTakes text, vision, returns video.——$720$720/M————
wan2.6-t2iTakes text, vision, returns vision.——$720$720/M————
wan2.6-t2vTakes text, returns video.——$720$720/M————
wan2.7-i2vTakes text, vision, audio, returns video.——$720$720/M————
wan2.7-r2vTakes text, vision, audio, video, returns video.——$720$720/M————
wan2.7-t2vTakes text, audio, returns video.——$720$720/M————
wan2.7-videoeditTakes text, vision, video, returns video.——$720$720/M————

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

Qwen on AIHubMix

Which Qwen model should I start with?

qwen3-vl-flash at $0.0206/M input — the cheapest entry here that declares tool calling, and it carries a 262K context. Move up to happyhorse-1.0-i2v when answer quality matters more than cost, or to qwen3.6-plus-preview-free for long-form reasoning.

Which of these models reason before answering?

25 of the 148 models here declare a reasoning phase — they work through the problem before producing an answer, which helps on multi-step problems at the cost of extra output tokens. Use the Reasoning filter above the table to see them. The catalog does not record anything further about how they differ, so this page does not sort them into families.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), some are the open-weight repository form (Qwen/…), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

How is cached input billed?

The Cache read column is the rate for input tokens served from the prompt cache — for example qwen3.5-flash bills cache hits at 10% of the input rate and qwen3.5-plus bills cache hits at 10% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.

Do I need a separate Qwen account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling Qwen in one line

One key, one endpoint, 908 models across 41 model authors.