OpenAI Models

141 modelsGeneral models free to startUp to 1.05M context

Usage

Last 30 days · 2026-08-26 to 2026-09-24

Tokens

1139B

Requests

74.3M

Models in use

101 of 141

Tokens per day, stacked by model

032.9B65.8B08-2609-0209-0909-1609-232026-08-26 — 50,061,293,925 tokens gpt-5.6-luna: 17,886,729,770 gpt-5.6-sol: 10,575,946,460 omni-moderation-latest: 5,985,594,420 68 more models: 5,374,132,430 gpt-4o-mini: 3,590,992,920 gpt-5.6-terra: 2,919,577,285 gpt-5.4-mini: 2,473,716,815 gpt-5.5: 1,254,603,8252026-08-27 — 42,457,561,595 tokens gpt-5.6-luna: 12,030,529,890 gpt-5.6-sol: 7,224,017,885 68 more models: 5,584,290,280 omni-moderation-latest: 5,401,546,400 gpt-4o-mini: 4,306,524,590 gpt-5.4-mini: 3,246,010,690 gpt-5.6-terra: 3,121,793,900 gpt-5.5: 1,542,847,9602026-08-28 — 31,844,035,350 tokens gpt-5.6-luna: 9,539,313,300 gpt-5.6-sol: 6,553,281,695 gpt-4o-mini: 4,210,950,825 68 more models: 3,885,064,745 omni-moderation-latest: 3,218,460,175 gpt-5.6-terra: 2,621,238,600 gpt-5.5: 1,000,492,940 gpt-5.4-mini: 815,233,0702026-08-29 — 20,434,916,100 tokens gpt-5.6-luna: 7,361,729,420 gpt-4o-mini: 3,315,014,415 68 more models: 3,035,220,230 gpt-5.6-sol: 2,835,756,100 omni-moderation-latest: 1,783,618,130 gpt-5.4-mini: 848,581,450 gpt-5.5: 794,088,680 gpt-5.6-terra: 460,907,6752026-08-30 — 20,001,907,215 tokens gpt-5.6-luna: 6,123,519,480 68 more models: 4,417,656,470 gpt-4o-mini: 2,695,687,130 gpt-5.6-sol: 2,435,225,610 omni-moderation-latest: 1,929,593,135 gpt-5.6-terra: 1,012,717,870 gpt-5.4-mini: 896,549,430 gpt-5.5: 490,958,0902026-08-31 — 38,857,820,740 tokens gpt-5.6-luna: 9,816,841,965 gpt-5.6-sol: 8,732,840,320 68 more models: 6,830,664,570 gpt-5.6-terra: 4,446,110,735 gpt-4o-mini: 3,583,158,430 omni-moderation-latest: 2,897,572,065 gpt-5.5: 1,386,135,625 gpt-5.4-mini: 1,164,497,0302026-09-01 — 43,685,093,910 tokens gpt-5.6-luna: 13,879,890,225 gpt-5.6-sol: 11,180,256,740 gpt-4o-mini: 7,015,598,660 68 more models: 4,173,805,225 gpt-5.6-terra: 3,238,286,090 omni-moderation-latest: 1,764,455,460 gpt-5.5: 1,653,594,145 gpt-5.4-mini: 779,207,3652026-09-02 — 38,222,420,860 tokens gpt-5.6-luna: 11,397,495,330 gpt-5.6-sol: 9,338,840,605 68 more models: 5,418,345,070 gpt-4o-mini: 4,770,885,020 gpt-5.5: 2,893,338,470 gpt-5.6-terra: 2,538,607,895 gpt-5.4-mini: 999,082,765 omni-moderation-latest: 865,825,7052026-09-03 — 42,991,346,590 tokens gpt-5.6-luna: 13,326,507,300 gpt-5.6-sol: 10,238,582,745 68 more models: 7,762,885,595 gpt-4o-mini: 4,073,560,535 gpt-5.6-terra: 2,937,210,240 gpt-5.5: 2,033,487,040 gpt-5.4-mini: 1,824,574,105 omni-moderation-latest: 794,539,0302026-09-04 — 34,343,138,745 tokens gpt-5.6-sol: 9,599,622,210 gpt-5.6-luna: 6,896,668,845 68 more models: 6,558,806,555 gpt-4o-mini: 3,548,708,695 gpt-5.6-terra: 2,860,377,160 gpt-5.4-mini: 2,306,338,105 gpt-5.5: 1,729,450,510 omni-moderation-latest: 843,166,6652026-09-05 — 19,002,576,490 tokens gpt-5.6-luna: 5,442,461,930 gpt-4o-mini: 4,548,632,440 68 more models: 3,314,553,650 gpt-5.6-sol: 1,655,025,830 gpt-5.4-mini: 1,575,443,960 gpt-5.6-terra: 1,091,145,705 omni-moderation-latest: 750,348,610 gpt-6-astra: 319,056,305 gpt-5.5: 305,908,0602026-09-06 — 19,322,362,185 tokens gpt-5.6-luna: 6,122,995,690 68 more models: 4,774,988,950 gpt-4o-mini: 2,381,496,365 gpt-5.4-mini: 1,904,111,250 gpt-5.6-sol: 1,526,206,650 gpt-5.6-terra: 849,602,895 gpt-6-astra: 747,594,890 omni-moderation-latest: 536,732,495 gpt-5.5: 478,633,0002026-09-07 — 39,499,954,290 tokens gpt-5.6-sol: 8,749,795,590 68 more models: 7,734,713,805 gpt-5.6-luna: 7,137,263,325 gpt-4o-mini: 4,248,041,170 gpt-6-astra: 3,859,903,605 gpt-5.5: 2,505,995,225 gpt-5.6-terra: 2,376,104,105 gpt-5.4-mini: 1,995,376,050 omni-moderation-latest: 892,761,4152026-09-08 — 38,624,887,870 tokens gpt-5.6-sol: 10,644,835,915 gpt-5.6-luna: 8,709,075,780 gpt-4o-mini: 5,449,317,490 68 more models: 4,107,500,555 gpt-5.6-terra: 2,561,729,025 gpt-6-astra: 2,497,865,420 gpt-5.5: 1,869,281,760 gpt-5.4-mini: 1,682,110,825 omni-moderation-latest: 1,103,171,1002026-09-09 — 39,498,966,685 tokens gpt-5.6-luna: 9,964,021,385 gpt-5.6-sol: 9,577,093,945 gpt-4o-mini: 5,602,063,810 68 more models: 4,639,622,370 gpt-6-astra: 2,832,395,730 gpt-5.5: 2,015,511,040 gpt-5.6-terra: 1,993,715,985 omni-moderation-latest: 1,694,143,660 gpt-5.4-mini: 1,180,398,7602026-09-10 — 35,350,859,115 tokens gpt-5.6-sol: 7,309,535,010 gpt-5.6-luna: 7,214,112,490 gpt-4o-mini: 4,843,444,700 gpt-6-astra: 4,308,521,105 68 more models: 4,168,805,120 gpt-5.6-terra: 2,687,330,215 omni-moderation-latest: 1,835,226,460 gpt-5.5: 1,670,790,730 gpt-5.4-mini: 1,313,093,2852026-09-11 — 30,326,375,965 tokens gpt-5.6-luna: 9,026,348,975 gpt-5.6-sol: 5,380,311,410 gpt-4o-mini: 4,832,847,825 68 more models: 3,858,413,660 gpt-6-astra: 1,966,558,480 gpt-5.4-mini: 1,639,874,880 omni-moderation-latest: 1,623,597,340 gpt-5.6-terra: 1,403,794,275 gpt-5.5: 594,629,1202026-09-12 — 22,139,402,950 tokens gpt-5.6-luna: 5,214,649,845 68 more models: 4,967,004,620 gpt-4o-mini: 4,559,452,365 gpt-5.6-sol: 3,304,867,035 gpt-5.4-mini: 1,317,107,995 omni-moderation-latest: 1,134,010,660 gpt-6-astra: 877,343,550 gpt-5.5: 407,178,700 gpt-5.6-terra: 357,788,1802026-09-13 — 22,016,795,645 tokens gpt-5.6-luna: 6,477,528,075 gpt-5.6-sol: 5,299,523,440 gpt-4o-mini: 2,821,375,020 68 more models: 2,514,371,090 gpt-6-astra: 1,514,862,405 gpt-5.4-mini: 1,511,840,295 omni-moderation-latest: 1,241,593,315 gpt-5.5: 368,401,650 gpt-5.6-terra: 267,300,3552026-09-14 — 36,871,744,245 tokens gpt-5.6-luna: 9,260,854,110 gpt-5.6-sol: 6,598,984,350 68 more models: 6,106,023,315 gpt-4o-mini: 4,376,445,645 gpt-6-astra: 4,254,931,125 gpt-5.5: 1,784,342,755 omni-moderation-latest: 1,630,247,910 gpt-5.6-terra: 1,495,935,040 gpt-5.4-mini: 1,363,979,9952026-09-15 — 42,729,050,550 tokens gpt-5.6-sol: 9,754,244,075 gpt-5.6-luna: 8,831,454,385 68 more models: 7,380,412,460 gpt-6-astra: 5,735,445,015 gpt-4o-mini: 4,883,648,745 gpt-5.6-terra: 2,200,890,745 gpt-5.4-mini: 1,934,534,385 omni-moderation-latest: 1,286,738,345 gpt-5.5: 721,682,3952026-09-16 — 51,201,319,050 tokens gpt-5.6-sol: 12,069,927,660 gpt-5.6-luna: 10,182,979,100 68 more models: 8,185,094,675 gpt-6-astra: 7,086,083,920 gpt-4o-mini: 5,312,493,435 gpt-5.6-terra: 4,286,463,615 omni-moderation-latest: 1,471,675,880 gpt-5.4-mini: 1,387,525,130 gpt-5.5: 1,219,075,6352026-09-17 — 65,766,861,945 tokens gpt-5.6-sol: 25,563,320,980 gpt-5.6-luna: 10,780,260,915 68 more models: 5,877,082,220 gpt-6-astra: 5,125,907,085 gpt-4o-mini: 4,878,062,215 gpt-5.4-mini: 4,446,371,540 gpt-5.6-terra: 4,029,780,920 gpt-5.5: 3,872,744,205 omni-moderation-latest: 1,193,331,8652026-09-18 — 46,712,525,145 tokens gpt-5.6-sol: 16,343,970,000 gpt-5.6-luna: 8,724,032,005 68 more models: 4,968,651,365 gpt-4o-mini: 4,765,069,765 gpt-5.6-terra: 4,457,009,010 gpt-6-astra: 3,688,307,085 gpt-5.4-mini: 1,934,787,435 omni-moderation-latest: 1,231,240,015 gpt-5.5: 599,458,4652026-09-19 — 27,984,373,070 tokens gpt-5.6-sol: 7,190,629,995 gpt-5.6-luna: 6,075,272,165 gpt-4o-mini: 4,666,013,530 gpt-6-astra: 3,434,063,075 68 more models: 3,234,976,375 gpt-5.6-terra: 1,793,728,100 omni-moderation-latest: 737,586,750 gpt-5.4-mini: 431,221,265 gpt-5.5: 420,881,8152026-09-20 — 38,980,904,505 tokens gpt-5.6-sol: 15,018,860,815 gpt-5.6-luna: 8,750,798,655 gpt-4o-mini: 4,139,917,520 gpt-5.6-terra: 3,385,081,020 68 more models: 3,242,539,290 gpt-5.5: 2,099,540,085 omni-moderation-latest: 1,013,786,270 gpt-6-astra: 831,141,775 gpt-5.4-mini: 499,239,0752026-09-21 — 39,663,350,525 tokens gpt-5.6-luna: 10,039,909,815 gpt-5.6-sol: 9,128,282,740 gpt-4o-mini: 5,961,223,105 gpt-6-astra: 4,920,742,960 68 more models: 3,560,891,915 gpt-5.5: 2,448,138,530 gpt-5.6-terra: 2,074,045,820 omni-moderation-latest: 918,111,800 gpt-5.4-mini: 612,003,8402026-09-22 — 41,900,101,100 tokens gpt-5.6-luna: 11,987,518,135 gpt-5.6-sol: 11,285,144,160 gpt-4o-mini: 6,395,635,390 68 more models: 4,100,974,890 gpt-6-astra: 3,953,115,385 gpt-5.6-terra: 1,651,740,235 omni-moderation-latest: 1,071,803,160 gpt-5.5: 776,450,745 gpt-5.4-mini: 677,719,0002026-09-23 — 54,074,601,505 tokens 68 more models: 19,901,659,645 gpt-4o-mini: 9,236,943,185 gpt-5.6-luna: 8,624,616,655 gpt-5.6-sol: 6,957,399,305 gpt-6-astra: 4,303,637,915 gpt-5.6-terra: 2,234,587,820 omni-moderation-latest: 1,546,223,510 gpt-5.4-mini: 689,199,910 gpt-5.5: 580,333,5602026-09-24 — 64,308,834,290 tokens 68 more models: 29,394,415,470 gpt-4o-mini: 10,087,879,690 gpt-6-astra: 9,110,270,565 gpt-5.6-sol: 5,982,148,530 gpt-5.6-luna: 5,432,003,050 omni-moderation-latest: 1,632,400,850 gpt-5.6-terra: 1,423,326,705 gpt-5.5: 634,344,735 gpt-5.4-mini: 612,044,695
  • gpt-5.6-luna
  • gpt-5.6-sol
  • gpt-4o-mini
  • gpt-6-astra
  • gpt-5.6-terra
  • omni-moderation-latest
  • gpt-5.4-mini
  • gpt-5.5
  • 68 more models

Which models that traffic went to

  1. GPT 5.6 Luna23.9%272B
  2. GPT 5.6 Sol22.7%258B
  3. GPT 4o Mini12.7%145B
  4. GPT 6 Astra6.3%71.4B
  5. GPT 5.6 Terra6.0%68.8B
  6. Omni Moderation4.4%50B
  7. GPT 5.4 Mini3.9%44.1B
  8. GPT 5.53.5%40.2B
  9. 68 more models16.6%189B

Share of 1139B tokens. 25 models with traffic report no token counts and cannot be ranked here, including tts-1 and gpt-4o-mini-tts — they are in the request view.

The two views disagree on purpose: a model can take a large share of the calls and a small share of the tokens — many short requests — or the reverse. Which one matters depends on whether your cost is driven by call volume or by prompt length. Measured on AIHubMix over the last 30 days, counting the 141 model IDs listed on this page; traffic routed through upstream-specific IDs that are not in the public catalog is not included.

All 141 OpenAI Models

Open in model list
OpenAI models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
gpt-5.5-freeTakes text, vision, PDF, returns text.1.05M128KFreeFree/MFree/M—66 tok/s19.67 s
gpt-6-lunaTakes text, vision, returns text.1.05M128K$0.1$0.5/M$0.01/M$0.125/M81 tok/s6.17 s
gpt-5.6-lunaTakes text, vision, returns text.1.05M128K$0.2$1.2/M$0.02/M$0.25/M48 tok/s4.27 s
gpt-5.6-terraTakes text, vision, returns text.1.05M128K$2$12/M$0.2/M$2.5/M59 tok/s3.82 s
gpt-6-solTakes text, vision, returns text.1.05M128K$2$10/M$0.2/M$2.5/M52 tok/s7.09 s
gpt-5.4Takes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—100 tok/s1.56 s
gpt-5.4-highTakes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—85 tok/s9.19 s
gpt-5.4-lowTakes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—86 tok/s1.85 s
gpt-5.6-solTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M41 tok/s6.16 s
gpt-5.6-sol-discTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M60 tok/s11.09 s
gpt-5.5Takes text, vision, PDF, returns text.1.05M128K$5$30/M$0.5/M—64 tok/s4.13 s
gpt-5.6-sol-proTakes text, vision, returns text.1.05M—$8$40/M$0.8/M—99 tok/s16.12 s
gpt-6-astraTakes text, vision, returns text.1.05M128K$10$50/M$1/M$12.5/M30 tok/s9.43 s
gpt-5.4-proTakes text, vision, returns text.1.05M128K$30$180/M——42 tok/s8.60 s
gpt-5.5-proTakes text, vision, returns text.1.05M128K$30$180/M——26 tok/s78.91 s
gpt-4.1-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—37 tok/s3.09 s
gpt-4.1-mini-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—24 tok/s1.53 s
gpt-4.1-nano-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—19 tok/s1.25 s
gpt-4.1-nanoTakes text, vision, PDF, returns text.1.05M33K$0.1$0.4/M$0.025/M—130 tok/s0.82 s
gpt-4.1-miniTakes text, vision, PDF, returns text.1.05M33K$0.4$1.6/M$0.1/M—58 tok/s0.74 s
gpt-4.1Takes text, vision, PDF, returns text.1.05M33K$2$8/M$0.5/M—69 tok/s1.13 s
autoTakes text, vision, audio, video, returns text.1M—FreeFree/M————
gpt-5-nanoTakes text, vision, returns text.400K128K$0.05$0.4/M$0.005/M—83 tok/s2.19 s
gpt-5.4-nanoTakes text, vision, returns text.400K128K$0.2$1.25/M$0.02/M—102 tok/s0.65 s
gpt-5-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M—80 tok/s2.84 s
gpt-5.1-codex-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M—168 tok/s0.70 s
gpt-5.4-miniTakes text, vision, returns text.400K128K$0.75$4.5/M$0.075/M—129 tok/s1.07 s
gpt-5Takes text, vision, returns text.400K128K$1.25$10/M$0.125/M—64 tok/s7.76 s
gpt-5-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—27 tok/s7.54 s
gpt-5.1Takes text, vision, PDF, returns text.400K128K$1.25$10/M$0.125/M—88 tok/s1.10 s
gpt-5.1-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—82 tok/s0.45 s
gpt-5.1-codex-maxTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—1 tok/s14.80 s
gpt-5.2Takes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—59 tok/s2.63 s
gpt-5.2-codexTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—82 tok/s0.66 s
gpt-5.2-highTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—48 tok/s0.94 s
gpt-5.2-lowTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—37 tok/s1.50 s
gpt-5.3-codexTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—30 tok/s8.29 s
gpt-chat-latestTakes text, vision, returns text.400K128K$5$30/M$0.5/M—119 tok/s3.29 s
gpt-5-proTakes text, vision, returns text.400K272K$15$120/M——9 tok/s313.10 s
gpt-5.2-proTakes text, vision, returns text.400K128K$21$168/M$2.1/M—33 tok/s23.97 s
o3-miniTakes text, vision, returns text.200K100K$1.1$4.4/M$0.55/M—474 tok/s10.90 s
o4-miniTakes text, vision, PDF, returns text.200K100K$1.1$4.4/M$0.275/M—97 tok/s5.80 s
codex-mini-latestTakes text, vision, returns text.200K—$1.5$6/M$0.375/M———
o3Takes text, vision, PDF, returns text.200K100K$2$8/M$0.5/M—66 tok/s5.06 s
o3-proTakes text, vision, returns text.200K100K$20$80/M$20/M—13 tok/s88.94 s
o1-proTakes text, returns text.200K—$170$680/M$170/M—19 tok/s96.00 s
gpt-oss-20b-freeTakes text, returns text.131K—FreeFree/M————
gpt-oss-120bTakes text, returns text.131K33K$0.18$0.9/M——1101 tok/s0.12 s
gpt-4o-freeTakes text, vision, PDF, returns text.128K16KFreeFree/MFree/M—18 tok/s6.84 s
gpt-realtime-2.1Takes text, vision, audio. Output modality not published.128K32KFreeFree/MFree/M———
gpt-oss-20bTakes text, returns text.128K33K$0.11$0.55/M——2619 tok/s0.12 s
gpt-4o-miniTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M—62 tok/s0.62 s
gpt-4o-mini-audio-previewTakes text, audio, returns text.128K16K$0.15$0.6/M————
gpt-4o-mini-search-previewTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M—189 tok/s1.57 s
gpt-5-chat-latestTakes text, vision, returns text.128K16K$1.25$10/M$0.125/M—77 tok/s0.85 s
gpt-5.1-chat-latestTakes text, vision, returns text.128K16K$1.25$10/M$0.125/M—107 tok/s0.94 s
gpt-5.2-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M—84 tok/s0.77 s
gpt-5.3-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M—100 tok/s0.82 s
gpt-4oTakes text, vision, PDF, returns text.128K16K$2.5$10/M$1.25/M—52 tok/s0.64 s
gpt-4o-2024-11-20Takes text, vision, returns text.128K16K$2.5$10/M$1.25/M—61 tok/s0.60 s
gpt-4o-audio-previewTakes text, audio, returns text.128K16K$2.5$10/M——10 tok/s2.49 s
gpt-4o-search-previewTakes text, vision, returns text.128K16K$2.5$10/M$1.25/M—121 tok/s2.33 s
gpt-audio-1.5Takes text, audio, returns text, audio.128K16K$2.5$10/M————
gpt-4o-2024-05-13128K4K$5$15/M$5/M—199 tok/s0.45 s
gpt-4o-transcribe-diarizeTakes text, audio, returns text.16K2K$2.5$10/M————
o1Takes text, returns text.0K—$15$60/M$7.5/M—104 tok/s59.85 s
dall-e-2Takes text, vision, returns vision.——FreeFree/M————
dall-e-3Takes text, vision, returns vision.——FreeFree/M————
gpt-image-1Takes text, vision, returns vision.——FreeFree/M————
gpt-image-1-miniTakes text, vision, returns vision.——FreeFree/M————
gpt-image-1.5Takes text, vision, returns vision.——FreeFree/M————
gpt-image-2Takes text, vision, returns vision.——FreeFree/MFree/M———
gpt-image-2-freeTakes text, vision, returns text, vision.——FreeFree/M————
gpt-image-2.5-flareTakes text, vision, returns vision.——FreeFree/MFree/M———
gpt-image-2.5-sunburstTakes text, vision, returns vision.——FreeFree/MFree/M———
gpt-live-transcribeTakes text, audio, returns text.——FreeFree/M————
sora-2Takes , returns video.——FreeFree/M————
sora-2-proTakes , returns video.——Free$720/M————
whisper-1Takes audio, returns text.——FreeFree/M————
whisper-large-v3Takes audio, returns text.——FreeFree/M————
whisper-large-v3-turboTakes audio, returns text.——FreeFree/M————
omni-moderation-latest——$0.02$0.02/M————
text-embedding-3-smallTakes text. Output modality not published.——$0.02$0.02/M————
text-embedding-ada-002Takes text. Output modality not published.——$0.1$0.1/M————
text-embedding-v1Takes text. Output modality not published.——$0.1$0.1/M————
GPT-OSS-20B——$0.11$0.55/M——2619 tok/s0.12 s
text-embedding-3-largeTakes text. Output modality not published.——$0.13$0.13/M————
gpt-4o-mini-2024-07-18Takes text, vision, returns text.——$0.15$0.6/M$0.075/M—80 tok/s1.09 s
gpt-4o-mini-global——$0.15$0.6/M$0.075/M———
text-moderation-007——$0.2$0.2/M————
text-moderation-latest——$0.2$0.2/M————
text-moderation-stable——$0.2$0.2/M————
aihubmix-routerTakes text, vision, returns text.——$0.4$1.6/M$0.1/M—46 tok/s3.01 s
text-ada-001——$0.4$0.4/M————
gpt-3.5-turbo——$0.5$1.5/M————
text-babbage-001——$0.5$0.5/M————
gpt-4o-mini-ttsTakes audio, returns audio.——$0.6$12/M———0.95 s
gpt-3.5-turbo-1106——$1$2/M————
o3-mini-global——$1.1$4.4/M$0.55/M———
gpt-3.5-turbo-0301——$1.5$1.5/M————
gpt-3.5-turbo-0613——$1.5$2/M————
gpt-3.5-turbo-instruct——$1.5$2/M————
davinci-002——$2$2/M————
o3-global——$2$8/M$0.5/M———
text-curie-001——$2$2/M————
gpt-4o-2024-08-06——$2.5$10/M$1.25/M—74 tok/s1.90 s
gpt-4o-2024-08-06-global——$2.5$10/M$1.25/M———
gpt-4o-zhTakes text, vision, returns text.——$2.5$10/M————
computer-use-preview——$3$12/M————
gpt-3.5-turbo-16k——$3$4/M————
gpt-3.5-turbo-16k-0613——$3$4/M————
o1-mini——$3$12/M$1.5/M———
o1-mini-2024-09-12——$3$12/M$1.5/M———
gpt-image-test——$5$40/M————
distil-whisper-large-v3-enTakes audio, returns text.——$5.556$5.556/M————
gpt-4-0125-preview——$10$30/M————
gpt-4-1106-preview——$10$30/M————
gpt-4-turbo——$10$30/M————
gpt-4-turbo-2024-04-09——$10$30/M————
gpt-4-turbo-preview——$10$30/M————
gpt-4-vision-preview——$10$30/M————
o3-deep-research——$10$40/M$2.5/M———
o1-2024-12-17Takes text, vision, returns text.——$15$60/M$7.5/M———
o1-previewTakes text, vision, returns text.——$15$60/M$7.5/M———
o1-preview-2024-09-12——$15$60/M$7.5/M———
tts-1Takes audio, returns audio.——$15$15/M————
tts-1-1106Takes audio, returns audio.——$15$15/M————
davinci——$20$20/M————
o3-pro-global——$20$80/M————
text-davinci-002——$20$20/M————
text-davinci-003——$20$20/M————
text-davinci-edit-001——$20$20/M————
text-search-ada-doc-001——$20$20/M————
gpt-4——$30$60/M————
gpt-4-0314——$30$60/M————
gpt-4-0613——$30$60/M————
tts-1-hdTakes audio, returns audio.——$30$30/M————
tts-1-hd-1106Takes audio, returns audio.——$30$30/M————
gpt-4-32k——$60$120/M————
gpt-4-32k-0314——$60$120/M————
gpt-4-32k-0613——$60$120/M————

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

OpenAI on AIHubMix

Which OpenAI model should I start with?

auto is free on input — the cheapest entry here that declares tool calling, and it carries a 1M context. Move up to o1-pro when answer quality matters more than cost, or to gpt-5.5-free for long-form reasoning.

Which of these models reason before answering?

43 of the 141 models here declare a reasoning phase — they work through the problem before producing an answer, which helps on multi-step problems at the cost of extra output tokens. Use the Reasoning filter above the table to see them. The catalog does not record anything further about how they differ, so this page does not sort them into families.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

How is cached input billed?

The Cache read column is the rate for input tokens served from the prompt cache — for example gpt-5 bills cache hits at 10% of the input rate and gpt-5-chat-latest bills cache hits at 10% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.

Do I need a separate OpenAI account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling OpenAI in one line

One key, one endpoint, 908 models across 41 model authors.