OpenAI Models

142 modelsGeneral models free to startUp to 1.05M context

Usage

Last 30 days · 2026-09-03 to 2026-10-02

Tokens

1155B

Requests

67.6M

Models in use

103 of 142

Tokens per day, stacked by model

032.9B65.8B09-0309-1009-1709-2410-012026-09-03 — 42,991,346,590 tokens gpt-5.6-luna: 13,326,507,300 69 more models: 11,620,946,740 gpt-5.6-sol: 10,238,582,745 gpt-4o-mini: 4,073,560,535 gpt-5.6-terra: 2,937,210,240 omni-moderation-latest: 794,539,0302026-09-04 — 34,343,138,745 tokens 69 more models: 10,594,595,170 gpt-5.6-sol: 9,599,622,210 gpt-5.6-luna: 6,896,668,845 gpt-4o-mini: 3,548,708,695 gpt-5.6-terra: 2,860,377,160 omni-moderation-latest: 843,166,6652026-09-05 — 19,002,576,490 tokens gpt-5.6-luna: 5,442,461,930 69 more models: 5,195,905,670 gpt-4o-mini: 4,548,632,440 gpt-5.6-sol: 1,655,025,830 gpt-5.6-terra: 1,091,145,705 omni-moderation-latest: 750,348,610 gpt-6-astra: 319,056,3052026-09-06 — 19,322,362,185 tokens 69 more models: 7,157,733,200 gpt-5.6-luna: 6,122,995,690 gpt-4o-mini: 2,381,496,365 gpt-5.6-sol: 1,526,206,650 gpt-5.6-terra: 849,602,895 gpt-6-astra: 747,594,890 omni-moderation-latest: 536,732,4952026-09-07 — 39,499,954,290 tokens 69 more models: 12,236,085,080 gpt-5.6-sol: 8,749,795,590 gpt-5.6-luna: 7,137,263,325 gpt-4o-mini: 4,248,041,170 gpt-6-astra: 3,859,903,605 gpt-5.6-terra: 2,376,104,105 omni-moderation-latest: 892,761,4152026-09-08 — 38,624,887,870 tokens gpt-5.6-sol: 10,644,835,915 gpt-5.6-luna: 8,709,075,780 69 more models: 7,658,893,140 gpt-4o-mini: 5,449,317,490 gpt-5.6-terra: 2,561,729,025 gpt-6-astra: 2,497,865,420 omni-moderation-latest: 1,103,171,1002026-09-09 — 39,498,966,685 tokens gpt-5.6-luna: 9,964,021,385 gpt-5.6-sol: 9,577,093,945 69 more models: 7,835,532,170 gpt-4o-mini: 5,602,063,810 gpt-6-astra: 2,832,395,730 gpt-5.6-terra: 1,993,715,985 omni-moderation-latest: 1,694,143,6602026-09-10 — 35,350,859,115 tokens gpt-5.6-sol: 7,309,535,010 gpt-5.6-luna: 7,214,112,490 69 more models: 7,152,689,135 gpt-4o-mini: 4,843,444,700 gpt-6-astra: 4,308,521,105 gpt-5.6-terra: 2,687,330,215 omni-moderation-latest: 1,835,226,4602026-09-11 — 30,326,375,965 tokens gpt-5.6-luna: 9,026,348,975 69 more models: 6,092,917,660 gpt-5.6-sol: 5,380,311,410 gpt-4o-mini: 4,832,847,825 gpt-6-astra: 1,966,558,480 omni-moderation-latest: 1,623,597,340 gpt-5.6-terra: 1,403,794,2752026-09-12 — 22,139,402,950 tokens 69 more models: 6,691,291,315 gpt-5.6-luna: 5,214,649,845 gpt-4o-mini: 4,559,452,365 gpt-5.6-sol: 3,304,867,035 omni-moderation-latest: 1,134,010,660 gpt-6-astra: 877,343,550 gpt-5.6-terra: 357,788,1802026-09-13 — 22,016,795,645 tokens gpt-5.6-luna: 6,477,528,075 gpt-5.6-sol: 5,299,523,440 69 more models: 4,394,613,035 gpt-4o-mini: 2,821,375,020 gpt-6-astra: 1,514,862,405 omni-moderation-latest: 1,241,593,315 gpt-5.6-terra: 267,300,3552026-09-14 — 36,871,744,245 tokens gpt-5.6-luna: 9,260,854,110 69 more models: 9,254,346,065 gpt-5.6-sol: 6,598,984,350 gpt-4o-mini: 4,376,445,645 gpt-6-astra: 4,254,931,125 omni-moderation-latest: 1,630,247,910 gpt-5.6-terra: 1,495,935,0402026-09-15 — 42,729,050,550 tokens 69 more models: 10,036,629,240 gpt-5.6-sol: 9,754,244,075 gpt-5.6-luna: 8,831,454,385 gpt-6-astra: 5,735,445,015 gpt-4o-mini: 4,883,648,745 gpt-5.6-terra: 2,200,890,745 omni-moderation-latest: 1,286,738,3452026-09-16 — 51,201,319,050 tokens gpt-5.6-sol: 12,069,927,660 69 more models: 10,791,695,440 gpt-5.6-luna: 10,182,979,100 gpt-6-astra: 7,086,083,920 gpt-4o-mini: 5,312,493,435 gpt-5.6-terra: 4,286,463,615 omni-moderation-latest: 1,471,675,8802026-09-17 — 65,766,861,945 tokens gpt-5.6-sol: 25,563,320,980 69 more models: 14,196,197,965 gpt-5.6-luna: 10,780,260,915 gpt-6-astra: 5,125,907,085 gpt-4o-mini: 4,878,062,215 gpt-5.6-terra: 4,029,780,920 omni-moderation-latest: 1,193,331,8652026-09-18 — 46,712,525,145 tokens gpt-5.6-sol: 16,343,970,000 gpt-5.6-luna: 8,724,032,005 69 more models: 7,502,897,265 gpt-4o-mini: 4,765,069,765 gpt-5.6-terra: 4,457,009,010 gpt-6-astra: 3,688,307,085 omni-moderation-latest: 1,231,240,0152026-09-19 — 27,984,373,070 tokens gpt-5.6-sol: 7,190,629,995 gpt-5.6-luna: 6,075,272,165 gpt-4o-mini: 4,666,013,530 69 more models: 4,087,079,455 gpt-6-astra: 3,434,063,075 gpt-5.6-terra: 1,793,728,100 omni-moderation-latest: 737,586,7502026-09-20 — 38,980,904,505 tokens gpt-5.6-sol: 15,018,860,815 gpt-5.6-luna: 8,750,798,655 69 more models: 5,841,318,450 gpt-4o-mini: 4,139,917,520 gpt-5.6-terra: 3,385,081,020 omni-moderation-latest: 1,013,786,270 gpt-6-astra: 831,141,7752026-09-21 — 39,663,350,525 tokens gpt-5.6-luna: 10,039,909,815 gpt-5.6-sol: 9,128,282,740 69 more models: 6,621,034,285 gpt-4o-mini: 5,961,223,105 gpt-6-astra: 4,920,742,960 gpt-5.6-terra: 2,074,045,820 omni-moderation-latest: 918,111,8002026-09-22 — 41,900,101,100 tokens gpt-5.6-luna: 11,987,518,135 gpt-5.6-sol: 11,285,144,160 gpt-4o-mini: 6,395,635,390 69 more models: 5,555,144,635 gpt-6-astra: 3,953,115,385 gpt-5.6-terra: 1,651,740,235 omni-moderation-latest: 1,071,803,1602026-09-23 — 54,074,601,505 tokens gpt-4o-mini: 9,236,943,185 gpt-5.6-luna: 8,624,616,655 gpt-6-luna: 7,342,767,345 gpt-6-sol: 7,197,460,450 gpt-5.6-sol: 6,957,399,305 69 more models: 6,630,965,320 gpt-6-astra: 4,303,637,915 gpt-5.6-terra: 2,234,587,820 omni-moderation-latest: 1,546,223,5102026-09-24 — 64,308,834,290 tokens gpt-6-luna: 13,140,074,350 gpt-6-sol: 11,268,384,490 gpt-4o-mini: 10,087,879,690 gpt-6-astra: 9,110,270,565 69 more models: 6,232,346,060 gpt-5.6-sol: 5,982,148,530 gpt-5.6-luna: 5,432,003,050 omni-moderation-latest: 1,632,400,850 gpt-5.6-terra: 1,423,326,7052026-09-25 — 32,988,886,660 tokens gpt-6-luna: 10,262,173,465 gpt-4o-mini: 5,418,504,725 gpt-6-sol: 4,872,203,970 69 more models: 3,927,466,560 gpt-5.6-luna: 3,827,627,950 gpt-6-astra: 2,402,203,500 gpt-5.6-sol: 1,083,245,115 omni-moderation-latest: 778,449,055 gpt-5.6-terra: 417,012,3202026-09-26 — 35,970,981,965 tokens gpt-6-luna: 15,799,681,740 gpt-6-sol: 5,480,506,785 gpt-4o-mini: 4,716,974,605 gpt-5.6-luna: 3,787,163,825 69 more models: 3,331,478,550 gpt-6-astra: 981,759,195 omni-moderation-latest: 929,996,455 gpt-5.6-sol: 690,789,175 gpt-5.6-terra: 252,631,6352026-09-27 — 30,000,198,490 tokens gpt-6-luna: 10,076,053,735 gpt-6-sol: 5,646,062,245 69 more models: 3,793,570,450 gpt-4o-mini: 2,999,777,345 gpt-5.6-luna: 2,346,485,270 gpt-6-astra: 2,155,384,260 omni-moderation-latest: 1,417,893,310 gpt-5.6-sol: 1,128,279,805 gpt-5.6-terra: 436,692,0702026-09-28 — 50,464,852,125 tokens gpt-6-sol: 10,552,480,555 gpt-6-luna: 9,346,855,530 gpt-5.6-luna: 5,404,109,730 69 more models: 5,213,718,975 gpt-4o-mini: 5,164,707,660 gpt-6-astra: 4,892,160,680 gpt-5.6-sol: 4,550,616,225 omni-moderation-latest: 3,001,907,605 gpt-5.6-terra: 2,338,295,1652026-09-29 — 52,459,056,035 tokens gpt-6-sol: 13,810,693,775 gpt-6-luna: 12,744,400,030 69 more models: 5,843,181,360 gpt-4o-mini: 5,424,727,350 gpt-6-astra: 4,843,759,465 gpt-5.6-sol: 3,645,032,735 gpt-5.6-luna: 2,832,435,655 gpt-5.6-terra: 1,736,284,935 omni-moderation-latest: 1,578,540,7302026-09-30 — 43,243,609,165 tokens gpt-6-luna: 10,732,764,310 69 more models: 9,948,539,740 gpt-6-sol: 7,283,011,500 gpt-4o-mini: 5,210,936,785 gpt-5.6-sol: 2,925,910,820 gpt-5.6-luna: 2,365,125,780 omni-moderation-latest: 2,182,773,510 gpt-6-astra: 1,741,124,875 gpt-5.6-terra: 853,421,8452026-10-01 — 27,934,521,335 tokens 69 more models: 9,588,991,425 gpt-6-luna: 8,902,499,370 gpt-4o-mini: 3,746,645,250 gpt-5.6-luna: 1,858,061,440 omni-moderation-latest: 1,596,889,350 gpt-6-sol: 998,371,005 gpt-5.6-sol: 840,575,490 gpt-6-astra: 266,649,800 gpt-5.6-terra: 135,838,2052026-10-02 — 28,877,273,890 tokens gpt-6-luna: 12,483,844,910 69 more models: 6,780,707,865 gpt-4o-mini: 3,428,238,365 gpt-5.6-luna: 2,722,282,915 omni-moderation-latest: 1,388,006,395 gpt-6-sol: 891,307,985 gpt-6-astra: 759,737,275 gpt-5.6-sol: 280,302,165 gpt-5.6-terra: 142,846,015
  • gpt-5.6-sol
  • gpt-5.6-luna
  • gpt-4o-mini
  • gpt-6-luna
  • gpt-6-astra
  • gpt-6-sol
  • gpt-5.6-terra
  • omni-moderation-latest
  • 69 more models

Which models that traffic went to

  1. GPT 5.6 Sol18.6%214B
  2. GPT 5.6 Luna18.1%209B
  3. GPT 4o Mini12.8%148B
  4. GPT 6 Luna9.6%111B
  5. GPT 6 Astra7.7%89.4B
  6. GPT 6 Sol5.9%68B
  7. GPT 5.6 Terra4.7%54.7B
  8. Omni Moderation3.4%39.1B
  9. 69 more models19.2%222B

Share of 1155B tokens. 26 models with traffic report no token counts and cannot be ranked here, including tts-1 and gpt-4o-mini-tts — they are in the request view.

The two views disagree on purpose: a model can take a large share of the calls and a small share of the tokens — many short requests — or the reverse. Which one matters depends on whether your cost is driven by call volume or by prompt length. Measured on AIHubMix over the last 30 days, counting the 142 model IDs listed on this page; traffic routed through upstream-specific IDs that are not in the public catalog is not included.

All 142 OpenAI Models

Open in model list
OpenAI models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
gpt-5.5-freeTakes text, vision, PDF, returns text.1.05M128KFreeFree/MFree/M—66 tok/s19.67 s
gpt-6-lunaTakes text, vision, returns text.1.05M128K$0.1$0.5/M$0.01/M$0.125/M98 tok/s2.24 s
gpt-5.6-lunaTakes text, vision, returns text.1.05M128K$0.2$1.2/M$0.02/M$0.25/M71 tok/s1.88 s
gpt-5.6-terraTakes text, vision, returns text.1.05M128K$2$12/M$0.2/M$2.5/M68 tok/s2.26 s
gpt-6-solTakes text, vision, returns text.1.05M128K$2$10/M$0.2/M$2.5/M61 tok/s4.96 s
gpt-6.1-solTakes text, vision, returns text.1.05M—$2$10/M$0.1/M$2.5/M37 tok/s11.74 s
gpt-5.4Takes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—48 tok/s1.39 s
gpt-5.4-highTakes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—53 tok/s10.76 s
gpt-5.4-lowTakes text, vision, PDF, returns text.1.05M128K$2.5$15/M$0.25/M—67 tok/s2.86 s
gpt-5.6-solTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M61 tok/s3.63 s
gpt-5.6-sol-discTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M93 tok/s6.96 s
gpt-5.5Takes text, vision, PDF, returns text.1.05M128K$5$30/M$0.5/M—73 tok/s4.13 s
gpt-5.6-sol-proTakes text, vision, returns text.1.05M—$8$40/M$0.8/M———
gpt-6-astraTakes text, vision, returns text.1.05M128K$10$50/M$1/M$12.5/M37 tok/s6.71 s
gpt-5.4-proTakes text, vision, returns text.1.05M128K$30$180/M——46 tok/s8.60 s
gpt-5.5-proTakes text, vision, returns text.1.05M128K$30$180/M——29 tok/s73.10 s
gpt-4.1-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—37 tok/s3.09 s
gpt-4.1-mini-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—24 tok/s1.53 s
gpt-4.1-nano-freeTakes text, vision, PDF, returns text.1.05M33KFreeFree/MFree/M—19 tok/s1.25 s
gpt-4.1-nanoTakes text, vision, PDF, returns text.1.05M33K$0.1$0.4/M$0.025/M—99 tok/s0.82 s
gpt-4.1-miniTakes text, vision, PDF, returns text.1.05M33K$0.4$1.6/M$0.1/M—58 tok/s0.74 s
gpt-4.1Takes text, vision, PDF, returns text.1.05M33K$2$8/M$0.5/M—62 tok/s1.04 s
autoTakes text, vision, audio, video, returns text.1M—FreeFree/M————
gpt-5-nanoTakes text, vision, returns text.400K128K$0.05$0.4/M$0.005/M—83 tok/s2.53 s
gpt-5.4-nanoTakes text, vision, returns text.400K128K$0.2$1.25/M$0.02/M—102 tok/s0.65 s
gpt-5-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M—103 tok/s4.38 s
gpt-5.1-codex-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M—168 tok/s0.70 s
gpt-5.4-miniTakes text, vision, returns text.400K128K$0.75$4.5/M$0.075/M—116 tok/s1.25 s
gpt-5Takes text, vision, returns text.400K128K$1.25$10/M$0.125/M—81 tok/s7.76 s
gpt-5-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—44 tok/s7.54 s
gpt-5.1Takes text, vision, PDF, returns text.400K128K$1.25$10/M$0.125/M—115 tok/s0.78 s
gpt-5.1-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—82 tok/s0.45 s
gpt-5.1-codex-maxTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M—84 tok/s5.95 s
gpt-5.2Takes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—82 tok/s2.28 s
gpt-5.2-codexTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—82 tok/s0.66 s
gpt-5.2-highTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—16 tok/s0.94 s
gpt-5.2-lowTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—18 tok/s0.62 s
gpt-5.3-codexTakes text, vision, PDF, returns text.400K128K$1.75$14/M$0.175/M—38 tok/s10.14 s
gpt-chat-latestTakes text, vision, returns text.400K128K$5$30/M$0.5/M—48 tok/s2.65 s
gpt-5-proTakes text, vision, returns text.400K272K$15$120/M——9 tok/s313.10 s
gpt-5.2-proTakes text, vision, returns text.400K128K$21$168/M$2.1/M—33 tok/s23.97 s
o3-miniTakes text, vision, returns text.200K100K$1.1$4.4/M$0.55/M—474 tok/s10.90 s
o4-miniTakes text, vision, PDF, returns text.200K100K$1.1$4.4/M$0.275/M—67 tok/s5.80 s
codex-mini-latestTakes text, vision, returns text.200K—$1.5$6/M$0.375/M———
o3Takes text, vision, PDF, returns text.200K100K$2$8/M$0.5/M—90 tok/s6.46 s
o3-proTakes text, vision, returns text.200K100K$20$80/M$20/M—13 tok/s88.94 s
o1-proTakes text, returns text.200K—$170$680/M$170/M—19 tok/s96.00 s
gpt-oss-20b-freeTakes text, returns text.131K—FreeFree/M————
gpt-oss-120bTakes text, returns text.131K33K$0.18$0.9/M——1101 tok/s0.12 s
gpt-4o-freeTakes text, vision, PDF, returns text.128K16KFreeFree/MFree/M—18 tok/s6.84 s
gpt-realtime-2.1Takes text, vision, audio. Output modality not published.128K32KFreeFree/MFree/M———
gpt-oss-20bTakes text, returns text.128K33K$0.11$0.55/M——2619 tok/s0.12 s
gpt-4o-miniTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M—52 tok/s0.62 s
gpt-4o-mini-audio-previewTakes text, audio, returns text.128K16K$0.15$0.6/M————
gpt-4o-mini-search-previewTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M—189 tok/s1.57 s
gpt-5-chat-latestTakes text, vision, returns text.128K16K$1.25$10/M$0.125/M—77 tok/s0.85 s
gpt-5.1-chat-latestTakes text, vision, returns text.128K16K$1.25$10/M$0.125/M—107 tok/s0.94 s
gpt-5.2-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M—97 tok/s0.50 s
gpt-5.3-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M—100 tok/s0.57 s
gpt-4oTakes text, vision, PDF, returns text.128K16K$2.5$10/M$1.25/M—52 tok/s0.64 s
gpt-4o-2024-11-20Takes text, vision, returns text.128K16K$2.5$10/M$1.25/M—61 tok/s0.60 s
gpt-4o-audio-previewTakes text, audio, returns text.128K16K$2.5$10/M——10 tok/s2.49 s
gpt-4o-search-previewTakes text, vision, returns text.128K16K$2.5$10/M$1.25/M—121 tok/s2.33 s
gpt-audio-1.5Takes text, audio, returns text, audio.128K16K$2.5$10/M————
gpt-4o-2024-05-13128K4K$5$15/M$5/M—199 tok/s0.45 s
gpt-4o-transcribe-diarizeTakes text, audio, returns text.16K2K$2.5$10/M————
o1Takes text, returns text.0K—$15$60/M$7.5/M—83 tok/s4.89 s
dall-e-2Takes text, vision, returns vision.——FreeFree/M————
dall-e-3Takes text, vision, returns vision.——FreeFree/M————
gpt-image-1Takes text, vision, returns vision.——FreeFree/M————
gpt-image-1-miniTakes text, vision, returns vision.——FreeFree/M————
gpt-image-1.5Takes text, vision, returns vision.——FreeFree/M————
gpt-image-2Takes text, vision, returns vision.——FreeFree/MFree/M———
gpt-image-2-freeTakes text, vision, returns text, vision.——FreeFree/M————
gpt-image-2.5-flareTakes text, vision, returns vision.——FreeFree/MFree/M———
gpt-image-2.5-sunburstTakes text, vision, returns vision.——FreeFree/MFree/M———
gpt-live-transcribeTakes text, audio, returns text.——FreeFree/M————
sora-2Takes , returns video.——FreeFree/M————
sora-2-proTakes , returns video.——Free$720/M————
whisper-1Takes audio, returns text.——FreeFree/M————
whisper-large-v3Takes audio, returns text.——FreeFree/M————
whisper-large-v3-turboTakes audio, returns text.——FreeFree/M————
omni-moderation-latest——$0.02$0.02/M————
text-embedding-3-smallTakes text. Output modality not published.——$0.02$0.02/M————
text-embedding-ada-002Takes text. Output modality not published.——$0.1$0.1/M————
text-embedding-v1Takes text. Output modality not published.——$0.1$0.1/M————
GPT-OSS-20B——$0.11$0.55/M——2619 tok/s0.12 s
text-embedding-3-largeTakes text. Output modality not published.——$0.13$0.13/M————
gpt-4o-mini-2024-07-18Takes text, vision, returns text.——$0.15$0.6/M$0.075/M—17 tok/s1.30 s
gpt-4o-mini-global——$0.15$0.6/M$0.075/M———
text-moderation-007——$0.2$0.2/M————
text-moderation-latest——$0.2$0.2/M————
text-moderation-stable——$0.2$0.2/M————
aihubmix-routerTakes text, vision, returns text.——$0.4$1.6/M$0.1/M—79 tok/s2.59 s
text-ada-001——$0.4$0.4/M————
gpt-3.5-turbo——$0.5$1.5/M————
text-babbage-001——$0.5$0.5/M————
gpt-4o-mini-ttsTakes audio, returns audio.——$0.6$12/M———0.95 s
gpt-3.5-turbo-1106——$1$2/M————
o3-mini-global——$1.1$4.4/M$0.55/M———
gpt-3.5-turbo-0301——$1.5$1.5/M————
gpt-3.5-turbo-0613——$1.5$2/M————
gpt-3.5-turbo-instruct——$1.5$2/M————
davinci-002——$2$2/M————
o3-global——$2$8/M$0.5/M———
text-curie-001——$2$2/M————
gpt-4o-2024-08-06——$2.5$10/M$1.25/M—61 tok/s2.31 s
gpt-4o-2024-08-06-global——$2.5$10/M$1.25/M———
gpt-4o-zhTakes text, vision, returns text.——$2.5$10/M————
computer-use-preview——$3$12/M————
gpt-3.5-turbo-16k——$3$4/M————
gpt-3.5-turbo-16k-0613——$3$4/M————
o1-mini——$3$12/M$1.5/M———
o1-mini-2024-09-12——$3$12/M$1.5/M———
gpt-image-test——$5$40/M————
distil-whisper-large-v3-enTakes audio, returns text.——$5.556$5.556/M————
gpt-4-0125-preview——$10$30/M————
gpt-4-1106-preview——$10$30/M————
gpt-4-turbo——$10$30/M————
gpt-4-turbo-2024-04-09——$10$30/M————
gpt-4-turbo-preview——$10$30/M————
gpt-4-vision-preview——$10$30/M————
o3-deep-research——$10$40/M$2.5/M———
o1-2024-12-17Takes text, vision, returns text.——$15$60/M$7.5/M—125 tok/s3.19 s
o1-previewTakes text, vision, returns text.——$15$60/M$7.5/M———
o1-preview-2024-09-12——$15$60/M$7.5/M———
tts-1Takes audio, returns audio.——$15$15/M————
tts-1-1106Takes audio, returns audio.——$15$15/M————
davinci——$20$20/M————
o3-pro-global——$20$80/M————
text-davinci-002——$20$20/M————
text-davinci-003——$20$20/M————
text-davinci-edit-001——$20$20/M————
text-search-ada-doc-001——$20$20/M————
gpt-4——$30$60/M————
gpt-4-0314——$30$60/M————
gpt-4-0613——$30$60/M————
tts-1-hdTakes audio, returns audio.——$30$30/M————
tts-1-hd-1106Takes audio, returns audio.——$30$30/M————
gpt-4-32k——$60$120/M————
gpt-4-32k-0314——$60$120/M————
gpt-4-32k-0613——$60$120/M————

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

OpenAI on AIHubMix

Which OpenAI model should I start with?

auto is free on input — the cheapest entry here that declares tool calling, and it carries a 1M context. Move up to o1-pro when answer quality matters more than cost, or to gpt-5.5-free for long-form reasoning.

Which of these models reason before answering?

43 of the 142 models here declare a reasoning phase — they work through the problem before producing an answer, which helps on multi-step problems at the cost of extra output tokens. Use the Reasoning filter above the table to see them. The catalog does not record anything further about how they differ, so this page does not sort them into families.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

How is cached input billed?

The Cache read column is the rate for input tokens served from the prompt cache — for example gpt-6.1-sol bills cache hits at 5% of the input rate and gpt-5 bills cache hits at 10% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.

Do I need a separate OpenAI account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling OpenAI in one line

One key, one endpoint, 911 models across 42 model authors.