OpenAI Models

134 modelsGeneral models free to startUp to 1.05M context

Usage

Last 29 days · 2026-08-10 to 2026-09-07

Tokens

1003B

Requests

74M

Models in use

90 of 134

Tokens per day, stacked by model

025B50.1B08-1008-1708-2408-3109-072026-08-10 — 32,048,569,365 tokens gpt-5.6-sol: 9,212,190,175 gpt-5.6-luna: 5,927,241,875 gpt-4o-mini: 4,363,708,580 72 more models: 2,855,875,350 gpt-5.6-terra: 2,529,833,790 gpt-5.5: 2,271,821,830 gpt-5.4: 2,060,391,350 gpt-5.4-mini: 1,484,172,485 omni-moderation-latest: 1,343,333,9302026-08-11 — 34,668,267,545 tokens gpt-5.6-sol: 9,224,441,820 gpt-5.6-luna: 6,548,111,845 gpt-4o-mini: 4,507,019,605 gpt-5.6-terra: 3,774,837,585 72 more models: 2,956,233,055 gpt-5.5: 2,667,531,560 gpt-5.4: 1,674,539,940 omni-moderation-latest: 1,671,747,210 gpt-5.4-mini: 1,643,804,9252026-08-12 — 36,158,085,580 tokens gpt-5.6-sol: 8,821,921,940 gpt-5.5: 6,860,207,495 gpt-5.6-luna: 5,592,114,045 72 more models: 3,533,406,130 gpt-4o-mini: 3,384,816,425 gpt-5.6-terra: 2,832,015,405 gpt-5.4: 1,948,077,565 omni-moderation-latest: 1,673,157,225 gpt-5.4-mini: 1,512,369,3502026-08-13 — 42,947,988,045 tokens gpt-5.6-sol: 9,136,245,545 gpt-5.6-luna: 9,068,074,350 72 more models: 5,978,739,375 gpt-5.5: 4,849,508,600 gpt-4o-mini: 4,057,242,245 gpt-5.6-terra: 3,602,383,020 omni-moderation-latest: 2,732,458,055 gpt-5.4: 1,838,364,180 gpt-5.4-mini: 1,684,972,6752026-08-14 — 36,449,501,915 tokens gpt-5.6-luna: 9,528,849,430 gpt-5.6-sol: 6,433,406,135 72 more models: 5,357,234,220 gpt-4o-mini: 3,985,265,990 gpt-5.5: 3,394,549,925 gpt-5.6-terra: 2,845,194,965 gpt-5.4: 2,525,769,330 omni-moderation-latest: 1,191,039,270 gpt-5.4-mini: 1,188,192,6502026-08-15 — 19,947,585,260 tokens gpt-5.6-sol: 3,916,210,040 gpt-5.6-luna: 3,613,505,950 72 more models: 3,097,720,890 gpt-4o-mini: 2,877,860,180 gpt-5.6-terra: 2,248,452,170 gpt-5.4-mini: 1,303,386,520 gpt-5.5: 1,164,548,550 omni-moderation-latest: 1,014,202,460 gpt-5.4: 711,698,5002026-08-16 — 26,554,229,840 tokens gpt-5.6-terra: 6,825,124,795 gpt-5.6-luna: 5,096,679,890 gpt-5.6-sol: 4,591,827,125 72 more models: 3,336,560,035 gpt-4o-mini: 2,309,900,020 omni-moderation-latest: 1,770,407,135 gpt-5.5: 1,171,518,985 gpt-5.4-mini: 728,555,725 gpt-5.4: 723,656,1302026-08-17 — 34,450,604,135 tokens gpt-5.6-sol: 8,667,199,180 gpt-5.6-luna: 7,345,805,010 gpt-5.6-terra: 5,294,374,820 gpt-4o-mini: 3,476,223,000 72 more models: 2,960,646,310 gpt-5.5: 2,783,496,590 gpt-5.4: 1,663,616,610 gpt-5.4-mini: 1,387,251,850 omni-moderation-latest: 871,990,7652026-08-18 — 39,124,814,300 tokens gpt-5.6-luna: 10,076,854,830 gpt-5.6-sol: 7,986,387,305 gpt-4o-mini: 4,884,259,185 72 more models: 4,409,843,895 omni-moderation-latest: 4,063,929,900 gpt-5.6-terra: 3,568,415,420 gpt-5.4-mini: 1,586,466,740 gpt-5.4: 1,394,153,350 gpt-5.5: 1,154,503,6752026-08-19 — 39,372,703,085 tokens gpt-5.6-luna: 12,143,529,870 gpt-5.6-sol: 6,647,256,060 omni-moderation-latest: 5,813,687,380 72 more models: 5,066,998,000 gpt-4o-mini: 3,273,939,105 gpt-5.6-terra: 2,566,660,730 gpt-5.4-mini: 1,489,597,035 gpt-5.4: 1,348,645,130 gpt-5.5: 1,022,389,7752026-08-20 — 45,075,006,735 tokens gpt-5.6-luna: 13,651,078,720 gpt-5.6-sol: 12,778,009,605 72 more models: 4,759,132,655 gpt-5.6-terra: 3,991,254,415 omni-moderation-latest: 3,311,346,030 gpt-4o-mini: 3,082,342,645 gpt-5.4: 1,211,523,685 gpt-5.4-mini: 1,203,433,505 gpt-5.5: 1,086,885,4752026-08-21 — 40,641,927,935 tokens gpt-5.6-luna: 13,417,320,110 gpt-5.6-sol: 7,615,098,190 gpt-5.4-mini: 6,195,608,595 gpt-4o-mini: 3,676,359,465 72 more models: 3,077,843,380 gpt-5.6-terra: 2,294,545,665 omni-moderation-latest: 1,725,068,975 gpt-5.5: 1,698,888,845 gpt-5.4: 941,194,7102026-08-22 — 23,369,927,180 tokens gpt-5.6-luna: 10,514,705,475 gpt-5.6-sol: 3,188,288,355 72 more models: 3,131,256,750 gpt-4o-mini: 2,801,579,210 gpt-5.4-mini: 1,129,654,790 gpt-5.6-terra: 950,451,865 omni-moderation-latest: 827,148,025 gpt-5.5: 574,465,105 gpt-5.4: 252,377,6052026-08-23 — 26,027,764,365 tokens gpt-5.6-luna: 11,731,746,685 gpt-5.6-sol: 5,108,590,390 72 more models: 2,186,040,175 gpt-5.6-terra: 2,042,484,545 gpt-4o-mini: 1,988,456,930 omni-moderation-latest: 1,383,186,260 gpt-5.4-mini: 731,735,375 gpt-5.4: 429,320,660 gpt-5.5: 426,203,3452026-08-24 — 40,102,269,660 tokens gpt-5.6-luna: 15,875,395,965 gpt-5.6-sol: 8,171,009,610 gpt-4o-mini: 3,642,828,010 gpt-5.6-terra: 3,129,841,775 72 more models: 2,817,270,045 omni-moderation-latest: 1,964,606,320 gpt-5.4-mini: 1,930,382,695 gpt-5.4: 1,362,372,775 gpt-5.5: 1,208,562,4652026-08-25 — 45,358,950,620 tokens gpt-5.6-luna: 13,375,812,305 gpt-5.6-sol: 10,544,890,995 gpt-5.6-terra: 4,786,624,565 omni-moderation-latest: 4,087,252,645 72 more models: 3,927,359,180 gpt-4o-mini: 3,538,045,825 gpt-5.4-mini: 2,597,553,615 gpt-5.5: 1,404,375,610 gpt-5.4: 1,097,035,8802026-08-26 — 50,061,293,925 tokens gpt-5.6-luna: 17,886,729,770 gpt-5.6-sol: 10,575,946,460 omni-moderation-latest: 5,985,594,420 72 more models: 4,346,306,500 gpt-4o-mini: 3,590,992,920 gpt-5.6-terra: 2,919,577,285 gpt-5.4-mini: 2,473,716,815 gpt-5.5: 1,254,603,825 gpt-5.4: 1,027,825,9302026-08-27 — 42,457,561,595 tokens gpt-5.6-luna: 12,030,529,890 gpt-5.6-sol: 7,224,017,885 omni-moderation-latest: 5,401,546,400 gpt-4o-mini: 4,306,524,590 72 more models: 3,612,227,975 gpt-5.4-mini: 3,246,010,690 gpt-5.6-terra: 3,121,793,900 gpt-5.4: 1,972,062,305 gpt-5.5: 1,542,847,9602026-08-28 — 31,844,035,350 tokens gpt-5.6-luna: 9,539,313,300 gpt-5.6-sol: 6,553,281,695 gpt-4o-mini: 4,210,950,825 omni-moderation-latest: 3,218,460,175 gpt-5.6-terra: 2,621,238,600 72 more models: 2,539,596,985 gpt-5.4: 1,345,467,760 gpt-5.5: 1,000,492,940 gpt-5.4-mini: 815,233,0702026-08-29 — 20,434,916,100 tokens gpt-5.6-luna: 7,361,729,420 gpt-4o-mini: 3,315,014,415 gpt-5.6-sol: 2,835,756,100 72 more models: 2,536,196,175 omni-moderation-latest: 1,783,618,130 gpt-5.4-mini: 848,581,450 gpt-5.5: 794,088,680 gpt-5.4: 499,024,055 gpt-5.6-terra: 460,907,6752026-08-30 — 20,001,907,215 tokens gpt-5.6-luna: 6,123,519,480 72 more models: 4,177,960,795 gpt-4o-mini: 2,695,687,130 gpt-5.6-sol: 2,435,225,610 omni-moderation-latest: 1,929,593,135 gpt-5.6-terra: 1,012,717,870 gpt-5.4-mini: 896,549,430 gpt-5.5: 490,958,090 gpt-5.4: 239,695,6752026-08-31 — 38,857,820,740 tokens gpt-5.6-luna: 9,816,841,965 gpt-5.6-sol: 8,732,840,320 72 more models: 6,009,790,215 gpt-5.6-terra: 4,446,110,735 gpt-4o-mini: 3,583,158,430 omni-moderation-latest: 2,897,572,065 gpt-5.5: 1,386,135,625 gpt-5.4-mini: 1,164,497,030 gpt-5.4: 820,874,3552026-09-01 — 43,685,093,910 tokens gpt-5.6-luna: 13,879,890,225 gpt-5.6-sol: 11,180,256,740 gpt-4o-mini: 7,015,598,660 72 more models: 3,402,126,035 gpt-5.6-terra: 3,238,286,090 omni-moderation-latest: 1,764,455,460 gpt-5.5: 1,653,594,145 gpt-5.4-mini: 779,207,365 gpt-5.4: 771,679,1902026-09-02 — 38,222,420,860 tokens gpt-5.6-luna: 11,397,495,330 gpt-5.6-sol: 9,338,840,605 gpt-4o-mini: 4,770,885,020 72 more models: 3,697,788,895 gpt-5.5: 2,893,338,470 gpt-5.6-terra: 2,538,607,895 gpt-5.4: 1,720,556,175 gpt-5.4-mini: 999,082,765 omni-moderation-latest: 865,825,7052026-09-03 — 42,991,346,590 tokens gpt-5.6-luna: 13,326,507,300 gpt-5.6-sol: 10,238,582,745 72 more models: 4,636,946,210 gpt-4o-mini: 4,073,560,535 gpt-5.4: 3,125,939,385 gpt-5.6-terra: 2,937,210,240 gpt-5.5: 2,033,487,040 gpt-5.4-mini: 1,824,574,105 omni-moderation-latest: 794,539,0302026-09-04 — 34,343,138,745 tokens gpt-5.6-sol: 9,599,622,210 gpt-5.6-luna: 6,896,668,845 72 more models: 3,744,965,620 gpt-4o-mini: 3,548,708,695 gpt-5.6-terra: 2,860,377,160 gpt-5.4: 2,813,840,935 gpt-5.4-mini: 2,306,338,105 gpt-5.5: 1,729,450,510 omni-moderation-latest: 843,166,6652026-09-05 — 19,002,576,490 tokens gpt-5.6-luna: 5,442,461,930 gpt-4o-mini: 4,548,632,440 72 more models: 2,761,886,445 gpt-5.6-sol: 1,655,025,830 gpt-5.4-mini: 1,575,443,960 gpt-5.6-terra: 1,091,145,705 gpt-5.4: 871,723,510 omni-moderation-latest: 750,348,610 gpt-5.5: 305,908,0602026-09-06 — 19,322,362,185 tokens gpt-5.6-luna: 6,122,995,690 72 more models: 4,687,936,370 gpt-4o-mini: 2,381,496,365 gpt-5.4-mini: 1,904,111,250 gpt-5.6-sol: 1,526,206,650 gpt-5.6-terra: 849,602,895 gpt-5.4: 834,647,470 omni-moderation-latest: 536,732,495 gpt-5.5: 478,633,0002026-09-07 — 39,499,954,290 tokens 72 more models: 9,082,159,295 gpt-5.6-sol: 8,749,795,590 gpt-5.6-luna: 7,137,263,325 gpt-4o-mini: 4,248,041,170 gpt-5.4: 2,512,458,115 gpt-5.5: 2,505,995,225 gpt-5.6-terra: 2,376,104,105 gpt-5.4-mini: 1,995,376,050 omni-moderation-latest: 892,761,415
  • gpt-5.6-luna
  • gpt-5.6-sol
  • gpt-4o-mini
  • gpt-5.6-terra
  • omni-moderation-latest
  • gpt-5.5
  • gpt-5.4-mini
  • gpt-5.4
  • 72 more models

Which models that traffic went to

  1. GPT 5.6 Luna28.0%280B
  2. GPT 5.6 Sol21.2%213B
  3. GPT 4o Mini10.8%108B
  4. GPT 5.6 Terra8.4%83.8B
  5. Omni Moderation6.3%63.1B
  6. GPT 5.55.2%51.8B
  7. GPT 5.4 Mini4.8%48.6B
  8. GPT 5.44.0%39.7B
  9. 72 more models11.4%115B

Share of 1003B tokens. 10 models with traffic report no token counts and cannot be ranked here, including auto and gpt-4o-audio-preview — they are in the request view.

The two views disagree on purpose: a model can take a large share of the calls and a small share of the tokens — many short requests — or the reverse. Which one matters depends on whether your cost is driven by call volume or by prompt length. Measured on AIHubMix over the last 29 days, counting the 134 model IDs listed on this page; traffic routed through upstream-specific IDs that are not in the public catalog is not included.

All 134 OpenAI Models

Open in model list
OpenAI models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
gpt-5.5-freeTakes text, vision, returns text.1.05M128KFreeFree/MFree/M42 tok/s3.58 s
gpt-5.6-lunaTakes text, vision, returns text.1.05M128K$0.2$1.2/M$0.02/M$0.25/M50 tok/s4.75 s
gpt-5.6-terraTakes text, vision, returns text.1.05M128K$2$12/M$0.2/M$2.5/M59 tok/s3.61 s
gpt-5.4Takes text, vision, returns text.1.05M128K$2.5$15/M$0.25/M70 tok/s1.69 s
gpt-5.4-highTakes text, vision, returns text.1.05M128K$2.5$15/M$0.25/M
gpt-5.4-lowTakes text, vision, returns text.1.05M128K$2.5$15/M$0.25/M
gpt-5.6-solTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M39 tok/s6.31 s
gpt-5.6-sol-discTakes text, vision, returns text.1.05M128K$4$20/M$0.4/M$5/M37 tok/s5.16 s
gpt-5.5Takes text, vision, returns text.1.05M128K$5$30/M$0.5/M64 tok/s5.13 s
gpt-6-astraTakes text, vision, returns text.1.05M128K$10$50/M$1/M$12.5/M21 tok/s17.02 s
gpt-5.4-proTakes text, vision, returns text.1.05M128K$30$180/M42 tok/s8.60 s
gpt-5.5-proTakes text, vision, returns text.1.05M128K$30$180/M14 tok/s19.20 s
gpt-4.1-freeTakes text, vision, returns text.1.05M33KFreeFree/MFree/M72 tok/s0.53 s
gpt-4.1-mini-freeTakes text, vision, returns text.1.05M33KFreeFree/MFree/M59 tok/s0.36 s
gpt-4.1-nano-freeTakes text, vision, returns text.1.05M33KFreeFree/MFree/M110 tok/s0.33 s
gpt-4o-freeTakes text, vision, returns text.1.05M33KFreeFree/MFree/M72 tok/s0.53 s
gpt-4.1-nanoTakes text, vision, returns text.1.05M33K$0.1$0.4/M$0.025/M73 tok/s0.82 s
gpt-4.1-miniTakes text, vision, returns text.1.05M33K$0.4$1.6/M$0.1/M58 tok/s0.74 s
gpt-4.1Takes text, vision, returns text.1.05M33K$2$8/M$0.5/M81 tok/s1.13 s
autoTakes text, vision, audio, video, returns text.1MFreeFree/M
gpt-5-nanoTakes text, vision, returns text.400K128K$0.05$0.4/M$0.005/M87 tok/s4.78 s
gpt-5.4-nanoTakes text, vision, returns text.400K128K$0.2$1.25/M$0.02/M102 tok/s0.65 s
gpt-5-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M94 tok/s4.07 s
gpt-5.1-codex-miniTakes text, vision, returns text.400K128K$0.25$2/M$0.025/M168 tok/s0.70 s
gpt-5.4-miniTakes text, vision, returns text.400K128K$0.75$4.5/M$0.075/M129 tok/s1.07 s
gpt-5Takes text, vision, returns text.400K128K$1.25$10/M$0.125/M69 tok/s7.76 s
gpt-5-chat-latestTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M77 tok/s0.85 s
gpt-5-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M64 tok/s6.24 s
gpt-5.1Takes text, vision, returns text.400K128K$1.25$10/M$0.125/M71 tok/s2.25 s
gpt-5.1-codexTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M82 tok/s0.62 s
gpt-5.1-codex-maxTakes text, vision, returns text.400K128K$1.25$10/M$0.125/M
gpt-5.2Takes text, vision, returns text.400K128K$1.75$14/M$0.175/M51 tok/s1.36 s
gpt-5.2-codexTakes text, vision, returns text.400K128K$1.75$14/M$0.175/M82 tok/s0.66 s
gpt-5.2-highTakes text, vision, returns text.400K128K$1.75$14/M$0.175/M
gpt-5.2-lowTakes text, vision, returns text.400K128K$1.75$14/M$0.175/M
gpt-5.3-codexTakes text, vision, returns text.400K128K$1.75$14/M$0.175/M75 tok/s11.73 s
gpt-chat-latestTakes text, vision, returns text.400K128K$5$30/M$0.5/M48 tok/s1.93 s
gpt-5-proTakes text, vision, returns text.400K128K$15$120/M9 tok/s313.10 s
gpt-5.2-proTakes text, vision, returns text.400K128K$21$168/M$2.1/M9 tok/s25.47 s
o3-miniTakes text, vision, returns text.200K100K$1.1$4.4/M$0.55/M474 tok/s6.65 s
o4-miniTakes text, vision, returns text.200K100K$1.1$4.4/M$0.275/M86 tok/s5.80 s
codex-mini-latestTakes text, vision, returns text.200K$1.5$6/M$0.375/M
o3Takes text, vision, returns text.200K100K$2$8/M$0.5/M40 tok/s6.46 s
o3-proTakes text, vision, returns text.200K100K$20$80/M$20/M13 tok/s88.94 s
o1-proTakes text, returns text.200K$170$680/M$170/M19 tok/s96.00 s
gpt-oss-20b-freeTakes text, returns text.131KFreeFree/M
gpt-oss-120bTakes text, returns text.131K33K$0.18$0.9/M1101 tok/s0.12 s
gpt-oss-20bTakes text, returns text.128K$0.11$0.55/M2619 tok/s0.12 s
gpt-4o-miniTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M49 tok/s0.62 s
gpt-4o-mini-search-previewTakes text, vision, returns text.128K16K$0.15$0.6/M$0.075/M189 tok/s1.57 s
gpt-5.1-chat-latestTakes text, vision, returns text.128K16K$1.25$10/M$0.125/M107 tok/s0.94 s
gpt-5.2-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M84 tok/s0.77 s
gpt-5.3-chat-latestTakes text, vision, returns text.128K16K$1.75$14/M$0.175/M100 tok/s0.82 s
gpt-4oTakes text, vision, returns text.128K16K$2.5$10/M$1.25/M52 tok/s0.64 s
gpt-4o-2024-11-20Takes text, vision, returns text.128K16K$2.5$10/M$1.25/M61 tok/s0.60 s
gpt-4o-audio-previewTakes text, audio, returns text.128K16K$2.5$10/M10 tok/s2.49 s
gpt-4o-search-previewTakes text, vision, returns text.128K16K$2.5$10/M$1.25/M121 tok/s2.33 s
gpt-audio-1.5Takes text, audio, returns text, audio.128K16K$2.5$10/M
gpt-4o-2024-05-13128K4K$5$15/M$5/M199 tok/s0.45 s
gpt-4o-transcribe-diarizeTakes text, audio, returns text.16K2K$2.5$10/M
o1Takes text, returns text.0K$15$60/M$7.5/M39 tok/s9.40 s
dall-e-2Takes text, vision, returns vision.FreeFree/M
dall-e-3Takes text, vision, returns vision.FreeFree/M
gpt-image-1Takes text, vision, returns vision.FreeFree/M
gpt-image-1-miniTakes text, vision, returns vision.FreeFree/M
gpt-image-1.5Takes text, vision, returns vision.FreeFree/M
gpt-image-2Takes text, vision, returns vision.FreeFree/MFree/M
gpt-image-2-freeTakes text, vision, returns text, vision.FreeFree/M
sora-2Takes , returns video.FreeFree/M
sora-2-proTakes , returns video.Free$720/M
whisper-1Takes audio, returns text.FreeFree/M
whisper-large-v3Takes audio, returns text.FreeFree/M
whisper-large-v3-turboTakes audio, returns text.FreeFree/M
omni-moderation-latest$0.02$0.02/M
text-embedding-3-smallTakes text. Output modality not published.$0.02$0.02/M
text-embedding-ada-002Takes text. Output modality not published.$0.1$0.1/M
text-embedding-v1Takes text. Output modality not published.$0.1$0.1/M
GPT-OSS-20B$0.11$0.55/M2619 tok/s0.12 s
text-embedding-3-largeTakes text. Output modality not published.$0.13$0.13/M
gpt-4o-mini-2024-07-18Takes text, vision, returns text.$0.15$0.6/M$0.075/M
gpt-4o-mini-audio-previewTakes text, audio, returns text.$0.15$0.6/M
gpt-4o-mini-global$0.15$0.6/M$0.075/M
text-moderation-007$0.2$0.2/M
text-moderation-latest$0.2$0.2/M
text-moderation-stable$0.2$0.2/M
aihubmix-routerTakes text, vision, returns text.$0.4$1.6/M$0.1/M
text-ada-001$0.4$0.4/M
gpt-3.5-turbo$0.5$1.5/M
text-babbage-001$0.5$0.5/M
gpt-4o-mini-ttsTakes audio, returns audio.$0.6$12/M0.95 s
gpt-3.5-turbo-1106$1$2/M
o3-mini-global$1.1$4.4/M$0.55/M
gpt-3.5-turbo-0301$1.5$1.5/M
gpt-3.5-turbo-0613$1.5$2/M
gpt-3.5-turbo-instruct$1.5$2/M
davinci-002$2$2/M
o3-global$2$8/M$0.5/M
text-curie-001$2$2/M
gpt-4o-2024-08-06$2.5$10/M$1.25/M
gpt-4o-2024-08-06-global$2.5$10/M$1.25/M
gpt-4o-zhTakes text, vision, returns text.$2.5$10/M
computer-use-preview$3$12/M
gpt-3.5-turbo-16k$3$4/M
gpt-3.5-turbo-16k-0613$3$4/M
o1-mini$3$12/M$1.5/M
o1-mini-2024-09-12$3$12/M$1.5/M
gpt-image-test$5$40/M
distil-whisper-large-v3-enTakes audio, returns text.$5.556$5.556/M
gpt-4-0125-preview$10$30/M
gpt-4-1106-preview$10$30/M
gpt-4-turbo$10$30/M
gpt-4-turbo-2024-04-09$10$30/M
gpt-4-turbo-preview$10$30/M
gpt-4-vision-preview$10$30/M
o3-deep-research$10$40/M$2.5/M
o1-2024-12-17Takes text, vision, returns text.$15$60/M$7.5/M
o1-previewTakes text, vision, returns text.$15$60/M$7.5/M
o1-preview-2024-09-12$15$60/M$7.5/M
tts-1Takes audio, returns audio.$15$15/M
tts-1-1106Takes audio, returns audio.$15$15/M
davinci$20$20/M
o3-pro-global$20$80/M
text-davinci-002$20$20/M
text-davinci-003$20$20/M
text-davinci-edit-001$20$20/M
text-search-ada-doc-001$20$20/M
gpt-4$30$60/M
gpt-4-0314$30$60/M
gpt-4-0613$30$60/M
tts-1-hdTakes audio, returns audio.$30$30/M
tts-1-hd-1106Takes audio, returns audio.$30$30/M
gpt-4-32k$60$120/M
gpt-4-32k-0314$60$120/M
gpt-4-32k-0613$60$120/M

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

OpenAI on AIHubMix

Which OpenAI model should I start with?

auto is free on input — the cheapest entry here that declares tool calling, and it carries a 1M context. Move up to o1-pro when answer quality matters more than cost, or to gpt-5.5-free for long-form reasoning.

Which of these models reason before answering?

43 of the 134 models here declare a reasoning phase — they work through the problem before producing an answer, which helps on multi-step problems at the cost of extra output tokens. Use the Reasoning filter above the table to see them. The catalog does not record anything further about how they differ, so this page does not sort them into families.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

How is cached input billed?

The Cache read column is the rate for input tokens served from the prompt cache — for example gpt-5 bills cache hits at 10% of the input rate and gpt-5-chat-latest bills cache hits at 10% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.

Do I need a separate OpenAI account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling OpenAI in one line

One key, one endpoint, 880 models across 38 model authors.