
GPT-6 Sol
探索 GPT-6 Sol API。

探索 GPT-6 Sol API。

GPT-6 Astra, the flagship model for complex reasoning and coding. Choose GPT-5.6 Terra to balance intelligence and cost,

OpenAI 最強大的圖像生成模型,具備跨多語言近乎完美的文字渲染、最高可達 4K 解析度,以及由推理驅動的 Thinking Mode。專為對準確性、速度與符合品牌調性的視覺輸出有嚴格要求的生產級工作流程而打造。

探索 gpt-oss-20b(free) API。

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

探索 GPT Image 2.5(sunburst) API。
GPT-Realtime-2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

seedance-2-0

The best voice model for audio in, audio out.

OpenAI 最強大的模型,專為最艱鉅的任務與長時間運行的代理式工作流程打造。GPT-5.5 Pro 在複雜程式設計、電腦操作、深度研究、資料分析與科學推理方面表現卓越—以 GPT-5.4 的延遲並具備更高的 token 效率,提供前沿級智能。非常適合對準確性與自主任務執行提出最高標準要求的企業與專業應用場景。

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments. GPT-5.3-Codex supports low, medium, high, and xhigh reasoning effort settings.

GPT Image 2 ALL is a comprehensive image generation model designed to handle a wide range of creative and professional visual tasks. It combines high-quality image creation, advanced prompt understanding, and versatile style support to deliver exceptional results across diverse use cases.

GPT-Image-1.5 is OpenAI’s image model in the GPT Image family . It is a natively multimodal GPT model designed to generate images from text prompts and to perform high-fidelity edits of input images while following user instructions closely.

gpt-5.2-pro is the highest-capability, production-oriented member of OpenAI’s GPT-5.2 family, exposed through the Responses API for workloads that demand maximal fidelity, multi-step reasoning, extensive tool use and the largest context/throughput budgets OpenAI offers.
GPT-5.1 是一款通用的指令微調語言模型,專注於跨產品工作流程的文本生成與推理。它支援多輪對話、結構化輸出格式,以及以程式碼為導向的任務,例如撰寫、重構與說明。典型用例包括聊天助理、檢索增強式問答、資料轉換,以及在支援的情況下透過工具或 API 進行代理式自動化。技術亮點包括以文本為中心的模態、指令遵循、JSON 風格輸出,以及與常見編排框架中的函式呼叫相容。
GPT Image 1 的成本優化版本。它是一款原生多模態語言模型,可接受文字與影像輸入,並產生影像輸出。
GPT-5 是 OpenAI 迄今為止最強大的程式碼模型。它在複雜前端生成與大型程式碼庫的偵錯方面有顯著提升。它能以直觀且具美感的成果將想法化為現實,只需一則提示即可打造優美且響應式的網站、應用程式與遊戲,並展現敏銳的美學感知。早期測試者也注意到其設計選擇,對於間距、字體排印與留白等元素有更深刻的理解。
GPT-5 Nano 是由 OpenAI 提供的人工智慧模型。
GPT-5 mini 是 OpenAI 針對成本與延遲優化的 GPT-5 系列成員,旨在以顯著更低的成本,為大規模生產使用提供 GPT-5 在多模態與指令遵循方面的絕大部分優勢。它面向以吞吐量、可預測的每 token 定價與快速回應為主要約束的環境,同時仍提供強大的通用能力。
coming soon
gpt-4o-image generate images as output, optionally using images as input

Sora 2 Pro 是我們最先進且最強大的媒體生成模型,能生成帶有同步音訊的影片。它可以從自然語言或圖像創建細節豐富、動態的影片片段。

GPT-5.2 is a multi-flavored model suite (Instant, Thinking, Pro) engineered for better long-context understanding, stronger coding and tool use, and materially higher performance on professional “knowledge-work” benchmarks.

探索 GPT Image 2.5 Flare API。
OpenAI o3‑pro 是 o3 推理模型的「pro」變體,經過工程化設計,以進行更長程的思考並輸出最可靠的回應,藉由採用私有思維鏈強化學習,並在科學、程式設計與商業等領域樹立全新的最先進基準——同時可在 API 中自主整合如網路搜尋、檔案分析、Python 執行與視覺推理等工具。

OpenAI 文字轉語音

超強大的影片生成模型,具備音效,支援對話格式。

語音轉文字,生成翻譯
.jpeg&w=3840&q=75)
OpenAI 最智慧且最直觀的旗艦模型,專為複雜程式設計、代理式工作流程、電腦操作、資料分析與深入研究而設計。以與 GPT-5.4 相同的低延遲提供前沿級智慧,同時以更高的 Token 效率完成任務。是面向高要求的專業與企業級工作負載的首選模型。
探索 tts-1 API。
GPT-4o mini TTS 是一款神經網路文字轉語音模型,旨在於面向使用者的應用程式中實現自然、低延遲的語音生成。它可將文字轉換為自然聽感的語音,提供可選語音、多種格式輸出與串流合成,帶來反應迅速的體驗。典型用例包括語音助理、IVR 與聯絡流程、產品內容朗讀與媒體旁白。技術亮點包括基於 API 的串流,以及匯出為 MP3 與 WAV 等常見音訊格式。
探索 o1-pro-2025-03-19 API。
GPT-4o Transcribe 是一款用於多語言、低延遲語音辨識的語音轉文字模型。它支援即時串流與批次轉錄,能從常見音訊格式產生帶有標點與句子切分的轉錄結果。典型用途包括即時字幕、語音助理輸入、會議筆記,以及媒體或通話錄音的轉錄。技術亮點包括音訊模態支援、長篇內容處理,以及適用於互動式與伺服端工作流程的 API。

GPT-5.4 is the frontier model for complex professional work. Reasoning.effort supports: none (default), low, medium, high and xhigh.
GPT-4o mini Audio Preview 是一款用於構建會話式音訊應用的輕量多模態模型。它在文字之外同時支援語音輸入與輸出,從而實現語音辨識、語音合成,以及文字與音訊混合對話,並透過工具/函式呼叫執行結構化操作。典型用例包括語音助理、附帶摘要的串流式轉錄、IVR 與呼叫機器人工作流程,以及具備音訊功能的應用程式內助理。技術亮點包括音訊 I/O、串流回應、指令遵循,以及透過聊天與工具 API 的整合。
GPT-4o mini Realtime Preview 是一款用於互動式語音與視覺體驗的即時多模態模型。它支援串流式輸入與輸出,可處理語音、文字與圖片,並能透過工具/函式呼叫執行實際操作。典型用例包括語音助理、即時通話處理、即時字幕,以及針對相機或螢幕內容的視覺問答。技術亮點包括雙向音訊、視覺理解、串流式回應,以及透過函式產生的結構化輸出。
探索 o1-2024-12-17 API。

GPT-5.3 Instant model used in ChatGPT

GPT-4o mini Audio 是一個用於語音與文字互動的多模態模型。它能執行語音辨識、翻譯與文字轉語音,遵循指令,並能呼叫工具以進行結構化操作,提供串流式回應。典型用途包括即時語音助理、即時字幕與翻譯、通話摘要,以及語音控制的應用程式。技術亮點包括音訊輸入與輸出、串流回應、函式呼叫,以及結構化 JSON 輸出。
探索 omni-moderation-2024-09-26 API。
探索 omni-moderation-latest API。
探索 tts-1-hd-1106 API。

The best voice model for audio in, audio out with Chat Completions.
探索 tts-1-1106 API。
GPT-Realtime-2.1 updates GPT-Realtime-2 with improved alphanumeric recognition, silence and noise handling, and interruption behavior. It supports speech-to-speech interactions with configurable reasoning effort, instruction following, and tool use for complex voice-agent workflows.
探索 tts-1-hd API。

Start with GPT-5.6 Sol for complex reasoning and coding, choose GPT-5.6 Terra to balance intelligence and cost, or use GPT-5.6 Luna for cost-sensitive, high-volume workloads.
An Ada-based text embedding model optimized for various NLP tasks.
A small text embedding model for efficient processing.
A large text embedding model for a wide range of natural language processing tasks.
An advanced AI model for generating images from text descriptions.
New version of DALL-E for image generation.

Coming Soon

Coming Soon

coming soon
O4-mini 是由 OpenAI 提供的人工智慧模型。
O3-mini 是由 OpenAI 提供的人工智慧模型。
O3 是由 OpenAI 提供的人工智慧模型。
O1-pro is an artificial intelligence model provided by OpenAI.
O1 is an artificial intelligence model provided by OpenAI.
gpt-oss-20b is an artificial intelligence model provided by cloudflare-workers-ai.
gpt-oss-120b is an artificial intelligence model provided by cloudflare-workers-ai.
GPT-4o mini 是由 OpenAI 提供的人工智慧模型。
<div>GPT-4o 是 OpenAI 最先進的多模態模型,速度更快、價格更低,且視覺能力更強,比 GPT-4 Turbo 更出色。此模型具備 128K 的上下文長度,知識截止於 2023 年 10 月。1106 系列及以上的模型支援 tool_calls 和 function_call。</div> 此模型支援的最大上下文長度為 128,000 個 token。
GPT-4.1 nano 是由 OpenAI 提供的人工智慧模型。 gpt-4.1-nano: 具備更大的上下文視窗—支援最多 1 million 個上下文 token,並能透過改進的長上下文理解更好地利用該上下文。 知識截止時間更新為 2024 年 6 月。 此模型支援的最大上下文長度為 1,047,576 個 token。
GPT-4.1 mini 是由 OpenAI 提供的人工智慧模型。gpt-4.1-mini:在小型模型效能上實現重大躍進,在許多基準測試中甚至超越 GPT-4o。它在智慧評估上達到或超越 GPT-4o,同時將延遲降低近一半,成本降低 83%。此模型支援的最大上下文長度為 1,047,576 個 token。
GPT-4.1 是由 OpenAI 提供的人工智慧模型。gpt-4.1-nano:具備更大的上下文視窗—支援多達 1 million 個上下文 token,並能透過改進的長上下文理解更好地利用該上下文。知識截止時間已更新為 2024 年 6 月。此模型支援的最大上下文長度為 1,047,576 個 token。