Unlock exclusive introductory pricing for the newly launched Gemini 3.5 Flash.
O

gpt-realtime-1.5

อินพุต:$3.2/M
เอาต์พุต:$12.8/M
บริบท:32,000
เอาต์พุตสูงสุด:4,096
วันที่เผยแพร่:Feb 24, 2026

The best voice model for audio in, audio out.

ใหม่
ใช้งานเชิงพาณิชย์

Technical Specifications of gpt-realtime-1.5

Itemgpt-realtime-1.5 (public positioning)
Model familyGPT Realtime 1.5 (voice-optimized variant)
Primary modalitySpeech-to-speech (S2S)
Input typesAudio (streaming), text
Output typesAudio (streaming), text, structured tool calls
APIRealtime API (WebRTC / persistent streaming sessions)
Latency profileOptimized for low-latency, live conversational interaction
Session modelStateful streaming sessions
Tool useFunction calling and tool integrations supported
Target use caseLive voice agents, assistants, interactive systems

Note: Exact token limits and context window sizes are not prominently documented in public summaries; the model is positioned for realtime responsiveness rather than extremely long context sessions.


What is gpt-realtime-1.5?

gpt-realtime-1.5 is a low-latency, speech-to-speech optimized model designed for live conversational systems. Unlike traditional request-response models, it operates through persistent streaming sessions, enabling natural turn-taking, interruption handling, and dynamic voice interaction.

It is purpose-built for applications where conversational flow speed matters more than maximum context length.


Main Features

  1. True speech-to-speech interaction — Accepts live audio input and streams spoken responses in real time.
  2. Low-latency architecture — Designed for sub-second conversational responsiveness in voice agents.
  3. Streaming-first design — Works via persistent sessions (WebRTC or streaming protocols).
  4. Natural turn-taking — Supports interruption handling and dynamic conversation flow.
  5. Tool calling support — Can trigger structured function calls during a realtime session.
  6. Production-ready voice agent foundation — Built specifically for interactive assistants, kiosks, and embedded devices.

Benchmark & Performance Positioning

OpenAI positions gpt-realtime-1.5 as an evolution of earlier realtime models with improved instruction-following, stability during extended voice sessions, and more natural prosody compared to earlier releases.

Unlike coding-focused models (e.g., Codex variants), performance is measured more by conversational latency, voice naturalness, and session stability than by leaderboard-style benchmarks.


Featuregpt-realtime-1.5gpt-audio-1.5
Primary goalLive voice interactionAudio-enabled chat workflows
LatencyOptimized for minimal delayBalanced quality/speed
Session typePersistent streaming sessionStandard Chat Completions flow
Context sizeOptimized for responsivenessLarger context support
Best use caseRealtime voice agentsConversational assistants with audio

When to Choose Each

  • Choose gpt-realtime-1.5 for call centers, kiosks, AI receptionists, or live embedded assistants.
  • Choose gpt-audio-1.5 for voice-enabled chat apps that require longer conversation memory or multimodal workflows.

Representative Use Cases

  • AI call center agents
  • Smart device assistants
  • Interactive kiosks
  • Live tutoring systems
  • Real-time language practice tools
  • Voice-controlled applications
  • How to access GPT realtime 1.5 API

Step 1: Sign Up for API Key

Log in to cometapi.com. If you are not our user yet, please register first. Sign into your CometAPI console. Get the access credential API key of the interface. Click “Add Token” at the API token in the personal center, get the token key: sk-xxxxx and submit.

cometapi-key

Step 2: Send Requests to GPT realtime 1.5 API

Select the “gpt-realtime-1.5” endpoint to send the API request and set the request body. The request method and request body are obtained from our website API doc. Our website also provides Apifox test for your convenience. Replace <YOUR_API_KEY> with your actual CometAPI key from your account. base url is Chat Completions

Insert your question or request into the content field—this is what the model will respond to . Process the API response to get the generated answer.

Step 3: Retrieve and Verify Results

Process the API response to get the generated answer. After processing, the API responds with the task status and output data.

คำถามที่พบบ่อย

ราคาสำหรับ gpt-realtime-1.5

สำรวจราคาที่แข่งขันได้สำหรับ gpt-realtime-1.5 ที่ออกแบบมาให้เหมาะสมกับงบประมาณและความต้องการการใช้งานที่หลากหลาย แผนการบริการที่ยืดหยุ่นของเรารับประกันว่าคุณจะจ่ายเฉพาะสิ่งที่คุณใช้เท่านั้น ทำให้สามารถขยายขนาดได้ง่ายเมื่อความต้องการของคุณเพิ่มขึ้น ค้นพบว่า gpt-realtime-1.5 สามารถยกระดับโปรเจกต์ของคุณได้อย่างไรในขณะที่ควบคุมต้นทุนให้อยู่ในระดับที่จัดการได้

Comet Price (USD / M Tokens)Official Price (USD / M Tokens)Discount
อินพุต:$3.2/M
เอาต์พุต:$12.8/M
อินพุต:$4/M
เอาต์พุต:$16/M
-20%

โค้ดตัวอย่างและ API สำหรับ gpt-realtime-1.5

เข้าถึงโค้ดตัวอย่างที่ครอบคลุมและทรัพยากร API สำหรับ gpt-realtime-1.5 เพื่อปรับปรุงกระบวนการผสานรวมของคุณ เอกสารประกอบที่มีรายละเอียดของเราให้คำแนะนำทีละขั้นตอน ช่วยให้คุณใช้ประโยชน์จากศักยภาพเต็มรูปแบบของ gpt-realtime-1.5 ในโครงการของคุณ