GPT-Live

Voice conversations with tools and reasoning built in

GPT-Live is an OpenAI voice model that can listen and speak at the same time. A caller can clarify a request while the assistant is talking, ask a question during a lookup, or ask it to slow down. This is full-duplex conversation: audio can flow in both directions at once.

Vapi connects that conversation to your tools and phone numbers, and provides the call records, analysis, and monitoring around it.

GPT-Live must be enabled for your Vapi organization. If you don’t have access yet, join the waitlist. We’re admitting users from the waitlist every day.

How it works

A GPT-Live assistant uses a speaker for the spoken interaction and a reasoner for tasks that need tools or detailed procedures. Delegation is the speaker asking the reasoner to handle one of those tasks.

For a scheduling assistant, the speaker can guide the caller toward a suitable appointment. The reasoner checks availability through your service and returns the options. The speaker can keep listening while that work happens.

GPT-Live handles speech directly. You don’t configure a separate speech-to-text transcriber or text-to-speech provider for the conversation.

What changes with GPT-Live

You can design the interaction around the caller’s goal, with room for questions and changes along the way. Detailed procedures belong with the reasoner. The speaker’s instructions focus on how to guide the conversation.

You also have more to design than the words alone. Set a voice, give the speaker a delivery brief, and use personality packs to explore different styles. A concise assistant can slow down for a date, explain one step at a time, or give the short version when asked.

Design your assistant explains these choices with examples, including how to rethink an existing squad.

What Vapi provides

CapabilityWhere to learn more
Browser, phone, and WebSocket callsConnect a call
Tools that look up information or take actionsConnect your tools
Speaker prompts, reasoner settings, and personality packsModel settings
22 preset voices, with audio previewsChoose a voice
Transcripts, available recordings, and post-call analysisReview calls
Structured outputs, scorecards, Boards, and MonitoringTrack quality

GPT-Live has a different set of supported features from Vapi’s other voice architectures. Check Limitations and FAQs if your application depends on squads, live call control, simulations, or a specific transfer flow.

Get started

For the underlying model, see OpenAI’s GPT-Live documentation. This section describes Vapi’s integration and its supported configuration.