Model Intelligence for transcribers, models, and voices
Model Intelligence is a Vapi feature that helps you choose the transcriber, model, and voice for your assistant. Choose a preset for a strong default combination or compare every option using weekly refreshed performance metrics.
Key concepts
How it works
When building a voice agent, Vapi orchestrates the transcriber, model, and voice. You can swap each component for any supported provider and model, which gives you flexibility when building an assistant.
Model Presets bundle the three components into a curated combination. Choose Balanced, High Intelligence, Ultra Fast, or Cost Saver based on your goal. Presets provide a dependable setup without requiring you to tune each component.
Performance metrics show latency, cost, and quality data for every transcriber, model, and voice. The data appears on component panels and the dropdown menu for each component. Use it to compare options and build the combination your assistant needs. See the performance metrics reference for how each metric is sourced and calculated.
If you change one component from its preset, the assistant moves to Customized. Every other preset setting remains unchanged, and you can continue editing any component.
Start with presets and optimize with performance metrics
Model Presets give you a dependable starting point when you do not know which component to choose. Performance metrics help you compare options and optimize your configuration with data instead of guessing.