
Developer-first voice AI platform for building and scaling enterprise voice agents with sub-500ms latency.
Enterprise deployments anchor the platform's market position. Amazon Ring moved all inbound support calls to the platform after evaluating more than 40 vendors and reported improved customer satisfaction scores; Kavak reports growth in sales and NPS in the Latin American used-car market, and Instawork runs more than a million minutes of monthly voice screening on the platform. Published case studies span automotive, freight, healthcare, insurance, logistics, and marketplace customers.
The architecture separates model choice from platform operation: teams bring their own transcription, language, and speech models or use managed integrations, while the platform supplies telephony control over custom SIP with warm transfers and voicemail detection, guardrails against model hallucination, and compliance certifications including SOC 2 and HIPAA. Package tiers add contractual uptime commitments up to a 99.9% SLA at the top tier.
The platform emerged from a pivot by the team behind Superpowered, a bootstrapped AI meeting-notes product the founders had grown to profitability after founding it in 2020. In November 2023 the company opened its API with Vapi Phone and Vapi Web products, letting developers create voice-based assistants through prompts and put them behind a phone number, with the assistant runtime stringing together third-party services for telephony, transcription, streaming, model responses, and speech synthesis.
The following three years moved the offering from a developer API into enterprise deployment: published case studies now include Amazon Ring, which routes all inbound support calls through the platform, alongside Kavak, UnityAI, Spring Venture Group, Instawork, Fleetworks, and Ancile Services, and the company added success packages with contractual uptime commitments on top of usage-based pricing.
The platform prices in two separable parts: usage-based call costs and subscription packages layered on top. Usage starts at a $0.05 per minute hosting fee with $5 in free credits, while transcription, language model, and speech providers bill at pass-through cost without markup and transport fees depend on the telephony provider; a sample 1,000-minute deployment using Deepgram transcription, OpenAI intelligence, and ElevenLabs voices through Vapi SIP displays a monthly estimate of $82–$129 beyond the $50 hosting fee.
Subscription packages scale by concurrency and organization count: Core at $29 monthly for builders with 10 concurrent calls, Pro for production teams at 10% of the hosting fee with a $999 monthly minimum including a 99% uptime SLA, and Premier through direct sales with a 99.9% uptime SLA, custom security controls, and a named support team. Add-ons cover HIPAA-eligible data handling at $2,000 monthly, additional concurrency lines, and additional organizations.
Vapi's platform provides the build and runtime layer for production voice agents: conversations configured through code or a dashboard UI, with sub-500ms response targets, natural turn-taking, and custom pronunciation handling. Builders can bring their own transcription, LLM, and speech models or choose from a managed catalog that spans more than 200 model integrations, and every capability is exposed through a REST API alongside the UI.
The platform covers telephony controls including custom SIP, warm transfers, voicemail detection, phone number management, and DTMF support, plus live-call tool access through external data connections and MCP. Supporting features include multilingual operation across 100+ languages, automated testing suites to catch hallucination risk before release, A/B experiments across prompts and voices, model fallbacks across providers, and reporting dashboards with exports for conversation quality.