Total time is roughly 20 minutes, most of it Docker pulling images. You need Docker and one AI provider API key. Every stack runs on your own infrastructure — no call audio, transcript, or credential ever leaves your machine unless you choose a cloud model provider.
The path
Once the stack is up, build your first agent and place your first call from the dashboard — no terminal needed:
Prefer the API? See Recipes.
Before you start
You do not need a telephony account to try VoicEra. Awebsocket agent runs entirely in the browser and needs only an STT, TTS, and LLM key. Add telephony when you want real phone numbers.
Where next
Once a call works:- Architecture — what you just started
- Running a campaign — outbound at volume
- Security hardening — before anyone else can reach it