Skip to main content
The Voice Agents category includes 7 projects that demonstrate real-time speech processing, conversational voice AI, and streaming pipelines. These projects show how to build agents that listen, speak, and reason in real time, powered by frameworks such as LiveKit, Pipecat, Gradium, and Sarvam.

Voice Agent Projects

Getting Started with Voice Agents

1

Navigate into a voice project

2

Configure your environment

Voice projects typically require STT/TTS keys (e.g., Deepgram, ElevenLabs, Sarvam) and LLM keys (e.g., OpenAI, Nebius, Gemini).
3

Install and run

Some projects may also expose a web UI for testing.
Voice projects often require WebSocket or WebRTC transport. For projects using LiveKit or Daily, follow the transport setup instructions in the individual READMEs.