This is an early release preview. You may encounter bugs.
Pipecat logo
Unclaimed

Framework voice

Pipecat

Python framework for real-time voice and multimodal conversational agents, with composable pipelines over WebSockets or WebRTC

A 88/100 GitHub score ? This grade is derived from GitHub signals, not user votes. Open for the full breakdown.
No votes yet

01 / About

What Pipecat is.

Pipecat is a Python framework for building real-time voice and multimodal conversational agents. It orchestrates audio and video streams, AI services, transports, and conversation logic as composable pipelines, so a single voice assistant or a multi-agent system can be assembled from modular processors.

Each pipeline is an agent. Pipelines can hand off to specialists, fan out in parallel, run as sidecar workers over a shared bus, or be distributed across processes and machines. Real-time interaction runs over transports such as WebSockets or WebRTC. Pipecat Flows, built into the framework, adds predefined or dynamic conversation paths with state management.

The surrounding tooling includes client SDKs for JavaScript, React, React Native, Swift, Kotlin, C++, and ESP32; a CLI (pipecat init) that scaffolds a project and can hand it to a coding assistant such as Claude Code or Codex, then monitors and deploys the agent; Whisker, a real-time debugger; Tail, a terminal dashboard; a Voice UI Kit of components and hooks; and Claude Code skills for scaffolding and deploying to Pipecat Cloud.

Service category Examples
Speech-to-text AssemblyAI, AWS, Azure, Cartesia, Deepgram, ElevenLabs, Gladia, Google, Groq (Whisper), Mistral, NVIDIA, OpenAI (Whisper), Speechmatics, Whisper, xAI
LLMs Anthropic, AWS, Azure, Baseten, Cerebras, DeepSeek, Fireworks AI, Gemini, Grok, Groq, and others
Other categories Text-to-speech, speech-to-speech, transport, serializers, video, memory, audio processing, community integrations

Features

  • Composable pipelines: builds agent behaviour from modular processors for audio, video, and AI services
  • Multi-agent composition: handoff, parallel fan-out, sidecar workers, and distributed deployment over a shared bus
  • Real-time transports: streaming interaction over WebSockets or WebRTC
  • Pipecat Flows: structured conversations with predefined or dynamic paths and state management
  • Client SDKs: JavaScript, React, React Native, Swift, Kotlin, C++, and ESP32
  • CLI: pipecat init scaffolds a project, and the CLI monitors and deploys agents to production
  • Debugging and monitoring: Whisker for real-time pipeline debugging and Tail for a terminal dashboard
  • Pluggable services: speech-to-text, LLM, text-to-speech, speech-to-speech, video, memory, and audio-processing providers, plus community integrations
  • Coding-agent skills: Claude Code skills for project scaffolding and deployment to Pipecat Cloud

02 / Discussion CREDIBILITY-GATED

Discussion

Reading is open to everyone. Posting and voting need a verified identity or a GitHub grade of B or higher.

  • No discussions yet.

03 / Build

Build with Pipecat.

Browse the catalogue for frameworks, tools, and harnesses, each scored on real GitHub credibility.

Get Pipecat →

Browse the catalogue