Framework for real-time multimodal voice agents
A Python framework that unifies speech-to-text, language models, text-to-speech, real-time APIs, telephony, MCP tools, and testing for conversational voice applications.
From ManuAGI - AutoGPT Tutorials — Trending Open-Source GitHub Projects : Kaneo, Rome, Paseo, Proliferate & Screendrop #288 at 02:56
Problem: The integration burden of speech, language, telephony, and real-time systems.
For: Builders of conversational voice applications.
Products from this video
CesiumJS Claude Code Claude Code Plugins Directory Cloudflare D1 Cloudflare R2 Cloudflare Workers Docker Compose Doop Doop ego lite Electron Git God's Eye View Google 3D Tiles hayamimi (早耳) Kaneo Kubernetes LatticeDB learn LiveKit LiveKit Agents MAX Framework Model Context Protocol (MCP) Modular Platform Mojo Dialer Munder Difflin Munder Difflin Needle 2 OCR It OpenAI Realtime API OpenAI Realtime API Paseo Paseo Proliferate Proliferate Python Rome Rome Screendrop SwarmForge Tailcat Tailscale Tesseract OCR vphone-cli