Speech-to-speech AI assistant with natural conversation flow, mid-speech interruption, vision capabilities and AI-initiated follow-ups. Features low-latency audio streaming, dynamic visual feedback, and works with local LLM/TTS services via OpenAI-compatible endpoints.
Catalog snapshot
Repository health snapshot: 51 forks and 5 open issues at review time. Platform support is inferred conservatively from repository topics.
Speech-to-speech AI assistant with natural conversation flow, mid-speech interruption, vision capabilities and AI-initiated follow-ups. Features low-latency audio streaming, dynamic visual feedback, and works with local LLM/TTS services via OpenAI-compatible endpoints. Fossfits includes Vocalis as an actively maintained speech ai option: its public repository was updated on 2025-04-14, uses the APACHE-2.0 license, and had 316 stars when reviewed. Check its documentation and release notes before production adoption.
artificial intelligence · conversational ai · speech to speech · visionprocessing