Projects with this topic
-
Self-hosted speech server in one Docker image. OpenAI-compatible /v1/audio/transcriptions and /v1/audio/speech across 12 ASR models (Whisper, Parakeet, Canary, Sherpa-ONNX, Vosk) and 3 TTS engines (Kokoro, Qwen3-TTS voice cloning, Chatterbox Turbo). Live WebSocket ASR, file staging, MCP built in. CPU + CUDA images.
Updated -
Axis Intelligence is an open-source, modular ambient AI platform designed to function as the central intelligence layer of a modern smart home. Built with a local-first approach, it combines real-time voice interaction, AI reasoning, voice cloning, and smart home automation into a unified system that users can fully control, extend, and customize. https://roxanneardary.com/axis-intelligence/
Updated -
Give your AI agent a voice. Local speech stack for Apple Silicon: TTS, ASR, forced alignment, voices, daemon, and MCP bridge.
Updated