kesha-voice-kit
Développement27 installation(s)v1.6.1Local multilingual voice toolkit — speech-to-text (STT), text-to-speech (TTS), speaker diarization, and language detection, over a CLI or an MCP server. Runs entirely offline on Apple Silicon, Linux, and Windows. No API keys, no cloud. NVIDIA Parakeet TDT for STT across 25 European languages, Kokoro-82M + Vosk-TTS for TTS in 9 languages, plus macOS AVSpeechSynthesizer for ~180 system voices with zero install.
Ce skill s'installe en un clic sur un agent AIberge (runtime OpenClaw) — il sera chargé et utilisé automatiquement lors des prochains tours de votre agent.