Projects with this topic
Sort by:
-
[mirror] OpenAI-compatible audio server in Docker. 7 ASR backends (Whisper, Distil-Whisper, Parakeet, Canary, Canary-Qwen) + 2 TTS engines (Kokoro, Qwen3-TTS voice cloning). Single /v1/audio/{transcriptions,speech,voices} surface. CPU + CUDA images. Hot model swap. MCP server built in.
Updated -
System-wide voice-to-text for macOS. Hold a hotkey, speak, text appears at your cursor. Supports OpenAI, Gemini, Apple Speech & local models. No subscription — bring your own API keys.
Updated -
Provides easy access to the whisper.cpp application on snap-enabled OS distributions.
Updated -
web components that record audio and transcribe it to text using openai api's https://attention1.gitlab.io/ai-interface
Updated