Projects with this topic
-
A self-hosted AI platform — inference, tool use, browser automation, image generation, speech synthesis, transcription, object storage, agentic code execution, and more — behind a single OpenAI-compatible endpoint. One docker-compose up.
Updated -
This is a fully client-side, JavaScript-only Text-to-Speech (TTS) service using eSpeakNG. No server-side bullshit. Just clone this repo into your desired directory, and you're all set.
Updated -
Qwen3-TTS over SSH. Pick a voice, clone a voice, design a voice - all through a YAML config piped via stdin. Models run locally, no API keys, no cloud bullshit.
Updated -
A PWA that provides a method to access chat like messaging functionality by using Rest API integrations from the Rhea Generative Framework.
The Rhea client app is component within the Rhea Generative Framework.
Originally intended as a way to demonstrate functionality found within the Rhea Generative Framework. The client app was designed to both demonstrate functionality and provide a foundation to build other components such as live chat (embeddable, etc.).
It has evolved over time to include additional functions for demonstration:
Persona management (role activation and management) Speech-to-text and text-to-speech (browser independent, both part of Rhea's Generative Framework server-side components and available for local hosting) STT captures audio for x seconds, transcribes and offers to either continue transcribing, send as a message, or manually edit.Updated -
A manimgl plugin to add narration.
UpdatedUpdated -
High-performance Text-to-Speech (TTS) REST API service for the Oremi ecosystem.
Updated -
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation
Updated -
Give your AI agent a voice. Local speech stack for Apple Silicon: TTS, ASR, forced alignment, voices, daemon, and MCP bridge.
Updated -
Flite text-to-speech synthesis bindings for Kit
Updated -
LittleCode is an open-source AI-powered design canvas that transforms natural language prompts—text, voice, or visual—into full, production-ready front-end code. Accessible to users of any age or skill level, it supports multiple frameworks, real-time previews, collaboration, and AI-driven design intelligence. With advanced features including accessibility tools, UX optimization, gamified learning, VR/AR support, and futuristic experimental capabilities like sentient-style AI and IoT integration, LittleCode enables anyone to see their design concepts come to life.
Updated -
A comprehensive Flutter mobile application developed for the "Mobile Application Development" course, Computer Science. This app demonstrates the integration of TMDB API to fetch upcoming movies, featuring a clean UI, and dynamic routing
Updated -
Curated NVIDIA text-to-speech (TTS) models, SDKs, and voice AI resources.
Updated -
Experimental text-to-speech utility via keyboard shortcut, powered by Large Language Models (LLM).
Updated -
This project provides a client package and example scripts for python to access the alphaspeech pro ASR APIs.
Updated -
Calculadora con voz, orientado para Linux
Updated -
This project provides a client package and example scripts for TypeScript to access the alphaspeech pro ASR stream API.
Updated -
🚛 ✈️ An advanced Streamlit dashboard designed as an AI-powered assistant for World Movers Phils Inc. This application leverages Google Gemini for multimodal interactions, enabling users to get information, request quotes, marketing, analyze documents/images, use voice commands, and more, all within a custom-themed interface.Updated -
The Python script utilizes the win32com library to interact with the Windows Speech API (SAPI), prompting the user to input text to be spoken aloud. It continuously speaks the input text using the default system text-to-speech engine until the user inputs "-1" to terminate the program.
Updated -
Simple module that helps you to create aloud report with numeric values.
Updated -
Python package to normalize text for speech-language models using different libraries.
Updated