AI-powered audio transcription tool using OpenAI Whisper. Convert speech to text with support for 8+ formats, batch processing, and multilingual transcription. Python CLI tool for developers.
-
Updated
Aug 10, 2025 - HTML
AI-powered audio transcription tool using OpenAI Whisper. Convert speech to text with support for 8+ formats, batch processing, and multilingual transcription. Python CLI tool for developers.
Windows GUI for building better Kokoro TTS voice outputs. Combines Kokoro voice random-walk search, target-audio scoring, RVC model training, and automatic post-generation voice conversion into a single queue-based workflow. Bootstraps its own Python environment from one executable.
ElevenLabs-Desktop - AI voice cloning and text-to-speech with unlimited characters and all voices.
A curated list of AI audio generation APIs, SDKs, and tools including text-to-speech, speech synthesis, music generation, voice cloning, sound design, and generative AI platforms. Covers commercial services, open source models with APIs, and production-ready infrastructure for developers building audio applications.
Seed Audio AI Generator: Turn text scripts into polished multi-track voiceovers, character dialogues, sound effects (SFX), and music beds. Powered by Doubao Seed Audio 1.0 workflows & Gemini TTS. Edit timeline, mix tracks, and download high-fidelity WAV files instantly in-browser. Fast SEO traffic generator.
Autovideo v2
This a on testing project for video dubbing. Feel free to contribute on this and fix some issues like "add more idioms", "felling wise translation".
🌟 Introduce Your Self 🌟 Created a visual self-introduction video using text, visuals, and motion — without showing face. Highlighted my academic journey, tech interests, and passion for Data Analytics. Expressed gratitude to Internee.pk and shared my professional goals for this internship. A creative blend of storytelling and analytics.
local text-to-speech and voice AI app for Apple Silicon Macs. > Everything runs on-device. No cloud. No subscriptions
Curate and access APIs, SDKs, and tools for AI audio generation, including text-to-speech, music creation, and sound design.
ElevenLabs-Desktop — AI voice cloning and text-to-speech with unlimited characters and all voices.
Stop paying monthly fees. High-fidelity voice cloning on your own GPU. No limits, no credits.
no need to use CODE_OF_CONDUCT.md , LICENSE ,SECURITY.md
Voyicer is a fan-made companion app collection that reads the screen out loud in a voice cloned from the real character. Clone any voice from 4 words.
Clone voices and generate unlimited text-to-speech in 29 languages with all premium features unlocked.
🌟 Showcase your data-driven journey with a visual introduction to your skills, passions, and ambitions as a Data Analyst and Software Engineer.
To associate your repository with the ai-voice-generator topic, visit your repo's landing page and select "manage topics."