Fast inference engine for whisper in C++ using CTranslate2.
Directory
Search results
Published directory entries matching your search.
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
FreeTTS uses AI for speech, voice, transcription, music, or other audio workflows.
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Google Flow Music uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Hume uses AI for speech, voice, transcription, music, or other audio workflows.
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
High-quality text-to-speech and voice recognition.
KikiVoice uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Kokoro TTS uses AI for speech, voice, transcription, music, or other audio workflows.
Deploy and scale machine learning models on Kubernetes. Built for LLMs, embeddings, and speech-to-text.
Kyutai TTS uses AI for speech, voice, transcription, music, or other audio workflows.
LazyPy uses AI for speech, voice, transcription, music, or other audio workflows.
Tts for Lojban using VITS TTS models.
LOVO uses AI for speech, voice, transcription, music, or other audio workflows.
Implementation of MusicLM, Google's new SOTA model for music generation using attention networks, in Pytorch.
"transforming the future of music creation".
Mazmazika uses AI for speech, voice, transcription, music, or other audio workflows.
MMAudio uses AI for speech, voice, transcription, music, or other audio workflows.
Moe TTS uses AI for speech, voice, transcription, music, or other audio workflows.