AI Dictation uses AI for speech, voice, transcription, music, or other audio workflows.
Directory
Search results
Published directory entries matching your search.
A free AI voice generator that generates natural sounding text-to-speech voice overs.
AI Wedding Toast uses AI for speech, voice, transcription, music, or other audio workflows.
AllVoiceLab uses AI for speech, voice, transcription, music, or other audio workflows.
An early look our AI Music experiment - YouTube Blog uses AI for speech, voice, transcription, music, or other audio workflows.
🔥🔥🔥自定义Android相机(仿抖音 TikTok),其中功能包括视频人脸识别贴纸,美颜,分段录制,视频裁剪,视频帧处理,获取视频关键帧,视频旋转,添加滤镜,添加水印,合成Gif到视频,文字转视频,图片转视频,音视频合成,音频变声处理,SoundTouch,Fmod音频处理。 Android camera(imitation Tik Tok), which includes video editor,audio editor,video face recognition stickers, segment recording,video cropping, video frame processing, get the first video frame, key frame, video rotation, add filter Mirror ,add watermark ,add gif to video, add text to video, picture to video, audio and video synthesis, audio change processing
AnyVoiceLab uses AI for speech, voice, transcription, music, or other audio workflows.
(USA) A API company for advanced Speech-to-Text, offering highly accurate transcription, summarization, and audio intelligence.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Audio.Z.AI uses AI for speech, voice, transcription, music, or other audio workflows.
AudioArena uses AI for speech, voice, transcription, music, or other audio workflows.
AudioCraft is a single-stop code base for all your generative audio needs: music, sound effects, and compression after training on raw audio signals.
You can also find more comprehensive list on and There's an AI AI Voice Cloning list.
Balabolka is a text-to-speech application (freeware).
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Bespoke Sounds uses AI for speech, voice, transcription, music, or other audio workflows.
LEGO recognition tool for identifying parts, minifigures, and sets from images.
Browser extension that helps solve audio CAPTCHA challenges using speech recognition.
Text-to-speech solutions with character.
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
CMU Sphinx uses AI for speech, voice, transcription, music, or other audio workflows.
Free speech-to-text tool for content creators that accurately transcribes audio & video files up to 2GB.
Course materials and notes for Stanford class CS231n: Deep Learning for Computer Vision.