Search results
Published directory entries matching your search.
Search results
4,896 listingsdaspartho/prompt-extend
Extending stable diffusion prompts with suitable style cues using text generation.
disco-diffusion
Frankensteinian amalgamation of notebooks, models and techniques for the generation of AI Art and Animations.
DreamStudio
DreamStudio is an easy-to-use interface for creating images using the Stable Diffusion image generation model.
Edge TTS Text To Speech
Works with speech, voice, music or other audio using machine-learning models.
Edge TTS Text To Speech
Works with speech, voice, music or other audio using machine-learning models.
Endel
Personalized soundscapes to help you focus, relax, and sleep. Backed by neuroscience.
EspNet
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
Faster Whisper
Fast inference engine for whisper in C++ using CTranslate2.
Free Tegami Tools JP
結婚式スピーチ・弔辞・年賀状・退職挨拶・お礼状・お詫び文など、冠婚葬祭と日常の手紙・挨拶文をAIが作成します。シチュエーションを選んで項目を埋めるだけで、そのまま使える文面が完成。82種類のツールを登録不要・完全無料で今すぐ使えます。.
FunASR
Speech recognition toolkit with streaming ASR, VAD, punctuation, speaker diarization, and OpenAI-compatible serving for voice AI applications.
Gempix2
Free production platform for text-to-image generation using Nano Banana V2 model.
ggerganov/whisper.cpp
Port of OpenAI's Whisper model in C/C++. It can be executed locally.
Good Tape
Good Tape uses AI for speech, voice, transcription, music, or other audio workflows.
Harmonai
We are a community-driven organization releasing open-source generative audio tools to make music production more accessible and fun for everyone.
Image Generation & Editing
1. Create Images: Generate images from text prompts using Gemini 2.0 Flash.
Image To Image Generation
Creates or edits images with generative models and visual controls.
Image to text
Showcase streaming text generation using the @huggingface/inference JS lib.
Image-based soundtrack generation
Creates or edits images with generative models and visual controls.
Img To Music
Img To Music uses AI for speech, voice, transcription, music, or other audio workflows.
Infinite Music
Distribution, publishing, funding, marketing, and a hands-on team for independent artists. We.
Kaldi
Open-source project that uses AI for speech, voice, transcription, music, or other audio workflows.
KangweiiLiu/Awesome Audio-driven Talking-Face-Generation
A curated list of resources of audio-driven talking face generation.