ImageBind One Embedding Space to Bind Them All.
Directory
Search results
Published directory entries matching your search.
Greg (Grzegorz) Surma - Portfolio; Machine Learning, Computer Vision, Self-Driving Cars, AI.
Official repository to use and implement LLaMA 3.1 models, Meta's state-of-the-art large language models.
Just like the IKEA Effect, but for the AI products. Try not to overenginner when working with LLMs.
](https://emu-video.metademolab.com/demo /demo): state-of-the-art text-to-video generation.
A text-based adventure-story game you direct (and star in) while the AI brings it to life.
Helps teams build, deploy, observe or operate machine-learning systems.
A free AI voice generator that generates natural sounding text-to-speech voice overs.
25th Aug, 2023 Feature: Adding TextArea for custom prompts.
Create Flashcards 10x faster. Generate Anki Flashcards from any File or Text with AI.
(USA) A API company for advanced Speech-to-Text, offering highly accurate transcription, summarization, and audio intelligence.
Text-to-Audio Generation with Latent Diffusion Models - Speech Research.
Balabolka is a text-to-speech application (freeware).
Open-source project that applies AI to workplace tasks, meetings, sales, marketing, or productivity.
Text-to-speech solutions with character.
Remove unwanted objects from photos, people, text, and defects from any picture for free. It.
Plugin for CMS Adobe Experience Manager (AEM) or Composum Pages helping the editor to create / edit / translate texts.
Free speech-to-text tool for content creators that accurately transcribes audio & video files up to 2GB.
Access full-text academic articles: J-STAGE is an online platform for Japanese academic journals.
Translate texts & full document files instantly. Accurate translations for individuals and Teams. Millions translate with DeepL every day.
Works with speech, voice, music or other audio using machine-learning models.
Works with speech, voice, music or other audio using machine-learning models.
Foundational Models for State-of-the-Art Speech and Text Translation.
High-level wrapper built on the top of Pytorch which supports vision, text, tabular data and collaborative filtering.