TextaVoice
TextaVoice (https://www.textavoice.com/) converts text to an audio file. The text (2000 characters maximum) is associated with a language and a...
Vibes
Vibes (https://vibes.ai/) generates images and videos from an instruction.Integrated with the Meta AI app (browse, remix, publish) and accessible from...
GPT-Live
GPT-Live (https://openai.com/index/introducing-gpt-live/) is OpenAI’s new voice model that powers ChatGPT Voice, real-time conversations with simultaneous listening and speech. GPT-Live processes...
Longscribe
Longscribe (https://longscribe.com) transcribes videos and long recordings (YouTube, Vimeo, TikTok, Instagram, Facebook, Twitch, X, Dailymotion, SoundCloud or imported audio/video files)....
Kyutai Pocket TTS
Kyutai Pocket TTS (https://kyutai.org/tts/ — documentation: https://github.com/kyutai-labs/pocket-tts) is an open-source text-to-speech model. Choice of language (French, English, German, Spanish, Italian,...
Lolaby
Lolaby (https://huggingface.co/spaces/build-small-hackathon/lolaby and documentation: https://huggingface.co/build-small-hackathon/lolaby-llama-3b) generates personalised lullabies based on an image (drawing or photo), a child’s first name, age...