Explore
Discover the best AI tools in one place
20 tools found
SpeechBrain is an open-source conversational AI toolkit designed for everyone, from researchers to developers. It provides a comprehensive suite of tools for building speech and audio processing applications. The platform supports a wide range of tasks including speech recognition and speaker identification.
Harmonai is an open-source AI platform dedicated to music generation and creative audio production. It enables musicians and creators to generate original compositions using advanced machine learning models. The platform supports various genres and styles, making AI-assisted music creation accessible to all skill levels.
DADABOTS uses neural networks to autonomously generate death metal music around the clock. The AI continuously produces new compositions without human intervention, creating an endless stream of algorithmic metal music.
Edge Dance uses AI to analyze music tracks and automatically generate synchronized dance routines that match the rhythm and style of the input audio. It creates choreographies for various dance styles and difficulty levels, making dance creation accessible to everyone. The tool is perfect for dancers, choreographers, and fitness enthusiasts looking for new routines.
Soundwise.ai is a free AI-powered transcription tool that converts audio and video files into accurate text transcripts. It supports multiple languages and file formats, making it easy to generate captions, subtitles, and meeting notes. The platform is designed for creators, journalists, and professionals who need fast, reliable transcription.
Gnod is an AI-powered discovery engine that helps users find new music, art, and other cultural content based on their preferences. It uses intelligent algorithms to map connections between different pieces of media. The platform aims to provide a serendipitous browsing experience.
Voice Isolator is an AI-powered audio tool that removes background noise and isolates voice from audio recordings. It is designed for podcasters, content creators, and professionals who need clean audio output. The tool uses advanced machine learning algorithms to separate voice from ambient sounds effectively.
Speechnotes is a fast and accurate speech-to-text tool that works directly in your browser. It supports multiple languages and offers a clean, distraction-free interface. Ideal for transcription, note-taking, and content creation from audio.
Freeway is a free transcription tool that lets users talk to their Mac for instant transcription. It provides real-time speech-to-text capabilities built specifically for macOS. The app focuses on simplicity and privacy with on-device processing.
UVR Online is a web-based tool that uses AI to remove vocals from any song, allowing users to create instrumental and acapella tracks quickly. It leverages advanced source separation models to isolate vocals from music with high accuracy. The tool is designed for musicians, DJs, content creators, and karaoke enthusiasts.
HappySRT is an open-source AI tool that handles transcription, translation, and summarization of audio and video content. It makes it easy to convert spoken language into text and translate it across multiple languages. The tool is designed for creators, researchers, and teams who need fast and accurate media processing.
AISong.Fun lets users generate original music tracks using artificial intelligence at no cost. Users can specify genre, mood, tempo, and instrumentation to create custom compositions. It is designed for content creators, hobbyists, and anyone looking for royalty-free background music.
Riffusion is an AI music generation tool that creates music in real-time from text prompts. It uses advanced neural networks to transform text descriptions into unique audio tracks. Ideal for creators looking for instant music inspiration.
Infinite Drum Machine is an experimental AI-powered music tool by Google that lets users create beats using sounds from everyday life. It uses machine learning to map real-world sounds to a drum machine interface, enabling unique rhythm creation without any musical expertise. The tool is part of Google's AI Experiments platform, showcasing creative applications of AI in music.
MusicLM is Google's AI music generation tool that allows users to create music from text descriptions. It represents Google's latest innovations in AI-powered audio generation.
Free Text-To-Speech converts written text into natural-sounding speech with customizable voice options. It supports multiple languages and voice styles, making it accessible for users who need audio output from text. The tool is entirely web-based and requires no installation.
TextaVoice offers a free text-to-speech service using advanced AI to generate natural-sounding human-like voices. It supports multiple languages and voice styles for various applications like podcasts and videos. The tool is designed for creators and educators looking for accessible audio content generation.
SayTxT by PDFgear is a free text-to-speech tool that converts PDFs, web pages, URLs, and even paper books into natural-sounding audio. It supports multiple languages and voice options, making it accessible for a wide range of users. The tool is designed for students, professionals, and anyone who prefers listening to reading.
Whisper Web is a free, browser-based transcription tool powered by OpenAI's Whisper model. It converts audio and video files into accurate text transcriptions with optional speaker diarization to identify different speakers. No installation or account is required — everything runs directly in the browser.
NVIDIA Broadcast leverages RTX GPU-powered AI to provide noise removal, virtual background, and auto-framing for streaming and video calls. It transforms your room into a professional studio.