Explore
Discover the best AI tools in one place
21 tools found
SpeechBrain is an open-source conversational AI toolkit designed for everyone, from researchers to developers. It provides a comprehensive suite of tools for building speech and audio processing applications. The platform supports a wide range of tasks including speech recognition and speaker identification.
Harmonai is an open-source AI platform dedicated to music generation and creative audio production. It enables musicians and creators to generate original compositions using advanced machine learning models. The platform supports various genres and styles, making AI-assisted music creation accessible to all skill levels.
DADABOTS uses neural networks to autonomously generate death metal music around the clock. The AI continuously produces new compositions without human intervention, creating an endless stream of algorithmic metal music.
Edge Dance uses AI to analyze music tracks and automatically generate synchronized dance routines that match the rhythm and style of the input audio. It creates choreographies for various dance styles and difficulty levels, making dance creation accessible to everyone. The tool is perfect for dancers, choreographers, and fitness enthusiasts looking for new routines.
Intervo AI is an open-source platform for building and deploying conversational AI agents. It provides tools for creating voice-enabled AI assistants that can handle complex interactions. Developers can customize and self-host their AI agent solutions.
Gnod is an AI-powered discovery engine that helps users find new music, art, and other cultural content based on their preferences. It uses intelligent algorithms to map connections between different pieces of media. The platform aims to provide a serendipitous browsing experience.
Voice Isolator is an AI-powered audio tool that removes background noise and isolates voice from audio recordings. It is designed for podcasters, content creators, and professionals who need clean audio output. The tool uses advanced machine learning algorithms to separate voice from ambient sounds effectively.
008 Agent is an AI-powered open-source softphone that transforms VoIP communications with intelligent call handling. It provides advanced call summarization and transcription capabilities for business communications. The platform is designed to make professional telephony accessible and intelligent.
MicroMusic is an AI-powered tool that generates Vital synthesizer presets from audio samples. Upload any sound and get unique, production-ready presets instantly. Perfect for music producers looking for inspiration.
Speechnotes is a fast and accurate speech-to-text tool that works directly in your browser. It supports multiple languages and offers a clean, distraction-free interface. Ideal for transcription, note-taking, and content creation from audio.
Create original music from text descriptions using AI. Generate royalty-free tracks for personal and commercial use. Experiment with different genres and styles instantly.
Voicebox is an open source voice cloning tool powered by Qwen3-TTS. It enables users to clone voices with high fidelity using advanced text-to-speech technology. Ideal for creators and developers who need realistic voice synthesis.
UVR Online is a web-based tool that uses AI to remove vocals from any song, allowing users to create instrumental and acapella tracks quickly. It leverages advanced source separation models to isolate vocals from music with high accuracy. The tool is designed for musicians, DJs, content creators, and karaoke enthusiasts.
HappySRT is an open-source AI tool that handles transcription, translation, and summarization of audio and video content. It makes it easy to convert spoken language into text and translate it across multiple languages. The tool is designed for creators, researchers, and teams who need fast and accurate media processing.
VoooAI creates multimedia workflows from simple text prompts. It supports various formats and integrates with OpenClaw for enhanced functionality.
Amical is an open-source AI application that transcribes speech into text, captures meeting notes, and helps with dictation. It supports multiple languages and integrates with various note-taking apps.
Riffusion is an AI music generation tool that creates music in real-time from text prompts. It uses advanced neural networks to transform text descriptions into unique audio tracks. Ideal for creators looking for instant music inspiration.
TextaVoice offers a free text-to-speech service using advanced AI to generate natural-sounding human-like voices. It supports multiple languages and voice styles for various applications like podcasts and videos. The tool is designed for creators and educators looking for accessible audio content generation.
Talat is a meeting notes application that uses on-device AI to transcribe and summarize meetings privately. It processes all audio locally on the user's device, ensuring complete data privacy and security. The tool is ideal for professionals who handle sensitive information and require confidential meeting documentation.
Paraspeech is a speech-to-text dictation tool that prioritizes user privacy by offering on-device transcription with optional cloud processing. It provides accurate voice-to-text conversion for professionals and individuals who need reliable dictation without compromising data security.
Whisper Web is a free, browser-based transcription tool powered by OpenAI's Whisper model. It converts audio and video files into accurate text transcriptions with optional speaker diarization to identify different speakers. No installation or account is required — everything runs directly in the browser.