Find AI tools and agents that support audio as input. Compare capabilities, output types, and use cases.
Conversational AI assistant by OpenAI for text, image, voice and code tasks.
Multimodal AI assistant by Google DeepMind integrated across Google products.
Alibaba's multimodal AI assistant based on the Qwen model family.
AI-powered audio and video editing with transcript-based workflows.
AI voice generator with realistic text-to-speech and voice cloning.
AI music generator that creates full songs from text prompts.
AI music generator founded by former Google DeepMind researchers.
Speech-to-text API platform with speaker diarization and audio intelligence.
Open-source speech recognition model by OpenAI for transcription.
Revenue intelligence platform that analyzes sales calls with AI.
Video interview and hiring intelligence platform with AI assessments.
AI meeting assistant that transcribes, summarizes and chats with notes.
AI notetaker that records, transcribes and summarizes meetings.
Free AI notetaker for Zoom, Meet and Teams with instant summaries.
AI meeting recorder for Google Meet, Zoom and Teams with insights.
Personalized AI assistant that records everything you've seen and heard.
AI video generator for stylized, music-reactive animations.
AI video creation from long-form text and videos for marketing.
AI that turns long videos into viral short clips with captions.
AI video repurposing platform for short-form social clips.
Online video editor with AI features for creators and teams.
AI voice generator with ultra-realistic TTS and voice cloning.
Voice cloning and AI speech platform for enterprise.
AI music composition for soundtracks and original scores.
AI music creation platform that publishes to streaming services.
AI music mastering, distribution, and samples platform.
AI podcast recording and editing studio in the browser.
Studio-quality podcast and video recording with AI editing.
AI audio tool that removes filler words and mouth sounds.
Speech-to-text API with real-time streaming and high accuracy.