Discover AI tools that convert audio input into text output. Compare features, quality, and pricing.
Conversational AI assistant by OpenAI for text, image, voice and code tasks.
Multimodal AI assistant by Google DeepMind integrated across Google products.
Alibaba's multimodal AI assistant based on the Qwen model family.
AI-powered audio and video editing with transcript-based workflows.
Speech-to-text API platform with speaker diarization and audio intelligence.
Open-source speech recognition model by OpenAI for transcription.
Revenue intelligence platform that analyzes sales calls with AI.
Video interview and hiring intelligence platform with AI assessments.
AI meeting assistant that transcribes, summarizes and chats with notes.
AI notetaker that records, transcribes and summarizes meetings.
Free AI notetaker for Zoom, Meet and Teams with instant summaries.
AI meeting recorder for Google Meet, Zoom and Teams with insights.
Personalized AI assistant that records everything you've seen and heard.
AI video creation from long-form text and videos for marketing.
AI that turns long videos into viral short clips with captions.
AI video repurposing platform for short-form social clips.
Online video editor with AI features for creators and teams.
AI podcast recording and editing studio in the browser.
Studio-quality podcast and video recording with AI editing.
Speech-to-text API with real-time streaming and high accuracy.
AI audio stem separation for music and film.
AI meeting copilot that analyzes engagement and sentiment.
AI meeting assistant and conversation intelligence for sales.
AI meeting assistant that records, transcribes, and analyzes meetings.
AI meeting notetaker without bots that runs locally.
AI voice assistant for clinicians to document patient visits.
Microsoft's AI clinical documentation assistant for ambient recording.
AI medical scribe that generates clinical notes from consultations.
AI for emergency call analysis and medical triage.
AI interview copilot that provides real-time interview assistance.