AI tool intelligence

Find the right AI products for the work.

Our live catalog contains 2,444 researched AI products. Each record keeps current product positioning, fit, limitations, pricing signals, and official destinations together for a clearer decision.

Search products, capabilities, and workflowsBrowse the index ↓
2,444AI products in the catalog
DatedAvailability and evidence checks
IndependentFit, limitations, and alternatives

Public AI tool catalog

Search this public selection by the work you need to improve.

Search the complete catalog by product, capability, workflow, category, or tag. Every result opens a profile built from the same live record.

AS
Catalog evaluation

Developer tools

AssemblyAI

AssemblyAI provides APIs for pre-recorded, real-time, and synchronous speech-to-text, plus models that analyze transcript content and power voice agents. Developers can add transcription, speaker identification, summaries, sentiment, safety controls, and other audio intelligence to product workflows.

Speech-to-textReal-time transcriptionVoice agentsSpeaker diarizationCall analytics
Review
AU
Catalog evaluation

Note-taking

Audionotes

Audionotes is a voice-first note-taking app that records or imports voice notes, audio and video files, images, text, and YouTube links. It transcribes and summarizes them into structured outputs, lets users chat with notes, and runs on iOS, Android, web, and Mac beta.

Voice notesAudio transcriptionMeeting notesVideo transcriptionNotion
Review
AU
Catalog evaluation

Audio editing tools

AudioShake

AudioShake provides APIs and tools for separating music, dialogue, effects, and individual speakers from audio and video. It also offers lyric transcription and alignment, speech recovery, and music compliance workflows for media companies, rights holders, developers, and audio production teams.

Stem separationDialogue isolationLyric transcriptionAudio APICopyright compliance
Review
AA
Catalog evaluation

Transcription tools

AudioTranscription.ai

AudioTranscription.ai converts uploaded audio and video files or audio URLs into downloadable transcripts. It supports more than 70 languages, timestamped output formats, dashboard-based transcript management, and a beta speaker-identification option. Customers buy transcription-hour credits instead of subscribing to a recurring plan.

Audio transcriptionVideo transcriptionSpeaker identificationMultilingual transcriptionREST API
Review
CA
Catalog evaluation

Podcast editing

Cleanvoice AI

Cleanvoice AI is an audio and video podcast editor that removes background noise, filler words, pauses, mouth sounds, breaths, and stutters. It also provides transcription, summaries, social content, multitrack editing, timeline exports, and an API for automating recorded-media post-production.

Background noise removalFiller word removalTranscriptionPodcast productionAudio API
Review
DE
Catalog evaluation

Speech recognition

Deepgram

Deepgram provides APIs for speech-to-text, text-to-speech, voice agents, and audio intelligence. Developers can use its real-time or batch services in the cloud, or explore self-hosted deployment options for voice-enabled applications.

Speech-to-textText-to-speechVoice agentsAudio intelligenceAPI
Review
EP
Catalog evaluation

Content creation

Easy-Peasy.AI

Easy-Peasy.AI is an all-in-one platform for generating text, images, videos, audio, and transcriptions. It also provides MARKy agents, custom bots trained on user data, visual workflows, templates, and an API for businesses and creators.

Text generationImage generationVideo generationCustom chatbotsTranscription
Review
GL
Catalog evaluation

Speech-to-text

Gladia

Gladia provides an API for real-time and asynchronous audio transcription, plus audio intelligence features that turn speech into structured data. Developers can process live or recorded audio and video with multilingual transcription, speaker diarization, timestamps, translation, and analysis capabilities.

Real-time transcriptionMultilingual transcriptionSpeaker diarizationAudio analysisVoice AI
Review
IA
Catalog evaluation

Developer tools

Inworld AI

Inworld AI is a developer platform for live voice and conversational AI applications. Its APIs provide text-to-speech, speech-to-text, speech-to-speech, and LLM routing, with custom voices, voice design, tool calling, and a multi-provider interface for real-time conversational experiences.

Real-time voice AIText to speechSpeech to textLLM routingAPIs
Review
RA
Catalog evaluation

AI tools

Rewind.ai

Rewind.ai bundles more than 400 AI tools and models for chat, images, video, writing, code, speech, music, and transcription. Users can work without an account within daily limits or sign up for more tokens and history, while developers can access the tools through one OpenAI-compatible API.

AI chatImage generationVideo generationSpeech generationOpenAI-compatible API
Review
RY
Catalog evaluation

Transcription tools

Rythmex

Rythmex converts uploaded audio and video into editable text transcripts. It supports more than 140 languages, transcription exports, an API, and an enterprise call analytics product for teams that need summaries, categorization, and performance insights from recorded calls.

Audio transcriptionVideo transcriptionSpeech analyticsTranscription APITeam management
Review
SA
Catalog evaluation

Cloud computing

Salad

Salad provides a distributed cloud platform for deploying production AI and machine learning workloads on consumer and data-center GPUs. Customers can run containerized inference, transcription, computer vision, rendering, and batch processing workloads without managing individual virtual machines.

Distributed GPU cloudAI inferenceContainer deploymentTranscription APIComputer vision
Review
SO
Catalog evaluation

Transcription

Sonix

Sonix is a speech-to-text platform for turning audio and video into searchable transcripts. It supports transcription and translation in more than 54 languages, provides an in-browser editor with speaker labels and timestamps, and exports transcripts and subtitles for sharing or production workflows.

Speech-to-textSubtitlesAudio translationSpeaker diarizationAPI
Review
SA
Catalog evaluation

Transcription

Speak AI

Speak AI captures meetings and processes uploaded audio, video, and text into searchable transcripts, summaries, and structured insights. Teams can analyze conversation libraries, connect the platform through APIs and MCP, or commission a branded voice AI application.

Audio transcriptionVideo transcriptionMeeting notesSentiment analysisAPI
Review
SP
Catalog evaluation

Voice generators

SpeechGen

SpeechGen generates downloadable speech from text with more than 5,000 voices in 150 languages. The browser-based tool supports voiceovers, multiple speakers, audio production controls, document conversion, transcription, and API workflows without a monthly subscription.

Text to speechVoiceoversMultilingualAPIAudio production
Review
SA
Catalog evaluation

Developer tools

Symbl.ai

Symbl.ai provides models and APIs for turning voice, video, chat, and text conversations into knowledge, events, and insights. Its developer platform supports real-time AI agents, multimodal experiences, conversation analytics, and prebuilt tools for product, revenue, and data teams.

Real-time AI agentsConversation analyticsTranscriptionSentiment analysisCall intelligence
Review
TA
Catalog evaluation

Note-taking

Talknotes

Talknotes records or accepts audio, then transcribes and formats it into notes, transcripts, task lists, and other text formats. It supports more than 50 languages, custom styles, file uploads, exports, Zapier, webhooks, and iOS and Android apps.

Voice notesAudio transcriptionContent creationZapierMobile apps
Review
TA
Catalog evaluation

Research assistants

Tapesearch

Tapesearch is a podcast-transcript search platform for locating discussions, mentions, and audio timestamps across podcasts. It supports full-text, semantic, and hybrid search, episode chat, keyword alerts, trend analysis, transcript downloads, and API access for research and monitoring workflows.

Podcast transcriptsSemantic searchBrand monitoringAI chatAPI
Review
TE
Catalog evaluation

Image generators

Textunbox

Textunbox provides browser-based and REST API tools for extracting text from images, generating images from text or voice, transcribing audio, translating text, and removing image backgrounds. Browser tools require an activation key, while the standardized API supports custom implementations.

OCRImage generationBackground removalAudio transcriptionREST API
Review
WA
Catalog evaluation

Meeting assistants

Wave

Wave is a cross-platform AI note-taking and recording app for meetings, phone calls, lectures, and other conversations. It records audio, creates transcripts and summaries, identifies speakers, supports 76 languages, and lets users search, share, import, or access recording data through its API.

Meeting recordingSpeech to textAI summariesZoomGoogle Meet
Review
WA
Catalog evaluation

Video editors

WayinVideo

WayinVideo is an AI video platform that turns long videos into short clips, locates moments through transcript-aware search, and generates summaries, transcripts, captions, and reframed outputs. It supports uploads and video URLs, social-ready exports, and an API for automated video-processing workflows.

AI video clippingVideo searchVideo summariesTranscriptionCaptioning
Review
WM
Catalog evaluation

Transcription

Whisper Memos

Whisper Memos is an iPhone and Apple Watch voice recorder that transcribes recordings and sends formatted emails with summaries. It supports custom summary prompts, audio-file imports, reminders, and automated delivery to notes, task, and productivity apps.

Voice MemosApple WatchAI SummariesZapierNotion
Review
PL
Discontinued record

Audio generation

PlayHT

PlayHT is an AI voice platform for generating text-to-speech audio, cloning voices, and dubbing content. It also provides voice changing, speech transcription, music and sound-effect generation, with a web studio and API for producing and integrating audio.

Text to speechVoice cloningAuto dubbingSpeech-to-textAPI
Review
VA
Discontinued record

Text to speech

Voiser AI

Voiser AI provides text-to-speech voiceovers, audio and video transcription, and AI video generation. Creators, developers, and businesses can use its web platform to produce multilingual audio and video content, with paid plans and an API option for eligible subscriptions.

AI voiceoverSpeech to textVideo generationMultilingualAPI
Review

How the index works

A catalog built for decisions, not rankings.

Commercial relationships can fund the research. They never change product status, fit, limitations, evidence, or the alternatives we show.

Verify what exists

We check availability, product direction, pricing signals, and when the supporting evidence was last reviewed.

Evaluate operational fit

We assess the work it supports, integration requirements, ownership, risk, and where human judgment remains necessary.

Keep the comparison honest

Commercial relationships are disclosed, while credible limitations and alternatives stay visible.

Evaluating a product for a live workflow?

Bring the decision to our engineers