AI tool intelligence

Find the right AI products for the work.

Our live catalog contains 2,444 researched AI products. Each record keeps current product positioning, fit, limitations, pricing signals, and official destinations together for a clearer decision.

Search products, capabilities, and workflowsBrowse the index ↓
2,444AI products in the catalog
DatedAvailability and evidence checks
IndependentFit, limitations, and alternatives

Public AI tool catalog

Search this public selection by the work you need to improve.

Search the complete catalog by product, capability, workflow, category, or tag. Every result opens a profile built from the same live record.

EL
Catalog evaluation

Voice generators

ElevenLabs

ElevenLabs provides web tools and APIs for generating speech, cloning voices, transcribing audio, dubbing media, creating music and sound effects, and deploying conversational agents. Its creative platform serves content production, while ElevenAgents supports voice and chat automation across customer channels.

Text to speechVoice cloningDubbingSpeech to textVoice agents
Review
AS
Catalog evaluation

Developer tools

AssemblyAI

AssemblyAI provides APIs for pre-recorded, real-time, and synchronous speech-to-text, plus models that analyze transcript content and power voice agents. Developers can add transcription, speaker identification, summaries, sentiment, safety controls, and other audio intelligence to product workflows.

Speech-to-textReal-time transcriptionVoice agentsSpeaker diarizationCall analytics
Review
AU
Catalog evaluation

Audio production tools

AudioStack

AudioStack is an audio production platform for media and advertising teams. Its console and API turn briefs or source content into ads, video audio, and long-form narration, combining script generation, voice, music, sound design, mixing, and mastering for high-volume or embedded workflows.

Audio advertisingText-to-speechAudio APIDynamic creativeAudio mastering
Review
BE
Catalog evaluation

Text to speech

BeyondWords

BeyondWords helps publishers turn written articles into audio using pre-made or cloned AI voices. Its platform provides publishing workflows, an embeddable player, distribution tools, analytics, and CMS or API integrations for editorial audio content.

Audio articlesVoice cloningAudio analyticsCMS integrationsAPI
Review
CO
Catalog evaluation

AI infrastructure

Cohere

Cohere provides enterprise AI products including generative models, retrieval models, speech-to-text, and the North AI workplace. Organizations can deploy models in a virtual private cloud, on premises, or through Cohere's managed Model Vault, with options for customization using proprietary data.

LLM APIPrivate deploymentAI agentsSemantic searchSpeech-to-text
Review
DE
Catalog evaluation

Speech recognition

Deepgram

Deepgram provides APIs for speech-to-text, text-to-speech, voice agents, and audio intelligence. Developers can use its real-time or batch services in the cloud, or explore self-hosted deployment options for voice-enabled applications.

Speech-to-textText-to-speechVoice agentsAudio intelligenceAPI
Review
DU
Catalog evaluation

Video generators

Dubverse

Dubverse is an AI video localization platform for dubbing videos, creating subtitles, and generating voiceovers. Creators and teams can translate video into multiple languages, select AI speakers, edit output in a web studio, and use a text-to-speech API for voice generation in applications and workflows.

AI DubbingVideo localizationSubtitlesVoice cloningText to speech API
Review
FA
Catalog evaluation

Voice generators

FakeYou

FakeYou is a creator platform for text-to-speech, voice conversion, voice design, and lip-synced video. Users can select available voices or submit a custom voice-clone request, generate audio, and animate source media. Free features are available, while paid memberships add faster processing, longer limits, and private-model options.

Text to speechVoice conversionVoice cloningLip syncAPI
Review
GA
Catalog evaluation

Image generators

getimg.ai

getimg.ai is a creative platform for generating images, videos, music, speech, and sound effects, then editing visual assets. It provides multiple models in one workspace, plus collaboration features, image and video upscaling, smart resizing, background removal, and an image and video API.

Text to ImageImage to VideoMusic GenerationBackground RemovalTeam Collaboration
Review
GL
Catalog evaluation

Speech-to-text

Gladia

Gladia provides an API for real-time and asynchronous audio transcription, plus audio intelligence features that turn speech into structured data. Developers can process live or recorded audio and video with multilingual transcription, speaker diarization, timestamps, translation, and analysis capabilities.

Real-time transcriptionMultilingual transcriptionSpeaker diarizationAudio analysisVoice AI
Review
HA
Catalog evaluation

Voice AI

Hume AI

Hume AI provides data, speech models, evaluations, and feedback systems for teams building voice AI. Its platform supports real-time speech interaction, expressive text-to-speech, voice creation, human feedback, and tools for measuring how voice systems perform in real-world conversations.

Text-to-SpeechSpeech-to-SpeechVoice evaluationVoice cloningHuman feedback
Review
IA
Catalog evaluation

Developer tools

Inworld AI

Inworld AI is a developer platform for live voice and conversational AI applications. Its APIs provide text-to-speech, speech-to-text, speech-to-speech, and LLM routing, with custom voices, voice design, tool calling, and a multi-provider interface for real-time conversational experiences.

Real-time voice AIText to speechSpeech to textLLM routingAPIs
Review
KA
Catalog evaluation

Music generators

Kits AI

Kits AI is a music-production platform for creating, changing, cloning, and processing voices. Musicians and creators can work with royalty-free singing voices, build custom digital voices, generate harmonies, separate stems, master audio, and use instrument models or text-to-speech.

Voice cloningSinging voicesVocal processingMusic masteringAudio API
Review
LI
Catalog evaluation

Content creation

ListenHub

ListenHub is a content creation platform that turns prompts, documents, and links into explainer videos, narrated slides, podcasts, speech, and images. It also offers voice cloning and an API for generating media programmatically, making it useful for repurposing information into audience-ready formats.

Explainer videosAI podcastsText to speechVoice cloningSlide decks
Review
LA
Catalog evaluation

Text to speech

Listnr AI

Listnr AI creates speech from text using a library of more than 1,000 voices across 142 languages. Creators and teams can generate voiceovers, clone voices, create multi-speaker audio, host podcasts, make text-to-video content, and integrate voice generation through an API.

Text to speechVoice cloningVoiceoversPodcast hostingText to video
Review
LO
Catalog evaluation

Voice generators

LOVO

LOVO's Genny platform combines text-to-speech voices, video editing, auto subtitles, voice cloning, AI image generation, and script writing. It supports marketing, training, explainer, social, audiobook, and podcast production, while an API lets developers use LOVO voices in their own applications.

Text to speechVoice cloningVideo creationAuto subtitlesDeveloper API
Review
MA
Catalog evaluation

Text to speech

Murf AI

Murf AI is a voice platform for creating studio-quality voiceovers, localizing video content, and building voice-enabled applications. Its Studio converts text into customizable speech, while Murf Dubbing and the API support multilingual content and production voice-agent workflows.

AI VoiceoversVoice CloningAI DubbingText to Speech APIVoice Agents
Review
NA
Catalog evaluation

Text to speech

Narakeet

Narakeet turns scripts, speaker notes, presentations, and subtitles into narrated audio and video. It supports text-to-speech voices in 100 languages, lets users edit projects as text, and offers automation APIs for commercial accounts.

Voice-overVideo automationPowerPointText to speechDeveloper API
Review
RE
Catalog evaluation

Audio generation

Respeecher

Respeecher provides text-to-speech, speech-to-speech, and voice-cloning tools for media, entertainment, and voice-based products. Its Voice Marketplace offers AI voices and a Pro Tools plugin, while its APIs support programmatic audio generation and real-time text-to-speech integrations.

Text-to-speechSpeech-to-speechVoice APIPro ToolsVoice marketplace
Review
RY
Catalog evaluation

Transcription tools

Rythmex

Rythmex converts uploaded audio and video into editable text transcripts. It supports more than 140 languages, transcription exports, an API, and an enterprise call analytics product for teams that need summaries, categorization, and performance insights from recorded calls.

Audio transcriptionVideo transcriptionSpeech analyticsTranscription APITeam management
Review
SO
Catalog evaluation

Transcription

Sonix

Sonix is a speech-to-text platform for turning audio and video into searchable transcripts. It supports transcription and translation in more than 54 languages, provides an in-browser editor with speaker labels and timestamps, and exports transcripts and subtitles for sharing or production workflows.

Speech-to-textSubtitlesAudio translationSpeaker diarizationAPI
Review
SP
Catalog evaluation

Voice generators

SpeechGen

SpeechGen generates downloadable speech from text with more than 5,000 voices in 150 languages. The browser-based tool supports voiceovers, multiple speakers, audio production controls, document conversion, transcription, and API workflows without a monthly subscription.

Text to speechVoiceoversMultilingualAPIAudio production
Review
SP
Catalog evaluation

Speech to text

Speechllect

Speechllect is a beta platform for real-time speech-to-text and text-to-speech workflows. It analyzes spoken words alongside emotion and tone, generates expressive synthetic speech, and combines both capabilities for automated business scenarios such as contact-center interactions.

Speech recognitionSpeech synthesisSentiment analysisCall centersAPI
Review
SU
Catalog evaluation

Audio generators

Supertone

Supertone provides voice AI products for creators and developers, including text-to-speech, voice cloning, real-time voice conversion, voice separation, and audio cleanup. Its Play web app creates spoken audio from text, while its API lets teams add speech synthesis to their own services.

Text to speechVoice cloningVoice conversionAudio cleanupAPI integration
Review

How the index works

A catalog built for decisions, not rankings.

Commercial relationships can fund the research. They never change product status, fit, limitations, evidence, or the alternatives we show.

Verify what exists

We check availability, product direction, pricing signals, and when the supporting evidence was last reviewed.

Evaluate operational fit

We assess the work it supports, integration requirements, ownership, risk, and where human judgment remains necessary.

Keep the comparison honest

Commercial relationships are disclosed, while credible limitations and alternatives stay visible.

Evaluating a product for a live workflow?

Bring the decision to our engineers