AI tool intelligence

Find the right AI products for the work.

Our live catalog contains 2,444 researched AI products. Each record keeps current product positioning, fit, limitations, pricing signals, and official destinations together for a clearer decision.

Search products, capabilities, and workflowsBrowse the index ↓
2,444AI products in the catalog
DatedAvailability and evidence checks
IndependentFit, limitations, and alternatives

Public AI tool catalog

Search this public selection by the work you need to improve.

Search the complete catalog by product, capability, workflow, category, or tag. Every result opens a profile built from the same live record.

EL
Catalog evaluation

Voice generators

ElevenLabs

ElevenLabs provides web tools and APIs for generating speech, cloning voices, transcribing audio, dubbing media, creating music and sound effects, and deploying conversational agents. Its creative platform serves content production, while ElevenAgents supports voice and chat automation across customer channels.

Text to speechVoice cloningDubbingSpeech to textVoice agents
Review
AV
Catalog evaluation

Audio tools

A.V. Mapping

A.V. Mapping analyzes uploaded videos, photos, or YouTube links to recommend matching music and sound effects. The platform supports music discovery, licensing, project workflows, and musician distribution, with plans for creators and an API for integrating its audio-visual matching capabilities.

Music matchingVideo analysisMusic licensingSound effectsAPI
Review
AC
Catalog evaluation

Content creation

AI Content Labs

AI Content Labs is a no-code content-generation platform for connecting AI providers and building visual workflows. Users can combine text, image, audio, video, search, and scraping tools, then use templates or create customized flows to generate content at scale.

Multimodal contentAI integrationsVisual workflowsCustom templatesGoogle Sheets
Review
AM
Catalog evaluation

Developer tools

AI/ML API

AI/ML API provides developers with a unified API for chat, reasoning, image, video, audio, voice, search, and other AI models. Its OpenAI- and Anthropic-compatible interfaces, playground, and MCP server let teams test models and integrate them into applications under usage-based billing.

APIModel RoutingModel Context ProtocolOpenAI CompatibilityServerless Inference
Review
AI
Catalog evaluation

Music generation

Aimi

Aimi Sync analyzes a video to generate a synchronized soundtrack, vocals, and voice-overs. It uses licensed samples from artists, offers arrangement and scene controls, and supports downloadable audio stems for post-production. The service is aimed at creators producing copyright-safe video audio.

Video SoundtracksAudio SyncVoice-oversRoyalty-free MusicAPI
Review
AB
Catalog evaluation

APIs

API.box

API.box provides developer APIs for generating music from text and handling related audio workflows. Its API collection includes music and lyric generation, music extension, vocal separation, WAV conversion, timestamped lyrics, and music video generation.

Music generationLyrics generationAudio processingMusic videoWebhooks
Review
AS
Catalog evaluation

Developer tools

AssemblyAI

AssemblyAI provides APIs for pre-recorded, real-time, and synchronous speech-to-text, plus models that analyze transcript content and power voice agents. Developers can add transcription, speaker identification, summaries, sentiment, safety controls, and other audio intelligence to product workflows.

Speech-to-textReal-time transcriptionVoice agentsSpeaker diarizationCall analytics
Review
AU
Catalog evaluation

Note-taking

Audionotes

Audionotes is a voice-first note-taking app that records or imports voice notes, audio and video files, images, text, and YouTube links. It transcribes and summarizes them into structured outputs, lets users chat with notes, and runs on iOS, Android, web, and Mac beta.

Voice notesAudio transcriptionMeeting notesVideo transcriptionNotion
Review
AU
Catalog evaluation

Audio editing tools

AudioShake

AudioShake provides APIs and tools for separating music, dialogue, effects, and individual speakers from audio and video. It also offers lyric transcription and alignment, speech recovery, and music compliance workflows for media companies, rights holders, developers, and audio production teams.

Stem separationDialogue isolationLyric transcriptionAudio APICopyright compliance
Review
AU
Catalog evaluation

Audio production tools

AudioStack

AudioStack is an audio production platform for media and advertising teams. Its console and API turn briefs or source content into ads, video audio, and long-form narration, combining script generation, voice, music, sound design, mixing, and mastering for high-volume or embedded workflows.

Audio advertisingText-to-speechAudio APIDynamic creativeAudio mastering
Review
AA
Catalog evaluation

Transcription tools

AudioTranscription.ai

AudioTranscription.ai converts uploaded audio and video files or audio URLs into downloadable transcripts. It supports more than 70 languages, timestamped output formats, dashboard-based transcript management, and a beta speaker-identification option. Customers buy transcription-hour credits instead of subscribing to a recurring plan.

Audio transcriptionVideo transcriptionSpeaker identificationMultilingual transcriptionREST API
Review
AS
Catalog evaluation

Audio editing tools

Audo Studio

Audo Studio is a browser-based audio-cleaning service that removes background noise and adjusts speech volume in uploaded recordings. It offers a free Starter plan for individual creators, paid usage options, and a developer API for batch or streaming speech noise removal.

Noise removalSpeech enhancementAudio masteringAudio APIPodcasting
Review
BA
Catalog evaluation

Music generators

Beatoven.ai

Beatoven.ai generates customizable background music and sound effects for videos, podcasts, games, and other projects. Its maestro model accepts prompts to create audio, while the platform provides MP3 or WAV downloads and a license for using downloaded tracks with audiovisual content.

Background musicSound effectsRoyalty-free musicText-to-musicMusic generation API
Review
BE
Catalog evaluation

Text to speech

BeyondWords

BeyondWords helps publishers turn written articles into audio using pre-made or cloned AI voices. Its platform provides publishing workflows, an embeddable player, distribution tools, analytics, and CMS or API integrations for editorial audio content.

Audio articlesVoice cloningAudio analyticsCMS integrationsAPI
Review
CA
Catalog evaluation

Audio generators

CassetteAI

CassetteAI provides generative audio models for music and sound effects through a metered API and live playground. Developers can prompt its models to create stereo WAV music tracks or short, loopable sound effects for games, creator applications, and real-time audio workflows.

Music GenerationSound EffectsAudio APIEdge InferenceJavaScript
Review
CA
Catalog evaluation

Podcast editing

Cleanvoice AI

Cleanvoice AI is an audio and video podcast editor that removes background noise, filler words, pauses, mouth sounds, breaths, and stutters. It also provides transcription, summaries, social content, multitrack editing, timeline exports, and an API for automating recorded-media post-production.

Background noise removalFiller word removalTranscriptionPodcast productionAudio API
Review
DE
Catalog evaluation

Speech recognition

Deepgram

Deepgram provides APIs for speech-to-text, text-to-speech, voice agents, and audio intelligence. Developers can use its real-time or batch services in the cloud, or explore self-hosted deployment options for voice-enabled applications.

Speech-to-textText-to-speechVoice agentsAudio intelligenceAPI
Review
EP
Catalog evaluation

Content creation

Easy-Peasy.AI

Easy-Peasy.AI is an all-in-one platform for generating text, images, videos, audio, and transcriptions. It also provides MARKy agents, custom bots trained on user data, visual workflows, templates, and an API for businesses and creators.

Text generationImage generationVideo generationCustom chatbotsTranscription
Review
FA
Catalog evaluation

Voice generators

FakeYou

FakeYou is a creator platform for text-to-speech, voice conversion, voice design, and lip-synced video. Users can select available voices or submit a custom voice-clone request, generate audio, and animate source media. Free features are available, while paid memberships add faster processing, longer limits, and private-model options.

Text to speechVoice conversionVoice cloningLip syncAPI
Review
GL
Catalog evaluation

Speech-to-text

Gladia

Gladia provides an API for real-time and asynchronous audio transcription, plus audio intelligence features that turn speech into structured data. Developers can process live or recorded audio and video with multilingual transcription, speaker diarization, timestamps, translation, and analysis capabilities.

Real-time transcriptionMultilingual transcriptionSpeaker diarizationAudio analysisVoice AI
Review
HI
Catalog evaluation

Content moderation

Hive

Hive provides cloud-based APIs and software for content moderation, AI-generated media detection, visual search, classification, and generation. Developers can integrate pre-trained models for text, image, video, and audio workflows, while enterprises can use its platform for brand protection, compliance, and platform integrity.

Content moderationDeepfake detectionMedia searchImage generationAPI
Review
KA
Catalog evaluation

Developer tools

KIE.ai

KIE.ai provides a unified API for video, image, audio, and language models from multiple providers. Developers can test models in a playground, use consistent request flows and webhooks, manage usage through credits, and access model-specific integration documentation.

Model APIsImage generationVideo generationAudio generationLarge language models
Review
KA
Catalog evaluation

Music generators

Kits AI

Kits AI is a music-production platform for creating, changing, cloning, and processing voices. Musicians and creators can work with royalty-free singing voices, build custom digital voices, generate harmonies, separate stems, master audio, and use instrument models or text-to-speech.

Voice cloningSinging voicesVocal processingMusic masteringAudio API
Review
LA
Catalog evaluation

Audio editing

LALAL.AI

LALAL.AI is an audio-processing suite for splitting vocals and instruments, cleaning noisy recordings, reducing echo and reverb, and creating or transforming voices. It is available on the web and mobile devices, with desktop, VST plugin, and API options for eligible plans.

Stem separationVocal removalNoise reductionVoice cloningAudio API
Review

How the index works

A catalog built for decisions, not rankings.

Commercial relationships can fund the research. They never change product status, fit, limitations, evidence, or the alternatives we show.

Verify what exists

We check availability, product direction, pricing signals, and when the supporting evidence was last reviewed.

Evaluate operational fit

We assess the work it supports, integration requirements, ownership, risk, and where human judgment remains necessary.

Keep the comparison honest

Commercial relationships are disclosed, while credible limitations and alternatives stay visible.

Evaluating a product for a live workflow?

Bring the decision to our engineers