AI tool intelligence

Find the right AI products for the work.

Our live catalog contains 2,444 researched AI products. Each record keeps current product positioning, fit, limitations, pricing signals, and official destinations together for a clearer decision.

Search products, capabilities, and workflowsBrowse the index ↓
2,444AI products in the catalog
DatedAvailability and evidence checks
IndependentFit, limitations, and alternatives

Public AI tool catalog

Search this public selection by the work you need to improve.

Search the complete catalog by product, capability, workflow, category, or tag. Every result opens a profile built from the same live record.

EA
Catalog evaluation

Developer tools

Evidently AI

Evidently AI is an open-source framework for evaluating, testing, and monitoring LLM applications, RAG systems, AI agents, and predictive machine-learning models. Teams can use built-in or custom metrics, generate test cases, run evaluations locally or in the cloud, and track quality over time.

LLM evaluationML monitoringData driftRAG evaluationOpen source
Review
FA
Catalog evaluation

AI observability

Fiddler AI

Fiddler AI is an enterprise AI control plane for monitoring, evaluating, securing, and governing agents, LLM applications, and traditional ML models. It gives teams execution context and decision lineage, applies guardrails, and supports continuous evaluation throughout the AI lifecycle.

Agentic observabilityLLM monitoringGuardrailsAI governanceModel evaluation
Review
GA
Catalog evaluation

AI development platforms

Gooey.AI

Gooey.AI is a low-code orchestration platform for creating, testing, and deploying collaborative AI workflows. Teams can combine prompts, knowledge files, code, and evaluations, run models from multiple providers, and deploy agents to web, WhatsApp, voice, SMS, Slack, and Facebook.

Low-code workflowsAI agentsMultilingual supportWhatsAppAPI
Review
HU
Catalog evaluation

AI development tools

Humanloop

Humanloop was acquired by Anthropic, and its platform was sunset on September 8, 2025. Before the sunset, it provided enterprise tooling for LLM evaluation, prompt management, and production observability, helping technical and non-technical teams collaborate on AI product development.

Prompt managementLLM evaluationsObservabilityModel monitoringDeveloper SDKs
Review
HA
Catalog evaluation

Voice AI

Hume AI

Hume AI provides data, speech models, evaluations, and feedback systems for teams building voice AI. Its platform supports real-time speech interaction, expressive text-to-speech, voice creation, human feedback, and tools for measuring how voice systems perform in real-world conversations.

Text-to-SpeechSpeech-to-SpeechVoice evaluationVoice cloningHuman feedback
Review
LA
Catalog evaluation

Data labeling

Labelbox

Labelbox is a platform and labeling service for creating high-quality training and evaluation data for AI systems. Teams can manage multimodal datasets, configure annotation projects, use expert labelers, and run workflows for RLHF, model evaluation, fine-tuning, and agent tasks.

Training dataRLHFModel evaluationMultimodal dataHuman feedback
Review
MA
Catalog evaluation

Developer tools

Martian

Martian is an AI research lab that makes selected tools available to developers. Its Gateway provides a single API for more than 200 models with OpenAI and Anthropic API compatibility, while its ARES and K-Steering projects support agent evaluation and inference-time behavior control.

LLM routingAPIMulti-model accessOpenAI compatibilityAgent evaluation
Review
OS
Catalog evaluation

Data analysis

Oda Studio

Oda Studio builds custom AI pipelines that turn complex, unstructured data into actionable insights. Its current services focus on construction and engineering drawings, financial and consulting documents, and customized generative-media models, combining data preparation, compute environments, evaluations, and deployment work.

Vision-language modelsKnowledge graphsDocument processingData annotationRAG
Review
SA
Catalog evaluation

Data labeling

Scale AI

Scale AI provides training data, model evaluations, and full-stack AI systems for enterprises, governments, and model developers. Its products support data annotation and management, generative AI application development, red-teaming, benchmarking, and the deployment of AI systems with human review.

Data annotationHuman-in-the-loopRLHFModel evaluationGenerative AI
Review
SU
Catalog evaluation

Data labeling

SuperAnnotate

SuperAnnotate is a data platform for creating, curating, annotating, and evaluating multimodal training data. It combines customizable annotation workflows, quality review, managed expert talent, and integrations to support AI model training, fine-tuning, reinforcement learning, RAG evaluation, and agent development.

Multimodal dataRLHFFine-tuningRAG evaluationHuman-in-the-loop
Review
VO
Catalog evaluation

Chatbot builders

Voiceflow

Voiceflow is a collaborative platform for building, testing, deploying, and monitoring chat and voice agents. Teams can combine knowledge bases, workflows, integrations, and model providers, then launch agents on web chat, phone, mobile, or custom interfaces.

Conversational AIKnowledge basesOmnichannel deploymentAPI integrationsAgent analytics
Review
AN
Discontinued record

Machine learning

Anote

Anote is a human-centered AI platform for creating labeled datasets, improving domain-specific models, and evaluating AI systems. It combines annotation workflows and human feedback with fine-tuning, benchmarks, synthetic data, and developer tools for teams building reliable AI applications.

Data annotationLLM evaluationFine-tuningHuman feedbackSynthetic data
Review
QV
Discontinued record

Prompt engineering

Query Vary

Query Vary's official documentation describes tools for designing, testing, and refining LLM prompts. Its main website could not be reached and its signup app returned an error during verification, so current product availability and commercial terms could not be confirmed.

LLM TestingPrompt EngineeringModel Comparison
Review
TE
Discontinued record

Developer tools

Terracotta

Terracotta's official terra-cotta.ai website could not be reached during verification. Its current product availability, capabilities, pricing, documentation, and company details could therefore not be confirmed from current official sources.

LLM fine-tuningModel evaluationDataset upload
Review

How the index works

A catalog built for decisions, not rankings.

Commercial relationships can fund the research. They never change product status, fit, limitations, evidence, or the alternatives we show.

Verify what exists

We check availability, product direction, pricing signals, and when the supporting evidence was last reviewed.

Evaluate operational fit

We assess the work it supports, integration requirements, ownership, risk, and where human judgment remains necessary.

Keep the comparison honest

Commercial relationships are disclosed, while credible limitations and alternatives stay visible.

Evaluating a product for a live workflow?

Bring the decision to our engineers