RXed AI News

AI to the bone.

AI Audits 128 tools audited

Every tool scored against the RXed AI Periodic Table v1.0 — 20 elements across 5 families, shown on two honest axes: Scope (how much of the AI stack the tool plays in) × Quality (how well it scores on what it covers). No single flattering number — a brilliant narrow tool isn't punished, a shallow platform can't hide. Independent, no affiliate links, every audit fully sourced. How we audit →

Audited 2026-08-19 · table v1.0
“The agent workspace for intelligent outbound. More pipeline. Less busywork.” — the vendor’s own words

The protocol work is the real story — a 37-tool MCP server and an MCP client, both directions, permission-scoped — wrapped around a €4,600-a-seat dialer whose core trick still puts two seconds of dead air in front of every prospect who answers.

Scope17/20
Quality7/10
Enterprise platformSalesAutomation & AgentsVoice & SpeechProductivityPaid
Vendor
Nooks Communications, Inc.
Origin
US — San Francisco
Pricing
Quote only ~$4,000-5,000/user/yr
Users (official only)
Not disclosed by Nooks. The vendor's own site cites a G2 rating of 4.8 across 1,500+ reviews and names customers including Rippling, ZoomInfo, Deel, Notion, ClickUp, Intercom, Pendo, Akamai, CyberArk, Cursor and Airtable. An independent directory put the base at 200+ companies in December 2025; that figure is not a company statement.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em4
Embeddings
Cx9
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg9
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft4
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Phone-first B2B outbound teams of five or more SDRs with a large enough addressable market to survive parallel dialing, who want dialing, sequencing and coaching consolidated in one workspace — and especially teams that want their sales data reachable from Claude, ChatGPT or their own agents, where the MCP server is a real differentiator. Wrong fit for small or budget-constrained teams, for anyone whose phone is a side channel, and for organisations with a limited TAM that list burnout would exhaust.
Full audit passport →Visit nooks.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-19 · table v1.0
“Record in-person salespeople and coach them 100x faster with Rilla.” — the vendor’s own words

Nobody records the kitchen-table conversation better, and the grounding is genuinely excellent — but at roughly €3,700-4,600 per seat with a five-seat floor, no public API and no MCP, you rent the best dataset in field sales and can never take it home.

Scope15/20
Quality5/10
SpecialistSalesVoice & SpeechLearning & TrainingFreemium
Vendor
Rilla (Rillavoice)
Origin
US — New York
Pricing
Quote only ~$4,000-5,000/seat/yr
Users (official only)
Not disclosed by Rilla. Independent analyst Sacra estimates the customer base crossed 2,000 accounts in mid-2025 and roughly $70M ARR by April 2026 — estimates, not company statements. Rilla's own published figure is dataset scale: outcome metadata across more than 1,300 companies.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx6
Context
Tr6
Tracing
Lg6
LLM
Compositions
Fc3
Function calling
Vx
Vector store
Rg8
RAG
Gr4
Guardrails
Mm7
Multimodal
Deployment
Ag4
Agents
Ft3
Fine-tuning
Fw2
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc1
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Field-sales and field-service organisations with five or more in-person reps and average tickets big enough that one extra close per rep per quarter pays the seat — home services, remodelling, home building, multifamily, senior living, dental and auto repair. Wrong tool for teams under five reps, for anyone who needs the conversation data in their own systems, and for phone or inside-sales coverage, which it does not do.
Full audit passport →Visit rilla.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-18 · table v1.0
“The AI platform for healthcare & life-science developers” — the vendor’s own words

The most complete healthcare AI stack you can buy as an API, and the only clinical vendor that speaks both MCP and A2A properly. Check the eval and fine-tuning gaps before you build your product on it.

Scope19/20
Quality7/10
Enterprise platformHealthcareAutomation & AgentsVoice & SpeechSecurity & ComplianceFreemium
Vendor
Corti ApS
Origin
Denmark — Copenhagen
Pricing
Free credits $50 · Speech to Text $0.0065/min · Medical Coding / Agents / Text Generation $4 in / $16 out per mTok +4 more
Users (official only)
Over 100 million patients served annually across health systems including the NHS; 250,000 patient interactions per day
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx8
Context
Tr9
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx3
Vector store
Rg8
RAG
Gr9
Guardrails
Mm7
Multimodal
Deployment
Ag9
Agents
Ft4
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm8
Small models
Emerging
Ma9
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In6
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Healthcare software vendors and EHR platforms that want to ship clinical AI features without building the model, compliance and audit infrastructure themselves, and European buyers who need inference to stay inside the EU.
Full audit passport →Visit corti.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-18 · table v1.0
“Every word captured. Every patient seen. The clinical AI partner to NHS systems - safe, regulated, proven at scale.” — the vendor’s own words

The only ambient voice technology certified UKCA Class IIa, with published safety science to back it. Narrow by design: no MCP, no A2A, desktop-only, and one independent NHS pilot found 20% of licensed users never logged in once.

Scope14/20
Quality6/10
SpecialistHealthcareVoice & SpeechProductivitySecurity & ComplianceFreemium
Vendor
TORTUS AI LTD
Origin
United Kingdom — London
Pricing
Per user £100–£200/user/mo · Enterprise Custom
Users (official only)
Over 2.5 million consultations processed across 1,000+ NHS organisations, 4 years live, 50+ clinical settings
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx6
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr9
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft3
Fine-tuning
Fw6
Frameworks & harnesses
Ev7
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In4
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: NHS trusts and GP federations that need an ambient scribe which will survive a clinical safety review on someone else's evidence rather than their own, and that are willing to invest in licence-utilisation management after go-live.
Full audit passport →Visit tortus.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-17 · table v1.0
“Ambient Clinical Intelligence engineered to transform your workflow” — the vendor’s own words

The only ambient scribe with a real developer platform behind it, and the only one that published research saying the industry's own quality yardstick is broken. Independently validated at $1,223 incremental revenue per provider per month, and honest enough to report a 31% hallucination rate in its own peer-reviewed study.

Scope17/20
Quality5/10
Enterprise platformHealthcareVoice & SpeechProductivityAutomation & AgentsPaid
Vendor
Suki AI, Inc.
Origin
US — Redwood City
Pricing
Documentation tier (formerly Suki Compose) ~$299/mo · Full assistant (formerly Suki Assistant) ~$399/mo · Enterprise $350-$500+/mo +1 more
Users (official only)
400+ healthcare systems and partners; ambient encounters across Suki for Clinicians and Partners grew 88% in six months
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx8
Context
Tr5
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr5
Guardrails
Mm7
Multimodal
Deployment
Ag6
Agents
Ft4
Fine-tuning
Fw9
Frameworks & harnesses
Ev5
Evaluations
Sm2
Small models
Emerging
Ma2
Multi-agent
Sy
Synthetic data
Pc5
Protocols
In3
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: US medical groups and health systems already standardised on Epic, Oracle Health, athenahealth or MEDITECH that want coding and order staging in the same contract as documentation, and can run a real pilot plus an adoption push. Also the first call for any healthtech company that has decided to partner rather than build ambient AI, because the developer platform has no serious rival in this category. Not for solo clinicians, small practices, or anyone who needs a posted price this week.
Full audit passport →Visit suki.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-17 · table v1.0
“The AI scribe and Front Desk giving time back to independent clinics” — the vendor’s own words

The one ambient scribe a solo clinician can actually buy this afternoon: posted prices from $39, a 7-day trial, 26,000 clinicians and 32.6 million visits transcribed in 2025. It scores low here because it deliberately ships none of the platform layer, and because nobody outside Freed has ever peer-reviewed its accuracy.

Scope17/20
Quality4/10
SMB toolHealthcareProductivityVoice & SpeechVoice AgentsFreemium
Vendor
Freed Inc.
Origin
US — San Francisco
Pricing
Starter $39/mo · Core $79/mo · Premier $119/mo or $104/mo annual +2 more
Users (official only)
26,000+ clinicians, 1,300+ clinics, 32,589,627 patient visits transcribed in 2025, 5.9M hours returned per year
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx5
Context
Tr5
Tracing
Lg6
LLM
Compositions
Fc4
Function calling
Vx
Vector store
Rg5
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag6
Agents
Ft3
Fine-tuning
Fw1
Frameworks & harnesses
Ev2
Evaluations
Sm1
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc1
Protocols
In2
Interpretability
Th2
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo clinicians and independent practices of roughly 2 to 50 people doing outpatient primary care, family medicine, internal medicine or mental health, who want a good note today without a procurement cycle and whose EHR runs in a browser. Also the cleanest way to find out whether ambient documentation helps you at all, because the trial costs nothing and takes minutes. Not for health systems, not for surgical or allied-health specialties, not for Canada, and not for anyone who needs to build on top of it.
Full audit passport →Visit getfreed.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-16 · table v1.0
“The GenAI evaluation and observability platform” — the vendor’s own words

The only eval platform we have scored that also ships a serious open-source gateway — Bifrost, Apache-2.0, 7,300 GitHub stars, 1,000+ models and an MCP gateway — wrapped around per-seat billing and three-day log retention on the free tier, so the bill tracks your headcount rather than your traffic.

Scope19/20
Quality7/10
SpecialistAutomation & AgentsCodingProductivityOpen sourceFreemium
Vendor
Maxim AI (H3 Labs Inc.)
Origin
US — San Francisco
Pricing
Developer Free · Professional $29/seat/mo · Business $49/seat/mo +1 more
Users (official only)
Not disclosed — no user or customer count published. The trust center names EY, Mindtickle, Atomicwork, Clinc and Qure among customers, and the Bifrost site claims 1,000+ teams use the gateway
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em5
Embeddings
Cx7
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft3
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy8
Synthetic data
Pc8
Protocols
In6
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Small to mid-sized product and engineering teams running multi-agent systems who want simulation, evals and tracing in one workflow and can live with per-seat billing — and anyone who wants Bifrost as a free LLM and MCP gateway regardless of whether they buy the platform.
Full audit passport →Visit getmaxim.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-16 · table v1.0
“Automated QA for Voice AI and Chat AI Agents” — the vendor’s own words

The lowest-friction way into voice-agent QA — $0.25 a test minute, no sales call, an MCP server and Claude Skills on every plan — but it grades transcripts rather than audio, its testing is state-blind, and most of the public evidence for how good its evaluations are was written by Cekura.

Scope18/20
Quality6/10
SpecialistAutomation & AgentsVoice & SpeechSecurity & ComplianceFreemium
Vendor
Cekura (Tatva Labs Inc.)
Origin
US — San Francisco
Pricing
Pay as you go $0 to start · Startup $500/mo · Enterprise Custom
Users (official only)
75+ customers across healthcare, BFSI, logistics, recruitment and retail
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em4
Embeddings
Cx6
Context
Tr7
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg4
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft4
Fine-tuning
Fw7
Frameworks & harnesses
Ev8
Evaluations
Sm4
Small models
Emerging
Ma5
Multi-agent
Sy8
Synthetic data
Pc7
Protocols
In7
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Small teams shipping a voice or chat agent who want structured simulation and CI gating running this week for tens of euros, and who will calibrate the scores themselves before trusting them.
Full audit passport →Visit cekura.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-15 · table v1.0
“The open-source platform for agent development and evaluation” — the vendor’s own words

The fastest self-hosted eval platform to stand up — one pip install, one process, Postgres when you need it — and the best RAG evaluation in the category, but the server is Elastic License 2.0, not open source, and "fully open source, no feature gates" is the vendor's own overstatement.

Scope17/20
Quality7/10
Open sourceAutomation & AgentsCodingProductivityFreemium
Vendor
Arize AI, Inc.
Origin
US — Berkeley, California
Pricing
Phoenix self-hosted $0 · Phoenix Cloud $0 · Arize AX Free $0 +2 more
Users (official only)
10,000+ GitHub stars and 3M+ monthly downloads per the vendor (11,062 stars measured on GitHub 15/08/2026), 22M+ monthly OpenTelemetry instrumentation downloads, 7k+ community members. Arize AI has raised $135M total, including a $70M Series C in February 2025
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx8
Context
Tr8
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr5
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy4
Synthetic data
Pc9
Protocols
In5
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams doing serious evaluation work — RAG quality, prompt iteration, agent trajectory testing — who want it running today without provisioning a database they did not previously operate. It is the right first pick for a small team, for anyone whose stack is Java or polyglot, and for European teams that need traces to stay on their own infrastructure with no sales cycle. It is also the strongest choice if you want your coding agent to drive observability through MCP. Look elsewhere if OSI-approved open source is a hard procurement line, if you need always-on production monitoring with alerting and an uptime SLA rather than a debugging and evaluation workbench, or if you are betting on graduating from the free tool to the vendor's paid platform without re-instrumenting — that path does not exist.
Full audit passport →Visit phoenix.arize.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-15 · table v1.0
“AI Gateway & LLM Observability Platform for AI Engineers” — the vendor’s own words

Mintlify bought it on 03/03/2026 and put it in maintenance mode, and the repository shows exactly that: the last GitHub release was August 2025 and the most recent commit removes an incident banner. Do not start here in 2026, however good the two-line integration is.

Scope17/20
Quality5/10
Open sourceAutomation & AgentsCodingProductivityFreemium
Vendor
Helicone, Inc., acquired by Mintlify on 03/03/2026
Origin
US — San Francisco, California
Pricing
Self-hosted $0 · Hobby $0 · Pro $79/mo +2 more
Users (official only)
16,000 organisations, 14.2 trillion tokens processed, 33 million end users tracked, over three years — figures published by the founders in the acquisition post. GitHub: 6,073 stars, 653 forks, 168 open issues (measured 15/08/2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em
Embeddings
Cx6
Context
Tr8
Tracing
Lg7
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg3
RAG
Gr7
Guardrails
Mm4
Multimodal
Deployment
Ag3
Agents
Ft5
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm5
Small models
Emerging
Ma4
Multi-agent
Sy3
Synthetic data
Pc5
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Existing users with production traffic already flowing, who should read this as a migration-planning document rather than a purchase decision: the service is live, your data is safe, and Mintlify has committed to helping you move. For anyone still choosing, there is one narrow case where Helicone remains the honest answer — you want a gateway, not an observability SDK, and you specifically need caching, per-user rate limiting or runtime prompt-injection blocking in the request path, on OpenAI models, and you accept a frozen roadmap for it. Everyone else should start somewhere with a live roadmap. Emphatically not the choice for a new production system, for anyone who needs deep evaluation or agent tooling, or for a self-hosted deployment on Vertex, Bedrock or Azure, which the open-source build does not support.
Full audit passport →Visit helicone.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-14 · table v1.0
“The agent engineering platform and open source frameworks developers need to ship great agents faster” — the vendor’s own words

The broadest agent platform we have scored, and the one most likely to surprise you on the invoice: Engine rescans every 6 hours at 10-15 LCU a run, which is 60 to 90 dollars a day per project if you leave the spend limit blank.

Scope16/20
Quality7/10
Enterprise platformAutomation & AgentsCodingProductivityFreemium
Vendor
LangChain, Inc.
Origin
US — San Francisco
Pricing
Developer $0 · Plus $39/seat/mo · Enterprise Custom +1 more
Users (official only)
35% of the Fortune 500, over 1 billion open-source downloads, over 1 billion events ingested per day on LangSmith
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr9
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm6
Small models
Emerging
Ma8
Multi-agent
Sy7
Synthetic data
Pc9
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already building on LangChain or LangGraph who want observability, evaluation, deployment and an autonomous fix loop from one vendor, and who have someone willing to set the spend limits before the first production trace lands. Also a fair choice for teams not using LangChain at all, because OTLP ingestion and fanout make it framework-agnostic in practice. Less convincing if you need to self-host without a sales cycle, if you are a large team where $39 a seat multiplies badly, or if you want the vendor's AI features to run on your own model keys.
Full audit passport →Visit langchain.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-14 · table v1.0
“Open-Source LLM Observability, Evaluation & AI Agent Tracing” — the vendor’s own words

Apache-2.0 with the whole product in the free self-host and the cheapest managed tier in the category at 19 dollars, and the only platform here that will block a PII leak before it reaches a user, as long as you self-host, because the guardrails run nowhere else.

Scope15/20
Quality7/10
Open sourceAutomation & AgentsCodingProductivityFreemium
Vendor
Comet ML, Inc.
Origin
US — New York
Pricing
Open source (self-hosted) $0 · Free Cloud $0 · Pro Cloud $19/user/mo +2 more
Users (official only)
150,000+ users, 10,000+ teams, 21,000+ GitHub stars (21,358 verified via the GitHub API on 14/08/2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx7
Context
Tr9
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg5
RAG
Gr8
Guardrails
Mm8
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy7
Synthetic data
Pc7
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams that want to own their evaluation and observability stack outright under a permissive licence, and small teams in particular, because the self-host has no span cap, no retention cap and no seat cap. It is the strongest pick if regression testing and automated prompt or tool-schema optimization are the actual job rather than dashboards, and the only pick here if you need to block PII or off-topic output before it reaches a user without buying a separate guardrail vendor. Also a sensible default for teams already using Comet for classical ML experiment tracking, since both run on the same platform. Less convincing if nobody can carry five stateful services in production, if you need SSO or fine-grained RBAC without an enterprise contract, or if you want the vendor's AI assistant and refuse to leave your own infrastructure to get it.
Full audit passport →Visit comet.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“The AI-Ready Governance Platform” — the vendor’s own words

14,000 customers, 300+ patents and the only AI governance vendor here that already had the enterprise installed base before the category existed. Runtime guardrail enforcement shipped in March 2026 and that is the real move. Read the release notes though: AI Policy Manager and guardrail enforcement were still public preview in the May 2026 release, and the median contract of $11,835 tells you nothing about what an AI governance deployment costs.

Scope15/20
Quality6/10
Enterprise platformSecurity & ComplianceAutomation & AgentsProductivityPaid
Vendor
OneTrust, LLC
Origin
US — Atlanta, Georgia
Pricing
Single module (e.g. Consent & Preferences) Quote only · AI Governance module Quote only · Multi-module enterprise Quote only
Users (official only)
More than 14,000 customers globally, including over half of the Fortune 500
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx8
Context
Tr8
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx4
Vector store
Rg7
RAG
Gr8
Guardrails
Mm
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large enterprises already running OneTrust for privacy, third-party risk or GRC, where AI governance is a module purchase rather than a new vendor and the registry can inherit an existing data map. Also multinationals that need EU AI Act, NIST AI RMF and ISO 42001 work backed by DataGuidance regulatory research across 300+ jurisdictions. Wrong fit for SMBs, for lean teams without dedicated GRC resource, and for anyone needing value inside weeks rather than a configuration project.
Full audit passport →Visit onetrust.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“Prompt, run, and deploy AI agents in seconds” — the vendor’s own words

The shortest path from an OAuth account to an agent-callable tool: 10,000+ tools across 3,000+ APIs behind one static OAuth MCP URL, with read/write/destructive annotations, real Node/Python/Go/Bash in every step and version history with rollback. Workday now owns it, and a free tier of 100 credits a month is not the customer they bought.

Scope13/20
Quality6/10
SpecialistAutomation & AgentsCodingProductivityFreemium
Vendor
Pipedream, Inc. (a Workday company)
Origin
US — San Francisco
Pricing
Free $0 · Basic $29/mo · Advanced $49/mo +2 more
Users (official only)
5,000+ customers and tens of thousands of users at acquisition (founder statement); 1,000,000+ developers claimed on the pricing page
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx6
Context
Tr8
Tracing
Lg6
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg4
RAG
Gr6
Guardrails
Mm4
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Developers and product teams who want an agent to call thousands of authenticated third-party tools over MCP without building auth or hosting a server, and who can accept managed-only hosting and a compute-time meter.
Full audit passport →Visit pipedream.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“The trusted leader in AI governance. Helping enterprises govern agentic AI systems at scale.” — the vendor’s own words

The deepest regulation-to-control translation in the category, and the sharpest thinking about where agent governance actually belongs. But the product that acts on that thinking, Agent Governor, is a research preview on one harness with no production SLA, and the open-source assessment framework Credo AI built its early reputation on has been archived as unmaintained.

Scope15/20
Quality6/10
SpecialistSecurity & ComplianceAutomation & AgentsResearchPaid
Vendor
Credo.AI Corp.
Origin
US — Palo Alto, California
Pricing
Platform (modular) Quote only · Advisory Services Quote only, separate engagement · Agent Governor No charge, design partners only
Users (official only)
Not disclosed. No customer count is published; named enterprise customers include Mastercard, Booz Allen Hamilton, Chevron and Madrigal Pharmaceuticals
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em4
Embeddings
Cx8
Context
Tr8
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx4
Vector store
Rg8
RAG
Gr7
Guardrails
Mm
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Regulated enterprises with a dedicated AI governance function and an in-house ML or data science team, where the job is proving EU AI Act, NIST AI RMF or ISO 42001 conformity across a portfolio of systems. Also the right call for organisations standardising agent work on Claude Code who want to help shape runtime enforcement rather than wait for it. Wrong fit for SMBs, for teams whose AI is third-party SaaS rather than models they build, and for anyone who needs self-hosted deployment.
Full audit passport →Visit credo.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“AI Teammates that power your compliance program” — the vendor’s own words

The only GRC platform I have audited that is itself ISO 42001 certified, processes AI personal data in Frankfurt under its own DPA, and built its MCP so an agent structurally cannot delete your compliance records. Nine named agents do real work. The hole in the middle is that after three years Scrut still will not name the models running any of it.

Scope17/20
Quality6/10
SMB toolSecurity & ComplianceAutomation & AgentsProductivityFreemium
Vendor
Scrut Automation Inc.
Origin
US — San Francisco Bay Area (engineering in Bengaluru)
Pricing
Compliance Automation (AWS Marketplace) $15,000 per 12-month contract · Small team, quote Quote only · Mid-market, quote Quote only +1 more
Users (official only)
2,500+ customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx7
Context
Tr6
Tracing
Lg4
LLM
Compositions
Fc7
Function calling
Vx5
Vector store
Rg9
RAG
Gr8
Guardrails
Mm3
Multimodal
Deployment
Ag8
Agents
Ft4
Fine-tuning
Fw6
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In6
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: European teams that need a contractual commitment on where AI processing of their personal data happens, mid-market companies running two or more frameworks where the flat bundle beats per-framework pricing, and anyone who wants agents doing the chasing but refuses to let an assistant delete a compliance record.
Full audit passport →Visit scrut.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“Automate compliance. Improve security. Reduce risk.” — the vendor’s own words

The best protocol surface in GRC and the most honest model disclosure I have found in the category: 100+ MCP tools across 38 categories with full read and write, and a support page that names OpenAI GPT-5 generation and says plainly that Secureframe trains no models of its own. The AI itself is thin behind that. There is no eval surface beyond a thumbs-up button, no agent layer, and the hosted MCP runs in the US and UK only.

Scope17/20
Quality6/10
Enterprise platformSecurity & ComplianceAutomation & AgentsProductivityFreemium
Vendor
Secureframe, Inc.
Origin
US — San Francisco
Pricing
Fundamentals From $5,000/year · Complete Quote only · Defense Quote only
Users (official only)
6000+ customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx7
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc9
Function calling
Vx5
Vector store
Rg8
RAG
Gr6
Guardrails
Mm4
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In6
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering-led teams who want their compliance program queryable and writable from Claude Code or Cursor, defense contractors working CMMC 2.0 and FedRAMP through the Defense tier, and anyone who needs the vendor to state on the record which models touch their data.
Full audit passport →Visit secureframe.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“The Enterprise Platform for the AI-era workforce, unifying skills intelligence, learning, and knowledge in one closed loop” — the vendor’s own words

The most transparent AI stack we have scored in this category, with every model provider named, three vendor guardrail layers switched on and a working MCP server that writes as well as reads. The catch is that the headline agent product was announced in April and still is not shipped, the AI credit meter went live in January, and customer support scores 1.9 out of 10 on TrustRadius against 8.0 for ease of use.

Scope19/20
Quality6/10
Enterprise platformLearning & TrainingAutomation & AgentsSearchProductivityPaid
Vendor
Docebo Inc. (NASDAQ: DCBO; TSX: DCBO)
Origin
CA — Toronto, Ontario
Pricing
Elevate Quote only · Enterprise Quote only · Add-ons Quote only +2 more
Users (official only)
30,000,000+ learners worldwide; $255.1M Annual Recurring Revenue and $74,800 Average Contract Value as at 30/06/2026, implying roughly 3,400 active customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx7
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx4
Vector store
Rg7
RAG
Gr9
Guardrails
Mm9
Multimodal
Deployment
Ag5
Agents
Ft2
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm4
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In5
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprises of 250-plus active learners running multi-audience training across employees, customers and partners, especially in regulated sectors where the ISO, SOC 2, PCI, GxP and FedRAMP paperwork does real procurement work and where naming every model provider satisfies an AI governance policy that vaguer vendors would fail. It is the strongest choice in this category for anyone who wants learning data inside Claude, ChatGPT, Copilot or Gemini today rather than eventually. Be careful if your buying case rests on AgentHub or Enterprise Knowledge, because both are unshipped and should be contracted with delivery dates rather than assumed. Not for organisations under 250 learners, not for teams without a dedicated L&D administrator to absorb the setup curve, not for anyone who needs guaranteed responsive support without negotiating an explicit SLA, and not for buyers who need talent management, performance reviews or succession planning, which Docebo deliberately does not do.
Full audit passport →Visit docebo.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“An all-in-one AI Meeting Assistant, Conversation and Revenue Intelligence platform” — the vendor’s own words

The best MCP implementation we have scored on any SMB tool — read and write, OAuth, org dashboard — bolted onto a notetaker whose advertised $19 seat becomes $77 the moment you want the features you came for.

Scope17/20
Quality5/10
SMB toolMeeting NotetakersSalesProductivityAutomation & AgentsFreemium
Vendor
Avoma, Inc.
Origin
US — Palo Alto
Pricing
Startup $19/mo annual, $29/mo monthly · Organization $29/mo annual, $39/mo monthly · Enterprise $39/mo annual only +4 more
Users (official only)
1000+ organisations (vendor claim, 'Trusted by 1000+ large and small next generation modern organizations across the globe'); no verified user count published
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx7
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag6
Agents
Ft2
Fine-tuning
Fw6
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma2
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Small and mid-sized revenue teams — roughly 5 to 100 recording seats — who want one tool covering notes, scheduling, coaching and forecasting instead of four, and especially teams already running agents in Claude or ChatGPT who want meeting data readable and writable from chat. Do the add-on arithmetic first and budget the Organization plan, because Startup cannot reach the MCP server. Not for EU organisations that need data resident in Europe, and not for anyone whose calls run in heavily accented or highly technical English without a transcript test first.
Full audit passport →Visit avoma.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“The most widely used medical AI among verified U.S. physicians” — the vendor’s own words

The best grounded-answer engine we have scored — and you cannot use it. Free to 915,000 verified U.S. clinicians, blocked in the EU and UK by OpenEvidence's own choice, with no API, no SDK and no MCP for anyone.

Scope18/20
Quality5/10
SpecialistHealthcareResearchChatSearchFreemium
Vendor
OpenEvidence
Origin
US — Miami
Pricing
Verified U.S. clinician $0 · EU / UK clinician Not available · Health system Epic integration Not published
Users (official only)
915,000 licence-verified U.S. clinicians, including 690,000+ licence-verified U.S. physicians; 100M+ clinical consultations completed; ~18M consultations in December 2025 alone
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em8
Embeddings
Cx7
Context
Tr5
Tracing
Lg8
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg10
RAG
Gr6
Guardrails
Mm5
Multimodal
Deployment
Ag8
Agents
Ft2
Fine-tuning
Fw2
Frameworks & harnesses
Ev3
Evaluations
Sm3
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc1
Protocols
In8
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: U.S.-licensed clinicians with an NPI who want fast, cited literature answers between patients, and U.S. health systems already running Epic who can negotiate the patient-context integration. Useless to anyone in the EU or UK — not hard, not expensive, unavailable. Developers and RevOps teams should read this as a closed consumer-grade product with no build surface at all.
Full audit passport →Visit openevidence.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“The AI-driven LMS built to drive proprietary expertise” — the vendor’s own words

The best AI guardrail console we have scored in this category: admins set culture, pedagogy and compliance prompts at group level, and a course that contradicts the compliance prompt is stopped before it generates. The AI behind it is text-to-text only by the vendor's own admission, runs one Azure OpenAI model family with no model choice, and MCP is a Q3 2026 promise rather than a URL you can point Claude at.

Scope19/20
Quality5/10
Enterprise platformLearning & TrainingAutomation & AgentsProductivitySearchFreemium
Vendor
360Learning SA
Origin
France — Paris
Pricing
Team $8/user/mo · Business Custom · Enterprise Custom
Users (official only)
2,500+ teams on the platform (vendor claim). Learner counts not disclosed.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx7
Context
Tr4
Tracing
Lg6
LLM
Compositions
Fc6
Function calling
Vx4
Vector store
Rg8
RAG
Gr8
Guardrails
Mm5
Multimodal
Deployment
Ag6
Agents
Ft2
Fine-tuning
Fw6
Frameworks & harnesses
Ev2
Evaluations
Sm3
Small models
Emerging
Ma2
Multi-agent
Sy
Synthetic data
Pc4
Protocols
In5
Interpretability
Th2
Thinking models
Tap or hover any element to see why it got that score.
Best for: Mid-market and enterprise L&D teams in the EU that want subject-matter experts authoring courses under centrally-set AI guardrails, and that can accept text-only generation, one model family and reporting that needs a BI tool bolted on.
Full audit passport →Visit 360learning.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-13 · table v1.0
“Salesloft AI agents work as an extension of your team, proactively handling tedious tasks, surfacing critical insights, and anticipating your next move” — the vendor’s own words

The cadence engine is still the best in the category and the MCP server is real, but you pay $125-165 per user per month for AI that independent reviewers call helper-grade, you cannot see which model writes your emails, and this is the vendor whose GitHub account was open from March to June 2025 and ended up costing hundreds of Salesforce tenants their data.

Scope19/20
Quality5/10
Enterprise platformSalesAutomation & AgentsProductivityVoice & SpeechPaid
Vendor
Salesloft, Inc., part of Clari + Salesloft following the December 2025 merger
Origin
US — Atlanta, Georgia (combined company also operates from Sunnyvale, California)
Pricing
Advanced ~$100-140/user/mo (reported) · Premier / Elite ~$140-185+/user/mo (reported) · Dialer add-on ~$200-400/user/yr +3 more
Users (official only)
4,000+ sales teams (Salesloft alone); more than 5,000 organisations across the combined Clari + Salesloft, including Adobe, IBM, 3M and Zoom
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em3
Embeddings
Cx8
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx3
Vector store
Rg6
RAG
Gr4
Guardrails
Mm7
Multimodal
Deployment
Ag7
Agents
Ft2
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm2
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In5
Interpretability
Th2
Thinking models
Tap or hover any element to see why it got that score.
Best for: Outbound organisations of 50-plus reps with a dedicated RevOps function, a Salesforce-centric workflow and procurement discipline strong enough to strip the escalators and auto-renewal out of the order form. It is a genuinely good execution layer for structured multi-channel cadences, and Conversations is worth having if you are not already paying a specialist for call intelligence. It is the wrong tool below roughly 20 reps, where seat minimums, implementation cost and admin burden destroy the return, and the wrong tool for anyone who needs to know which accounts are in market, because Salesloft ships no intent data, no visitor identification and no contact enrichment at any price. If model transparency or an eval surface is a hard requirement for your AI governance policy, this platform does not currently meet it.
Full audit passport →Visit salesloft.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-12 · table v1.0
“An AI Care Partner that expands clinical capacity by automating administrative work” — the vendor’s own words

The most transparent clinical AI vendor on the market and the only one that publishes a EUR price — but it is walking away from the small practices that made it.

Scope17/20
Quality6/10
SpecialistHealthcareProductivityVoice & SpeechResearchFreemium
Vendor
Heidi Health Pty Ltd
Origin
Australia — Melbourne
Pricing
Free · Clinician $150/mo · Practice $180/mo +1 more
Users (official only)
2 million+ consults per week across 116 countries and 110 languages, 200+ specialties; 73 million patient consults and 3.1 billion minutes transcribed since launch
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg8
RAG
Gr8
Guardrails
Mm8
Multimodal
Deployment
Ag7
Agents
Ft4
Fine-tuning
Fw7
Frameworks & harnesses
Ev7
Evaluations
Sm7
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In6
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individual clinicians and mid-sized practices across the EU, UK, Canada and Australia who want a real free tier, a published price in their own currency, multilingual coverage and the most legible safety testing in the category — and who can live with the risk that the vendor's attention is moving upmarket. Not the pick if you need MCP, customer-run evals, or a new API integration in Australia or New Zealand.
Full audit passport →Visit heidihealth.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-12 · table v1.0
“The leading ambient AI assistant, reducing practitioner burn-out and improving patient care” — the vendor’s own words

The only ambient scribe that beat a control group in a randomised trial, and the only one with a real public API — but it stopped publishing its price.

Scope17/20
Quality6/10
SpecialistHealthcareProductivityVoice & SpeechAutomation & AgentsFreemium
Vendor
Nabla Technologies
Origin
France — Paris
Pricing
Free · Pro ~$119/mo · Enterprise / Nabla Connect Custom
Users (official only)
190+ healthcare organizations, 100,000 clinicians, 40 million patient encounters annually
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag7
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm2
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc4
Protocols
In7
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Health systems and EHR vendors that want ambient documentation they can build on rather than be locked into, and buyers who want efficacy evidence that survives a sceptical clinical governance committee. Also the strongest option on this list for an EU buyer, since it is Paris-based, GDPR-compliant and works in 35+ languages. Individual clinicians can still start free — just budget for a sales conversation you did not used to need.
Full audit passport →Visit nabla.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-11 · table v1.0
“The trust platform for voice and chat agents. Test before launch. Monitor in production. Red-team risky behavior before it reaches customers.” — the vendor’s own words

The deepest voice-agent testing loop we have scored on tool-call contracts, DTMF and IVR behaviour, and turning a production failure back into a regression test, sold by a company that will not publish a single price and grades other people's agents on a judge model it refuses to name.

Scope18/20
Quality6/10
SpecialistAutomation & AgentsVoice & SpeechSecurity & CompliancePaid
Vendor
Hamming AI, founded 2024 by Sumanyu Sharma (CEO) and Marius Buleandra (CTO), both ex-Tesla; Y Combinator Summer 2024
Origin
US — San Francisco
Pricing
Startup Contact us · Agency Contact us · Enterprise Contact us
Users (official only)
Not disclosed as a user count. Hamming states 10K+ agents monitored, 4M+ calls tested and 10M+ minutes protected, and names Bland Labs, Booked AI, Grove AI, Synthpop, Maven AGI, Basata, Anima, Serve, Opalite Health and mdhub in customer testimonials and case studies. $3.8M seed led by Mischief announced 18/12/2024; aggregators put total funding at $4.3M-$4.5M and headcount at 17-22 in mid-2026.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em4
Embeddings
Cx7
Context
Tr9
Tracing
Lg5
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg4
RAG
Gr8
Guardrails
Mm9
Multimodal
Deployment
Ag6
Agents
Ft3
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm3
Small models
Emerging
Ma5
Multi-agent
Sy8
Synthetic data
Pc6
Protocols
In6
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams running a phone-based voice agent in production on Vapi, Retell, LiveKit, Pipecat, ElevenLabs or Bland, where the agent navigates an IVR, calls booking or CRM tools with real side effects, and sits under HIPAA or PCI. Not for a team that needs to price the tool before talking to sales, or one that wants to audit the judge before trusting its verdicts.
Full audit passport →Visit hamming.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-11 · table v1.0
“Superhuman Docs is the first surface that brings teams, data, and AI together. Not in a chat window, or dozens of threads. Right where work actually happens and with the power to build with you.” — the vendor’s own words

Renamed Superhuman Docs on 08/07/2026, this is still the best database-backed doc on the market and now has the best-built MCP server of any workspace tool we have scored, but its AI runs on models the vendor will not name, cannot read a PDF, has no quality measurement of any kind, and its enterprise-search product was wound down into a sibling you have to buy separately.

Scope16/20
Quality5/10
Enterprise platformProductivityApp BuildersAutomation & AgentsFreemium
Vendor
Superhuman, Inc. (formerly Grammarly, Inc.; renamed 29/10/2025). Coda Project, Inc. founded 2014 by Shishir Mehrotra and Alex DeNeui, acquired by Grammarly 17/12/2024; Mehrotra became CEO of the combined company
Origin
US — San Francisco
Pricing
Free $0 · Pro $12/mo per Doc Maker annual ($15 monthly) · Business $33/mo per Doc Maker annual ($40 monthly) +1 more
Users (official only)
Not disclosed for Coda specifically. Superhuman states 40M+ people, 50K+ organizations and 3,000 educational institutions, but that figure covers the whole suite (Grammarly, Docs, Mail, Go) and Grammarly's base dominates it. Coda's own pre-acquisition boilerplate claimed 'over 50,000 teams', and CEO Mehrotra said 'over 40,000 enterprises' in a 2024 interview; the two numbers were never reconciled. Named customers from acquisition-era marketing: Figma, DoorDash, Square, The New York Times.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em4
Embeddings
Cx8
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm4
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev2
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In4
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: Ops, finance and programme teams who already run their operating cadence on relational tables and formulas rather than prose, and who want an AI that can act on that structure and an MCP endpoint that lets Claude or ChatGPT write into the same source of truth. Right for a team that can live with unnamed models and no quality measurement. Wrong for anyone whose main input is PDFs, or who needs cross-tool enterprise search without buying Superhuman Go as well.
Full audit passport →Visit docs.superhuman.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-10 · table v1.0
“The active observability platform for agents” — the vendor’s own words

The deepest eval-to-CI loop in the category, and the only platform here that bills every byte of your traces with no spending cap, so set billing alerts before the first production trace.

Scope16/20
Quality7/10
SpecialistAutomation & AgentsCodingProductivityPaid
Vendor
Braintrust Data, Inc.
Origin
US — San Francisco
Pricing
Starter $0 · Pro $249/mo · Enterprise Custom
Users (official only)
Not disclosed. Braintrust names Notion, Stripe, Vercel, Zapier, Airtable and Instacart as production users, and reports Notion moved issue triage from 3 to 30 issues per day on the platform.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx7
Context
Tr9
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm8
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy7
Synthetic data
Pc8
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering and product teams who already score outputs systematically and want the regression gate in CI rather than a dashboard they check afterwards. Especially good value for larger teams, because unlimited users at every tier beats per-seat pricing badly at 10 people and up. Skip it if what you actually want is traces and cost tracking, where an observability-first tool does the job for a fraction of the price, or if you need to own the whole stack, where an MIT-licensed self-host is the honest answer instead.
Full audit passport →Visit braintrust.dev Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-10 · table v1.0
“Open Source AI Engineering Platform” — the vendor’s own words

The open-source default for LLM tracing, MIT-licensed with 79 MCP tools out of the box, and you pay for it in operations: v3 means Postgres plus ClickHouse plus Redis plus S3, and ClickHouse now owns the company.

Scope15/20
Quality7/10
Open sourceAutomation & AgentsCodingProductivityPaid
Vendor
Langfuse, a ClickHouse, Inc. company
Origin
DE — Berlin
Pricing
Self-hosted Open Source $0 · Cloud Hobby $0 · Cloud Core $29/mo +4 more
Users (official only)
2,300+ companies, 10+ billion observations processed per month, 50M+ SDK installs per month, 32,700+ GitHub stars, 16+ full-time employees
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em
Embeddings
Cx8
Context
Tr9
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg5
RAG
Gr5
Guardrails
Mm8
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev8
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy4
Synthetic data
Pc9
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams that want to own their observability stack and have someone who can run four services in production, or European teams that need data to stay in their own infrastructure without an enterprise sales cycle. Also the right default for anyone tracing high volumes on a small team, because self-hosting removes per-trace cost entirely and the Cloud tiers carry unlimited users from $29 up. Less convincing if you want evaluation to be turnkey with a merge gate on day one, if you need the in-product AI assistant and refuse to use Cloud, or if you cannot carry the operational load, in which case a managed platform is the honest choice.
Full audit passport →Visit langfuse.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-09 · table v1.0
“ClickUp is the AI work platform that not only helps manage and orchestrate work, but also does the work for you.” — the vendor’s own words

The most feature-dense AI layer among work-management platforms — real memory, real multi-model choice, real agents — but every bit of it rides on a mandatory per-seat Brain add-on, and ClickUp's headline agent benchmark is self-graded.

Scope15/20
Quality6/10
SMB toolProductivityAutomation & AgentsMeeting NotetakersFreemium
Vendor
ClickUp
Origin
US — San Diego, CA
Pricing
Free Forever · Unlimited $7/user/mo · Business $12/user/mo +3 more
Users (official only)
Over 20 million users, 5M+ teams, surpassed $300M in annual recurring revenue
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx8
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm5
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams that want the deepest AI feature set inside a single work-management tool and are comfortable paying the mandatory Brain add-on across every seat, not just the ones using AI.
Full audit passport →Visit clickup.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-09 · table v1.0
“The AI work platform that turns strategy into execution, at scale” — the vendor’s own words

A genuinely rebuilt agent platform bolted onto a mature work-management core — strong governance and MCP support, but the AI still runs on borrowed frontier models and monday's own Q1 2026 earnings call described real agent/MCP usage as 'still not very significant numbers.'

Scope15/20
Quality6/10
Enterprise platformAutomation & AgentsProductivityMeeting NotetakersFreemium
Vendor
monday.com Ltd.
Origin
Israel — Tel Aviv (dual HQ: New York)
Pricing
Free · Basic $9/seat/mo · Standard $12/seat/mo +2 more
Users (official only)
Over 250,000 customers worldwide; 65,016 paid customers with more than 10 users, up 7% year-over-year
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em3
Embeddings
Cx8
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg7
RAG
Gr8
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already running work management on monday who want to add lightweight, governed AI agents without adopting a separate AI platform — not a fit for teams that want to build or fine-tune their own models.
Full audit passport →Visit monday.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-08 · table v1.0
“A global market leader in conversational and agentic AI” — the vendor’s own words

Enterprise-grade agentic CX that runs live at Lufthansa and Mercedes-Benz scale — but you'll pay six figures a year for it, and since September 2025 NICE owns the roadmap, not Cognigy.

Scope18/20
Quality6/10
Enterprise platformSupport AgentsAutomation & AgentsVoice & SpeechProductivityPaid
Vendor
Cognigy GmbH (operating as NiCE Cognigy)
Origin
Germany — Düsseldorf
Pricing
Enterprise Custom
Users (official only)
Not disclosed. NICE names Lufthansa, Mercedes-Benz and Nestlé as enterprise customers and cites an estimated 80% ARR growth for Cognigy within NICE in 2026.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx6
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx5
Vector store
Rg8
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev8
Evaluations
Sm5
Small models
Emerging
Ma3
Multi-agent
Sy8
Synthetic data
Pc7
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large contact centers already running Genesys, Avaya or Amazon Connect that need a proven, high-volume agentic CX layer and can absorb a six-figure annual contract and a 2-4 month implementation.
Full audit passport →Visit cognigy.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-08 · table v1.0
“Multiplayer AI for human-agent collaboration” — the vendor’s own words

The fastest way to wire agents into Slack, Drive and Notion for €29 a seat — until the 1GB/user data cap and Enterprise-only governance push you into a custom quote.

Scope17/20
Quality6/10
Enterprise platformAutomation & AgentsProductivitySearchFreemium
Vendor
Dust
Origin
France — Paris
Pricing
Trial Free · Pro ~$31/user/mo · Enterprise Custom
Users (official only)
300,000+ agents deployed, 3,000+ organizations running on Dust (self-reported live homepage stats); funding coverage of the May 2026 Series B additionally cites ~70% weekly active usage and zero churn in 2025
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx7
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx5
Vector store
Rg8
RAG
Gr6
Guardrails
Mm4
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Slack-centric teams and mid-to-large enterprises that want employees building and running their own agents against company knowledge without waiting on an engineering team to ship each one.
Full audit passport →Visit dust.tt Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“AI coding assistant” — the vendor’s own words

The widest model catalogue and the deepest agent surface in any coding tool — but since 1 June 2026 the $10 sticker price is only the entry fee on a token meter, and heavy agent users are burning a month of credits in a week.

Scope18/20
Quality8/10
Enterprise platformCodingAutomation & AgentsProductivityFreemium
Vendor
GitHub, Inc. (Microsoft subsidiary since 2018)
Origin
US — San Francisco
Pricing
Free $0 · Pro $10/mo · Pro+ $39/mo +3 more
Users (official only)
Over 26 million users (Satya Nadella, Microsoft Q1 FY2026 earnings call); 4.7 million paid subscribers as of Q2 FY2026, up ~75% year over year, per Microsoft's 28/01/2026 earnings disclosure
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em7
Embeddings
Cx9
Context
Tr7
Tracing
Lg10
LLM
Compositions
Fc9
Function calling
Vx6
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag9
Agents
Ft4
Fine-tuning
Fw9
Frameworks & harnesses
Ev5
Evaluations
Sm9
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering organisations already on GitHub that want one governed AI layer across every IDE, the CLI and the PR pipeline — and that will actively manage model choice and credit burn. Less suited to teams needing predictable flat costs, to regulated environments that require exclusion to hold in agent mode, or to anyone wanting to adapt a model to a proprietary codebase.
Full audit passport →Visit github.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“The voice AI evaluation platform for agents at scale. Simulate, observe, and review all in one place.” — the vendor’s own words

The most complete voice-agent evaluation stack we have scored: audio-native metrics, OpenTelemetry tracing, a hosted MCP server and an A2A connection type, all from a 10-person company that has raised $31M and grades other people's agents without publishing any validation of its own judges.

Scope18/20
Quality7/10
SpecialistAutomation & AgentsVoice & SpeechSecurity & ComplianceFreemium
Vendor
Coval (coval.dev), founded 2024 by Brooke Hopkins, previously eval infrastructure lead at Waymo
Origin
US — San Francisco
Pricing
Starter $100/mo · Growth $500/mo · Enterprise From $4,500/mo
Users (official only)
Not disclosed as a user count. Coval states more than 60 enterprises use the platform, including Zoom and Deepgram, and reports 10x year-over-year revenue growth and tens of millions of evaluations run. Team size 10 per its Y Combinator profile.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em5
Embeddings
Cx7
Context
Tr9
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg5
RAG
Gr7
Guardrails
Mm9
Multimodal
Deployment
Ag7
Agents
Ft6
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm6
Small models
Emerging
Ma5
Multi-agent
Sy8
Synthetic data
Pc9
Protocols
In6
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already running a voice agent in production on Vapi, LiveKit, Pipecat, Retell or an in-house stack, who need regression testing in CI and audio-level failure evidence before a release, and who can justify $500 a month or more.
Full audit passport →Visit coval.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“The leading Agentic Trust Management Platform” — the vendor’s own words

8,500 organisations and the only inline agent kill switch in GRC: Drata's MCP Proxy evaluates a tool call before it executes, where every competitor still dashboards it afterwards. Read the fine print though. It is Limited Availability, Anthropic-only, and an anonymised slice of your data trains a model shared with other customers.

Scope16/20
Quality7/10
Enterprise platformSecurity & ComplianceAutomation & AgentsProductivityPaid
Vendor
Drata Inc.
Origin
US — San Diego
Pricing
Foundation Quote only · Advanced Quote only · Enterprise Quote only +1 more
Users (official only)
8,500+ organizations worldwide
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em6
Embeddings
Cx7
Context
Tr8
Tracing
Lg6
LLM
Compositions
Fc8
Function calling
Vx5
Vector store
Rg8
RAG
Gr9
Guardrails
Mm
Multimodal
Deployment
Ag8
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Security teams that already run agents on Anthropic and need to prove to an auditor what those agents did, plus growth-stage and enterprise companies wanting SOC 2 or ISO 27001 with an MCP server their engineers will actually use from Claude or Cursor.
Full audit passport →Vanta vs Drata vs Sprinto →Visit drata.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“Accelerate work with AI agents that collaborate, automate, and think alongside your teams” — the vendor’s own words

The most open enterprise knowledge agent we have scored: real model choice, 100+ connectors, an MCP client and a genuinely usable free tier. It is also now a Workday subsidiary with no MCP server of its own, no eval surface, and a shared-integration permission model that can show people files they cannot open in SharePoint.

Scope19/20
Quality6/10
Enterprise platformAutomation & AgentsProductivityLearning & TrainingSearchFreemium
Vendor
Sana Labs AB, a Workday, Inc. company
Origin
SE — Stockholm
Pricing
Sana Agents Free $0 · Sana Agents Team $30/user/mo · Sana Agents Civic $0 +3 more
Users (official only)
1,000,000+ employees across hundreds of enterprises
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx9
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx5
Vector store
Rg9
RAG
Gr6
Guardrails
Mm8
Multimodal
Deployment
Ag8
Agents
Ft2
Fine-tuning
Fw8
Frameworks & harnesses
Ev4
Evaluations
Sm4
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In6
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Mid-to-large enterprises, especially Workday customers, that want one governed AI layer over a sprawling document estate and are prepared to configure shared-integration scopes carefully. Also worth a look for any team of five or fewer who simply want a free, model-agnostic, citation-backed assistant over Drive and SharePoint, because the free tier is one of the better ones on the market. Not for organisations under 300 people shopping the Learn side, not for compliance-heavy training programmes, and not for anyone who needs to call Sana from another agent.
Full audit passport →Visit sanalabs.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“the system of action where humans and AI run work together” — the vendor’s own words

The Work Graph is a genuinely good substrate for agents, and Asana is betting the company on it — but AI is still under 1% of a $790.8M revenue base, credit accounting is opaque enough that one customer burned 4 million credits in a month, and you cannot pick a model.

Scope16/20
Quality6/10
Enterprise platformProductivityAutomation & AgentsChatFreemium
Vendor
Asana, Inc. (NYSE: ASAN; CEO Dan Rogers since 21/07/2025)
Origin
US — San Francisco
Pricing
Personal Free · Starter $10.99/user/mo · Advanced $24.99/user/mo +4 more
Users (official only)
25,928 core customers spending $5,000+ annually and 817 customers spending $100,000+ annually as of Q4 FY2026; FY2026 revenue $790.8M, up 9% year over year; AI offerings exited FY2026 at over $6M ARR
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx8
Context
Tr5
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm4
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm3
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: Cross-functional teams of roughly 20 to 500 people who already run their work in Asana and want intake, routing and status reporting handled automatically — and who have someone willing to own credit budgeting. Not the right buy for a team of two (the free tier no longer supports one), for anyone who needs model choice or measurable AI quality, or for a small business expecting the headline seat price to be the total cost of the AI.
Full audit passport →Visit asana.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“The AI Agent Platform for Revenue teams” — the vendor’s own words

The best MCP server any SaaS vendor ships right now, 41 tools with real write actions, bolted onto a platform that gives you no model choice, no eval surface, and a quote-only contract that lands between $100 and $300 a seat before the five-figure implementation fee.

Scope18/20
Quality6/10
Enterprise platformSalesAutomation & AgentsProductivityVoice & SpeechFreemium
Vendor
Outreach Corporation
Origin
US — Seattle
Pricing
Amplify Essentials Quote only · Amplify Core Quote only · Amplify Plus Quote only +2 more
Users (official only)
Not disclosed on Outreach's own About page at audit date. Vendor press boilerplate carried by third parties says 'more than 5,500 companies'; getLatka estimates 6,000 customers, ~$300.8M ARR and 1,400 employees, all third-party estimates rather than official figures.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em4
Embeddings
Cx8
Context
Tr6
Tracing
Lg6
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft2
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm2
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprise revenue teams of 20 seats and up, already deep in Salesforce, with a dedicated RevOps function and, ideally, someone who wants Outreach data flowing into Claude, Copilot or a custom agent through MCP. That last case is where it beats everything else in the category. Not for small teams, not for cold-email-at-volume shops, and not for anyone who needs to see the price before booking a demo.
Full audit passport →Visit outreach.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“The world's first Autonomous Trust Platform” — the vendor’s own words

A $15,000 median contract buys the cheapest full-featured GRC platform in the category and the only no-code agent builder in it. What it does not buy is an MCP server, a named model, or one independently verified number behind the word autonomous.

Scope14/20
Quality6/10
SMB toolSecurity & ComplianceAutomation & AgentsProductivityPaid
Vendor
Sprinto Inc.
Origin
US — San Francisco (primary engineering in Bengaluru, India)
Pricing
Foundation Quote only · Growth Quote only · Add-on modules Quote only
Users (official only)
3,000+ companies across 75 countries
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx7
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In4
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Cloud-native startups and mid-market SaaS teams chasing their first SOC 2 or ISO 27001 on a real budget, especially two frameworks at once, and compliance leads who want to build their own agents without waiting for a vendor roadmap.
Full audit passport →Vanta vs Drata vs Sprinto →Visit sprinto.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“Go is the proactive AI assistant that knows what you know and offers help without you even having to ask.” — the vendor’s own words

The best distribution in the assistant market — Grammarly's 40M daily users and a million-plus app surface — wrapped around an agent platform that is still labelled beta ten months after launch, will not say which models it runs, and gives you no way to test whether its agents got the job right.

Scope15/20
Quality6/10
SMB toolProductivityAutomation & AgentsCopywritingFreemium
Vendor
Superhuman Platform Inc. (formerly Grammarly, Inc.; renamed 29/10/2025 after acquiring Coda and Superhuman Mail)
Origin
US — San Francisco
Pricing
Free $0 · Pro $12/mo · Business $33/mo +1 more
Users (official only)
Not disclosed for Superhuman Go specifically. The Chrome Web Store lists ~3,000 users for the standalone Superhuman Go extension. Parent company reports 40M+ daily active users, 50,000 organizations and one-third of the Fortune 500 across the platform, with Go also reachable by toggling it on inside the existing Grammarly extension.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em6
Embeddings
Cx8
Context
Tr5
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr8
Guardrails
Mm3
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc4
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and small teams already living inside the Grammarly extension who want cross-app context and one-click actions in Gmail, Calendar, Drive and Jira for free or €10 a month, and who are not yet ready to let an assistant act unsupervised.
Full audit passport →Visit superhuman.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“Generative AI for clinical conversations” — the vendor’s own words

The best-evidenced ambient scribe in healthcare and the most closed: deep Epic context and line-by-line source traceability, zero developer surface, and a peer-reviewed benefit of 16 minutes a day.

Scope17/20
Quality5/10
Enterprise platformHealthcareProductivityVoice & SpeechAutomation & AgentsPaid
Vendor
Abridge AI, Inc.
Origin
US — Pittsburgh
Pricing
Enterprise licence ~$208/mo · Full implementation $250-$500/mo · Individual / small practice Not sold
Users (official only)
300+ health systems live, 100M+ clinical conversations annually, partners serving 250M+ patients
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx9
Context
Tr9
Tracing
Lg8
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag7
Agents
Ft4
Fine-tuning
Fw2
Frameworks & harnesses
Ev4
Evaluations
Sm2
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc1
Protocols
In7
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large US health systems already deep in Epic that can run an enterprise procurement, mandate clinician review, and drive adoption past the 50%-of-visits mark where the measured benefit actually shows up. Not for individual clinicians, small practices, or anyone in the EU shopping today.
Full audit passport →Visit abridge.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-07 · table v1.0
“The AI Notetaker built for Team Collaboration” — the vendor’s own words

The best EU-hosting story in meeting AI and an MCP server at €18 a seat, sold with an intrusive bot and transcription that gets shaky the moment someone has an accent.

Scope18/20
Quality5/10
SMB toolMeeting NotetakersProductivitySalesAutomation & AgentsVoice & SpeechFreemium
Vendor
tldx Solutions GmbH
Origin
DE — Aachen
Pricing
Free €0 · Pro €18/mo · Business €29/mo +1 more
Users (official only)
2M+ users, 21,000+ customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em5
Embeddings
Cx7
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag6
Agents
Ft2
Fine-tuning
Fw6
Frameworks & harnesses
Ev2
Evaluations
Sm2
Small models
Emerging
Ma2
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In4
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: EU sales and customer-success teams who need meeting intelligence with data staying in Europe, want CRM sync and an MCP server without a Gong-sized contract, and can live with a visible bot in the call.
Full audit passport →Granola vs Fathom vs Otter vs Fireflies vs tl;dv →Visit tldv.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“A voice AI platform for deploying production-grade AI agents across phone, SMS, and chat” — the vendor’s own words

Bland owns its whole voice stack (speech-to-text, inference, text-to-speech) and has built the testing rigour to match: evals with LLM judges, node-level regression tests, scenario suites with production gates. That engineering depth is the reason to pick it, and the reason a non-technical team should not.

Scope17/20
Quality7/10
Enterprise platformVoice AgentsVoice & SpeechAutomation & AgentsText-to-SpeechFreemium
Vendor
Bland (Bland AI, Inc.)
Origin
US — San Francisco, CA; founded 2023 by Isaiah Granet (CEO) and Sobhan Nejad, Y Combinator alumni
Pricing
Start Free · $0.14/min talk time, $0.05/min transfer · Build $299/month · $0.12/min talk time, $0.04/min transfer · Scale $499/month · $0.11/min talk time, $0.03/min transfer +1 more
Users (official only)
250+ enterprise customers, hundreds of thousands of self-service users, over 175 million AI phone calls handled in the past year, and more than 3.5 million calls per week
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em
Embeddings
Cx8
Context
Tr9
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg8
RAG
Gr9
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft7
Fine-tuning
Fw9
Frameworks & harnesses
Ev10
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy5
Synthetic data
Pc8
Protocols
In3
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering teams in regulated industries (healthcare, financial services, insurance) automating high-volume phone work where the conversation is long and non-linear, data residency and sub-processor count are audit questions, and someone on staff will own agent quality the way they would own a service in production.
Full audit passport →Bland vs Retell vs Vapi →Visit bland.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“ASAPP is the agentic Customer Experience Platform (CXP) purpose-built for enterprise customer service. CXP orchestrates AI agents, human expertise, and enterprise systems to resolve customer issues faster and more accurately across voice and digital channels.” — the vendor’s own words

The strongest safety and observability story in contact-center AI — 100% of responses monitored in production, continuous adversarial red-teaming against 50+ vulnerability types, and a rare published rate card at $0.70 a chat and $1.00 a call. The risk here is the company, not the product: a third CEO in three years landed on 4 August 2026, and there has been no new funding round since April 2021.

Scope15/20
Quality7/10
Enterprise platformSupport AgentsAutomation & AgentsVoice & SpeechChatProductivityPaid
Vendor
ASAPP, Inc.
Origin
US — New York, New York
Pricing
GenerativeAgent — Interaction (Digital) $0.70/interaction · GenerativeAgent — Interaction (Call) $1.00/interaction · Longer AWS contracts up to 92-96% off +2 more
Users (official only)
Not disclosed. GenerativeAgent stated as live across Fortune 100 enterprise contact centres; reference customers American Airlines, DISH, Ernst & Young and JetBlue
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx8
Context
Tr9
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr9
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft4
Fine-tuning
Fw7
Frameworks & harnesses
Ev9
Evaluations
Sm
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In6
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Regulated, high-volume voice and chat operations — airlines, telecoms, insurers, banks — that need auditable AI with a human approval gate more than they need an open stack, and that have enough interaction volume for containment to actually pay back $0.70 a chat.
Full audit passport →Visit asapp.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“The system for product development” — the vendor’s own words

€9-15 a seat a month buys real MCP access and a workspace where ten-plus AI agents sit as named teammates on your issues, but every coding session draws from a separate USD credit meter Linear only shipped in June 2026 and already renamed once — the agent platform is ahead of the billing model.

Scope17/20
Quality7/10
SMB toolAutomation & AgentsProductivityCodingFreemium
Vendor
Linear
Origin
US — San Francisco
Pricing
Free $0 · Basic $10/user/mo · Business $16/user/mo +1 more
Users (official only)
40,000+ companies
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em7
Embeddings
Cx7
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg7
RAG
Gr7
Guardrails
Mm4
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm5
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Software teams already running Linear for issue tracking who want AI coding agents and triage automation wired directly into the same workspace, and who are comfortable managing a separate AI-credit budget alongside seat licenses.
Full audit passport →Visit linear.app Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“SOC 2, HIPAA, ISO 27001, PCI, and GDPR Compliance” — the vendor’s own words

16,000 customers and $300M ARR buy the most auditor-familiar compliance platform there is, and an agent that is better engineered than most AI products twice its price — then year two arrives and G2 reviewers report 30-50% uplifts.

Scope14/20
Quality7/10
Enterprise platformSecurity & ComplianceAutomation & AgentsProductivityPaid
Vendor
Vanta Inc.
Origin
US — San Francisco
Pricing
Essentials Get personalized pricing · Plus Get personalized pricing · Professional Get personalized pricing +1 more
Users (official only)
More than 16,000 companies; $300M ARR crossed April 2026
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em5
Embeddings
Cx8
Context
Tr7
Tracing
Lg6
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr8
Guardrails
Mm
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In4
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Growth-stage and enterprise teams that need SOC 2 or ISO 27001 to close deals, run a standard cloud stack, and will use the MCP server to pull compliance context into the coding agent they already have open.
Full audit passport →Vanta vs Drata vs Sprinto →Visit vanta.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“Agentic AI that helps you organize, write, search, schedule, and more with just a prompt” — the vendor’s own words

€28-110 a seat a month for the most open MCP integration in email AI, but it still won't send, archive or delete anything without you clicking first — a very good assistant you drive by hand, not yet the autonomous one the marketing implies.

Scope14/20
Quality6/10
SMB toolProductivityAutomation & AgentsFreemium
Vendor
Shortwave Communications, Inc.
Origin
US — San Francisco
Pricing
Business $30/seat/mo · Premier $45/seat/mo · Max $120/seat/mo +1 more
Users (official only)
20,000 monthly active users
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx7
Context
Tr3
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx5
Vector store
Rg8
RAG
Gr7
Guardrails
Mm5
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw
Frameworks & harnesses
Ev
Evaluations
Sm6
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Gmail-native founders, small teams, and support/ops functions who want an agentic email layer wired into their existing stack via MCP, and who are comfortable approving actions rather than delegating them outright.
Full audit passport →Visit shortwave.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“AI calendar assistant that helps you plan, protect, and adapt your time” — the vendor’s own words

The June 2026 2.0 release turned a background calendar blocker into a real agent product with the best approval gate in its category: every AI write, including ones coming from ChatGPT or Claude, lands in Preview Mode before it touches your live calendar. Score it as a scheduling tool that got MCP right, not as an AI platform.

Scope11/20
Quality6/10
SMB toolProductivityAutomation & AgentsFreemium
Vendor
Reclaim.ai, Inc. (a Dropbox company)
Origin
US — Portland, OR (10115 NW Ash St., Portland, OR 97229, per the official DPA); founded 2019 by Henry Shapiro and Patrick Lightbody, acquired by Dropbox in August 2024
Pricing
Lite Free forever · Starter $10/seat/mo (annual) or $12/seat/mo (monthly) · Business $15/seat/mo (annual) or $18/seat/mo (monthly) +1 more
Users (official only)
More than 500,000 people across 60,000 companies active with Reclaim
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx6
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg5
RAG
Gr8
Guardrails
Mm
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and teams up to about 100 seats who already live in Google Calendar or Outlook, want their own priorities defended automatically, and want to drive their calendar from ChatGPT or Claude without letting an assistant silently rewrite their week.
Full audit passport →Visit reclaim.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“AI for Enterprise Revenue. The only Revenue Orchestration Platform with Revenue Context built to manage the complexity of modern enterprise revenue.” — the vendor’s own words

First revenue platform to ship a real MCP server, and still the forecast a board believes — but you pay $200-310 per user per month for two products eight months into a merger, and the group's last agentic security incident took hundreds of Salesforce tenants with it.

Scope15/20
Quality6/10
Enterprise platformSalesAutomation & AgentsVoice & SpeechProductivityPaid
Vendor
Clari + Salesloft (Clari, Inc.)
Origin
US — Sunnyvale, California
Pricing
Clari Forecast (base) $100-120/user/mo · Clari Copilot $60-110/user/mo · Salesloft engagement layer $50-80/user/mo +1 more
Users (official only)
More than 5,000 companies across the combined Clari + Salesloft, including Adobe, IBM, 3M and Zoom (Clari alone previously reported 1,500+ organisations and $5 trillion in revenue managed)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em5
Embeddings
Cx8
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr5
Guardrails
Mm7
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev6
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In4
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprise revenue teams with 100+ reps, dedicated RevOps capacity, and an existing AI assistant they want fed live pipeline data through MCP rather than another closed AI surface to log into.
Full audit passport →Visit clari.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-06 · table v1.0
“Automation AI that helps housing and healthcare organizations streamline communications and improve operational efficiency.” — the vendor’s own words

The best conversational AI in US housing by a distance, and $200M ARR across one in six American apartments says the market agrees — but this is a vertical product, not an AI platform: no public API, no MCP, no published fair-housing testing, and a quote-only price with a reported €23,000 floor that walls out anyone under roughly 700 units.

Scope14/20
Quality6/10
Enterprise platformAutomation & AgentsVoice & SpeechChatProductivityFreemium
Vendor
Elise A.I. Technologies Corp.
Origin
US — New York, New York
Pricing
Enterprise (only tier) Quote only · Reported per-unit rate $3-6/unit/mo · Reported annual minimum ~$25,000/yr
Users (official only)
$200M ARR after five consecutive years of 100% growth; powers one in six apartments in the United States; 100M+ patient conversations processed across dermatology, women's health, orthopedics and ophthalmology
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx8
Context
Tr7
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw4
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In3
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: US multifamily operators above roughly 700 units, and specialty medical groups on athenahealth, ModMed, Nextech or AdvancedMD, who want conversations handled rather than a platform to build on — and who will write vendor audit rights and fair-housing testing into the contract themselves, because EliseAI does not publish them.
Full audit passport →Visit eliseai.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-05 · table v1.0
“Retell is a platform to build, test, deploy, and monitor AI voice and chat agents with telephony, prompts, tools, and analytics built in.” — the vendor’s own words

The developer-first voice stack: real MCP server, real fine-tuning, a live pricing calculator on the homepage — just remember $0.07/min is the floor, not the bill.

Scope16/20
Quality7/10
SpecialistVoice AgentsAutomation & AgentsVoice & SpeechSalesFreemium
Vendor
Retell AI
Origin
US — Redwood City, California
Pricing
Pay as you go $0.07-$0.31/min · Enterprise Custom
Users (official only)
3,000+ businesses use Retell's voice agents; 55+ million AI phone calls routed monthly, including customers such as the San Antonio Spurs, Motorola, Lenovo and Anker
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em
Embeddings
Cx7
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft7
Fine-tuning
Fw8
Frameworks & harnesses
Ev8
Evaluations
Sm8
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering teams that want direct, code-level control over a voice agent stack — model choice, prompt iteration, MCP-based tooling — without an enterprise sales cycle or managed-service overhead.
Full audit passport →Bland vs Retell vs Vapi →Visit retellai.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-05 · table v1.0
“The leading vendor of enterprise AI agents for customer service” — the vendor’s own words

Best-in-class enterprise voice AI, closed and priced accordingly: no public rate card, budget $150K+/year, and most changes still route through PolyAI's own team.

Scope14/20
Quality6/10
Enterprise platformVoice AgentsAutomation & AgentsVoice & SpeechChatFreemium
Vendor
PolyAI
Origin
UK — London
Pricing
Enterprise Custom
Users (official only)
100+ enterprise customers; 2,000+ live deployments across 45 languages and 25+ countries
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev8
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large enterprises with high call volume — contact centers, hospitality, banking, insurance — that want a white-glove, best-in-class voice experience and can absorb a six-figure annual contract plus a multi-week sales cycle.
Full audit passport →Visit poly.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-04 · table v1.0
“Your company's AI Source of Truth” — the vendor’s own words

Real citations, daily-run evals and a working MCP server, but self-serve pricing is gone: budget a sales call before you get a floor price.

Scope16/20
Quality6/10
Enterprise platformProductivitySearchAutomation & AgentsPaid
Vendor
Guru Technologies, Inc.
Origin
US — Philadelphia, PA
Pricing
Self-Serve (historical, third-party) $25/seat/mo, 10-seat min ($250/mo floor) · Enterprise Custom
Users (official only)
2,500+ companies
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx7
Context
Tr8
Tracing
Lg7
LLM
Compositions
Fc5
Function calling
Vx5
Vector store
Rg9
RAG
Gr7
Guardrails
Mm3
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev8
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th3
Thinking models
Tap or hover any element to see why it got that score.
Best for: Mid-size to large orgs (51-1000+ employees) centralizing knowledge scattered across Slack, Salesforce, Confluence and Drive who need cited, permission-aware answers and can absorb an enterprise sales cycle instead of a self-serve signup.
Full audit passport →Visit getguru.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-04 · table v1.0
“A faster way to create professional presentations” — the vendor’s own words

Genuinely fast, brand-locked slide design with a real MCP connector now live, but the AI layer itself is thin: no disclosed model, no evals, no free tier.

Scope12/20
Quality4/10
SMB toolPresentationsProductivityMarketingCopywritingFreemium
Vendor
Beautiful Slides, Inc.
Origin
US — San Francisco, CA
Pricing
Pro $12/mo · Team $40/user/mo · Enterprise Custom
Users (official only)
3M+ users, 35M+ slides created, 195+ countries, 50K+ companies
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx4
Context
Tr3
Tracing
Lg5
LLM
Compositions
Fc3
Function calling
Vx
Vector store
Rg4
RAG
Gr4
Guardrails
Mm5
Multimodal
Deployment
Ag3
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev2
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and small-to-mid teams (sales, consulting, marketing) who want fast, brand-consistent decks without a designer and don't need RAG, evals, or model choice.
Full audit passport →Gamma vs Canva Magic Studio vs Beautiful.ai →Visit beautiful.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-03 · table v1.0
“WRITER is where the world's leading enterprises orchestrate AI-powered work.” — the vendor’s own words

A real full-stack enterprise agent platform with a genuine MCP-based connector spine and a 1M-token model — the catch is the price tag: no mid-tier, and Enterprise pricing plus deployment effort put it out of reach for anyone below a few hundred seats.

Scope17/20
Quality8/10
Enterprise platformAutomation & AgentsCopywritingProductivityPaid
Vendor
Writer, Inc.
Origin
US — San Francisco, CA
Pricing
Starter ~$29–39/user/month (third-party reported) · Enterprise Custom
Users (official only)
Not disclosed precisely. Writer's own newsroom/about page says 'hundreds of the world's leading enterprises' and 'hundreds of customers'; third-party press (referenced in Writer's own newsroom) cites '300 companies' as of April 2025.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em4
Embeddings
Cx10
Context
Tr8
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg9
RAG
Gr8
Guardrails
Mm7
Multimodal
Deployment
Ag9
Agents
Ft8
Fine-tuning
Fw8
Frameworks & harnesses
Ev5
Evaluations
Sm7
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large enterprises (regulated industries especially — healthcare, finance) that want one governed platform for agentic workflows across teams and are prepared to go through a sales-led Enterprise deployment rather than self-serve.
Full audit passport →Visit writer.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-03 · table v1.0
“The AI Powered SuperApp for Work” — the vendor’s own words

A genuinely good AI calendar-and-task app that has bolted an 'AI Employees' agent suite on top — score it as a scheduling tool with an agentic side project, not as an AI platform, because most of the periodic table simply doesn't apply here.

Scope10/20
Quality6/10
SMB toolProductivityAutomation & AgentsPaid
Vendor
Motion
Origin
US — San Francisco, CA (Y Combinator W20 batch; founded 2019 in the Bay Area)
Pricing
Pro AI $19/seat/mo (annual) or $29/seat/mo (monthly) · Business AI $29/seat/mo (annual) or $49/seat/mo (monthly) · AI Employee ~$49/mo (per third-party trackers; not listed as a distinct SKU on the current official pricing page)
Users (official only)
Over 1 million top performers and teams (current homepage claim); separately, 10,000+ SMB customers specifically on the 'AI Employees' agentic tier as of the Sept 2025 Series C announcement
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx6
Context
Tr
Tracing
Lg5
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr4
Guardrails
Mm5
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and small-to-mid-size teams who want their to-do list and calendar actually merged and auto-managed, and SMBs curious about bolting lightweight autonomous 'AI Employee' agents onto that same workspace without hiring more headcount.
Full audit passport →Visit usemotion.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-02 · table v1.0
“The most productive inbox ever made.” — the vendor’s own words

The fastest inbox with the most convincing personal-voice AI on the market — Ask AI and Auto Drafts genuinely ground answers in your own mail, calendar and the web — but it's a premium per-seat tool from a company that renamed itself after the product eight months ago.

Scope11/20
Quality6/10
SMB toolProductivityAutomation & AgentsSalesPaid
Vendor
Superhuman, Inc. (formerly Grammarly; renamed after acquiring Superhuman Mail in July 2025)
Origin
US — San Francisco
Pricing
Starter $25/mo · Business $33/mo · Enterprise Custom
Users (official only)
Not disclosed specifically for Superhuman Mail. Parent company Superhuman states 40M+ daily active users across the full Suite (Grammarly, Docs, Mail, Go), per CEO Rahul Vohra
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em6
Embeddings
Cx7
Context
Tr5
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw2
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Sales and business-development professionals on Gmail or Outlook who process high email volume and can justify the Business-tier price for CRM-integrated Auto Drafts.
Full audit passport →Visit superhuman.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-02 · table v1.0
“Webflow AI streamlines workflows, delivers personalized experiences at scale, and improves performance for both humans and machines.” — the vendor’s own words

A visual site builder with real AI polish and now the most complete MCP surface in no-code — but it's an assistive layer bolted onto Webflow, not an AI platform in its own right.

Scope12/20
Quality6/10
SMB toolWebsite BuildersProductivityMarketingAutomation & AgentsFreemium
Vendor
Webflow, Inc.
Origin
US — San Francisco
Pricing
Basic (Site) $15/mo · Premium (Site) $25/mo · Core (Workspace) $28/mo +3 more
Users (official only)
More than 60,000 websites published using Webflow's AI site builder
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx5
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg5
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev2
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Marketing teams and agencies that want a fast, governed way to spin up and iterate on production websites, especially those already driving builds through Claude or another MCP-compatible agent.
Full audit passport →v0 vs Framer vs Webflow →Visit webflow.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-01 · table v1.0

The best test-before-deploy discipline in voice AI — simulation, LLM-judge evals and red-teaming are the actual product — priced for enterprises that spend $300K a year to replace an IVR.

Scope15/20
Quality8/10
Enterprise platformVoice AgentsAutomation & AgentsVoice & SpeechChatPaid
Vendor
Parloa
Origin
DE — Berlin
Pricing
Not published — custom enterprise quotes only; independent reviews report typical deployments from ~$300K (~€275K) per year
Users (official only)
Not disclosed; company-stated ARR above $50M, 380 employees, customers include Allianz, Booking.com, SAP, Swiss Life
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr9
Tracing
Lg8
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg7
RAG
Gr9
Guardrails
Mm8
Multimodal
Deployment
Ag9
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm
Small models
Emerging
Ma7
Multi-agent
Sy7
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprises with high-volume voice contact centers replacing legacy IVRs, especially in regulated industries where the simulation and red-team lifecycle justifies a six-figure annual contract.
Full audit passport →Visit parloa.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-08-01 · table v1.0

Fine-tunes on your own conversations and coaches human agents in real time like nobody else, but the integration surface is closed — no MCP — and the tuning bill is paid in months of internal effort.

Scope16/20
Quality7/10
Enterprise platformSupport AgentsAutomation & AgentsChatVoice & SpeechProductivityPaid
Vendor
Cresta Intelligence, Inc.
Origin
US — Palo Alto
Pricing
Not published — custom enterprise quotes only; per-seat plus usage licensing reported by independent reviews
Users (official only)
Over $100M ARR; customers include United Airlines, Cox Communications, Marriott; 500+ employees
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx8
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr8
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft9
Fine-tuning
Fw8
Frameworks & harnesses
Ev8
Evaluations
Sm7
Small models
Emerging
Ma5
Multi-agent
Sy6
Synthetic data
Pc3
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprises that keep humans in the loop — large contact centers in travel, financial services and telecom where real-time agent coaching and conversation intelligence on 100% of calls matter more than full automation.
Full audit passport →Visit cresta.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“Build, customize, test, deploy, and monitor AI-powered agents” — the vendor’s own words

The broadest enterprise agent platform on paper — Atlas, MCP, A2A, real observability — but $800M ARR and CIOs saying 'it isn't there yet' are both true; your agents will only be as good as your Data Cloud.

Scope19/20
Quality7/10
Enterprise platformAutomation & AgentsSalesChatPaid
Vendor
Salesforce
Origin
US — San Francisco
Pricing
Salesforce Foundations Free · Flex Credits $500 per 100k credits · Conversations $2 per conversation +3 more
Users (official only)
29,000+ Agentforce deals closed since launch; $800M Agentforce ARR, up 169% Y/Y
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em7
Embeddings
Cx8
Context
Tr9
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx7
Vector store
Rg8
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag7
Agents
Ft3
Fine-tuning
Fw9
Frameworks & harnesses
Ev7
Evaluations
Sm4
Small models
Emerging
Ma8
Multi-agent
Sy5
Synthetic data
Pc9
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Salesforce-first enterprises with clean Data Cloud foundations that want governed, observable agents inside their CRM perimeter — not mid-market teams hoping to bolt AI onto messy data at a predictable monthly cost.
Full audit passport →Visit salesforce.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“The highest performing Customer Agent” — the vendor’s own words

The most deployed customer agent and the only one whose price you can model before a sales call — $0.99 per outcome is honest, but budget on the 38–72% resolution independent tests find, not the 76% headline.

Scope18/20
Quality7/10
SMB toolSupport AgentsChatAutomation & AgentsSalesFreemium
Vendor
Fin (formerly Intercom; Salesforce acquisition pending)
Origin
US — San Francisco
Pricing
~€0.85 ($0.99) per outcome (resolution, procedure handoff, disqualification); ~€8.60 ($9.99) per sales qualification. Standalone on other helpdesks: $49/month base incl. 50 outcomes. Inside Intercom: seats from ~€25 ($29)/seat/month on top. Copilot $35/seat. Fin Voice custom. 14-day free trial, unlimited outcomes
Users (official only)
12,000+ customers, 2 million weekly resolutions, 76% average resolution rate (vendor-reported)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg9
RAG
Gr8
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft3
Fine-tuning
Fw8
Frameworks & harnesses
Ev8
Evaluations
Sm6
Small models
Emerging
Ma6
Multi-agent
Sy5
Synthetic data
Pc7
Protocols
In5
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already on Intercom, or any support team with 300+ monthly conversations and a well-maintained help center that wants a modellable per-outcome price and same-week deployment instead of an enterprise sales cycle.
Full audit passport →Fin vs Decagon vs Sierra →Visit fin.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“The collaborative AI powering lawyers to review and research faster, draft smarter, and advise with precision” — the vendor’s own words

Best-in-class tabular review and 12-jurisdiction research on an $866M war chest — but there is no eval surface, an independent rubric ranks it last of the big three, and the price only exists behind a sales call.

Scope17/20
Quality6/10
SpecialistResearchAutomation & AgentsProductivityPaid
Vendor
Legora (formerly Leya)
Origin
US — New York
Pricing
Not published — sales-led. Third-party estimates ~€260–740/user/month (reported $300–800), ~10-seat minimum ≈ €26k+/year floor; Agent Pro moved to consumption-based pricing 06/2026
Users (official only)
100,000+ lawyers at 1,500+ law firms and in-house legal teams across 50+ markets
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx8
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg9
RAG
Gr8
Guardrails
Mm5
Multimodal
Deployment
Ag8
Agents
Ft3
Fine-tuning
Fw8
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Mid-size firms (10–100 lawyers) doing high-volume, multilingual or cross-border document review — you get the best review surface in legal AI at a reported third of Harvey's price, if you accept a sales-led buy and verify output yourself.
Full audit passport →Visit legora.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“AI that's built for the way your business actually works” — the vendor’s own words

Airtable's agent stack is the deepest in this category outside a frontier lab - Field Agents, Omni, an official MCP server, and a new multi-agent Superagent - but three separate billing meters (seats, records, automation runs) mean the real cost shows up well after the $20/seat sticker price.

Scope13/20
Quality6/10
Enterprise platformAutomation & AgentsProductivityResearchFreemium
Vendor
Airtable (Formagrid, Inc.)
Origin
US — San Francisco
Pricing
Free $0 · Team $20/user/mo · Business $45/user/mo +1 more
Users (official only)
500,000+ organizations, including 80% of the Fortune 100
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg6
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Operations, marketing and product teams that want AI acting directly inside a real structured database rather than a chat sidebar, and who are disciplined about tracking seats, records and automation runs against their plan.
Full audit passport →Visit airtable.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“The CRM that does the work” — the vendor’s own words

The best dialer-CRM for call-heavy SMB teams now ships a real AI layer — Chloe plus an official MCP server punch above the category — but the sticker price is not the real price.

Scope12/20
Quality6/10
SMB toolSalesAutomation & AgentsProductivityPaid
Vendor
Close
Origin
US — Austin, TX (100% remote team)
Pricing
Solo $9/mo · Essentials $35/mo · Growth $99/mo +1 more
Users (official only)
11,500+ sales teams
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr7
Tracing
Lg5
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Inside-sales teams of roughly 2–30 reps living on the phone who want CRM, dialer and an AI teammate on one bill — and who budget the usage costs, not the sticker price.
Full audit passport →Visit close.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-31 · table v1.0
“AI website builder for designers and teams” — the vendor’s own words

The AI agent genuinely edits a live, publishable site instead of spitting out a mockup — rare in this category — but there's no open MCP, no code export, and the $10 headline creeps fast once seats and credit tiers stack up.

Scope14/20
Quality5/10
SMB toolWebsite BuildersAutomation & AgentsProductivityMarketingFreemium
Vendor
Framer B.V.
Origin
NL — Amsterdam
Pricing
Free $0 · Basic $10/mo · Pro $30/mo +1 more
Users (official only)
Over 500,000 monthly active users (self-reported at Series D)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg8
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg4
RAG
Gr6
Guardrails
Mm5
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw4
Frameworks & harnesses
Ev4
Evaluations
Sm5
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Design-led teams and startups that want an AI agent editing a real, publishable site rather than generating a mockup, and who are comfortable with vendor lock-in (no code export) in exchange for design quality and speed.
Full audit passport →v0 vs Framer vs Webflow →Visit framer.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-30 · table v1.0
“Glean is the Work AI platform connected to your enterprise's data. Find, create, and automate anything.” — the vendor’s own words

$300M ARR in 15 months and the only enterprise platform shipping both MCP and A2A — but nobody publishes a price, buyers report six figures past 500 seats, and the same connector breadth that makes it powerful is, per independent security researchers, an exfiltration channel standard DLP can't see.

Scope18/20
Quality8/10
Enterprise platformSearchProductivityAutomation & AgentsPaid
Vendor
Glean Technologies, Inc.
Origin
US — Palo Alto
Pricing
Glean Enterprise (legacy seats) ~$50–75/user/mo (buyer-reported) · Glean Enterprise Flex Usage-based FlexCredits · Enterprise Custom
Users (official only)
$300M ARR, up from $100M ~15 months earlier; Fortune 500 customer count 'nearly doubled' year over year; 85%+ of customers use Glean across 5+ departments
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em8
Embeddings
Cx9
Context
Tr9
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx8
Vector store
Rg10
RAG
Gr8
Guardrails
Mm7
Multimodal
Deployment
Ag9
Agents
Ft3
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm8
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc10
Protocols
In
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large enterprises (500+ employees) with knowledge scattered across dozens of SaaS tools and the IT capacity to run a multi-week connector rollout, who want one governed layer for search, assistant and agents instead of point solutions per department.
Full audit passport →Visit glean.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-30 · table v1.0
“The AI concierge for every customer” — the vendor’s own words

80% of Decagon's own inference now runs on models it fine-tuned in-house, and case studies show real actions, not just deflection — one OpenAI-cited customer resolves 91% of tickets with zero human involvement. But it's a $4.5B valuation against a third-party estimate of $35M ARR, roughly 130x, and reviewers keep hitting the same wall: nobody can tell you why the agent did what it did.

Scope18/20
Quality7/10
SpecialistSupport AgentsAutomation & AgentsChatVoice & SpeechPaid
Vendor
Decagon AI, Inc.
Origin
US — San Francisco
Pricing
Platform fee (est.) ~$50,000/yr · Usage (est.) ~$0.99/conversation or ~$0.50/resolution · Median enterprise ACV (est.) ~$386,000/yr
Users (official only)
10M+ customers served, 80% deflection rate (undated, Decagon's own site); separately, 100+ new global enterprise customers added in fiscal year 2025
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em
Embeddings
Cx8
Context
Tr7
Tracing
Lg9
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag9
Agents
Ft7
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm9
Small models
Emerging
Ma5
Multi-agent
Sy4
Synthetic data
Pc6
Protocols
In2
Interpretability
Th4
Thinking models
Tap or hover any element to see why it got that score.
Best for: Technical teams at high-volume support operations (15K+ tickets/month) already on Salesforce, Zendesk or Intercom, who want workflow control over chat, voice and email resolution and can resource a sales-led, multi-week implementation rather than a self-serve setup.
Full audit passport →Fin vs Decagon vs Sierra →Visit decagon.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-29 · table v1.0
“AI agent platform for customer experience” — the vendor’s own words

Outcome pricing took Sierra from $100M to $200M ARR in a year, but the receipt is a black box: no rate card, a ~€130K+ platform floor, and roughly €1.30 per resolved conversation negotiated behind an NDA.

Scope18/20
Quality7/10
Enterprise platformSupport AgentsAutomation & AgentsChatSalesFreemium
Vendor
Sierra
Origin
US — San Francisco
Pricing
Annual platform minimum ~$150,000-750,000+/yr · Implementation/setup ~$50,000-200,000 · Per-resolution fee ~$1.00-2.50 (reported ~$1.50 typical)
Users (official only)
Serves over 40% of the Fortune 50 and one in three of the world's largest banks; 30%+ of customers have annual revenue over $10B, 50%+ over $1B.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx9
Context
Tr9
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx6
Vector store
Rg9
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag10
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev10
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In4
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Large consumer or B2B enterprises (Fortune 500-scale, high ticket volume) that can absorb a six-figure sales cycle and want an AI layer bolted onto existing CX infrastructure rather than a full helpdesk replacement.
Full audit passport →Fin vs Decagon vs Sierra →Visit sierra.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-29 · table v1.0
“AI software for legal and professional services” — the vendor’s own words

Harvey won legal-AI distribution — 100,000+ lawyers, majority of the AmLaw 100 — but a rival's own published benchmark has it scoring below raw Claude on answer quality; you're paying ~€1,050+/seat/month for the workflow and security wrapper, not the smartest model in the room.

Scope18/20
Quality7/10
Enterprise platformResearchAutomation & AgentsProductivityPaid
Vendor
Harvey (Counsel AI Corporation)
Origin
US — San Francisco
Pricing
Small/specialized firm (25-50 attorneys) $1,500-2,000+/user/mo · Mid-market (50-200 attorneys) $1,200-1,500/user/mo · AmLaw 100 (200+ attorneys) $100-200/user/mo
Users (official only)
100,000+ lawyers across 1,300+ organizations in 60+ countries; majority of the AmLaw 100, 500+ in-house legal teams, 50 asset management firms.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx8
Context
Tr8
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx5
Vector store
Rg9
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag9
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev5
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Mid-to-large law firms and corporate legal departments doing high-volume US federal/major-state document work, with the budget and governance discipline to fact-check every citation before it reaches a filing.
Full audit passport →Visit harvey.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-28 · table v1.0
“The CRM for agentic revenue.” — the vendor’s own words

The first CRM to ship a native, write-gated MCP server — Universal Context and always-on revenue agents make it the sharpest AI-native CRM under €70/seat, if you budget for the credit meter.

Scope15/20
Quality6/10
SMB toolSalesAutomation & AgentsProductivityFreemium
Vendor
Attio
Origin
UK — London
Pricing
Free $0 · Plus $29/user/mo annual ($36 monthly) · Pro $69/user/mo annual ($86 monthly) +1 more
Users (official only)
30,000+ customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em4
Embeddings
Cx9
Context
Tr7
Tracing
Lg5
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg7
RAG
Gr7
Guardrails
Mm5
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: SaaS startups and PLG teams that want AI baked into a flexible, Notion-like data model and are already comfortable wiring Claude or ChatGPT into their CRM via MCP — less suited to teams that want a simple flat-rate seat price with no credit-metering surprises.
Full audit passport →Visit attio.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-28 · table v1.0
“Europe's most trusted B2B data for growing pipeline.” — the vendor’s own words

The most GDPR-defensible phone-verified contact database in Europe — the AI on top researches and summarises, it doesn't act, and there's no native protocol for agents to plug into it.

Scope14/20
Quality4/10
Enterprise platformSalesMarketingResearchPaid
Vendor
Cognism
Origin
UK — London
Pricing
Standard Custom · Pro Custom · CRM Enrichment Custom add-on +1 more
Users (official only)
4,000+ revenue teams / companies worldwide
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr4
Prompts
Em4
Embeddings
Cx6
Context
Tr4
Tracing
Lg4
LLM
Compositions
Fc3
Function calling
Vx
Vector store
Rg5
RAG
Gr8
Guardrails
Mm
Multimodal
Deployment
Ag3
Agents
Ft
Fine-tuning
Fw3
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In5
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: European and UK-focused outbound teams (5+ sellers) where phone-verified mobiles and defensible GDPR/CCPA sourcing matter more than agentic automation — pair it with a separate sequencing or agent tool if you want the AI to act, not just inform.
Full audit passport →Visit cognism.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-27 · table v1.0
“Your All-in-One AI Workspace” — the vendor’s own words

2 million monthly active users, a $2.6B valuation reached in 14 months, and 6,000+ business clients say the multi-agent bet is working. Orchestrating 80+ frontier models plus 700+ MCP integrations is genuinely broader than anything else in this category. But credits burn unpredictably across chat, slides, image, video and phone calls sharing one pool, and independent testing puts phone-call task success around 80%, not the effortless demo.

Scope15/20
Quality7/10
SMB toolAutomation & AgentsProductivityResearchChatFreemium
Vendor
Genspark (MainFunc, Inc.)
Origin
US — Palo Alto, CA
Pricing
Free $0/mo · Plus from $24.99/mo ($19.99/mo annual) · Pro from $249.99/mo ($199.99/mo annual) +2 more
Users (official only)
More than 2 million monthly active users
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx9
Context
Tr5
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg7
RAG
Gr6
Guardrails
Mm9
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev3
Evaluations
Sm8
Small models
Emerging
Ma9
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and small teams who want one subscription across research, documents, design and real-world tasks like phone calls, and who can tolerate variable credit costs in exchange for breadth.
Full audit passport →Visit genspark.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-27 · table v1.0
“v0 is an AI agent that helps anyone create real code and full-stack apps and agents.” — the vendor’s own words

v0's own benchmark: 93.9% error-free code generation vs 78.4% for raw Claude Opus on the same task. That's a real edge for Next.js scaffolding, not a vibe-coding gimmick. But it's still credit-metered pricing with no hard spending cap, and chat history counts as input tokens on every turn, so cost per feature climbs as a session grows.

Scope14/20
Quality7/10
SMB toolWebsite BuildersCodingAutomation & AgentsProductivityFreemium
Vendor
Vercel Inc.
Origin
US — San Francisco, CA
Pricing
Free $0/mo · Plus $30/user/mo · Business $100/user/mo +1 more
Users (official only)
More than 4 million people have used v0 to build apps
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx9
Context
Tr5
Tracing
Lg9
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm7
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already building on Next.js/Vercel who want a screenshot-to-production pipeline with real deploy-time security guardrails, not just a demo generator.
Full audit passport →v0 vs Framer vs Webflow →Visit vercel.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-26 · table v1.0
“AI-powered camera control and one-click video creation for creators, brands, agencies and marketing teams” — the vendor’s own words

30+ video and image models behind one MCP connector is genuinely useful plumbing, wrapped in a company Forbes caught passing off stock footage as AI output and distributing non-consensual celebrity deepfakes -- the tech works, the trust doesn't.

Scope15/20
Quality6/10
SpecialistVideo GenerationImage GenerationAutomation & AgentsFreemium
Vendor
Higgsfield AI, Inc.
Origin
US — San Francisco
Pricing
Free $0 · Starter $19/mo · Plus $47-49/mo +3 more
Users (official only)
$200M annualized revenue run rate (company-reported); a later Forbes investigation cited the company's own internal claim of $300M ARR across 300,000 paying users by February 2026, though that figure was never issued as a clean press release
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr5
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg3
RAG
Gr2
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft6
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams that want the broadest model selection for social/ad video production and can tolerate a volatile, controversy-prone vendor -- budget for account disruption, not just credits.
Full audit passport →Visit higgsfield.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-26 · table v1.0
“AI-powered vibe-coding platform that builds fully-functional apps and websites from plain-language descriptions -- frontend, backend, database, auth and payments included” — the vendor’s own words

The fastest sentence-to-working-app path on the market -- $100M ARR nine months after Wix's $80M cash-out proves the demand -- but a critical SSO-bypass disclosure and rising complaints about support and pricing since the acquisition are the tax on that speed.

Scope14/20
Quality5/10
SMB toolApp BuildersCodingAutomation & AgentsProductivityFreemium
Vendor
Base44 (owned by Wix.com Ltd.)
Origin
Israel — Tel Aviv
Pricing
Free $0 · Starter $16/mo · Builder $40/mo +3 more
Users (official only)
2 million users (Wix Q3 2025 earnings, Nov 2025); $100M annualized recurring revenue (Wix Q4 2025 earnings, Mar 2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx6
Context
Tr3
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg3
RAG
Gr3
Guardrails
Mm5
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev3
Evaluations
Sm5
Small models
Emerging
Ma2
Multi-agent
Sy
Synthetic data
Pc5
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Non-developers and small teams who need a working internal tool or MVP fast and can tolerate a still-evolving security posture and a support experience that's gotten slower since the Wix acquisition.
Full audit passport →Visit base44.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-25 · table v1.0
“Speak human to every customer” — the vendor’s own words

Real infrastructure behind the $50M raise and the Amazon Ring deployment — but budget 3-7x the advertised $0.05/minute once STT, LLM, TTS and telephony all send separate invoices.

Scope18/20
Quality6/10
SpecialistVoice AgentsAutomation & AgentsVoice & SpeechSalesPaid
Vendor
Vapi, Inc.
Origin
US — San Francisco
Pricing
Build $0.05/min · Scale Custom
Users (official only)
1M+ developers, 2.7M+ unique agents created, 1B+ calls processed
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em4
Embeddings
Cx7
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc9
Function calling
Vx3
Vector store
Rg6
RAG
Gr6
Guardrails
Mm5
Multimodal
Deployment
Ag9
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev7
Evaluations
Sm6
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Engineering teams that want full programmatic control over a production voice-agent stack and are willing to own a multi-vendor billing and ops surface to get it.
Full audit passport →Bland vs Retell vs Vapi →Visit vapi.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-25 · table v1.0
“The AI creative platform” — the vendor’s own words

One subscription, 30+ third-party image/video/audio models and a genuine MCP layer on top — real breadth for around €13-30/month, but 'unlimited' only covers a rotating subset of models and Reddit's complaint pattern is thin support once something breaks.

Scope14/20
Quality6/10
SpecialistImage GenerationVideo GenerationAutomation & AgentsMarketingFreemium
Vendor
Freepik Company S.L.U. (rebranded product name: Magnific)
Origin
Spain — Málaga
Pricing
Premium $20/mo · Premium+ $45/mo · Business $69/mo +2 more
Users (official only)
1M+ paying subscribers, $230M ARR, 290+ enterprise clients (BBC, Puma, Amazon Prime Video), ~4M images generated daily; ranked #1 generative AI web company in Europe by users (Andreessen Horowitz)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr5
Tracing
Lg3
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm10
Multimodal
Deployment
Ag7
Agents
Ft6
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm6
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Marketing and creative teams that want every current image/video model plus a stock library under one subscription, and don't need enterprise-grade evaluation, audit trail or interpretability tooling.
Full audit passport →Midjourney vs Ideogram vs Recraft vs Freepik →Visit freepik.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“Build AI agents for work” — the vendor’s own words

The no-code agent builder non-engineers actually keep using: every model included, deep MCP, a real audit trail. You price your workflows, not your seats, and the credit meter is where the surprises live.

Scope15/20
Quality7/10
SMB toolAutomation & AgentsProductivityResearchFreemium
Vendor
Gumloop, Inc. (formerly AgentHub)
Origin
US — San Francisco
Pricing
Pro $37/mo · Enterprise Custom
Users (official only)
Not publicly disclosed (private WAU chart shown, no figure); ~24 employees as of Mar 2026. Named customers: Shopify, Ramp, Gusto, Samsara, Instacart, Opendoor.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg7
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev6
Evaluations
Sm7
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Marketing, sales, ops and support teams that want to automate AI-heavy, multi-step work without waiting on engineering — and are willing to watch the credit meter as workflows scale.
Full audit passport →Visit gumloop.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“#1 AI OS for revenue teams” — the vendor’s own words

The best-governed AI platform in sales — ISO 42001, scoped agents, full audit trails — sold behind a mandatory platform fee that makes it irrational under about 30 reps.

Scope16/20
Quality7/10
Enterprise platformSalesAutomation & AgentsVoice & SpeechProductivityPaid
Vendor
Gong.io
Origin
US — San Francisco (parent Gong.io Ltd. is Israel-based; personal data is processed in the US, Europe and Israel)
Pricing
Sales-led Custom
Users (official only)
More than 4,000 companies (most recent official figure; not restated since)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx9
Context
Tr8
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg9
RAG
Gr9
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft4
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In6
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Revenue organisations of roughly 30+ reps running complex, high-value deals with a RevOps or enablement function to drive adoption — and specifically for regulated or EU-based buyers who need AI governance they can hand to a DPO. Below 20 reps the platform fee makes the arithmetic indefensible.
Full audit passport →Visit gong.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“The open model for visual intelligence” — the vendor’s own words

Still the model that renders text correctly — and with 4.0 it added open weights, a native MCP server and a $0.03/image API. Buy it for on-brand posters, logos and packaging copy; skip it for video, vectors or anything that needs run-level observability.

Scope11/20
Quality7/10
SpecialistImage GenerationMarketingFreemium
Vendor
Ideogram AI Inc.
Origin
Canada — Toronto
Pricing
Free · Plus $20/mo · Pro $60/mo +2 more
Users (official only)
Not disclosed recently; independent trackers reported ~1.2M monthly active and ~150k paying subscribers (~$20M ARR) as of Q3 2024
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr4
Tracing
Lg
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg
RAG
Gr6
Guardrails
Mm7
Multimodal
Deployment
Ag4
Agents
Ft8
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm8
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Designers, marketers and automation builders who need brand-consistent logos, posters, packaging and ad creative with legible in-image text at a predictable per-image cost — especially agent pipelines that want image generation over MCP, or enterprises that need to fine-tune and run the model behind their own firewall.
Full audit passport →Midjourney vs Ideogram vs Recraft vs Freepik →Visit ideogram.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“Build apps and sites with AI” — the vendor’s own words

The fastest way for a non-engineer to ship a working app, and it deploys where it is built. But effort-based billing has no default cap, and July 2025 proved the agent could wipe a production database during a code freeze.

Scope16/20
Quality7/10
SMB toolApp BuildersCodingAutomation & AgentsProductivityFreemium
Vendor
Replit, Inc.
Origin
US — San Francisco
Pricing
Starter Free · Replit Core $25/mo · Replit Pro $100/mo +1 more
Users (official only)
50M+ registered users; 150,000 paying customers; users at 85% of the Fortune 500
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx7
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev4
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo builders, PMs and small teams who want to prototype and ship full-stack apps or internal tools fast, in the browser — provided they keep humans in the loop on anything production and set spending alerts.
Full audit passport →Visit replit.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“AI sales platform for revenue growth” — the vendor’s own words

The cheapest way to put a 230M-contact database, a sequencer and a first-class MCP server in one seat — but the data is roughly 65-70% accurate in practice, and worse in Europe.

Scope15/20
Quality6/10
SMB toolSalesAutomation & AgentsMarketingResearchFreemium
Vendor
Apollo.io
Origin
US — San Francisco
Pricing
Free · Basic $59/mo · Professional $99/mo +2 more
Users (official only)
Millions of sellers at over 600,000 companies; 230M+ verified B2B contacts
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em4
Embeddings
Cx7
Context
Tr6
Tracing
Lg5
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm5
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In3
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: US-focused small and mid-market teams that want database, sequencer and dialer in one cheap seat and will verify contacts before sending. EU-targeting teams should price in a second, GDPR-native data source rather than trusting Apollo's European records.
Full audit passport →Visit apollo.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“Next-Gen AI video & image generator” — the vendor’s own words

The revenue and quality leader in AI video — Kling 3.0 does native 4K, 15-second multi-shot clips with synced audio and standout character consistency. But it's a walled media generator: no official MCP, a thin developer harness, and moderation strict enough to flag innocent prompts.

Scope11/20
Quality5/10
SpecialistVideo GenerationImage GenerationMarketingFreemium
Vendor
Kuaishou Technology (Kling AI)
Origin
China — Beijing
Pricing
Basic Free · Standard $10/mo · Pro $37/mo +2 more
Users (official only)
60M+ registered creators, ~12M monthly active users, 30,000+ enterprise/API clients, 600M+ videos generated; ARR ~$240M (company-disclosed Dec 2025), ~$500M (Sacra estimate, May 2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr5
Tracing
Lg
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr6
Guardrails
Mm9
Multimodal
Deployment
Ag3
Agents
Ft4
Fine-tuning
Fw6
Frameworks & harnesses
Ev3
Evaluations
Sm7
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Creators, marketers and e-commerce or short-drama studios who need frontier-quality video with strong character consistency and synced audio at a low per-second cost, and who can work within a 15-second cap, strict content filters, and an API-only developer surface with no native agent integration.
Full audit passport →Runway vs Kling vs Luma vs Sora →Visit klingai.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-24 · table v1.0
“Transform text and images into videos” — the vendor’s own words

The video quality is real. The shutdown date is realer: the API dies 24/09/2026, 62 days after this audit. Don't build here.

Scope10/20
Quality5/10
SpecialistVideo GenerationSocial MediaMarketingPaid
Vendor
OpenAI
Origin
US — San Francisco
Pricing
Free · Go $8/mo · Plus $20/mo +3 more
Users (official only)
Not disclosed. The consumer app that carried Sora's user base was discontinued on 26/04/2026, so no current figure exists.
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr7
Guardrails
Mm9
Multimodal
Deployment
Ag
Agents
Ft3
Fine-tuning
Fw3
Frameworks & harnesses
Ev3
Evaluations
Sm7
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Nobody starting today. This passport exists as a warning, not a recommendation. If you already have Sora in production, use the remaining window for final backfills and export your library — OpenAI says data will be permanently deleted after any final export window.
Full audit passport →Runway vs Kling vs Luma vs Sora →Visit openai.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-23 · table v1.0
“Build systems to grow revenue” — the vendor’s own words

The strongest AI-native GTM platform in the market, but the entry ticket is €170/month and the dual-credit meter punishes teams who skip the free test runs.

Scope15/20
Quality7/10
SMB toolAutomation & AgentsResearchMarketingFreemium
Vendor
Clay Labs, Inc.
Origin
US — New York
Pricing
Free · Launch $185/mo · Growth $495/mo +1 more
Users (official only)
10,000+ customers (incl. Anthropic, Intercom, Notion); past 1 billion Claygent agent runs
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm5
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev7
Evaluations
Sm7
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: GTM engineering teams of 2-5 who run multi-provider enrichment waterfalls and custom AI research plays, and who will actually use the free test loop before deploying at scale.
Full audit passport →Visit clay.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-23 · table v1.0
“AI notetaking out of this world” — the vendor’s own words

The most generous free AI notetaker there is, and the MCP hook into ChatGPT and Claude is the smartest thing it shipped in 2026, but everything advanced still gates behind paid and your data lives in the US.

Scope9/20
Quality6/10
SpecialistMeeting NotetakersProductivityResearchVoice & SpeechFreemium
Vendor
Fathom Video Inc.
Origin
US — San Francisco
Pricing
Free · Team $19/mo · Premium $20/mo +2 more
Users (official only)
Vendor states use at 300,000+ companies; specific active-user count not disclosed
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em4
Embeddings
Cx6
Context
Tr5
Tracing
Lg6
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg7
RAG
Gr5
Guardrails
Mm7
Multimodal
Deployment
Ag
Agents
Ft
Fine-tuning
Fw
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Individuals and small teams who live in Zoom, Google Meet or Teams and want zero-friction meeting notes, searchable history, and their transcripts available inside ChatGPT or Claude without paying up front.
Full audit passport →Granola vs Fathom vs Otter vs Fireflies vs tl;dv →Visit fathom.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-22 · table v1.0
“The truth-seeking AI assistant” — the vendor’s own words

Frontier tokens at a tenth of peer prices with the widest modality spread after Google — and a guardrails record that forces you to bring your own safety layer.

Scope18/20
Quality6/10
Frontier labChatAutomation & AgentsSearchVideo GenerationFreemium
Vendor
xAI (SpaceXAI) — wholly owned SpaceX subsidiary since 02/02/2026
Origin
US — Palo Alto
Pricing
Free · SuperGrok Lite $10/mo · SuperGrok $30/mo +1 more
Users (official only)
117M monthly active users of Grok AI features (of 550M combined Grok+X MAU), per SpaceX's IPO filing
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em4
Embeddings
Cx8
Context
Tr5
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx7
Vector store
Rg7
RAG
Gr2
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft2
Fine-tuning
Fw7
Frameworks & harnesses
Ev2
Evaluations
Sm8
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Cost-sensitive builders of high-volume agents, search and multimodal workloads who bring their own safety layer and eval harness — not for customer-facing use without one, and not for regulated EU contexts without a DSA review.
Full audit passport →Visit x.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-22 · table v1.0
“Don't type, just speak” — the vendor’s own words

Category-best AI cleanup on a cloud-only pipeline with a documented reliability problem — 75+ logged outages in six months, so trial it two weeks past payment before you standardise on it.

Scope8/20
Quality6/10
SpecialistVoice & SpeechProductivityFreemium
Vendor
Wispr AI
Origin
US — San Francisco
Pricing
Flow Basic Free · Flow Pro $15/mo · Flow Enterprise Custom
Users (official only)
Not disclosed — no official user count; company cited a 19% paid-conversion rate at the June 2025 Series A. Last official financial marker: $700M valuation at the November 2025 $25M raise ($81M total funding).
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx8
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr5
Guardrails
Mm8
Multimodal
Deployment
Ag
Agents
Ft
Fine-tuning
Fw
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Dictation-heavy workers on multiple platforms — email, chat, AI prompting — who accept cloud processing and will switch Privacy Mode on from day one. Not for offline work, spotty networks, or screens showing confidential content.
Full audit passport →Visit wisprflow.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-21 · table v1.0
“AI for designers, creatives, teams” — the vendor’s own words

Buy it for production SVG vectors, brand styles and a €0.03-per-image API — the one image platform that ships a real MCP server; skip it for video-first or document-grounded work.

Scope11/20
Quality7/10
SpecialistImage GenerationVideo GenerationMarketingFreemium
Vendor
Recraft, Inc.
Origin
US — San Francisco
Pricing
Free · Basic $12/mo · Pro $20/mo +2 more
Users (official only)
4M+ users
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg
RAG
Gr7
Guardrails
Mm9
Multimodal
Deployment
Ag7
Agents
Ft6
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm8
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Designers, marketers and automation builders who need brand-consistent logos, icons, vectors and ad creative at predictable per-image cost — especially agent pipelines that want image generation over MCP instead of another proprietary SDK.
Full audit passport →Midjourney vs Ideogram vs Recraft vs Freepik →Visit recraft.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-21 · table v1.0
“The ultimate AI executive assistant” — the vendor’s own words

The easiest on-ramp for SMB email, meeting and back-office agents — real evals and 2,500+ integrations, but zero MCP support and a default model that follows Lindy's cost curve, not yours.

Scope17/20
Quality7/10
SMB toolAutomation & AgentsProductivityChatFreemium
Vendor
Lindy AI, Inc.
Origin
US — San Francisco
Pricing
Plus $49.99/mo · Pro $99.99/mo · Max $199.99/mo +2 more
Users (official only)
400K+ users (vendor-stated)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx8
Context
Tr7
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg7
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev7
Evaluations
Sm6
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo operators and small teams who want one agent running inbox triage, scheduling, meeting notes and follow-ups in an afternoon — not builders who need MCP, model pinning or complex branching logic.
Full audit passport →Visit lindy.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-20 · table v1.0
“The AI notepad for back-to-back meetings” — the vendor’s own words

The best bot-free notetaker is now a context layer — MCP on every plan is the move that matters. Budget for the 25-note free cap and Enterprise-gated privacy controls.

Scope13/20
Quality7/10
SMB toolMeeting NotetakersProductivityChatAutomation & AgentsFreemium
Vendor
Granola Inc.
Origin
GB — London
Pricing
Basic Free · Business $14/mo · Enterprise $35/mo
Users (official only)
Not disclosed (no public user, revenue or retention figures as of the 03/2026 Series C)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx9
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg8
RAG
Gr7
Guardrails
Mm7
Multimodal
Deployment
Ag4
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo operators and small teams on Mac/Windows/iOS who type during calls and want bot-free notes that feed Claude/ChatGPT via MCP — not for large multi-speaker meetings or privacy-strict teams below the Enterprise tier.
Full audit passport →Granola vs Fathom vs Otter vs Fireflies vs tl;dv →Visit granola.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-20 · table v1.0
“World's most powerful creative AI suite” — the vendor’s own words

64+ models, a sub-50ms realtime canvas and an open-weights in-house foundation model for $9/month — the best value in creative AI, if you accept opaque compute units and Discord-only support.

Scope12/20
Quality7/10
SpecialistImage GenerationVideo GenerationFreemium
Vendor
Krea, Inc.
Origin
US — San Francisco
Pricing
Free · Basic $8.75/mo · Pro $35/mo +3 more
Users (official only)
40M users claimed on the official pricing page ('Trusted by 40M'); 20M stated at the April 2025 Series B
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx7
Context
Tr5
Tracing
Lg
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg
RAG
Gr5
Guardrails
Mm10
Multimodal
Deployment
Ag6
Agents
Ft9
Fine-tuning
Fw8
Frameworks & harnesses
Ev3
Evaluations
Sm8
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Designers, architects and small studios that want every frontier image and video model plus brand-locked LoRAs in one €8–32/month subscription — not for teams needing audit controls, eval tooling or vector output.
Full audit passport →Visit krea.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-19 · table v1.0
“Orchestrate intelligent agents for marketing workflows” — the vendor’s own words

The best brand-governance harness in marketing AI, but at $69/seat you buy the harness, not a better model: teams of 5+ win, solo marketers keep their €20 chatbot.

Scope13/20
Quality6/10
SpecialistMarketingCopywritingAutomation & AgentsFreemium
Vendor
Jasper AI, Inc.
Origin
US — Austin
Pricing
Pro $69/mo · Business Custom
Users (official only)
100,000+ businesses
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx8
Context
Tr5
Tracing
Lg7
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg7
RAG
Gr7
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Marketing teams of 5+ producing brand-sensitive content at volume, where governance and consistency are the daily bottleneck. Solo marketers and small teams get better prose per euro from a frontier chatbot at €20/month.
Full audit passport →Visit jasper.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-19 · table v1.0
“Your AI meeting agent” — the vendor’s own words

Capable cross-channel meeting intelligence with the most hostile onboarding in the category: a 1.5/5 Trustpilot full of 'malware' reviews and EU AI Act exposure on sentiment scoring make this a governance decision before a productivity one.

Scope11/20
Quality5/10
SMB toolMeeting NotetakersProductivityAutomation & AgentsSearchFreemium
Vendor
Read AI, Inc.
Origin
US — Seattle
Pricing
Free · Pro $19.75/mo · Enterprise $29.75/mo +1 more
Users (official only)
5M+ monthly active users, ~1M new users/month, teams at 90%+ of Fortune 500 (company-reported)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em6
Embeddings
Cx7
Context
Tr4
Tracing
Lg6
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg6
RAG
Gr3
Guardrails
Mm6
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: US-based individuals and sales teams who voluntarily adopt it, want meeting notes connected to CRM and email, and accept US data residency. EU organizations, regulated industries and anyone whose counterparties did not consent to recording should treat it as a compliance risk first — the intentional-user experience is decent, the bystander experience is the problem.
Full audit passport →Visit read.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-18 · table v1.0
“Visual Suite for Everyone” — the vendor’s own words

The most AI any SMB already owns: 265M users got an agentic suite bolted onto their templates with Canva AI 2.0 in April 2026 — unmatched breadth, metered in opaque Standard/Premium/Ultra credits, and no single Magic tool beats its dedicated competitor.

Scope12/20
Quality6/10
SMB toolPresentationsImage GenerationMarketingProductivityCopywritingFreemium
Vendor
Canva Pty Ltd
Origin
AU — Sydney
Pricing
Free · Pro $18/mo · Business $25/mo +1 more
Users (official only)
265M+ monthly active users, 31M+ paying subscribers, $4B annualized revenue
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg6
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm8
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: SMBs and marketing teams already living in Canva templates who want one subscription covering design, copy, image gen and ads. Not for anyone whose bar is best-in-class output on a single modality — the dedicated tool wins each head-to-head.
Full audit passport →Gamma vs Canva Magic Studio vs Beautiful.ai →Visit canva.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-18 · table v1.0
“Edit shorts 10x faster with AI” — the vendor’s own words

Best animated captions in the category — independent testing puts accuracy near 99% vs ~95% for Opus Clip — but it's a polish layer, not a platform: the long-to-short clipper is a €12–19/mo add-on, the API is metered in minutes, and there's no MCP.

Scope10/20
Quality4/10
SpecialistVideo GenerationSocial MediaMarketingFreemium
Vendor
Submagic
Origin
FR — Paris
Pricing
Starter $19/mo · Pro $39/mo · Business + API $69/mo +1 more
Users (official only)
4M+ businesses (vendor homepage claim)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr4
Prompts
Em
Embeddings
Cx5
Context
Tr3
Tracing
Lg5
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr3
Guardrails
Mm8
Multimodal
Deployment
Ag5
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo creators, podcasters and 1–2 seat teams publishing short-form daily who already know which clip they want and need it looking finished in minutes. Pair it with a clipper like Opus Clip if your workflow starts from 60-minute recordings.
Full audit passport →Visit submagic.co Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-17 · table v1.0
“AI notetaker, transcription, insights” — the vendor’s own words

35 million users and $100M ARR on meeting transcription, now betting on a two-way MCP hub — feeding meeting data to Claude/ChatGPT and pulling from Notion, Salesforce and Gmail — but Pro's minute cap got cut 5x with no price drop, and the sales-agent feature everyone wants is Enterprise-only.

Scope13/20
Quality6/10
SMB toolMeeting NotetakersProductivityAutomation & AgentsVoice & SpeechFreemium
Vendor
Otter.ai, Inc.
Origin
US — Mountain View
Pricing
Basic Free · Pro $16.99/mo · Business $30/mo +1 more
Users (official only)
35M+ users, $100M+ ARR, 1B+ meetings processed, $1B+ customer ROI generated
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em5
Embeddings
Cx9
Context
Tr6
Tracing
Lg6
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg8
RAG
Gr6
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams running high meeting volume on Zoom/Teams/Meet who want a searchable cross-meeting knowledge base and are willing to go Business or Enterprise for admin controls, unlimited minutes and Sales Notetaker.
Full audit passport →Granola vs Fathom vs Otter vs Fireflies vs tl;dv →Visit otter.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-17 · table v1.0
“AI video clipping and editing” — the vendor’s own words

ClipAnything genuinely reads any video genre for the moment you want, and ships MCP + API access from the $29/mo tier — but it bills by source-video minutes, not output clips, so a 90-minute podcast burns your month's allowance on one export or fifteen.

Scope10/20
Quality6/10
SpecialistVideo GenerationSocial MediaAutomation & AgentsFreemium
Vendor
OpusClip Inc.
Origin
US — Mountain View
Pricing
Free · Starter $15/mo · Pro $29/mo +1 more
Users (official only)
16M+ creators and businesses, 440M+ clips created, 160B+ total views from posted clips
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx6
Context
Tr4
Tracing
Lg6
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr4
Guardrails
Mm9
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Creators and media teams publishing daily from long-form video (podcasts, webinars, streams) who want a fast first-draft clipper and are willing to pay a premium for the ClipAnything multimodal engine and built-in MCP/API access.
Full audit passport →Visit opus.pro Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-14 · table v1.0
“The #1 AI assistant for meetings” — the vendor’s own words

The most widely-adopted AI meeting notetaker — strong transcription, 200+ integrations and a solid MCP connector, but AI credits and per-seat pricing erode the value at team scale.

Scope12/20
Quality5/10
SMB toolMeeting NotetakersProductivityAutomation & AgentsVoice & SpeechFreemium
Vendor
Fireflies.ai Corp
Origin
US — Miami
Pricing
Free · Pro $18/mo · Business $29/mo +1 more
Users (official only)
20+ million people across 500,000+ organisations; used by 75% of Fortune 500
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr5
Prompts
Em
Embeddings
Cx6
Context
Tr5
Tracing
Lg5
LLM
Compositions
Fc4
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm6
Multimodal
Deployment
Ag6
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Sales, marketing and recruiting teams that run many video meetings and want automatic transcription, CRM logging and searchable meeting intelligence without manual note-taking.
Full audit passport →Granola vs Fathom vs Otter vs Fireflies vs tl;dv →Visit fireflies.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-14 · table v1.0
“AI-editing for every kind of video” — the vendor’s own words

The best text-based editor for spoken-word content — Underlord is a genuinely useful AI co-editor, but AI credit caps and the cloud-dependent workflow make it expensive for heavy producers.

Scope13/20
Quality5/10
SpecialistVideo GenerationVoice & SpeechProductivityFreemium
Vendor
Descript
Origin
US — San Francisco
Pricing
Free · Hobbyist $24/mo · Creator $35/mo +2 more
Users (official only)
6 million creators and teams
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx6
Context
Tr4
Tracing
Lg5
LLM
Compositions
Fc4
Function calling
Vx
Vector store
Rg5
RAG
Gr3
Guardrails
Mm8
Multimodal
Deployment
Ag6
Agents
Ft4
Fine-tuning
Fw5
Frameworks & harnesses
Ev2
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Podcasters, educators and content creators who produce speech-heavy video or audio and want to edit by text rather than timeline.
Full audit passport →Visit descript.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-14 · table v1.0
“Imagination just got a team” — the vendor’s own words

The best cinematic AI video generator on the market in 2026 — HDR pipeline and camera motion are class-leading, but it's a creative tool, not a developer platform, and the credit economics bite hard at 1080p.

Scope16/20
Quality5/10
SpecialistVideo GenerationImage GenerationMarketingFreemium
Vendor
Luma AI, Inc.
Origin
US — Palo Alto
Pricing
Plus $30/mo · Pro $90/mo · Ultra $300/mo +2 more
Users (official only)
30M+ users (reported, not officially confirmed by Luma on its own site)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg5
LLM
Compositions
Fc2
Function calling
Vx
Vector store
Rg3
RAG
Gr2
Guardrails
Mm9
Multimodal
Deployment
Ag6
Agents
Ft3
Fine-tuning
Fw7
Frameworks & harnesses
Ev2
Evaluations
Sm7
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc6
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Creative teams and marketers who need cinematic AI video with professional-grade HDR output and camera control, and who can absorb credit costs at production resolution.
Full audit passport →Runway vs Kling vs Luma vs Sora →Visit lumalabs.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-13 · table v1.0
“Less structure, more intelligence” — the vendor’s own words

The most capable general-purpose autonomous agent on the market — genuine multi-step task execution in a real sandbox VM — but opaque credit economics, unresolved Meta-acquisition turmoil and thin compliance make it a bet, not infrastructure.

Scope16/20
Quality6/10
SpecialistAutomation & AgentsResearchCodingFreemium
Vendor
Butterfly Effect Pte. Ltd. (now part of Meta, acquisition contested by Chinese regulators)
Origin
Singapore — Singapore
Pricing
Free · Pro $20/mo · Pro $40/mo +1 more
Users (official only)
Millions of active users; 2M+ waitlist within 7 days of launch; $100M ARR 8 months post-launch
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr4
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev3
Evaluations
Sm5
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc4
Protocols
In3
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Researchers, founders and operators who want an autonomous AI agent for multi-step web research, rapid prototyping and recurring admin tasks — and can absorb unpredictable credit costs and the Meta-ownership uncertainty.
Full audit passport →Visit manus.im Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-13 · table v1.0
“All-in-one AI video platform for business” — the vendor’s own words

The enterprise AI-video default — best-in-class avatar quality, compliance and LMS depth — but the AI platform surface is thin: no embeddings, no vector store, no fine-tuning, no agents.

Scope11/20
Quality5/10
SpecialistAvatar VideoVideo GenerationVoice & SpeechProductivityFreemium
Vendor
Synthesia Limited
Origin
UK — London
Pricing
Basic Free · Starter $29/mo · Creator $89/mo +1 more
Users (official only)
1M+ users, 50,000+ teams, 90%+ of Fortune 100
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr4
Tracing
Lg5
LLM
Compositions
Fc3
Function calling
Vx
Vector store
Rg6
RAG
Gr6
Guardrails
Mm8
Multimodal
Deployment
Ag4
Agents
Ft
Fine-tuning
Fw4
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Enterprise L&D, compliance training and internal communications teams that need repeatable, multilingual, governed AI video at scale — especially those with SCORM/LMS requirements and security review obligations.
Full audit passport →HeyGen vs Synthesia →Visit synthesia.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-12 · table v1.0
“The best AI for coding” — the vendor’s own words

Now Devin Desktop — the most agent-forward IDE with real enterprise compliance; you're betting on Cognition's Devin-first roadmap and living with its churn: two owners, two names and a retired agent inside one year.

Scope17/20
Quality8/10
SMB toolCodingAutomation & AgentsProductivityFreemium
Vendor
Cognition AI, Inc. (acquired Windsurf 14/07/2025; product renamed Devin Desktop 02/06/2026)
Origin
US — San Francisco
Pricing
Free · Pro $20/mo · Max $200/mo +2 more
Users (official only)
Not disclosed post-acquisition; at the July 2025 acquisition Cognition officially stated $82M ARR, 350+ enterprise customers and 'hundreds of thousands of daily active users'
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em7
Embeddings
Cx9
Context
Tr7
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx7
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm7
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Developers and teams who want maximum agent autonomy with minimal context babysitting — especially regulated environments needing FedRAMP/HIPAA/ITAR or JetBrains shops — and who accept a Devin-first roadmap over a stable standalone editor.
Full audit passport →Cursor vs Windsurf (Devin Desktop) →Visit cognition.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-12 · table v1.0
“Building real-world intelligence” — the vendor’s own words

The strongest independent AI media stack — frontier video plus a real developer platform — but it's a media specialist, not a full AI stack, and credits meter everything.

Scope14/20
Quality6/10
SpecialistVideo GenerationImage GenerationMarketingFreemium
Vendor
Runway AI, Inc.
Origin
US — New York
Pricing
Free · Standard $15/mo · Pro $35/mo +2 more
Users (official only)
50M+ creators (vendor claim on pricing page); active-user counts not independently disclosed
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx8
Context
Tr6
Tracing
Lg
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg5
RAG
Gr6
Guardrails
Mm10
Multimodal
Deployment
Ag4
Agents
Ft6
Fine-tuning
Fw8
Frameworks & harnesses
Ev3
Evaluations
Sm7
Small models
Emerging
Ma
Multi-agent
Sy7
Synthetic data
Pc3
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Creative teams, marketers and product developers who need frontier video/image/audio generation — hands-on in the Creative app or embedded via one API — and accept credit-metered costs that climb with ambition.
Full audit passport →Runway vs Kling vs Luma vs Sora →Visit runwayml.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-12 · table v1.0
“Create stunning apps & websites” — the vendor’s own words

The fastest path from prompt to working web app in the browser — but token economics break at scale and the generated code is prototype-grade, not production-ready.

Scope15/20
Quality6/10
SMB toolApp BuildersCodingAutomation & AgentsProductivityFreemium
Vendor
StackBlitz, Inc.
Origin
US — San Francisco
Pricing
Free · Pro $25/mo · Teams $30/mo +1 more
Users (official only)
4 million registered users
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx6
Context
Tr4
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr4
Guardrails
Mm5
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm6
Small models
Emerging
Ma3
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th5
Thinking models
Tap or hover any element to see why it got that score.
Best for: Rapid prototyping, hackathons, MVP validation, and landing pages — anyone who needs a working web app in hours, not months, and accepts that production hardening is a separate phase.
Full audit passport →Bolt vs Lovable →Visit bolt.new Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-12 · table v1.0
“Discover, create, and share music” — the vendor’s own words

Best-in-class audio fidelity and the only AI music tool with major-label licensing — but the 2025 download ban and ongoing licensing transition make it a gamble for any workflow that needs to export audio today.

Scope8/20
Quality4/10
SpecialistMusicVoice & SpeechFreemium
Vendor
Udio Inc.
Origin
US — New York
Pricing
Free · Standard $10/mo · Pro $30/mo
Users (official only)
Not disclosed
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx5
Context
Tr3
Tracing
Lg
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg4
RAG
Gr3
Guardrails
Mm8
Multimodal
Deployment
Ag
Agents
Ft4
Fine-tuning
Fw
Frameworks & harnesses
Ev2
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Creators who need the highest audio fidelity and vocal realism for ideation and streaming-only use, and who can tolerate iteration to get good results — not for teams that need reliable export-to-production workflows today.
Full audit passport →Suno vs Udio →Visit udio.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-11 · table v1.0
“AI research tool and thinking partner” — the vendor’s own words

The best grounded-research surface in consumer AI and generous for free — but the 2026 agentic leap is paywalled behind Ultra, and there is still no API.

Scope12/20
Quality7/10
SMB toolResearchProductivityChatFreemium
Vendor
Google
Origin
US — Mountain View
Pricing
Gemini Notebook Free · Gemini Notebook in Plus $4.99/mo · Gemini Notebook in Pro $19.99/mo +2 more
Users (official only)
Not disclosed — Google says 'millions of people and organizations' use NotebookLM
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em6
Embeddings
Cx9
Context
Tr6
Tracing
Lg9
LLM
Compositions
Fc6
Function calling
Vx5
Vector store
Rg10
RAG
Gr6
Guardrails
Mm9
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Researchers, students, analysts and small-business operators who want trustworthy answers grounded in their own documents — and are willing to live inside Google's ecosystem to get it.
Full audit passport →Visit notebooklm.google Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-11 · table v1.0
“Effortless AI design for presentations, websites” — the vendor’s own words

The fastest way for a non-designer to ship a professional deck, doc or site — and with a real API plus MCP server it is automation-ready; just watch the credit meter and the PPTX export.

Scope12/20
Quality6/10
SMB toolPresentationsProductivityMarketingSocial MediaFreemium
Vendor
Gamma Tech, Inc.
Origin
US — San Francisco
Pricing
Free · Plus $9/mo · Pro $18/mo +3 more
Users (official only)
100M users (company announcement); 70M users and $100M ARR at Series B
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx6
Context
Tr5
Tracing
Lg7
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg5
RAG
Gr5
Guardrails
Mm8
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev
Evaluations
Sm5
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Founders, marketers, educators and small teams who need polished visual content fast without a designer — and automation builders who want deck generation inside their workflows via API or MCP.
Full audit passport →Gamma vs Canva Magic Studio vs Beautiful.ai →Visit gamma.app Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“AI coding agent” — the vendor’s own words

The strongest agentic IDE money can buy — just know the sticker price is a floor on a usage meter, and its next owner is SpaceX/xAI with the deal closing Q3 2026.

Scope17/20
Quality8/10
SMB toolCodingAutomation & AgentsProductivityFreemium
Vendor
Anysphere, Inc. (acquisition by SpaceX announced 16/06/2026, expected to close Q3 2026 pending regulatory approval)
Origin
US — San Francisco
Pricing
Hobby Free · Pro $20/mo · Pro+ $60/mo +3 more
Users (official only)
User count not disclosed; vendor claims 'more than half of the Fortune 500'; $1B annualized revenue confirmed by company release 11/2025, $2B ARR by 02/2026 per Reuters/CNBC reporting
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em7
Embeddings
Cx9
Context
Tr7
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx7
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag9
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm7
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Professional developers and SMB engineering teams who live in an agentic workflow all day and will actively manage model choice and credit burn — less compelling for occasional coders (Hobby tier is evaluation-only) or regulated enterprises needing deep compliance certifications.
Full audit passport →Cursor vs Windsurf (Devin Desktop) →Visit cursor.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“AI voice generator and agents platform” — the vendor’s own words

The audio benchmark grew into a serious agents platform with a testing framework bigger vendors lack — just watch the shared credit meter, overages bite without warning.

Scope18/20
Quality7/10
SpecialistVoice & SpeechText-to-SpeechAutomation & AgentsMusicFreemium
Vendor
ElevenLabs
Origin
GB — London
Pricing
Free · Starter $6/mo · Creator $22/mo +3 more
Users (official only)
User counts not disclosed; official: $500M+ ARR (Q1 2026), 530 employees
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx7
Context
Tr8
Tracing
Lg7
LLM
Compositions
Fc8
Function calling
Vx7
Vector store
Rg8
RAG
Gr6
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft8
Fine-tuning
Fw8
Frameworks & harnesses
Ev9
Evaluations
Sm7
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams shipping voice — audiobooks, dubbing, agents for support/sales lines — who want frontier audio quality plus a tested, deployable agent stack, and who will set a calendar reminder to check credit usage before each billing cycle.
Full audit passport →Visit elevenlabs.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Vibe code apps and websites” — the vendor’s own words

The fastest idea-to-working-app path a non-coder can buy — treat every output as a first draft and budget a security review before real user data touches it.

Scope13/20
Quality6/10
SMB toolApp BuildersCodingAutomation & AgentsProductivityFreemium
Vendor
Lovable
Origin
SE — Stockholm
Pricing
Free · Pro $25/mo · Business $50/mo +1 more
Users (official only)
8M users and $500M annualized revenue run rate (company-reported June 2026, up from $400M in February); 1M+ new projects/week
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx7
Context
Tr6
Tracing
Lg7
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg6
RAG
Gr5
Guardrails
Mm6
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev4
Evaluations
Sm
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Solo founders and SMB teams validating MVPs, internal tools and demos who will harden the output (or pay someone to) before real user data arrives; not for regulated data, complex backend logic, or production apps nobody security-reviews.
Full audit passport →Bolt vs Lovable →Visit lovable.dev Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Turn your ideas into videos” — the vendor’s own words

The most complete avatar-video stack for small teams — real API, MCP and consent guardrails — as long as you budget for the credit meter, not the sticker price.

Scope14/20
Quality6/10
SpecialistAvatar VideoVideo GenerationVoice & SpeechMarketingTranslationFreemium
Vendor
HeyGen
Origin
US — Los Angeles
Pricing
Free · Creator $29/mo · Pro $49/mo +2 more
Users (official only)
30M+ users in 196 countries, 85% of the Fortune 100, $200M ARR
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx6
Context
Tr5
Tracing
Lg6
LLM
Compositions
Fc5
Function calling
Vx
Vector store
Rg6
RAG
Gr7
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft7
Fine-tuning
Fw7
Frameworks & harnesses
Ev2
Evaluations
Sm5
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: SMBs, L&D and marketing teams producing recurring multilingual spokesperson video without cameras or crews — and builders who want video generation inside agents via the API or MCP connector.
Full audit passport →HeyGen vs Synthesia →Visit heygen.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Free AI-powered answer engine” — the vendor’s own words

The best first-pass research layer for cited answers — treat citations as leads not proof, and expect quotas and models to shift under you.

Scope16/20
Quality6/10
SMB toolSearchResearchChatFreemium
Vendor
Perplexity AI
Origin
US — San Francisco
Pricing
Free · Pro $20/mo · Max $200/mo +1 more
Users (official only)
Not disclosed (third-party estimates 34–45M MAU, mid-2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em6
Embeddings
Cx7
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg7
RAG
Gr5
Guardrails
Mm7
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw5
Frameworks & harnesses
Ev3
Evaluations
Sm6
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc5
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Analysts and operators doing time-sensitive, source-verifiable research who will actually click the citations — as a retrieval layer feeding a writing assistant, not as a single source of truth.
Full audit passport →Visit perplexity.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Make any song you can imagine” — the vendor’s own words

The best consumer music generator by a wide margin — buy it for the songs, not for integration: there is no public API, and two of the three major labels are still suing.

Scope9/20
Quality6/10
SpecialistMusicVoice & SpeechFreemium
Vendor
Suno, Inc.
Origin
US — Cambridge, MA
Pricing
Free Plan · Pro Plan $8/mo · Premier Plan $24/mo
Users (official only)
2M+ paid subscribers, ~$300M ARR (company statements); '100M users' is a vendor claim without methodology
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx6
Context
Tr4
Tracing
Lg9
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr6
Guardrails
Mm8
Multimodal
Deployment
Ag
Agents
Ft8
Fine-tuning
Fw6
Frameworks & harnesses
Ev
Evaluations
Sm
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Content creators, podcasters and small businesses who need original, commercially-licensed music fast and work inside Suno's web app; not for developers needing programmatic generation or for high-stakes sync/label work until the remaining lawsuits resolve.
Full audit passport →Suno vs Udio →Visit suno.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Building the most beautiful AI models” — the vendor’s own words

Still the best-looking images in AI, sold as a walled garden — unbeatable for aesthetics, a dead end for pipelines, and carrying live copyright-lawsuit risk.

Scope9/20
Quality6/10
SpecialistImage GenerationVideo GenerationMarketingFreemium
Vendor
Midjourney, Inc.
Origin
US — San Francisco
Pricing
Basic $10/mo · Standard $30/mo · Pro $60/mo +1 more
Users (official only)
Not disclosed. CEO David Holz: revenue 'significantly above $200M' and unfunded; no official user count published
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em
Embeddings
Cx9
Context
Tr5
Tracing
Lg
LLM
Compositions
Fc
Function calling
Vx
Vector store
Rg
RAG
Gr6
Guardrails
Mm9
Multimodal
Deployment
Ag
Agents
Ft6
Fine-tuning
Fw2
Frameworks & harnesses
Ev
Evaluations
Sm7
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc1
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Designers, marketers and content teams who want the most beautiful hero images and concept art available and generate them by hand — not businesses that need API-driven pipelines, legible in-image text, or legal indemnification for commercial assets.
Full audit passport →Midjourney vs Ideogram vs Recraft vs Freepik →Visit midjourney.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-10 · table v1.0
“Unravel the mystery of AGI” — the vendor’s own words

Frontier-grade open-weight models at a tenth of US prices — but it's an engine, not a platform, and the hosted service's China data residency is a hard no for sensitive work.

Scope17/20
Quality5/10
Frontier labChatCodingResearchFreemium
Vendor
DeepSeek (Hangzhou DeepSeek Artificial Intelligence, backed by High-Flyer)
Origin
CN — Hangzhou
Pricing
DeepSeek Chat Free · deepseek-v4-flash Pay-per-token · deepseek-v4-pro Pay-per-token
Users (official only)
Not disclosed by vendor; third-party tracker AICPB estimates ~139M app MAU
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em2
Embeddings
Cx9
Context
Tr2
Tracing
Lg9
LLM
Compositions
Fc7
Function calling
Vx
Vector store
Rg4
RAG
Gr2
Guardrails
Mm3
Multimodal
Deployment
Ag6
Agents
Ft6
Fine-tuning
Fw4
Frameworks & harnesses
Ev2
Evaluations
Sm8
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc2
Protocols
In5
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Cost-sensitive builders and self-hosters who bring their own platform: run the MIT weights on EU/US infrastructure for frontier-class reasoning without the hosted-service data risk — and keep sensitive work away from the official app and API.
Full audit passport →Mistral vs DeepSeek →Visit deepseek.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-08 · table v1.0
“Frontier AI LLMs, assistants, agents” — the vendor’s own words

The EU-sovereign near-frontier stack — open weights, aggressive pricing and a real agent platform; what you trade is ecosystem depth and the absolute top end, not sovereignty.

Scope18/20
Quality7/10
Frontier labChatAutomation & AgentsCodingFreemium
Vendor
Mistral AI
Origin
FR — Paris
Pricing
Free · Education plan $5.99/mo · Pro $14.99/mo +2 more
Users (official only)
No global MAU disclosed. Company-disclosed: ~450,000 customers incl. 1,031 high-value customers (07/2025); ARR above $400M (02/2026), targeting $1B in 2026
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em7
Embeddings
Cx7
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg7
RAG
Gr7
Guardrails
Mm8
Multimodal
Deployment
Ag8
Agents
Ft8
Fine-tuning
Fw8
Frameworks & harnesses
Ev6
Evaluations
Sm9
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: EU SMBs and regulated teams that want near-frontier capability with EU data residency by default, open weights as a self-hosting exit route, and one European vendor covering chat, coding agents, document AI and custom models.
Full audit passport →Mistral vs DeepSeek →Visit mistral.ai Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-08 · table v1.0
“Industry leading, open-source AI” — the vendor’s own words

The most-deployed open weights in the West, now in managed decline — Meta's frontier work moved to proprietary Muse, the Llama API is dead, and the license was never open source.

Scope15/20
Quality5/10
Open sourceChatCodingResearchFree
Vendor
Meta Platforms (Meta Superintelligence Labs)
Origin
US — Menlo Park
Pricing
Free
Users (official only)
1.2 billion cumulative downloads (Meta-reported)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr6
Prompts
Em
Embeddings
Cx7
Context
Tr3
Tracing
Lg6
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg3
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag4
Agents
Ft8
Fine-tuning
Fw4
Frameworks & harnesses
Ev4
Evaluations
Sm6
Small models
Emerging
Ma
Multi-agent
Sy
Synthetic data
Pc3
Protocols
In
Interpretability
Th2
Thinking models
Tap or hover any element to see why it got that score.
Best for: US teams with existing, working Llama deployments that value stability over frontier progress. Anyone choosing a NEW open-weight stack in mid-2026 should shortlist actively developed alternatives (Qwen, DeepSeek, Gemma, Mistral) first — and EU companies must treat Llama 4's multimodal models as license-blocked.
Full audit passport →Visit ai.meta.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-07 · table v1.0
“Google's AI assistant” — the vendor’s own words

The broadest full-stack AI platform on the market — frontier models plus real data infrastructure at Google scale; the tax is Google Cloud complexity and a product surface that reshuffles fast.

Scope19/20
Quality9/10
Frontier labChatAutomation & AgentsVideo GenerationResearchCodingFreemium
Vendor
Google (Alphabet)
Origin
US — Mountain View
Pricing
Free · Google AI Plus $4.99/mo · Google AI Pro $19.99/mo +1 more
Users (official only)
900M+ monthly active users (Gemini app), 230 countries
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em9
Embeddings
Cx9
Context
Tr9
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx9
Vector store
Rg9
RAG
Gr9
Guardrails
Mm10
Multimodal
Deployment
Ag9
Agents
Ft9
Fine-tuning
Fw9
Frameworks & harnesses
Ev9
Evaluations
Sm9
Small models
Emerging
Ma9
Multi-agent
Sy6
Synthetic data
Pc9
Protocols
In
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Businesses already on Google Workspace or Google Cloud that want one vendor for frontier models, RAG/data infrastructure and multimodal generation — and can absorb Google's pace of product renames and previews.
Full audit passport →ChatGPT vs Claude vs Gemini →Visit gemini.google Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-07 · table v1.0
“Next generation AI assistant” — the vendor’s own words

The agent-era platform — best-in-class harness, context engineering and MCP — but you bring your own embeddings, vector store and fine-tuning.

Scope19/20
Quality7/10
Frontier labChatAutomation & AgentsCodingResearchFreemium
Vendor
Anthropic
Origin
US — San Francisco
Pricing
Free · Pro $20/mo · Max 5x $100/mo +3 more
Users (official only)
Consumer MAU not disclosed. Official: >$47B run-rate revenue (06/2026); >1,000 enterprise customers spending $1M+/yr (05/2026)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr9
Prompts
Em3
Embeddings
Cx10
Context
Tr8
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx2
Vector store
Rg7
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag10
Agents
Ft3
Fine-tuning
Fw9
Frameworks & harnesses
Ev7
Evaluations
Sm8
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc9
Protocols
In5
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams building serious agentic systems — coding, research, long-horizon knowledge work — who want the best harness and MCP ecosystem and are comfortable pairing Claude with external embedding/vector infrastructure.
Full audit passport →ChatGPT vs Claude vs Gemini →Visit anthropic.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-06 · table v1.0
“AI built for work” — the vendor’s own words

The deepest enterprise AI stack you can buy into your existing Office estate — but licensed seats are not used seats, and the ROI math only clears if you do the governance and adoption work Microsoft won't do for you.

Scope18/20
Quality7/10
Enterprise platformProductivityAutomation & AgentsChatResearchFreemium
Vendor
Microsoft
Origin
US — Redmond
Pricing
Microsoft 365 Copilot Chat Free · Microsoft 365 Copilot Business $25.20/mo · Microsoft 365 Business Standard with Copilot $28.20/mo +2 more
Users (official only)
20M+ paid M365 Copilot seats; 150M MAU across the Copilot family
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx9
Context
Tr8
Tracing
Lg9
LLM
Compositions
Fc8
Function calling
Vx6
Vector store
Rg8
RAG
Gr9
Guardrails
Mm7
Multimodal
Deployment
Ag8
Agents
Ft6
Fine-tuning
Fw8
Frameworks & harnesses
Ev7
Evaluations
Sm5
Small models
Emerging
Ma8
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th8
Thinking models
Tap or hover any element to see why it got that score.
Best for: Organizations already committed to Microsoft 365 with clean SharePoint governance and the will to run structured adoption — the more of your work that lives in the Graph, the more the $30 pays back; the messier your tenant, the more it doesn't.
Full audit passport →Visit microsoft.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-06 · table v1.0
“Meet your 24/7 AI team” — the vendor’s own words

Excellent AI glue if your work already lives in Notion — outside the workspace it reads far better than it writes, and the credits meter is the price of the agent dream.

Scope15/20
Quality6/10
SMB toolProductivityAutomation & AgentsChatResearchFreemium
Vendor
Notion Labs, Inc.
Origin
US — San Francisco
Pricing
Free · Plus $10/mo · Business $20/mo +1 more
Users (official only)
100M+ users
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em5
Embeddings
Cx8
Context
Tr6
Tracing
Lg8
LLM
Compositions
Fc6
Function calling
Vx
Vector store
Rg7
RAG
Gr7
Guardrails
Mm5
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw7
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma4
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th6
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams already running docs, projects and knowledge in Notion on a Business plan who want agents and search over their own context — not for teams whose critical data lives elsewhere or who mainly need best-in-class writing.
Full audit passport →Visit notion.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-05 · table v1.0
“Chat, work, create, code with AI” — the vendor’s own words

The deepest stack in AI and still the default — just budget for model churn and don't build on parts already marked for deprecation.

Scope18/20
Quality8/10
Frontier labChatAutomation & AgentsResearchFreemium
Vendor
OpenAI
Origin
US — San Francisco
Pricing
Free · Go $8/mo · Plus $20/mo +3 more
Users (official only)
900M weekly active users, 50M paying subscribers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr8
Prompts
Em8
Embeddings
Cx9
Context
Tr7
Tracing
Lg9
LLM
Compositions
Fc9
Function calling
Vx7
Vector store
Rg8
RAG
Gr7
Guardrails
Mm9
Multimodal
Deployment
Ag8
Agents
Ft9
Fine-tuning
Fw7
Frameworks & harnesses
Ev5
Evaluations
Sm8
Small models
Emerging
Ma7
Multi-agent
Sy
Synthetic data
Pc7
Protocols
In
Interpretability
Th9
Thinking models
Tap or hover any element to see why it got that score.
Best for: Teams that want one vendor covering the whole AI stack at frontier quality and accept vendor lock-in plus fast product churn.
Full audit passport →ChatGPT vs Claude vs Gemini →Visit openai.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-05 · table v1.0
“AI workflow automation platform” — the vendor’s own words

The deepest AI-native automation harness you can run on your own hardware — the fair-code license and steep cloud tiers are the fine print worth reading.

Scope17/20
Quality7/10
Open sourceAutomation & AgentsCodingProductivityFreemium
Vendor
n8n GmbH
Origin
DE — Berlin
Pricing
Community Edition Free · Starter €20/mo · Pro €50/mo +2 more
Users (official only)
1.7M monthly active builders; 1,400+ enterprise customers
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em7
Embeddings
Cx7
Context
Tr8
Tracing
Lg8
LLM
Compositions
Fc9
Function calling
Vx7
Vector store
Rg8
RAG
Gr8
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw9
Frameworks & harnesses
Ev8
Evaluations
Sm7
Small models
Emerging
Ma6
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: Technical SMBs and builders who want frontier-grade AI workflow depth — agents, RAG, evals, MCP — with full data control and near-zero license cost, and can spare the DevOps hours self-hosting demands.
Full audit passport →n8n vs Zapier vs Make →Visit n8n.io Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-05 · table v1.0
“The visual AI automation platform” — the vendor’s own words

The friendliest way to put AI agents to work across 3,500+ apps — as long as you watch the credit meter and don't need self-hosting or real evals.

Scope17/20
Quality6/10
SMB toolAutomation & AgentsProductivityFreemium
Vendor
Make (Celonis, Inc.)
Origin
CZ — Prague
Pricing
Free · Core $12/mo · Pro $21/mo +2 more
Users (official only)
400,000+ organizations (homepage claim); 3.1M registered users per official 2024 wrap-up
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em6
Embeddings
Cx7
Context
Tr7
Tracing
Lg8
LLM
Compositions
Fc9
Function calling
Vx6
Vector store
Rg7
RAG
Gr5
Guardrails
Mm6
Multimodal
Deployment
Ag8
Agents
Ft
Fine-tuning
Fw8
Frameworks & harnesses
Ev2
Evaluations
Sm6
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th7
Thinking models
Tap or hover any element to see why it got that score.
Best for: SMB and mid-market ops teams that want trustworthy AI agents acting across their existing SaaS stack with zero code and zero infrastructure — and are willing to pay per step for it.
Full audit passport →n8n vs Zapier vs Make →Visit make.com Prices and details change — this passport is re-verified at least quarterly.
Audited 2026-07-05 · table v1.0
“Any AI. Every tool. One system” — the vendor’s own words

The best connector in the business wearing an agent costume — per-task pricing punishes exactly the automations that work.

Scope13/20
Quality6/10
SMB toolAutomation & AgentsProductivityFreemium
Vendor
Zapier Inc.
Origin
US — San Francisco (remote-first)
Pricing
Agents Free · Agents Pro $33.33/mo · Agents Enterprise Custom
Users (official only)
3.4M+ businesses (company-reported)
All 20 element scores
Reactive
Retrieval & Memory
Orchestration
Validation
Models
Primitives
Pr7
Prompts
Em
Embeddings
Cx5
Context
Tr6
Tracing
Lg6
LLM
Compositions
Fc9
Function calling
Vx
Vector store
Rg4
RAG
Gr7
Guardrails
Mm4
Multimodal
Deployment
Ag7
Agents
Ft
Fine-tuning
Fw6
Frameworks & harnesses
Ev3
Evaluations
Sm
Small models
Emerging
Ma5
Multi-agent
Sy
Synthetic data
Pc8
Protocols
In
Interpretability
Th
Thinking models
Tap or hover any element to see why it got that score.
Best for: Non-technical SMB teams that need cross-app automation today and can live with cloud-only and per-task economics.
Full audit passport →n8n vs Zapier vs Make →Visit zapier.com Prices and details change — this passport is re-verified at least quarterly.