Skip to main content
Arthur logo

Arthur

Control plane to monitor, evaluate, secure, and govern enterprise AI and agents

Editorially reviewedChecked Sep 2026How we review
ai-governance#ai-governance#model-monitoring#ai-observability#llm-evaluation
Free plan Claimed API Self-hosted Teams
Toolglade editorial score

Our assessment — not a user rating. How we score

3.6/5
7.2/10 composite
Enterprise readiness
8.0
Compliance posture
7.0
Workflow depth
7.0
Integration surface
7.0
Transparency
7.0
Toolglade’s take

A strong, security-forward choice for enterprise ML and AI-platform teams that need to monitor, evaluate, and govern models and agents in production; too technical for non-engineering business users.

About Arthur

Arthur is an AI governance and monitoring platform offering one control plane to discover, observe, evaluate, and secure AI models and agents. It provides step-level agent observability, continuous evaluations, runtime guardrails and behavioral analytics, plus cost tracking and model rerouting. It maintains an open-source Arthur Engine for self-hosted evaluation and supports SaaS, VPC, and on-prem deployment. Plans start free, with a $60/mo Premium tier and custom Enterprise pricing.

Arthur (Arthur AI) provides a centralized control plane for trustworthy enterprise AI, spanning model monitoring, observability, evaluations, guardrails, and governance across the AI lifecycle. Its agent discovery finds every agent running across endpoints, cloud, and on-prem environments (including shadow, unregistered agents), then gives full step-level observability into each agent's reasoning, tool calls, retrieval, and handoffs so teams can see not just what an agent did but why. For runtime protection, Arthur's Agent Behavioral Analytics continuously monitors AI applications and alerts on anomalies such as prompt injection, anomalous tool sequences, and data-egress spikes, streaming findings to SOC tools like CrowdStrike Falcon, Elastic, Splunk, and Datadog. It also offers cost controls that track spend by team and can reroute requests that don't need an expensive model to cheaper or internally hosted open-source models. Arthur maintains an open-source Arthur Engine on GitHub for self-serve, self-hosted evaluation. Arthur is used by enterprise AI, security, audit, and governance teams and offers deployment as multi-tenant SaaS, single-tenant SaaS, and self-managed VPC/BYOCloud/on-prem. The honest limitation: it is a technical, ML/security-team-oriented platform (not a business-user tool), the richest deployment, SSO, SLA, and BAA options are gated to the custom-priced Enterprise tier, and the free and Premium tiers cap use cases, data retention, and volume.

Weighing your options?See how Arthur compares to the alternatives.

Pros & cons

Pros

  • Covers the full lifecycle: monitoring, evals, guardrails, and governance in one place
  • Agent discovery surfaces shadow/unregistered agents with step-level traces
  • Flexible deployment including self-managed VPC/BYOCloud/on-prem
  • Open-source Arthur Engine lets teams start free and self-host
  • Integrates with existing SOC tooling (CrowdStrike, Splunk, Datadog, Elastic)

Cons

  • Aimed at ML/security engineers, not business or non-technical users
  • SSO, SLAs, BAA, and dedicated VPC are Enterprise-tier only (custom pricing)
  • Free and Premium tiers cap use cases, data retention, and volume
  • Full-value deployment requires meaningful data/model integration work

Pricing plans

Free
$0 / month
  • Core model performance monitoring
  • Built-in cloud data connectors
  • Up to 4 use cases
  • Unlimited seats
  • 7-day data retention
Premium
$60 / month
  • Everything in Free
  • Customizable metrics and dashboards
  • Custom alerting and webhooks
  • Up to 100 use cases
  • 30-day data retention
Enterprise
Custom
  • Everything in Premium
  • Dedicated/managed VPC / on-prem
  • SSO, uptime SLAs, and BAA
  • Dedicated customer success manager
  • Custom data, jobs, traces, and evals

Key features

API
Self-hosted
Team collaboration
Integrations
CrowdStrike Falcon, Splunk, Datadog, Elastic, OpenTelemetry
Input types
ML model predictions, LLM/agent traces and spans, prompts, evaluation datasets
Output types
monitoring dashboards, evaluation scores, anomaly/security alerts, audit logs and lineage
Best For
AI model monitoring, agent security and discovery, LLM evaluation, governance and compliance reporting

Compare key features

View all alternatives →
Feature
Arthur
Fiddler AI
IBM watsonx.governance
Pricing
Freemium
Contact for pricing
Contact for pricing
Free plan
Yes
No
No
Free trial
No
No
No
API
Yes
Yes
Yes
Self-hosted
Yes
Yes
Yes
Team support
Yes
Yes
Yes

Frequently asked questions

Does Arthur have a free version?+

Yes. Arthur offers a Free plan at $0/mo with core monitoring for up to 4 use cases and unlimited seats, plus a free, open-source Arthur Engine you can self-host from GitHub.

Can Arthur be self-hosted or run on-prem?+

Yes. Enterprise supports self-managed VPC, BYOCloud, and on-prem deployment, and the open-source Arthur Engine runs inside your own stack.

What does Arthur do beyond model monitoring?+

It adds LLM/agent evaluations, agent discovery, runtime guardrails and behavioral analytics, cost tracking with model rerouting, and governance features like audit logs and lineage.

Reviews

Write a review

Pick a rating
Loading reviews…
Compare

Compare Arthur with other AI tools

Side-by-side pages for pricing, features, and best-fit use cases.

All comparisons →

Similar tools you may like