AI Dispatch/api/index.json
agent · Data Analysis

LangWatch

langwatch.aiupdated Aug 1, 2026view as JSON
30 fields evidenced · 25 with no public information. Every value below links to the document it came from and the date we checked it.
Runtime governance
2/3

Are scope, spend, and authority enforced while the agent runs?

Scored from 2 evidenced field(s): runtime_governance, permission_scopes. Automated score; capped at 2 pending review.

Audit trail
2/3

Is there a tamper-evident record of what it actually did?

Scored from 2 evidenced field(s): audit_trail, explainability. Automated score; capped at 2 pending review.

Compliance
2/3

Are compliance obligations attached per engagement?

Scored from 2 evidenced field(s): compliance_certs, regulatory_alignment. Automated score; capped at 2 pending review.

Outcome settlement
unscored

Does payment depend on a verified result?

Insurance & indemnity
unscored

Can the deployment be insured, and is the customer indemnified?

Assurance

Insurance available
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Insurance carriers
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Coverage limits
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Indemnification
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Liability cap
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Audit trail
Audit log → SIEM
Audit log → SIEM
langwatch.ai · checked Aug 1, 2026
Tamper-evident log
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Explainability
The judge reads the whole trace like you would, expanding each step, so a verdict comes with the reasoning behind it.
The judge reads the whole trace like you would, expanding each step, so a verdict comes with the reasoning behind it.
langwatch.ai · checked Aug 1, 2026
Runtime governance
RBAC + REST APIs SCIM + SSO Cost-center attribution
RBAC + REST APIs SCIM + SSO Cost-center attribution
langwatch.ai · checked Aug 1, 2026
Permission scopes
RBAC + REST APIs SCIM + SSO
RBAC + REST APIs SCIM + SSO
langwatch.ai · checked Aug 1, 2026
Compliance certifications
ISO 27001 Certified GDPR Compliant EU data Residency Monitored by Vanta
ISO 27001 Certified GDPR Compliant EU data Residency Monitored by Vanta
langwatch.ai · checked Aug 1, 2026
Regulatory alignment
GDPR Compliant EU data Residency
GDPR Compliant EU data Residency
langwatch.ai · checked Aug 1, 2026
Outcome-based pricing
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Settlement mechanism
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Dispute process
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
SLA terms
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Evaluation coverage
Evaluate Score everything, from a single output to a whole conversation, offline and live in production.
Evaluate Score everything, from a single output to a whole conversation, offline and live in production.
langwatch.ai · checked Aug 1, 2026

Agency

Autonomy level
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Human oversight
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Goal complexity
PM writes the goal Plain English. No code, no YAML. The brief is the spec.
PM writes the goal Plain English. No code, no YAML. The brief is the spec.
langwatch.ai · checked Aug 1, 2026
Action space
Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
langwatch.ai · checked Aug 1, 2026
Operating environment
Simulate real users Text and voice conversations from a simulated user that pushes your agent turn after turn, like the real world does.
Simulate real users Text and voice conversations from a simulated user that pushes your agent turn after turn, like the real world does.
langwatch.ai · checked Aug 1, 2026
Initiative
Langy turns a PM's goal into a full Scenario test plan, then turns the failures into pull requests.
Langy turns a PM's goal into a full Scenario test plan, then turns the failures into pull requests.
langwatch.ai · checked Aug 1, 2026

Safety

Safety evaluations
Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
langwatch.ai · checked Aug 1, 2026
Red teaming
Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
langwatch.ai · checked Aug 1, 2026
Safety policy
GDPR Compliant EU data Residency
GDPR Compliant EU data Residency
langwatch.ai · checked Aug 1, 2026
Usage restrictions
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Model or system card
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Incident reporting
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Third-party evaluations
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Data handling
Custom retention policy
Custom retention policy
langwatch.ai · checked Aug 1, 2026

Practicality

Pricing model
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Price point
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Availability
start free in 5 minutes
start free in 5 minutes
langwatch.ai · checked Aug 1, 2026
Deployment options
Cloud, self-hosted, or hybrid. Self-hosted Docker, Kubernetes/Helm, or in your VPC Hybrid Data plane on your infra, control plane on ours Cloud Managed multi-tenant SaaS
Cloud, self-hosted, or hybrid. Self-hosted Docker, Kubernetes/Helm, or in your VPC Hybrid Data plane on your infra, control plane on ours Cloud Managed multi-tenant SaaS
langwatch.ai · checked Aug 1, 2026
Integrations
OpenTelemetry native Full GenAI spec support, so your traces work with any framework and any OTel-compatible stack.
OpenTelemetry native Full GenAI spec support, so your traces work with any framework and any OTel-compatible stack.
langwatch.ai · checked Aug 1, 2026
Supported regions
EU / US / UK / APAC
EU / US / UK / APAC
langwatch.ai · checked Aug 1, 2026
Support model
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this

Foundation models

Base models
Claude Code, Codex, opencode and more
Claude Code, Codex, opencode and more
langwatch.ai · checked Aug 1, 2026
Model provider
Claude Code, Codex, opencode and more
Claude Code, Codex, opencode and more
langwatch.ai · checked Aug 1, 2026
Model swappable
Works with every agent framework, no rewrite required.
Works with every agent framework, no rewrite required.
langwatch.ai · checked Aug 1, 2026
Open weights
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Fine-tuning
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Context window
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this

Ecosystem

Protocols supported
Every tool call, skill, and MCP server is traced
Every tool call, skill, and MCP server is traced
langwatch.ai · checked Aug 1, 2026
Tool use
Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
langwatch.ai · checked Aug 1, 2026
Multi-agent
Langy drafts the plan Picks the simulator, generates the scenarios, writes the JudgeAgent rubric.
Langy drafts the plan Picks the simulator, generates the scenarios, writes the JudgeAgent rubric.
langwatch.ai · checked Aug 1, 2026
API access
RBAC + REST APIs
RBAC + REST APIs
langwatch.ai · checked Aug 1, 2026
Open source
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this
Marketplace presence
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this

Impact

User base
Trusted in production by AI agents are still tested by hand, breaking in production.
Trusted in production by AI agents are still tested by hand, breaking in production.
langwatch.ai · checked Aug 1, 2026
Deployment scale
Trusted in production by AI agents are still tested by hand, breaking in production.
Trusted in production by AI agents are still tested by hand, breaking in production.
langwatch.ai · checked Aug 1, 2026
Target sectors
Critical when you handle payments at scale.
Critical when you handle payments at scale.
langwatch.ai · checked Aug 1, 2026
High-risk domains
Critical when you handle payments at scale.
Critical when you handle payments at scale.
langwatch.ai · checked Aug 1, 2026
Documented incidents
no public information
langwatch.ai · checked Aug 1, 2026 · source did not state this