agent · Data Analysis
LangWatch
30 fields evidenced · 25 with no public information. Every value below links to the document it came from and the date we checked it.
Trust gap
How this is scoredRuntime governance
2/3
Are scope, spend, and authority enforced while the agent runs?
Scored from 2 evidenced field(s): runtime_governance, permission_scopes. Automated score; capped at 2 pending review.
Audit trail
2/3
Is there a tamper-evident record of what it actually did?
Scored from 2 evidenced field(s): audit_trail, explainability. Automated score; capped at 2 pending review.
Compliance
2/3
Are compliance obligations attached per engagement?
Scored from 2 evidenced field(s): compliance_certs, regulatory_alignment. Automated score; capped at 2 pending review.
Outcome settlement
unscored
Does payment depend on a verified result?
Insurance & indemnity
unscored
Can the deployment be insured, and is the customer indemnified?
Assurance
- Insurance available
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Insurance carriers
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Coverage limits
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Indemnification
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Liability cap
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Audit trail
- Audit log → SIEM
“Audit log → SIEM”
langwatch.ai · checked Aug 1, 2026 - Tamper-evident log
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Explainability
- The judge reads the whole trace like you would, expanding each step, so a verdict comes with the reasoning behind it.
“The judge reads the whole trace like you would, expanding each step, so a verdict comes with the reasoning behind it.”
langwatch.ai · checked Aug 1, 2026 - Runtime governance
- RBAC + REST APIs SCIM + SSO Cost-center attribution
“RBAC + REST APIs SCIM + SSO Cost-center attribution”
langwatch.ai · checked Aug 1, 2026 - Permission scopes
- RBAC + REST APIs SCIM + SSO
“RBAC + REST APIs SCIM + SSO”
langwatch.ai · checked Aug 1, 2026 - Compliance certifications
- ISO 27001 Certified GDPR Compliant EU data Residency Monitored by Vanta
“ISO 27001 Certified GDPR Compliant EU data Residency Monitored by Vanta”
langwatch.ai · checked Aug 1, 2026 - Regulatory alignment
- GDPR Compliant EU data Residency
“GDPR Compliant EU data Residency”
langwatch.ai · checked Aug 1, 2026 - Outcome-based pricing
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Settlement mechanism
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Dispute process
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- SLA terms
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Evaluation coverage
- Evaluate Score everything, from a single output to a whole conversation, offline and live in production.
“Evaluate Score everything, from a single output to a whole conversation, offline and live in production.”
langwatch.ai · checked Aug 1, 2026
Agency
- Autonomy level
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Human oversight
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Goal complexity
- PM writes the goal Plain English. No code, no YAML. The brief is the spec.
“PM writes the goal Plain English. No code, no YAML. The brief is the spec.”
langwatch.ai · checked Aug 1, 2026 - Action space
- Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
“Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.”
langwatch.ai · checked Aug 1, 2026 - Operating environment
- Simulate real users Text and voice conversations from a simulated user that pushes your agent turn after turn, like the real world does.
“Simulate real users Text and voice conversations from a simulated user that pushes your agent turn after turn, like the real world does.”
langwatch.ai · checked Aug 1, 2026 - Initiative
- Langy turns a PM's goal into a full Scenario test plan, then turns the failures into pull requests.
“Langy turns a PM's goal into a full Scenario test plan, then turns the failures into pull requests.”
langwatch.ai · checked Aug 1, 2026
Safety
- Safety evaluations
- Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
“Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.”
langwatch.ai · checked Aug 1, 2026 - Red teaming
- Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.
“Red teaming Adversarial simulations probe for jailbreaks, policy breaks, and unsafe tool calls before your users find them.”
langwatch.ai · checked Aug 1, 2026 - Safety policy
- GDPR Compliant EU data Residency
“GDPR Compliant EU data Residency”
langwatch.ai · checked Aug 1, 2026 - Usage restrictions
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Model or system card
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Incident reporting
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Third-party evaluations
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Data handling
- Custom retention policy
“Custom retention policy”
langwatch.ai · checked Aug 1, 2026
Practicality
- Pricing model
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Price point
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Availability
- start free in 5 minutes
“start free in 5 minutes”
langwatch.ai · checked Aug 1, 2026 - Deployment options
- Cloud, self-hosted, or hybrid. Self-hosted Docker, Kubernetes/Helm, or in your VPC Hybrid Data plane on your infra, control plane on ours Cloud Managed multi-tenant SaaS
“Cloud, self-hosted, or hybrid. Self-hosted Docker, Kubernetes/Helm, or in your VPC Hybrid Data plane on your infra, control plane on ours Cloud Managed multi-tenant SaaS”
langwatch.ai · checked Aug 1, 2026 - Integrations
- OpenTelemetry native Full GenAI spec support, so your traces work with any framework and any OTel-compatible stack.
“OpenTelemetry native Full GenAI spec support, so your traces work with any framework and any OTel-compatible stack.”
langwatch.ai · checked Aug 1, 2026 - Supported regions
- EU / US / UK / APAC
“EU / US / UK / APAC”
langwatch.ai · checked Aug 1, 2026 - Support model
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
Foundation models
- Base models
- Claude Code, Codex, opencode and more
“Claude Code, Codex, opencode and more”
langwatch.ai · checked Aug 1, 2026 - Model provider
- Claude Code, Codex, opencode and more
“Claude Code, Codex, opencode and more”
langwatch.ai · checked Aug 1, 2026 - Model swappable
- Works with every agent framework, no rewrite required.
“Works with every agent framework, no rewrite required.”
langwatch.ai · checked Aug 1, 2026 - Open weights
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Fine-tuning
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Context window
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
Ecosystem
- Protocols supported
- Every tool call, skill, and MCP server is traced
“Every tool call, skill, and MCP server is traced”
langwatch.ai · checked Aug 1, 2026 - Tool use
- Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.
“Every tool call, skill, and MCP server is traced, and mockable or fixtured for deterministic runs.”
langwatch.ai · checked Aug 1, 2026 - Multi-agent
- Langy drafts the plan Picks the simulator, generates the scenarios, writes the JudgeAgent rubric.
“Langy drafts the plan Picks the simulator, generates the scenarios, writes the JudgeAgent rubric.”
langwatch.ai · checked Aug 1, 2026 - API access
- RBAC + REST APIs
“RBAC + REST APIs”
langwatch.ai · checked Aug 1, 2026 - Open source
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
- Marketplace presence
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this
Impact
- User base
- Trusted in production by AI agents are still tested by hand, breaking in production.
“Trusted in production by AI agents are still tested by hand, breaking in production.”
langwatch.ai · checked Aug 1, 2026 - Deployment scale
- Trusted in production by AI agents are still tested by hand, breaking in production.
“Trusted in production by AI agents are still tested by hand, breaking in production.”
langwatch.ai · checked Aug 1, 2026 - Target sectors
- Critical when you handle payments at scale.
“Critical when you handle payments at scale.”
langwatch.ai · checked Aug 1, 2026 - High-risk domains
- Critical when you handle payments at scale.
“Critical when you handle payments at scale.”
langwatch.ai · checked Aug 1, 2026 - Documented incidents
- no public informationlangwatch.ai · checked Aug 1, 2026 · source did not state this