1. Home
  2. Companies
  3. Judgment Labs
Judgment Labs logoJL

Judgment Labs

About

Judgment Labs is an applied-research company focused on the continuous-improvement infrastructure for AI agents. The company builds Agent Behavior Monitoring (ABM) technology designed to help AI-native teams analyze production data and identify behavioral anomalies - such as instruction drifts and context retrieval loss - in deployed agents at scale.

The company's product suite includes Agent Search for behavioral-level trajectory querying, Agent Judge for trajectory-level evaluation, Behavior Discovery for surfacing failure modes from unlabeled production data, and AutoRubrics for automatically constructing evaluation rubrics from verifiable signals. Together, these tools form a stack aimed at turning production telemetry into actionable improvements for agent reliability.

Judgment Labs has raised $32 million across combined Seed and Series A funding. The company operates as an applied-research lab targeting last-mile agent reliability, serving AI-native teams and applied-research organisations.

Similar companies

Arize AI logoAA

Arize AI

Arize AI is the leading AI & Agent Engineering observability and evaluation platform, providing one place for development, observability, and evaluation of AI systems.

Variance logoVA

Variance

Variance builds AI agents that automate threat detection, policy enforcement, and investigations for risk and compliance teams at large institutions.

Guild.ai logoGU

Guild.ai

Guild.ai builds an independent control plane for deploying, governing, and observing AI agents in enterprise production environments across any model or framework.

Resolve AI logoRA

Resolve AI

Resolve AI builds autonomous AI agents that investigate and resolve production issues, manage on-call operations, and handle daily tasks to reduce incident resolution time for engineering teams.

AfterQuery logoAF

AfterQuery

AfterQuery is a San Francisco-based applied research lab that develops data solutions - including reasoning traces, agent environments, and computer use trajectories - for frontier foundation model developers.

Duku AI logoDA

Duku AI

Duku AI builds an autonomous testing platform that simulates real user journeys, self-heals as code changes, and runs continuously after every build to catch failures early.