Build agents that get the job done reliably. Turn your users’ requests into repeatable tests.

Backed by investors from

Y Combinator, Cursor, Clay, Vercel

Product

Turn a captured trace into a reviewed, rerunnable case with a pinned environment.

I was charged twicerefund the duplicate

Tasks from real usage

Real tasks include the user’s original context and starting data.

Tools that match production

Not a support catalog. Reviewed cases rerun selected observed calls in pinned environments.

Stateful environments

Later calls see earlier writes. Reruns use the same reviewed starting definition; missing or unsupported facts stay visible or inconclusive.

Features

From real usage to measured performance.

Painted blue sky with pale drifting clouds

Observability

Trace real requests, model calls, and tool activity

Eval Suggestions

Find tasks worth replaying from real usage and coverage gaps

Agent Replays

Rerun the task in a stateful environment with simulated tools

Evaluations

Score each rerun and compare performance across agent versions

Secure

Your agent. Your infrastructure. Your control.

  • Control over traces and context
  • Keys scoped per project and organization
  • SOC 2 Type II, ISO 27001, and HIPAA compliant
Project settings

Project key

Hue project key, redacted

Capture policy

Traces
Context
  • SOC 2 compliance badge
  • HIPAA compliance badge
  • ISO 27001:2022 compliance badge

Developers

Get started with your agent

Set up Hue in my project using docs.hue.run/guides/agent-setup

Compatibility

Keep the instrumentation you have

Traces from any OpenTelemetry tool

Hue’s SDK is one way in, and OTLP traces from any OpenTelemetry tool are another.

Language agnostic

Trace model calls and tool use with native SDKs supporting most languages.

Your existing agent stack

Spans from your framework are recognized in the convention it already uses.

Build agents that get the job done reliably.Replay a real task before you ship.

Two lily pads on still water
Hue