Live tracing, now GA

Ship AI features you can actually trust

Nimbus traces every token, request, and dollar across your LLM stack — in real time, with alerts that fire before your users ever notice.

Star on GitHub
production · us-east
1h24h7d
Tokens / min
p95 latency
312ms
spend today
$1,284
error rate
0.04%

Trusted by teams building on top of every model

VercelLinearRetoolSupabaseRampCursorClerkFly.io
VercelLinearRetoolSupabaseRampCursorClerkFly.io
One agent, full visibility

Everything you need to run AI in production

Drop in the SDK and get traces, costs, and guardrails in one place — no sampling, no blind spots.

Real-time tracing

Every LLM call, tool invocation, and retry — streamed as spans with sub-second latency.

Cost analytics

Attribute spend to features, customers, and prompts. Catch a runaway loop before the invoice does.

Latency & SLOs

Percentile dashboards and burn-rate alerts, so a slow model provider pages you — not your users.

Guardrails

PII redaction, prompt-injection detection, and evals on live traffic, wired in with one line.

99.98%
Ingest uptime
<40ms
Added overhead
4.2B
Spans / day
30s
To first trace

Simple, usage-based pricing

Start free. Scale when you do. No per-seat games.

Hobby

For side projects and kicking the tires.

$0/forever
Start free
  • 50k spans / mo
  • 7-day retention
  • 1 project
  • Community support
Most popular

Pro

For teams shipping AI to real users.

$49/mo
  • 10M spans / mo
  • 90-day retention
  • Unlimited projects
  • Cost & latency alerts
  • Guardrails + evals

Enterprise

SSO, on-prem, and a real human on Slack.

Custom
Talk to us
  • Volume pricing
  • Unlimited retention
  • SSO / SAML
  • On-prem deploy
  • 99.99% SLA

Start tracing in 30 seconds

Two lines of code. Your first spans stream in before your coffee's done.

Book a demo