Skip to content

Product2 publishersIndependently confirmed2 min readPublished

New Relic's Ground Truth CLI lets AI agents re-check whether a production fix held

New Relic's Ground Truth CLI, now in limited preview, lets developers and AI agents query production telemetry and verify fixes without a dashboard. On-call teams will mostly use its recovery check, which reruns against a condition that someone still has to write.

The Product Desk

How we use AISend a correction

Illustration accompanying New Relic's Ground Truth CLI lets AI agents re-check whether a production fix held
Generated illustration

What happened

  • Ground Truth CLI connects to New Relic Autopilot, NerdGraph and memory APIs, giving access to system telemetry and AI-powered analysis.
  • Its recovery checks repeat automatically and produce evidence of whether an application is running within specified recovery conditions.
  • The tool records investigation results, so engineers can revisit earlier findings or pass an unresolved incident to another team member.
  • It works in scripts and continuous integration pipelines, and AI coding tools can reach New Relic's operational data through it programmatically.

Compiled by The Product DeskSomething wrong?How this is made

Why it matters

  • constraint A recovery check covers only the period it ran, so a team that stops checking soon after deploy can sign off on a fix that an intermittent fault later reopens.
  • capability An incident an agent works overnight can go to the morning engineer with its findings attached, so agent investigations can slot into shift-based on-call rotations.
  • exposure Once coding agents read production telemetry inside CI pipelines, the gate that decides whether an agent's proposed change merges has to sit in that same pipeline.

An engineer patches a checkout service that has been throwing errors and closes the incident once the error rate falls. Devops.com's description of Ground Truth CLI uses nearly that scene: after an update, the engineer tests whether error rates have returned to acceptable levels [4]. The same article adds a caveat. A deployed fix does not necessarily mean the underlying problem is resolved, and intermittent failures can continue after a first test says the repair worked [7].

The pitch is headless observability, meaning agents and scripts pull telemetry programmatically with no graphical interface involved [9]. Devops.com frames the release around the idea that dashboards are designed for human users while agents need direct, programmatic access to production data [10]. What the tool actually does is smaller and more practical. A check runs against a recovery condition and repeats automatically, so the team gets evidence of whether the application is inside that condition [3].

As evidence that observability vendors are rebuilding around agents, the record is thin. This is one vendor's product [1], in limited preview [11]. New Relic also already offered Ground Truth to compatible AI tools through the Model Context Protocol. The CLI is a second route to the same capabilities, built for terminal workflows [12].

On an incident call, the comfortable assumption is that one green result after deploy means customers have stopped seeing errors. A shopper at checkout sees whatever the service does at the hour an intermittent fault comes back, and a single post-deploy test can miss that hour [7]. Repeated checks close that gap only as well as the condition they test. The article does not say who sets that condition or when the preview ends.

Devops.com says that in many cases engineers still need to evaluate findings and decide whether proposed changes are appropriate [6]. So the planning question for agent-driven troubleshooting is about authority, and it fits on a 2x2. One axis is who writes the recovery condition: a named person before the fix, or the agent after it. The other is what a passing check is allowed to do: report, or close the incident. A person-written condition with report-only is the safe corner and the slow one, because someone still reads every result. An agent-written condition with auto-close lets the agent grade its own repair, and I would refuse it. I'd start teams in the corner where a person writes the condition before the fix and the agent runs the checks but cannot close the ticket. The cost is a human sign-off on every incident until the checks have a record of catching faults that came back.

What to watch

  • New Relic publishing a price and a general-availability date for Ground Truth CLI, including whether recovery-check runs count toward billing.
  • Whether the preview lets a passing recovery check close an incident on its own, or only report the result to an engineer.
  • Another observability vendor shipping agent-facing command-line access; until one does, the pattern rests on New Relic alone.

Clarity's read

What the record supports and how the coverage leans. The claims behind it follow.

Reality

Evidence35
Adoption
Insufficient
Hype gap+30
Incentives70
Confidence60
Why these scores

Claim ledger

Ranked by verification strength, evidence, and original report placement.

  1. [1]

    New Relic introduced Ground Truth CLI, a command-line tool that enables software developers and AI agents to analyze system performance and verify whether technical issues have been resolved without using dashboards.

  2. [2]

    Ground Truth CLI connects to New Relic Autopilot, NerdGraph and memory APIs, providing access to system telemetry and AI-powered analysis.

  3. [3]

    The CLI supports automated recovery checks that can be repeated automatically, providing evidence about whether the application is operating within specified recovery conditions.

Sources

2 independent publishers whose own reporting we read for this story.

  1. devops.com

    1 article · October 8, 2026

    New Relic Debuts Command-Line Observability Tool for Developers and AI Agents
  2. helpnetsecurity.com

    1 article · October 6, 2026

    New Relic adds terminal-based investigation and recovery checks with Ground Truth CLI

Share your take

Let Clarity write the post for you.

Signed-in readers get a short post drafted on this story in the register they choose — narrative, analytical, or a direct position — editable to the last word before it goes anywhere. The share buttons at the top of this story work without an account.

Topics and entities

Follow any of these and your For You feed starts watching them — no settings page required.

Topics

Entities

Loading related stories