Product2 publishersIndependently confirmed2 min readPublished
New Relic's Ground Truth CLI lets AI agents re-check whether a production fix held
New Relic's Ground Truth CLI, now in limited preview, lets developers and AI agents query production telemetry and verify fixes without a dashboard. On-call teams will mostly use its recovery check, which reruns against a condition that someone still has to write.
The Product Desk

What happened
- Ground Truth CLI connects to New Relic Autopilot, NerdGraph and memory APIs, giving access to system telemetry and AI-powered analysis.
- Its recovery checks repeat automatically and produce evidence of whether an application is running within specified recovery conditions.
- The tool records investigation results, so engineers can revisit earlier findings or pass an unresolved incident to another team member.
- It works in scripts and continuous integration pipelines, and AI coding tools can reach New Relic's operational data through it programmatically.
Compiled by The Product DeskSomething wrong?How this is made
Why it matters
- constraint A recovery check covers only the period it ran, so a team that stops checking soon after deploy can sign off on a fix that an intermittent fault later reopens.
- capability An incident an agent works overnight can go to the morning engineer with its findings attached, so agent investigations can slot into shift-based on-call rotations.
- exposure Once coding agents read production telemetry inside CI pipelines, the gate that decides whether an agent's proposed change merges has to sit in that same pipeline.
An engineer patches a checkout service that has been throwing errors and closes the incident once the error rate falls. Devops.com's description of Ground Truth CLI uses nearly that scene: after an update, the engineer tests whether error rates have returned to acceptable levels [4]. The same article adds a caveat. A deployed fix does not necessarily mean the underlying problem is resolved, and intermittent failures can continue after a first test says the repair worked [7].
The pitch is headless observability, meaning agents and scripts pull telemetry programmatically with no graphical interface involved [9]. Devops.com frames the release around the idea that dashboards are designed for human users while agents need direct, programmatic access to production data [10]. What the tool actually does is smaller and more practical. A check runs against a recovery condition and repeats automatically, so the team gets evidence of whether the application is inside that condition [3].
As evidence that observability vendors are rebuilding around agents, the record is thin. This is one vendor's product [1], in limited preview [11]. New Relic also already offered Ground Truth to compatible AI tools through the Model Context Protocol. The CLI is a second route to the same capabilities, built for terminal workflows [12].
On an incident call, the comfortable assumption is that one green result after deploy means customers have stopped seeing errors. A shopper at checkout sees whatever the service does at the hour an intermittent fault comes back, and a single post-deploy test can miss that hour [7]. Repeated checks close that gap only as well as the condition they test. The article does not say who sets that condition or when the preview ends.
Devops.com says that in many cases engineers still need to evaluate findings and decide whether proposed changes are appropriate [6]. So the planning question for agent-driven troubleshooting is about authority, and it fits on a 2x2. One axis is who writes the recovery condition: a named person before the fix, or the agent after it. The other is what a passing check is allowed to do: report, or close the incident. A person-written condition with report-only is the safe corner and the slow one, because someone still reads every result. An agent-written condition with auto-close lets the agent grade its own repair, and I would refuse it. I'd start teams in the corner where a person writes the condition before the fix and the agent runs the checks but cannot close the ticket. The cost is a human sign-off on every incident until the checks have a record of catching faults that came back.
What to watch
- New Relic publishing a price and a general-availability date for Ground Truth CLI, including whether recovery-check runs count toward billing.
- Whether the preview lets a passing recovery check close an incident on its own, or only report the result to an engineer.
- Another observability vendor shipping agent-facing command-line access; until one does, the pattern rests on New Relic alone.
Clarity's read
What the record supports and how the coverage leans. The claims behind it follow.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+30
- Incentives70
- Confidence60
Claim ledger
Ranked by verification strength, evidence, and original report placement.
- [1]
New Relic introduced Ground Truth CLI, a command-line tool that enables software developers and AI agents to analyze system performance and verify whether technical issues have been resolved without using dashboards.
- [2]
Ground Truth CLI connects to New Relic Autopilot, NerdGraph and memory APIs, providing access to system telemetry and AI-powered analysis.
- [3]
The CLI supports automated recovery checks that can be repeated automatically, providing evidence about whether the application is operating within specified recovery conditions.
- [4]
In devops.com's example, an engineer troubleshooting errors in an online checkout application could use the tool to examine behavior and, following a software update, test whether error rates have returned to acceptable levels.
- [5]
The tool supports scripts and continuous integration pipelines, and through the CLI AI coding tools can access operational information and New Relic capabilities programmatically.
- [6]
According to devops.com, in many cases engineers still need to evaluate findings and determine whether proposed changes are appropriate.
- [7]
According to devops.com, successfully deploying a software fix does not necessarily mean the underlying problem has been resolved; applications may continue experiencing intermittent failures even after an initial test indicates a repair was successful.
- [8]
Ground Truth CLI records investigation results, enabling engineers to examine earlier findings or transfer an unresolved incident to another team member.
- [9]
Headless observability is an approach that enables developers, automation tools and AI agents to access system telemetry and performance data programmatically without traditional graphical interfaces.
- [10]
Devops.com says traditional observability dashboards are designed for human users, while AI agents need direct, programmatic access to production data to investigate problems and evaluate fixes.
- [11]
Ground Truth CLI is currently in limited preview.
- [12]
New Relic already offers Ground Truth connectivity through the Model Context Protocol (MCP), allowing compatible AI tools to access operational information; the CLI is another method of accessing those capabilities, designed for terminal-based workflows.
Sources
2 independent publishers whose own reporting we read for this story.
- devops.comNew Relic Debuts Command-Line Observability Tool for Developers and AI Agents
1 article · October 8, 2026
- helpnetsecurity.comNew Relic adds terminal-based investigation and recovery checks with Ground Truth CLI
1 article · October 6, 2026
Topics and entities
Follow any of these and your For You feed starts watching them — no settings page required.
Topics
- Headless observabilityFollow
- AI agents in site reliability engineeringFollow
- ObservabilityFollow