Cloudflare has put Traces into open beta, recording each supported step a request takes through its edge as an OpenTelemetry span. The spans export to any OTLP endpoint, so edge decisions can sit in the tracing backend a team already runs.
Reality
- Evidence40
- Adoption
- Insufficient
- Hype gap+10
- Incentives85
- Confidence45
Cloudflare is merging logs, traces, alerts, dashboards and exports into one observability platform, shipped as eight updates under one pricing model. With Logpush now on self-serve plans, teams that run their own Cloudflare log pipeline have a built-in alternative to cost out.
Reality
- Evidence50
- Adoption
- Insufficient
- Hype gap+20
- Incentives75
- Confidence55
AWS's CloudWatch Omni, generally available since September 23, lets Okta and Entra ID users investigate incidents without AWS console access. CloudWatch dashboard sharing has let outsiders view prebuilt graphs since 2020, so what Omni adds is the investigation itself, one of the reasons teams paid for third-party platforms.
Reality
- Evidence50
- Adoption
- Insufficient
- Hype gap+20
- Incentives60
- Confidence55
LiteLLM launched Lens on September 30, a tool that uses AI agents to find recurring failures across agent traces sent through its model gateway. Customers host the analyzer and its databases, and the 200,000-trace volume CTO Ishaan Jaffer cites is a future target Lens has not been measured against.
Reality
- Evidence40
- Adoption
- Insufficient
- Hype gap+25
- Incentives60
- Confidence40
Cloudflare's Workers Issues, now in open beta, groups production errors from one config line and sends them to Claude Code, Cursor or Devin. Detection and dispatch happen inside Cloudflare; the fix, and any gate before it ships, live in the agent's workflow.
Reality
- Evidence45
- Adoption10
- Hype gap+15
- Incentives85
- Confidence50
One SaaS team's rebuild of incident detection on Kafka and Flink cut telemetry lag from over 40 seconds to under 10, according to its CNCF post. Over eighteen months its recall ran from about 60% to 86% and back to 64%, so speed alone says little about which outages get caught.
Reality
- Evidence45
- Adoption20
- Hype gap+10
- Incentives40
- Confidence40
Spring Boot 4.2.0-M2 adds management.observations.conventions, a property that switches semantic conventions between Micrometer and OpenTelemetry. Teams moving to 4.2 now have to pick the naming their dashboards and alerts will match.
Reality
- Evidence55
- Adoption
- Insufficient
- Hype gap0
- Incentives
- Insufficient
- Confidence60
Atlassian rebuilt a metrics platform fed by about 100,000 hosts on OpenTelemetry while its applications kept sending StatsD packets to the same address. The platform team did the work alone. Moving thousands of services onto OpenTelemetry SDKs comes later.
Reality
- Evidence50
- Adoption60
- Hype gap+10
- Incentives45
- Confidence55
AWS Cost Anomaly Detection works from Cost Explorer data up to 24 hours old, so a dollar alarm on an agent fires after the money is spent. OpenTelemetry's GenAI spec has no cost attribute either, so teams must price each span from cache-split tokens and sum the trace tree.
Reality
- Evidence55
- Adoption20
- Hype gap+10
- Incentives
- Insufficient
- Confidence50
AWS made CloudWatch Omni generally available with 17 built-in evaluators that score an AI agent's answers on live production traffic. For operators, the job moves from confirming an agent is running to deciding whether an automated score is good enough to gate a release.
Reality
- Evidence35
- Adoption20
- Hype gap+30
- Incentives65
- Confidence35
Next.js 16's stable instrumentation.ts runs telemetry setup once per runtime before app code loads, a dev.to walkthrough says. Because an async register holds the server until it resolves, a three-second SDK connect becomes three seconds of cold start.
Reality
- Evidence35
- Adoption
- Insufficient
- Hype gap+20
- Incentives
- Insufficient
- Confidence40
Agent telemetry guidance on dev.to moves run IDs, prompts and tool arguments off metric labels and into traces, capping a sample counter at 16 series. Teams that adopt it review each new label value like a schema change and rely on sampled traces for per-run evidence.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+5
- Incentives
- Insufficient
- Confidence50
The new commands are the easy part. The change worth reviewing is that the token your deploy job already holds can now edit DNS records, renew a domain and add a project member, all with clean JSON output.
Reality
- Evidence50
- Adoption
- Insufficient
- Hype gap+10
- Incentives45
- Confidence55
AWS's new observability surface discovers services and adjusts alarms against the targets you set, and signs engineers in through IAM Identity Center at a URL of your own. Instrumentation coverage decides what it can see.
Reality
- Evidence35
- Adoption15
- Hype gap+30
- Incentives80
- Confidence55
A dev.to review of three AI coding tools says Copilot ships its hard spend stop switched off, leaving budgets as alerts. Claude Code needs an opt-in before it meters seats in dollars and Cursor bills in arrears, so a real cap takes different work on each.
Reality
- Evidence28
- Adoption
- Insufficient
- Hype gap+35
- Incentives80
- Confidence30
MCP's 2026-07-28 spec deprecated protocol-level Logging and left it at least twelve months of support. A dev.to guide says OpenTelemetry's MCP conventions replace it, and since they are still unstable, teams should pin instrumentation versions.
Reality
- Evidence45
- Adoption
- Insufficient
- Hype gap+10
- Incentives20
- Confidence45
Dataiku ships Agent Management in October with connectors into Salesforce, AWS, Microsoft and Google agent runtimes, metered per agent monitored. The bill grows with every agent the inventory turns up.
Reality
- Evidence44
- Adoption10
- Hype gap+34
- Incentives84
- Confidence54
LangChain's survey of 1,340 practitioners found 89% had observability on their agents and 37.3% ran online evaluations. The OpenTelemetry GenAI attribute registry explains why only the second number measures quality.
Reality
- Evidence58
- Adoption64
- Hype gap+10
- Incentives42
- Confidence52
The otelbridge adapter builds trace context first so the enablement check and the Emit call see the same values, because a processor downstream may filter on the sampled flag that a cheap probe would drop.
Reality
- Evidence58
- Adoption15
- Hype gap−20
- Incentives60
- Confidence55
AWS has added three skill-focused evaluators to Strands Evals and Bedrock AgentCore Evaluations. Two of them call a model once per invoked skill, and the third is a deterministic check that exists only in Strands.
Reality
- Evidence52
- Adoption10
- Hype gap+15
- Incentives80
- Confidence55
Earlier coverage
- Foundry's automated grading needs the trace to carry the content it grades
Build · September 21, 2026 · 1 publisher
- The opentelemetry-instrument launcher patches dagster.asset before definitions.py imports it
Build · September 21, 2026 · 1 publisher
- A Jackson Databind CVE fix ships in the Payara release that removes @Traced
Build · September 21, 2026 · 1 publisher
- Kubernetes v1.37 exposes its latency histograms twice over
Leadership · September 20, 2026 · 1 publisher
- Taabi's alerting agent compiles an English sentence into a rule document a plain engine executes
Build · September 18, 2026 · 1 publisher
- Sorting the last 100 review comments finds 75% of the feedback a rule or a test can catch
Build · September 18, 2026 · 1 publisher
- Arcjet's guard returns allow or deny one call before the refund goes out
Build · September 17, 2026 · 1 publisher
- Elastic Beanstalk's Cluster Mode shares one EKS cluster across every app in a subnet set
Build · September 17, 2026 · 1 publisher
- Google's Agent Anomaly Detection audits agent traces against four OWASP agentic risks
Security · September 17, 2026 · 1 publisher
- Five retrieved guidelines lift AppWorld's hard-task success by 14.2 points
Build · September 17, 2026 · 2 publishers
- Arcjet puts a policy check in front of every action an AI agent takes
Product · September 17, 2026 · 1 publisher
- Atlassian swapped its metrics engine behind the same StatsD address on 100,000 hosts
Product · September 17, 2026 · 1 publisher
- Swapping Backstage for an agent hands the platform team an index to keep current
Build · September 17, 2026 · 1 publisher
- Splunk is putting a log-trained LLM on Hugging Face under an open source license
Product · September 16, 2026 · 1 publisher
- A stalled login service took four of the five Apex inner-loop steps offline
Build · September 16, 2026 · 1 publisher
- JetBrains' Service Map draws the architecture from each span the moment it arrives
Build · September 16, 2026 · 1 publisher
- Red Hat clocks sandbox isolation at under 5 percent of an agent request's latency
Product · September 15, 2026 · 1 publisher
- Cilium 1.19 replaces the per-pod Envoy with one eBPF program per node
Build · September 14, 2026 · 1 publisher
- Agent-cache's tool cache returns the first ticket's ID when the tool writes instead of reads
Build · September 13, 2026 · 1 publisher
- Mixing self-hosted Qwen with Bedrock Claude costs you telemetry, not a rewrite
Build · August 14, 2026 · 1 publisher
- Mohdel 1.0 computes per-call cost from a price catalog you maintain yourself
Build · September 11, 2026 · 1 publisher
- A poorly scoped supervisor prompt sends 20 percent of requests to the wrong specialist
Build · September 11, 2026 · 1 publisher
- ADR-0001 demotes Langfuse to a projection of Kept's in-process trace
Build · September 10, 2026 · 1 publisher
- GitHub moves Copilot's sandbox lock into the JetBrains plugin itself
Product · September 9, 2026 · 1 publisher
- One org-level webhook turns every GitHub Actions run into a trace you can drill into
Product · September 9, 2026 · 1 publisher
- Seven MCP tool-server bugs billed Databricks $499K a year in retried tokens
Build · September 1, 2026 · 1 publisher
- Aurora DSQL's second writable endpoint bills every commit for the inter-Region round trip
Build · September 1, 2026 · 1 publisher
- Bedrock's managed agentic retrieval nests a second loop inside the call your RAG logs count as one
Build · August 31, 2026 · 1 publisher
- Broadcom folds private AI into an integrated VMware Cloud Foundation stack
Product · August 31, 2026 · 1 publisher
- OpenTelemetry reaches CNCF graduation, meeting governance and other criteria
Product · August 31, 2026 · 1 publisher
- A module-scope variable leaks user IDs across requests on Vercel's Fluid Compute
Build · August 31, 2026 · 1 publisher
- Hundreds of 200 OKs a day hid an agent inventing tracking numbers for three weeks
Build · August 29, 2026 · 1 publisher
- Experiential Labs bets its open-source router's traces will train cheaper replacements for rented models
Build · August 27, 2026 · 1 publisher
- Instrumentation is solved. The bill for storing what it collects is not.
Build · August 27, 2026 · 1 publisher
- AWS puts agent evaluation on OpenTelemetry, and on exactly three span roles
Build · August 26, 2026 · 2 publishers
- Agent traces became product data, and the write pattern now picks your storage
Build · August 26, 2026 · 1 publisher
- JetBrains ships its OpenTelemetry plugin to four more IDEs; instrumentation is still your problem
Build · August 26, 2026 · 1 publisher
- SWE-Bench does not measure the job: why agent scores are the wrong readiness signal
Build · August 25, 2026 · 1 publisher
- On-device inference turns a phone's temperature into a release-gating variable
Product · August 25, 2026 · 1 publisher
- OpenTelemetry's maintainers say the helper class you are about to write is the bug
Build · August 25, 2026 · 1 publisher