The town square

Feed

Enterprise technology news for the people who buy and run it. Our editors pick stories from company newsrooms and trusted trade press, and every one credits and links its source. You’re seeing Top: recent stories ranked by what readers open, save and discuss. Nobody pays to rank. How ranking works

Narrow it by category below, or follow companies and categories to get their stories in For you and the Monday brief.

Matching stories

Observability

Extend Datadog RUM and Product Analytics to Shopify and Salesforce

Use Datadog RUM and Product Analytics to monitor checkout journeys on Shopify and customer experiences in Salesforce Experience Cloud.

Datadog·via Datadog Blog
Observability

DevRel newsletter — September 2026

Hello from the Elastic DevRel team! In this newsletter, we cover jina-ocr-v1, the latest blogs and videos, and upcoming events like Elastic{ON}.

Elastic·via Elastic Blog
ObservabilityHow-to

What if your agent's hallucinations had a budget? How to start using SLOs for agent behavior

At Grafana Labs, observability is what we do. So as we started building AI agents, we naturally reached for the same instincts we bring to every system: measure it, set targets, and make reliability something you can reason about instead of hope for. That instinct led us somewhere unexpectedly useful. It turns out one of the oldest ideas in reliability engineering, the error budget, maps…

Grafana Labs·via Grafana Labs Blog
Observability

Cribl Stream and Databricks: A faster path to analytics-ready telemetry

The new Databricks Zerobus Destination for Cribl Stream helps organizations move analytics-ready telemetry into Delta Tables with fewer moving parts, less latency, and more control.

Cribl·via Cribl Blog
Observability

Using TypeSafe’s Jev for evals in Datadog Agent Observability

See how we wired Jev into online evals on live spans and offline evals inside Datadog experiments, using a single rubric for both.

Datadog·via Datadog Blog
Observability

Trust, but verify: Atomic claim checking against LLM hallucinations

Hallucinations kept an LLM out of our knowledge base cleanup for years. Splitting generation from atomic claim verification fixed it and cleared a four-year backlog of duplicate articles.

Elastic·via Elastic Blog
Observability

Elastic Cloud on Azure gets a speed boost: Compute-optimized instances on Azure Cobalt ARM

Starting September 24, 2026, Elastic Cloud on Azure adds compute-optimized (ARM) hardware profiles backed by Microsoft's Cobalt ARM processors, delivering up to 37% better throughput than the previous generation Ampere Altra VMs at lower cost.

Elastic·via Elastic Blog
Observability

AI is compressing attack timelines. Cyber resilience needs proof.

How banks can turn the ECB’s AI-cybersecurity expectations into an evidence-driven operating model

Cribl·via Cribl Blog
Observability

Dynatrace achieves ISO/IEC 42001 certification

Dynatrace has achieved the international standard certification for AI management systems, ISO/IEC 42001:2023. The certification applies to the AI management systems (AIMS) supporting the Dynatrace platform, including the AI-powered capabilities Dynatrace develops and embeds in Dynatrace SaaS and Dynatrace Managed. The post Dynatrace achieves ISO/IEC 42001 certification appeared first on…

Dynatrace·via Dynatrace News
Observability

From insight to innovation: How Kiro Crew, Dynatrace, and AWS are helping teams do more

Modern businesses are sitting on a goldmine of information. Your observability data can help you understand what is breaking, what is underperforming, and where your next improvement should come from. The gap isn’t data – it is the distance between seeing something and being able to do something about it That’s exactly what Kiro Crew, […] The post From insight to innovation: How Kiro Crew,…

Dynatrace·via Dynatrace News
Observability

Teaching a 9B model to investigate production alerts

Learn how we fine-tuned Qwen3.5-9B into a specialized agent for change attribution that achieved 87% of GLM-5.3’s Recall@5 roughly 5% of the investigation cost.

Datadog·via Datadog Blog
Observability

Find answers in your logs faster with Datadog’s Tap to Parse

Use Datadog’s Tap to Parse feature to extract searchable fields from unstructured logs across Log Explorer, Log Pipelines, and Observability Pipelines.

Datadog·via Datadog Blog
Observability

Configure RUM SDKs remotely from Datadog

Use RUM Remote Configuration to change SDK sampling rates, privacy settings, and data collection independently of your release cycle.

Datadog·via Datadog Blog
Observability

Elastic Stack 8.19.22 released

Version 8.19.22 of the Elastic Stack was released today. We recommend you upgrade to this latest version . We recommend 8.19.22 over the previous version 8.19.21 For details of the issues that have been fixed and a full list of changes for each product in this version, please refer to the release notes .

Elastic·via Elastic Blog
Observability

The people who help shape Cribl: Celebrating some of our early employees

Cribl·via Cribl Blog
Observability

Is your most expensive talent spending valuable time writing status updates?

Major incidents can turn expert engineers into coordinators and status writers. Shared context can reduce duplicated investigative work. The post Is your most expensive talent spending valuable time writing status updates? appeared first on BigPanda .

BigPanda·via BigPanda Blog
Observability

The 5 Stages of AI-Human Collaboration to Improve Operational Reliability by PagerDuty

AI adoption is now a measurable driver of both uptime and growth. According to PagerDuty’s 2026 State of AI-First Digital Operations report, 75% of organizations... The post The 5 Stages of AI-Human Collaboration to Improve Operational Reliability appeared first on PagerDuty .

PagerDuty·via PagerDuty Blog
Observability

Introducing New Relic Compound Alerts

Discover how New Relic Compound Alerts intelligently correlates related alerts into actionable operational issues to improve incident response and operational efficiency.

New Relic·via New Relic Blog
Observability

Announcing the 2026 Observability Forecast

The Observability Forecast 2026 offers insights from 2,575 IT and engineering leaders and practitioners worldwide on the future of observability.

New Relic·via New Relic Blog
ObservabilityHow-to

How to Build an SRE Agent That Actually Works (Without Blowing the Token Budget)

Learn how to build an SRE agent that delivers accurate, low-latency incident triage with multi-tiered memory and RAG—without blowing your token budget.

New Relic·via New Relic Blog
Observability

Enterprise Service Management Is for More Than IT: HR, Facilities, and Finance Use Cases

09/21/26 What Enterprise Service Management Actually Means Enterprise service management (ESM) is the practice of applying IT service management principles, such as service catalogs, structured intake, automated approvals, and ticket tracking, to departments outside IT. The label sounds abstract until you see it in context: ... The post Enterprise Service Management Is for More Than IT: HR,…

SolarWinds·via SolarWinds Blog
Observability

Grafana Alerting: Scale alert routing without scaling complexity using multiple notification policies

Alert routing often starts simple. A team creates a few contact points, adds some label matchers, and builds a notification policy tree that sends each alert to the right destination. But alerting configurations rarely stay simple. As an organization grows, its notification policy tree must accommodate more teams, services, and routing requirements. Changes for one team still require editing a…

Grafana Labs·via Grafana Labs Blog
Observability

How Adaptive Tail Sampling Works in the OpenTelemetry Collector

A technical deep dive into Honeycomb's adaptive tail sampling processor for the OpenTelemetry Collector: how decisions get made, the samplers available, how thresholds compose with the rest of a sampling pipeline, performance benchmarks, deployment limitations, and how it compares to Refinery.

Honeycomb·via Honeycomb Blog
Observability

Your pod may be requesting 25× more CPU than you think — and Kubernetes won’t tell you

You’ve run kubectl describe pod. Your dashboard shows 40 millicores requested for the api-server container. You probably assumed your monitoring had the full picture. The Kubernetes scheduler actually reserved 1000 millicores — a full CPU core — for it. How could both values be true? If you’re sizing clusters, building cost models, or tuning HPA […] The post Your pod may be requesting 25× more…

Dynatrace·via Dynatrace News
Observability

WebMCP Monitoring: Why Website Uptime Isn’t Enough for AI Agents

AI agents can fail even when your website looks healthy. See how WebMCP monitoring uses synthetic tests to validate the full agent journey, from tool discovery to backend response. The post WebMCP Monitoring: Why Website Uptime Isn’t Enough for AI Agents appeared first on LogicMonitor .

LogicMonitor·via LogicMonitor Blog
Observability

CriblCon 26: Can’t-miss sessions if you want to be a superstar analyst

Cribl·via Cribl Blog
Observability

What TypeSafe’s Jev means for telemetry

Cribl·via Cribl Blog
Observability

Centralized Log Management: A Comprehensive Guide for Engineers

Learn how centralized log management helps developers and engineers improve incident response, reduce costs, and gain better control over log data.

New Relic·via New Relic Blog
Observability

LLM Observability: The 8 Best Tools for Production AI Systems

Compare the top LLM observability tools for production AI — tracing, cost tracking, and evaluation — and see what to weigh before you add another tool to your stack.

New Relic·via New Relic Blog
Observability

Automate Incident Management with PagerDuty Slack by PagerDuty

Most organizations managing major incidents realized that every moment matters. Context-switching between different tools – with multiple web and chat surfaces having to be open... The post Automate Incident Management with PagerDuty Slack appeared first on PagerDuty .

PagerDuty·via PagerDuty Blog
Observability

September 15 SolarWinds Service Desk Release

09/15/26 Discover the new release with insight from our product team: Recap: Service Desk — September 15 New Feature Pre-Release. Newest product updates: Choose the right foundation for each Process Integration Plan Type: All Plans | Status: General Availability Service Desk now gives admins more flexibility ... The post September 15 SolarWinds Service Desk Release appeared first on SolarWinds…

SolarWinds·via SolarWinds Blog
Observability

Dynatrace Release Radar 08.26

This series covers recent Dynatrace releases and updates, focusing on what’s new, what’s changed, and how these recent enhancements can benefit you and your organization. Each post covers newly available capabilities and where to explore them. The post Dynatrace Release Radar 08.26 appeared first on Dynatrace news .

Dynatrace·via Dynatrace News
Observability

Investigate Smarter from Any Agent: Blast Radius and AI Post-Incident Reviews Come to the PagerDuty MCP Server by Antonella Hidalgo Moreno

PagerDuty’s new intelligent Model Context Protocol (MCP) tools, recently announced in Early Access, are a step toward autonomous operations. Here’s how they help customers get... The post Investigate Smarter from Any Agent: Blast Radius and AI Post-Incident Reviews Come to the PagerDuty MCP Server appeared first on PagerDuty .

PagerDuty·via PagerDuty Blog
Observability

AI can speed up development. Can operations keep up?

AI can accelerate software delivery. The operating model must also keep up with the context, ownership, and decisions that follow. The post AI can speed up development. Can operations keep up? appeared first on BigPanda .

BigPanda·via BigPanda Blog
Observability

Red Cards, Comeback Wins, and the Unsung Defense Keeping Modern Business Online

09/15/26 For IT Pro Day 2026, SolarWinds surveyed the global THWACK® community and discovered our IT professionals and sports pros aren’t all that different. Cut through the tech jargon and it’s clear modern IT pros are playing championship-level defense, navigating unforced errors, and quietly keeping ... The post Red Cards, Comeback Wins, and the Unsung Defense Keeping Modern Business Online…

SolarWinds·via SolarWinds Blog
Observability

A Better Way to Monitor Every Digital Journey with LogicMonitor Synthetics and Internet Performance Monitoring

LogicMonitor Synthetics and Internet Performance Monitoring helps ITOps teams catch digital experience issues earlier with outside-in visibility across apps, networks, APIs, and SaaS. The post A Better Way to Monitor Every Digital Journey with LogicMonitor Synthetics and Internet Performance Monitoring appeared first on LogicMonitor .

LogicMonitor·via LogicMonitor Blog
Observability

Digital Experience Monitoring with Grafana Cloud: Session Replay, synthetic checks, and faster investigations

When something breaks in production, the questions that matter most are also the toughest to answer from metrics alone: who was affected, what did they actually see, and is this worth waking someone up for? Answering those questions requires a fuller picture of the issue and its impact on your users. That’s where Digital Experience Monitoring (DEM) in Grafana Cloud comes in. By combining Frontend…

Grafana Labs·via Grafana Labs Blog
Observability

Outage Retrospective: The AWS Outages That Proved the Value of Independent Monitoring

Internet Performance Monitoring detected the AWS outages before AWS acknowledged them. Learn why independent, outside-in monitoring is critical to faster incident response. The post Outage Retrospective: The AWS Outages That Proved the Value of Independent Monitoring appeared first on LogicMonitor .

LogicMonitor·via LogicMonitor Blog
Observability

Why Government IT Needs AI-Powered Observability

Legacy systems, cloud growth, and AI are reshaping government IT. See how connected visibility helps teams respond faster and keep essential services running. The post Why Government IT Needs AI-Powered Observability appeared first on LogicMonitor .

LogicMonitor·via LogicMonitor Blog
ObservabilityHow-to

The AI Economy Has a Senior Engineer Problem. Here’s How to Solve It by PagerDuty

According to a 2025 report from Ravio, entry-level hiring (especially in engineering roles) has collapsed by more than 73% due to increasing AI capabilities. That... The post The AI Economy Has a Senior Engineer Problem. Here’s How to Solve It appeared first on PagerDuty .

PagerDuty·via PagerDuty Blog