Field Signal · Curated Reading
The Engineering
Signal.
High-signal field notes, architecture teardowns, and engineering writing from 50 publications. A focused reading queue for becoming a stronger end-to-end engineer. Updated 20 Aug.
Accelerate investigations with AI in Datadog Incident Response
Investigate incidents with Bits Investigation as a fellow responder, receive AI-generated summaries in chat, and capture critical decisions made on bridge calls.
10 Years of Meta’s Commitment to Python
This year marks Meta’s 10th consecutive year as a sponsor of the Python Software Foundation (PSF), the charitable organization dedicated to advancing, supporting, and protecting the open-source Python programming language and the community that sustains it. Python is one of the world’s most influent...
Discover, govern, and scale Azure infrastructure in the AI era
Learn how Terraform helps teams discover unmanaged Azure resources, reduce infrastructure drift, and establish governance as cloud and AI environments scale.
How to audit what your AI agents are accessing
AI agents do a lot. Aperture keeps the audit trail.
HCP Terraform powered by Infragraph is now in limited availability
Hybrid and multi-cloud estates create data silos. HCP Terraform powered by Infragraph provides a single source of truth to help optimize and secure infrastructure.
Datadog acquires Adaptive ML
Datadog has acquired Adaptive ML, a platform for building, owning, and deploying specialized AI agents and models.
Debug and evaluate your AI app from your coding agent with Datadog Agent Observability
Learn how to give your coding agent access to Datadog Agent Observability data to classify failures, run RCA, bootstrap evaluators, and generate fixes.
5 pitfalls to avoid when measuring DevEx in the AI era
Don’t mistake AI adoption for productivity. Learn how to avoid 5 common pitfalls when measuring DevEx, with practices from Datadog engineering
Lessons learned from scaling to 1 million Lambda functions
In this post, we share our journey and the lessons learned from building and running a fully serverless, multi-account software as a service (SaaS) platform at scale. We’ll explore why true scale-to-zero is critical, how we handle quota management, why engaging AWS service teams early saved us from ...
Preparing for OMB M-26-14: How Datadog supports federal logging maturity
Learn how Datadog helps federal agencies prepare for OMB M-26-14 by providing centralized telemetry data, threat detection, and automated incident response.
Datadog achieves GovRAMP High authorization
Datadog achieves GovRAMP High authorization, bringing unified observability and security to state and local government agencies for critical systems.
Open source maintainership in the age of AI
AI has really changed the game around software development. More people are leveraging AI than ever to contribute patches to projects they use. To me, this is a good thing as more folks will contribute patches rather than fork or not fix them. The main problem is that AI has made generating code fas...
A no-nonsense explainer to Agentic AI
Cut through the buzzwords with a clear explanation of the agentic AI stack.
Terraform MCP server: Four real-world AI infrastructure patterns
Discover how Terraform MCP Server helps AI agents make better infrastructure decisions using trusted organizational context and guardrails.
Reduce CDN log costs with searchable archives
Route high-volume CDN logs to low-cost object storage with Observability Pipelines and search them with Archive Search—without a second tool.
Privacy-Aware Infrastructure in the AI-Native Era: An Asset Classification Case Study
Privacy controls — systems that enforce retention, access, allowed-purpose, downstream-sharing, or anonymization policies — require a reliable understanding of data to function. Before such a control can operate effectively, it must know exactly what it is looking at. This can be complex, as demonst...
Deploy Boundary on Kubernetes with official Helm charts
HashiCorp Boundary now offers official Helm charts for deploying controllers and workers on Kubernetes. Learn which chart fits your deployment model and how to get started.
Introducing the Cluster API plugin for Headlamp
Headlamp is an open-source, extensible Kubernetes SIG UI project designed to let you explore, manage, and debug cluster resources directly from a browser. Cluster API (CAPI) is a Kubernetes sub-project that brings declarative, Kubernetes-style APIs to cluster lifecycle management. It lets platform t...
Inspect Volcano workloads faster with Headlamp
Volcano is a cloud native batch scheduler for Kubernetes, built for high-performance computing, AI/ML, and other batch workloads. Headlamp is an extensible Kubernetes web UI. With its plugin system, Headlamp can surface APIs and workflows beyond the built-in Kubernetes resources. The Volcano plugin ...
See your serverless: introducing the Headlamp plugin for Knative
Headlamp is an open-source, extensible Kubernetes SIG UI project designed to let you explore, manage, and debug cluster resources. Knative brings serverless workloads to Kubernetes, handling traffic routing, autoscaling, and revision management so teams can deploy and iterate without fighting infras...
How we used DSPy to turn AI evaluations into better responses in Dash chat
We used DSPy to improve LLM judges and optimize our chat experience, creating an evaluation-driven feedback loop that produced better outputs.
Stop sharing access secrets—try Border0 + Tailscale for free
Border0 ties every connection to a real person, securing databases, Kubernetes, SSH, and more.
Boundary 1.0 releases RDP session recording and improved management
Boundary 1.0 releases with support for RDP session recording and a preview ahead towards securing AI agent access in a brave new agentic world.
Scaling without friction: Aliases at project scope in Boundary
Boundary now supports aliases at project scope, aligning access with your organization's infrastructure while enabling teams to scale independently without conflicts.
How we saved over $3 million in idle compute costs with Datadog Kubernetes Autoscaling
See how how multidimensional autoscaling reduced overprovisioning and mitigated reliability risks at scale.