Applications | Datadog

What Are Feature Flag Best Practices for AI-Native Teams?

10 best practices for using feature flags as a control plane for safety, cost, and velocity in AI-native software delivery.

What Is AI Agent Observability?

The importance of AI engineering (LLMOps) for measuring success metrics, improving velocity, and resolving issues for agentic applications and systems.

What Is AI-Led Incident Response?

How AI transforms incident detection, coordination, and postmortem learning, and the maturity path teams follow to get there.

What Is Digital Experience Monitoring (DEM)?

How complete visibility across the user journey, from synthetic testing to real user monitoring to product analytics, connects technical performance to business outcomes.

What is Root Cause Analysis?

Root-cause analysis can help IT teams determine the root causes of a problem or incident rather than just addressing its symptoms.

What Is an MCP Server (Model Context Protocol Server)?

Learn about the advantages of MCP servers for making use of AI models, services, and external tools for back-end services, observability, and security.

What is Workflow Automation?

Monitor your technology and business metrics to enable data-driven decisions.

What Is Observability? Pillars, Tools, and Use Cases

How observability works, what signals it uses, and how to evaluate platforms.

What is Cloud Architecture Diagramming? How it Works & Use Cases

Use cloud architecture diagrams for collaboration, planning, and a shared view at scale for new architectures, cloud migrations, cost evaluations, and optimizations for existing environments.

What Are DORA Metrics?

Ingest, monitor, apply, and act on DevOps Research and Assessment (DORA) metrics to identify issues in software development, release processes, and continuous integration/continuous delivery (CI/CD) workflows.

What Is Static Analysis?

Discover how to leverage Static Analysis solutions to improve the quality of your organization's code.

What is Shift Left & Shift Left Testing?

Take control of your software development, and bring feature changes and improvements to production faster, using preemptive testing patterns.

What Is Pair Programming & How Does It Work?

Learn how pair programming can be used to improve outcomes in software development

OpenTelemetry Overview

Learn how OpenTelemetry standardizes observability data collection across distributed systems — and why it has become the default instrumentation standard for modern infrastructure.

Incident Management Overview

Learn how incident management can help your organization customize and streamline your incident response process.

What Is a Flame Graph? How to Read One, With Examples

How to read a flame graph, and how tracing and profiling flame graphs differ.

What Is Distributed Tracing? How It Works and Tools

How traces and spans give you end-to-end visibility into a request's path through your stack.

...
...