
Why Application Performance Management (APM) is critical for modern applications and production systems.
The Four Golden Signals of monitoring and how they help you detect performance issues early.
The benefits of a well-designed monitoring system, including faster troubleshooting, better reliability, and improved user experience.
What Datadog is and how it fits into modern observability and monitoring stacks.
Why Datadog is needed in real production environments and what problems it solves.
Key benefits of using Datadog, including faster troubleshooting, better visibility, and improved system reliability.
How Datadog compares to traditional monitoring tools and why teams choose it for modern applications.
What a host, agent, and tags are in Datadog and how they relate to each other.
How the Datadog Agent works and why it is required for collecting metrics and traces.
Why tags are important for filtering, grouping, and analyzing monitoring data.
How these basic concepts are used in real monitoring scenarios throughout the course.
Learn how the Datadog agent exposes host and system metrics, uses default cost dashboards, and filters costs with user defined tags like location, enabling targeted monitoring across infrastructure.
Learn how metrics summary and types work in Datadog, filtering by timeframe and tags, and understanding distribution, count, rate, and gauge for reporting and graphing.
Master Datadog APM services and the service map to monitor inter-service calls between web and database components, with traces, error tracking, and latency metrics P50, P90, P99.
Learn to send custom tags, including username, from a .NET app to Datadog using the Datadog trace NuGet package, with a helper logger for set tag and set exception.
Learn to use APM custom tags, search, filter, and facets in Datadog by tagging MVC requests with dog request ID and username, turning them into searchable facets for future calls.
Create and customize a Datadog dashboard with timeboard or screenboard layouts, add widgets such as timeseries, query values, and tables, and monitor latency, error rate, endpoints, and service maps.
Create and manage monitors in Datadog to prevent incidents by configuring host, metrics, APM, and watchdog monitors, then test notifications and review status and history.
You will learn how integrate Datadog with Pager Duty and receive an alert
You will understand what is SLA, SLO, SLI and Error Budget
1. You will be able to create SLO
2. To make Date Correlation on the SLO
3. Set up Alerts based on the SLO
This course will help you to:
Understand basic and advanced concepts of Application Performance Management (APM) and Datadog tool usage.
Build APM for your application from scratch using Datadog.
Visualize the entire request path and quickly identify where bottlenecks or errors occur.
Track application errors and slow queries with just a few clicks.
Install and configure the Datadog Agent and query essential system and application metrics.
Analyze Datadog APM services, trace searches, and code profiling, including .NET Core API monitoring with SQL service layer visibility.
Create dashboards using different widgets, including:
Time series
Query value
Top lists
Tables and pie charts
We will build dashboards together for:
Latency (time series & query value)
Error rate (time series & query value)
Top API endpoints
Host resource usage
Service map visualization
Applying formulas for advanced insights
Create monitors and alerts for hosts, services, endpoints, IIS, and Watchdog:
Request latency
Error rate
Specific API endpoints
SQL query duration
Host data reporting
Watchdog monitors
Understand SLA, SLO, SLI, and Error Budgets by creating Datadog SLOs for success rate tracking.
Integrate monitors with Slack and PagerDuty for real-time incident notifications.
Use Synthetic Monitoring with API and browser tests, including
GET/POST requests with Bearer token authorization.
Manage logs in Datadog, including saving and viewing logs using NLog.
Use Datadog Notebooks to:
Investigate incidents
Collaborate with team members using comments
Create and reuse custom templates
Monitor key application aspects such as availability, reliability, scalability, and performance duration.
Set up and monitor a Node.js application with MongoDB, including:
Service configuration
Custom metrics
Distributed tracing
Learn the latest Datadog updates and features used in modern production environments.