LINUX OBSERVABILITY AND PERFORMANCE
Monitor and Troubleshoot Servers with eBPF, perf, Prometheus, Grafana, System Metrics, Flame Graphs, Tracing, and Performance Profiling
When a Linux server slows down, the real challenge is not finding activity. It is identifying which resource is actually limiting useful work, which process is responsible, and what evidence proves the root cause.
LINUX OBSERVABILITY AND PERFORMANCE provides a practical, production-focused guide to diagnosing Linux performance problems using modern observability, profiling , tracing, and monitoring tools.
Rather than teaching isolated commands, this book develops a repeatable troubleshooting workflow:
Detect -> Measure -> Localize -> Trace -> Profile -> Fix -> Validate
1. You will begin with native Linux performance analysis and learn how to investigate CPU utilization, scheduler pressure, memory behavior, swapping, storage latency, filesystem exhaustion, network problems, processes, cgroups v2, and Pressure Stall Information.
2. From there, you will move into deeper profiling with perf, hardware counters, call stacks, scheduler analysis, and flame graphs. You will then use eBPF, BCC, and bpftrace to investigate scheduler latency, file access, slow storage operations, TCP retransmissions, system calls, kernel activity, application functions, and on-CPU and off-CPU behavior.
3. The book also shows how to build a persistent Linux observability platform with Node Exporter, Prometheus, PromQL, Grafana, and Alertmanager. You will create operator-focused dashboards, recording rules, actionable alerts, and multi-host monitoring workflows designed to support real incident response.
4. Modern telemetry is covered through OpenTelemetry, Grafana Alloy, and Pyroscope. You will learn how to collect and correlate metrics, traces, and continuous profiles so that a Grafana alert can lead to a slow trace, a historical flame graph, and ultimately the exact code path or system behavior responsible for degraded performance.
5. Hands-on labs throughout the book reinforce each major skill. You will deliberately create CPU saturation, scheduler contention, memory pressure, disk latency, network degradation, application regressions, cgroup throttling, and competing workloads , then diagnose and remediate them using measurable evidence.
The final capstone brings everything together into a complete Linux observability and performance platform using:
Native Linux performance tools perf and hardware counters Flame graphs eBPF, BCC, and bpftrace Node Exporter Prometheus and PromQL Grafana and Alertmanager OpenTelemetry Grafana Alloy Pyroscope Continuous profiling Metrics, traces, and profile correlation Failure injection and performance validation
By the end of the book, you will be able to move from vague symptoms such as "the server is slow" to an evidence-based conclusion that explains what changed, where the bottleneck exists, which workload caused it, how it was fixed, and how the improvement was validated.
This book is designed for Linux system administrators, DevOps engineers, SREs, platform engineers, infrastructure engineers, cloud engineers, performance engineers, and production support professionals who want a practical and modern approach to Linux observability and performance troubleshooting.
"synopsis" may belong to another edition of this title.
Seller: California Books, Miami, FL, U.S.A.
Condition: New. Print on Demand. Seller Inventory # I-9798177984773
Seller: PBShop.store US, Wood Dale, IL, U.S.A.
PAP. Condition: New. New Book. Shipped from UK. Established seller since 2000. Seller Inventory # L2-9798177984773
Seller: PBShop.store UK, Fairford, GLOS, United Kingdom
PAP. Condition: New. New Book. Shipped from UK. Established seller since 2000. Seller Inventory # L2-9798177984773
Quantity: Over 20 available