Disk Usage Monitoring
df -h shows free space per partition; du -sh * shows what is hogging space — together they keep your servers healthy.
Introduction
df -h shows free space per partition; du -sh * shows what is hogging space — together they keep your servers healthy.
Beginner analogy: think of Unix as a kitchen. The shell is the chef who reads your order, the kernel is the stove and fridge that actually cook and store, files are the ingredients, and pipes are the conveyor belts moving food from one chef to the next. Every Unix command you learn is one well-designed kitchen tool.
In this lesson we will walk through Disk Usage Monitoring step by step, see exactly how Linux handles it under the hood, look at the practical commands you will type every day on real servers, study a real-world DevOps scenario, and finish with the interview questions you will absolutely face when applying to AWS, Google Cloud, Red Hat, Netflix, Stripe and every modern infrastructure team.
Understanding the topic
Core concepts to understand:
- 🧠 Clear definition and mental model of disk usage monitoring.
- 🐧 How the Linux kernel and shell collaborate to make it happen.
- 📂 Where files, processes and configuration live on a standard Linux server.
- 🔁 How disk usage monitoring fits inside scripts, cron jobs and CI/CD pipelines.
- 🛡 Permissions, users, groups and least-privilege practices around disk usage monitoring.
- 🚧 Common pitfalls: missing quotes, unset variables, wrong exit codes, dangerous
rm -rf. - 🏢 Real production scenarios at AWS, Netflix, Stripe and modern SaaS DevOps teams.
Syntax reference
Visual workflow / architecture:
User Command|vShell Interpreter|vCommand Parsing|vKernel Interaction|vSystem Resources|vCommand Execution|vTerminal Output
df -h && free -m && uptimePull live metrics from /proc and standard tools.
Informative example
Hands-on commands you can copy-paste:
tar -czf creates a gzipped archive in one step. Always tag backups with $(date +%F) so rotation scripts can sort and prune them. du -sh + sort -h finds disk hogs in seconds.
# Backup with compressiontar -czf /backup/app-$(date +%F).tar.gz /opt/appls -lh /backup/# Restoremkdir -p /opt/app && tar -xzf /backup/app-2026-05-12.tar.gz -C /# Disk usage of foldersdu -sh /var/* 2>/dev/null | sort -h | tail -5
Sample terminal output:
-rw-r--r-- 1 root root 142M May 12 02:00 app-2026-05-12.tar.gz8.0K /var/local2.4M /var/cache1.2G /var/log4.7G /var/lib
Walk-through: notice how every Unix tool prints structured text and returns an exit code (0 = success, anything else = failure). That is the contract that lets you chain commands with &&, pipe them with |, and trust them inside automation. Reading these messages carefully is the difference between a senior Linux engineer and a junior one.
Real-world use
In production, Disk Usage Monitoring is part of every infrastructure engineer's daily flow at companies like AWS, Google Cloud, Netflix, Stripe, Shopify, GitHub and Red Hat. Engineers SSH into Linux servers, write small focused bash scripts, schedule them with cron or systemd timers, monitor them in Grafana and ship them through CI/CD. Mastering disk usage monitoring means safer deploys, faster incident response and dramatically fewer 3 AM pages.
Best practices
- Always start scripts with
#!/bin/bashandset -euo pipefailso they fail fast on errors and unset variables. - Quote variables:
"$file"not$file— protects against spaces and word-splitting bugs. - Use absolute paths in cron, scripts and systemd units —
$PATHis minimal in those environments. - Log to
/var/log/<app>/and rotate withlogrotateso disks never fill up. - Run as the least-privileged user; reserve
sudofor the few commands that truly need root.
Common mistakes
rm -rf $VAR/when$VARis empty — wipes the whole filesystem. Always quote and validate.- Cron jobs that run from a fresh shell with no
$PATH— your script works manually but fails at 2 AM. - Forgetting
2>&1on logs — silent failures because stderr was thrown away. - Editing config files without taking a backup (
cp file file.bak) — no way to roll back.
Hands-on exercise
Interview preparation — practice these questions:
- Q1. Explain Disk Usage Monitoring in one sentence as if to a junior teammate.
- Q2. Walk through the exact Linux commands you would run for disk usage monitoring on a production server.
- Q3. What is the difference between Unix and Linux, and where does disk usage monitoring live in the stack?
- Q4. How would disk usage monitoring behave inside a cron job vs an interactive shell, and why?
- Q5. Name two security or permission concerns around disk usage monitoring and how you would mitigate them.
- Q6. How does disk usage monitoring integrate with monitoring, logging and a CI/CD pipeline?
- Q7. Scenario: a 3 AM PagerDuty alert says disk usage monitoring failed in production. Walk me through your debugging.