⚡ Site Reliability Engineering (SRE) & Systems Security: Master Curriculum Index
Welcome to the Site Reliability Engineering (SRE) & Systems Security knowledge base. This master curriculum is organized into 3 specialized sub-tracks spanning quantitative reliability management, performance benchmarking & telemetry, and enterprise systems security.
Every guide across all 3 sub-curriculums adheres to a standardized two-part learning format:
- ⚡ Quick Dive: Architecture cheat sheets, calculation formulas, severity matrices, and trade-off summaries.
- 📖 Extended Guide: Deep operational runbooks, mathematical models, eBPF probes, postmortem templates, and security blueprints.
🧭 Sub-Curriculum Directory
| Track Directory | Domain | Guides Count | Highlights & Core Topics | Track Master Link |
|---|---|---|---|---|
01_site_reliability_engineering/ |
SRE Foundations & Operations | 6 Guides | SRE principles, SLIs/SLOs & Error Budgets (burn-rate alerting), SEV-1/2/3 Incident Command, Blameless Postmortems, Toil reduction, Capacity planning (k6), and Chaos Engineering (Chaos Mesh). | 📖 Open SRE Operations Index |
02_observability_and_performance/ |
Observability & Benchmarking | 6 Guides | Production caching failures (thundering herds, stampedes), network protocol dependencies, high-performance database benchmarking, Grafana visualization, Linux eBPF profiling (bpftrace), and distributed tracing tail latency. |
📖 Open Observability Index |
03_systems_security/ |
Systems Security & Data Protection | 5 Guides | Authentication (FIDO2/Argon2id), Encoding vs. Encryption vs. Tokenization, Sensitive PII data management, Zero Trust architecture (BeyondCorp/mTLS), and STRIDE threat modeling. | 📖 Open Systems Security Index |