Domain 3 is about keeping a network healthy, available and well understood over time — the difference between a network that just works and one that quietly falls apart.
Availability and redundancy
Business services are expected to stay up. High availability removes single points of failure with redundancy, load balancing and failover.
High Availability & Redundancy
Keeping services running through failure: load balancing and failover.
Tap or hover a part to learn more.
No single point of failure.
High availability starts by removing single points of failure — duplicate servers, links, power and paths so no one component can take the whole service down.
Check your understanding
1. What does a load balancer do?
2. Roughly how much downtime does 99.9% allow per year?
Monitoring
You can't manage what you can't see. Know the core tools: SNMP (device metrics), syslog (centralised logs), and flow data (NetFlow/sFlow) for traffic analysis. Baselines let you spot when something is abnormal.
Documentation
Network diagrams, IP address management (IPAM), VLAN and port maps, and change logs are what let a team operate and troubleshoot quickly. Undocumented networks are slow and dangerous to change.
Disaster recovery and business continuity
Understand backups, redundancy sites, and metrics like RTO (how quickly you must recover) and RPO (how much data you can afford to lose).
Why this domain matters
Operations is what separates a technician who can build a network from a professional who can run one reliably. Employers value people who monitor proactively, document diligently and plan for failure.
