Back to All Roles
OperationsMid Level

NOC Engineer

Monitors data center infrastructure and network operations 24/7, responding to alerts, troubleshooting issues, and ensuring maximum uptime. First line of defense for infrastructure incidents.

Experience

2-5 years

Growth

High Growth

Growing demand with the expansion of cloud infrastructure

Education: Associate or Bachelor's degree in IT or related field

Preferred Certs: CCNA, COMPTIA NETWORK+

Key Responsibilities

  1. Monitor infrastructure and network health 24/7
  2. Respond to and escalate incidents appropriately
  3. Perform routine maintenance tasks
  4. Document issues and resolutions
  5. Coordinate with on-site technicians

Skills & Expertise

Required Skills

  • Network monitoring
  • DCIM tools
  • Incident management
  • Linux/Windows administration
  • Ticketing systems

Preferred Skills

  • Scripting (Python, Bash)
  • Cloud platforms
  • Virtualization

Certifications

Preferred

CCNACompTIA Network+

A Day in the Life

What a typical workday looks like

My day begins at 7:15 AM with a strong coffee and logging into the Nagios monitoring dashboard. I quickly scan the overnight alerts from our Phoenix data center, noting a potential thermal anomaly near Rack 14 in the west corridor. The DCIM system shows slight temperature variations that didn't trigger critical thresholds, but warrant investigation. I prioritize a physical inspection of the CRAC unit supporting that zone and review the shift handover notes from the night team. By 10:30 AM, I'm deep in incident response mode. A Tier 1 ticket from our cloud services team requires immediate attention - a backbone switch in Cluster B is showing intermittent packet loss. I initiate remote diagnostics through our Cisco ACI platform, coordinating with the network engineering team to determine whether this is a hardware issue or routing configuration problem. Simultaneously, I begin drafting a preliminary incident report in ServiceNow, documenting each diagnostic step meticulously. Midday brings a scheduled maintenance window for our backup power systems. I join a conference bridge with facilities and infrastructure teams to coordinate a planned firmware upgrade on the UPS units in Modules 7 and 8. We carefully sequence the updates to ensure redundant power paths remain active, using our change management protocol to minimize potential service disruption. The team reviews rollback procedures and confirms backup generator readiness. Around 2:45 PM, I dive into a detailed root cause analysis for a recurring network performance issue we've tracked over the past week. Using Wireshark and our NetFlow analysis tools, I correlate traffic patterns and identify potential bottlenecks in our east-west network connectivity. I prepare a comprehensive report with recommended topology adjustments, annotating specific interface configurations that might require optimization. As my shift winds down, I meticulously prepare the handover brief for the evening team. I update our central tracking system with all open items, highlighting the switch diagnostics, thermal monitoring status near Rack 14, and the network performance investigation. A quick verbal briefing with the incoming NOC engineer ensures smooth continuity. I confirm all critical systems are stable, log my final observations, and begin my end-of-shift system checkout routine.

Career Progression

Related Roles