NOC Engineer
Thiruvananthapuram, India · On-site · Full-time
- Posted 1mo ago
- From Armada’s careers page
- Location
- Thiruvananthapuram, India
- Work mode
- On-site
- Type
- Full-time
- Experience
- 5+ years
- Department
- Information Technology
Opens the listing on job-boards.greenhouse.io
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.
About the role
Infrastructure Monitoring & Incident Response\n- Monitor critical physical infrastructure using PLC, BMS, and DCIM platforms such as Distech, Radix IoT/Mango, Schneider Electric, Siemens, or similar systems.\n- Monitor and respond to alerts involving power, cooling, network connectivity, environmental conditions, and facility infrastructure.\n- Provide Tier 2 support for escalations from L1 technicians and coordinate escalation to engineering, facilities, vendors, or other teams as required.\n- Own incidents through their lifecycle, from initial triage and troubleshooting through resolution, stakeholder communication, and documentation.\n- Support incident response and post-incident reviews, identifying opportunities to improve operational reliability and response procedures.\n- Develop and refine monitoring dashboards, alerting thresholds, and operational health indicators.\n\nData Center Operations\n- Perform and coordinate routine health checks of critical infrastructure, including UPS systems, PDUs, CRAC/CRAH units, backup generators, and environmental monitoring systems.\n- Troubleshoot and coordinate resolution of mechanical and electrical infrastructure issues using a working knowledge of MEP systems.\n- Read and interpret technical documentation, including electrical one-line diagrams, network diagrams, schematics, and equipment documentation.\n- Coordinate scheduled and emergency maintenance activities with internal teams, vendors, and remote hands personnel.\n- Participate in change management processes and assess the operational and infrastructure impact of proposed changes.\n- Ensure physical and logical security policies and procedures are followed within data center and edge environments.\n\nModular & Edge Infrastructure\n- Operate and support modular, containerized, micro, and distributed edge data center environments.\n- Help maintain availability, continuity, and resiliency across geographically distributed and remotely operated infrastructure.\n- Coordinate remote troubleshooting and hands-on support for edge deployments.\n- Implement and continuously improve operational practices for remote infrastructure monitoring, fault detection, escalation, and recovery.\n- Support environmental sensors and IoT-based monitoring integrations across remote infrastructure.\n\nTools, Automation & Documentation\n- Use operational platforms such as Grafana, Zoho Desk, Zenduty, ServiceNow, Jira, SolarWinds, or similar tools for monitoring, analysis, ticketing, and incident management.\n- Maintain accurate documentation of configurations, incidents, changes, maintenance activities, and recurring operational tasks.\n- Support automation and reporting initiatives focused on infrastructure performance, availability, capacity, and uptime.\n- Assist with gathering operational evidence and infrastructure data for compliance and audit requirements.\n- Identify opportunities to automate repetitive operational activities using scripting and monitoring integrations.\n\nCollaboration & Operational Excellence\n- Partner closely with Facilities, Infrastructure Engineering, Networking, IT, Security, and other operational teams to maintain reliable infrastructure.\n- Communicate clearly during incidents, maintenance activities, escalations, and operational handoffs.\n- Participate in incident reviews and root cause analysis and help drive corrective and preventive actions.\n- Contribute to the development of runbooks, standard operating procedures, escalation processes, and operational best practices.\n- Participate in an on-call or shift-based operating model as required to support 24x7 infrastructure.\n\nRequired Qualifications\n- 5+ years of experience in NOC, data center operations, critical infrastructure operations, or a related technical role.\n- Strong working knowledge of PLC, BMS, and/or DCIM platforms.\n- Hands-on experience monitoring or supporting power and cooling infrastructure in data center or other mission-critical environments.\n- Working knowledge of mechanical and electrical systems, including UPS, PDU, CRAC/CRAH, generators, and environmental monitoring systems.\n- Solid understanding of networking fundamentals, infrastructure monitoring, alerting, and monitoring protocols.\n- Experience with critical infrastructure monitoring platforms such as Distech, Radix IoT/Mango, Schneider Electric, Siemens, or equivalent systems.\n- Experience with incident management, troubleshooting, escalation, and operational change management.\n- Ability to read and interpret electrical one-lines, network schematics, technical diagrams, and equipment documentation.\n- Strong analytical, troubleshooting, communication, and documentation skills.\n- Ability to work effectively in a shift-based and/or on-call environment, including nights and weekends when required.\n\nPreferred Qualifications\n- Experience operating or supporting modular, containerized, micro, or edge data centers.\n- Experience supporting geographically distributed infrastructure and coordinating remote hands activities.\n- Familiarity with environmental sensors, telemetry, IoT integrations, and remote monitoring systems.\n- Working knowledge of Linux and Windows server environments.\n- Basic scripting and automation experience using Python, PowerShell, Bash, or similar technologies.\n- Experience with monitoring and operational platforms such as Grafana, ServiceNow, Jira, SolarWinds, Zenduty, or Zoho Desk.\n- Familiarity with compliance, audit evidence collection, and operational security requirements.\n- Certifications such as CDCP, CDCTP, CompTIA Server+, CCNA, or equivalent industry certifications.\n- Highly organized, detail-oriented, and comfortable operating in fast-paced, mission-critical environments.
Skills they ask for
Pick one to see other roles that ask for it.
About Armada
AI infrastructure for the edgeArmada builds compute infrastructure and software for deploying AI and data workloads at edge locations.
See all 16 roles at ArmadaMore roles at Armada
See all 16- Senior Mechanical Design Engineer – Liquid Cooling SystemsUnited States · Senior · RemoteResearch and Development (R&D) · Senior · RemoteUnited States1w
- Data Center Structural EngineerUnited States · Senior · RemoteEngineering · Senior · RemoteUnited States1w
- Vice President, People & TalentUnited States · Executive · RemoteHuman Resources · Executive · RemoteUnited States4w
- Vice President, People & TalentUnited States · Executive · RemoteHuman Resources · Executive · RemoteUnited States4w
Let the right jobs find you
In your inbox every Wednesday and SaturdayPersonalised suggestions from verified career pages, matched to your role, location, level and skills.