Critical Infrastructure Engineer
eClerx
Let the right jobs find you
Get personalised suggestions from verified company career pages, matched to your role, location, level, and skills.
Overview
Position Type
External
Experience
Not specified
Job Description
Role Purpose
The Critical Infrastructure Engineer is responsible for monitoring, managing and improving the reliability of critical infrastructure supporting telecom and data centre operations. The role focuses on ensuring high availability of power, cooling and environmental systems while proactively identifying infrastructure risks, recurring alarms and opportunities for permanent corrective action. The engineer is expected to move beyond reactive incident handling by driving backlog reduction, equipment refresh planning and long-term infrastructure reliability improvements. This aligns with the current CI direction of moving from repeat break-fix to asset-led preventive management.
Key Responsibilities
- Lead and manage a team of Critical Infrastructure Engineers/Technicians, ensuring effective shift coverage, operational governance, coaching, workload balancing and adherence to operational standards.
- Analyze infrastructure trends and proactively identify recurring alarms, noisy sites, aging infrastructure and operational gaps impacting service reliability.
- Drive Top Talker analysis, identify recurring failure patterns and develop permanent corrective action plans to eliminate repeat alarms and unnecessary dispatches.
- Review historical incidents, alarm trends and backlog tickets to identify systemic issues and recommend engineering improvements.
- Prepare and present analytical reports, dashboards and executive summaries highlighting recurring issues, infrastructure risks, equipment performance and operational improvements.
- Identify End-of-Life (EOL) / End-of-Support (EOS) equipment, evaluate operational risk and prepare structured equipment refresh and replacement plans for management review.
- Coordinate with Engineering, Procurement, Vendors and Field teams to drive permanent remediation, equipment replacement and closure of long-pending infrastructure issues.
- Maintain ownership of backlog reduction initiatives and ensure ageing tickets are continuously reviewed, tracked and driven to closure.
- Support continuous improvement initiatives by identifying opportunities for automation, operational efficiencies and proactive monitoring enhancements. This aligns with the CI objective of moving from repeat break-fix to asset-led preventive management.