Critical Infrastructure Engineer – Reliability & Infrastructure Analytics
eClerx
Full TimeNot specifiedPosted about 22 hours ago
Let the right jobs find you
Get personalised suggestions from verified company career pages, matched to your role, location, level, and skills.
Overview
Position Type
Full Time
Experience
Not specified
Job Description
Key Responsibilities
- Lead and manage a team of Critical Infrastructure Engineers/Technicians, ensuring effective shift coverage, operational governance, coaching, workload balancing and adherence to operational standards.
- Analyze infrastructure trends and proactively identify recurring alarms, noisy sites, aging infrastructure and operational gaps impacting service reliability.
- Drive Top Talker analysis, identify recurring failure patterns and develop permanent corrective action plans to eliminate repeat alarms and unnecessary dispatches.
- Review historical incidents, alarm trends and backlog tickets to identify systemic issues and recommend engineering improvements.
- Prepare and present analytical reports, dashboards and executive summaries highlighting recurring issues, infrastructure risks, equipment performance and operational improvements.
- Identify End-of-Life (EOL) / End-of-Support (EOS) equipment, evaluate operational risk and prepare structured equipment refresh and replacement plans for management review.
- Coordinate with Engineering, Procurement, Vendors and Field teams to drive permanent remediation, equipment replacement and closure of long-pending infrastructure issues.
- Maintain ownership of backlog reduction initiatives and ensure ageing tickets are continuously reviewed, tracked and driven to closure.
- Support continuous improvement initiatives by identifying opportunities for automation, operational efficiencies and proactive monitoring enhancements. This aligns with the CI objective of moving from repeat break-fix to asset-led preventive management.
Additional Required Skills
- Strong analytical skills with the ability to interpret operational data, identify trends and convert findings into actionable recommendations.
- Ability to identify recurring alarms, noisy sites and infrastructure risks using data analysis and operational reporting.
- Strong leadership and people-management skills with experience managing technical teams, conducting coaching sessions and driving operational governance.
- Excellent documentation skills with the ability to create and maintain detailed operational records, RCA reports, equipment history, maintenance documentation and knowledge articles that can be used for future troubleshooting, audits and continuous improvement.
- Ability to prepare executive-level reports and present infrastructure risks, backlog analysis, EOL/EOS recommendations and permanent-fix roadmaps to management.
- Strong stakeholder-management skills with the ability to coordinate effectively across Operations, Engineering, Vendors, OEMs and Field teams.
Preferred Experience
- Must have prior experience leading technical teams in a Telecom, Data Center or Critical Infrastructure environment.
- Experience preparing SOPs, MOPs, RCA documentation, equipment history records, maintenance logs and audit documentation.
- Experience working with BMS/DCIM/EMS platforms and infrastructure monitoring tools.
- Exposure to asset lifecycle management, equipment refresh programs and preventive maintenance planning.
- Experience analyzing recurring infrastructure incidents and implementing permanent corrective actions.