Site Reliability Engineer I
Backblaze
Let the right jobs find you
Get personalised suggestions from verified company career pages, matched to your role, location, level, and skills.
Overview
Position Type
Full Time
Experience
2+ years
Job Description
The Right Fit:\n- Must be located in Bangalore.\n- 2 - 4 years of relevant experience.\n- Knowledge of Sysadmin and Linux skills.\n- Desire to learn and develop all necessary technical skills.\n- Strong analytical thinking.\n- Strong skills in working with different teams and communication.\n- Knowledge of network cabling, network classification, and network topology.\n\nAt Backblaze, we value being fair and good to our customers, partners, and employees. That’s why diversity, equity, and inclusion are at the core of our values. We are committed to fostering a workforce where all employees feel a sense of belonging regardless of race, ethnicity, nationality, gender, sexual orientation, age, religion, socio-economic status, ability, veteran status, and education. We believe that our dedication to cultivating a diverse workspace not only allows us to better serve our customers in over 175 countries, but further reinforces our commitment to doing the right thing. We are proud to be an Equal Opportunity Employer.\n\nTo understand more about the data we collect and process as part of your application, please view our Backblaze Employee Privacy Notice.\n\nWhat You’ll Do:\n- Act as first point of contact for all customer affecting issues\n- Be a Key Driver for managing the resolution of technical problems\n- Ensure that incident management processes are following and that incident post-mortems are completed to capture process deviations and areas for improvement\n- Deliver consistent communication to Management\n- Respond to zabbix alerts/regular monitoring of zabbix, either by taking direct action on alerts or escalating. Acknowledge every alert if direct action taken, or with escalation point of contact.\n- Make sure escalations are handed off successfully.\n- Ensure health of pods across all sites (define pod alerts on zabbix).\n- Work through daily filesystem checks for pods.\n- Troubleshoot technical issues for DC Techs -> advanced pod questions, deployment questions, migration troubleshooting, and ansible playbook issues.\n- Identification and escalating any potential issues regarding the network.\n- Vault pre-deployment configuration and testing.\n- Start Vault Migrations, monitor migration pods, handle applicable migration pod health checks.\n- Document/Work on automating Daily Items.\n- Document/Provide Network IP's for upcoming deployments.\n- Monitor Releases/Updates to the Server Farm, escalate issues as they arise.\n- Engaging in on-call rotation shifts.\n- Assist fellow TechOps team members in handling tasks.\n- Making recommendations for improvements in organizational productivity.\n- Be able to work outside of normal business hours(weekend shift, holidays & evenings) as needed\n