zafinlabsamericasinc
Site Reliability Engineer I
At a Glance
- Location
- India - Trivandrum
- Experience
- 8+ years
- Department
- Cloud Engineering
- Posted
- 2026-05-14T02:38:33-04:00
Key Requirements
Required Skills
Domain Knowledge
- Engineering
- Regulatory
Requirements
8+ years of experience in cloud support, operations, or a related role.
Advanced expertise in Microsoft Azure (preferred) or equivalent cloud platforms.
Demonstrated experience in designing and scaling container orchestration systems like AKS or OpenShift.
Proven leadership in managing automated deployment pipelines, including Azure DevOps.
Mastery in enterprise monitoring platforms (e.g., Azure Insights, Grafana) and predictive analytics tools.
Advanced scripting skills with PowerShell, Python, or similar languages.
Responsibilities
Manage the resolution of complex technical issues involving Zafin’s products and Azure cloud environment.
Design and implement strategic operational enhancements to improve resiliency and system reliability.
Conduct in-depth Root Cause Analysis (RCA) for high-severity incidents and drive initiatives to reduce error recurrence.
Represent the organization in external client escalation calls, providing expert guidance and solutions.
Optimize cloud infrastructure for high performance, scalability, and cost-effectiveness.
Provide thought leadership in managing and scaling container orchestration platforms such as AKS and OpenShift.