
Search by job, company or skills
Experience : 5+ years
Notice Period: 0-30 Days
Key Responsibilities:
· Operational Stability & SRE Practices
· Maintain high system availability through proactive monitoring and incident prevention
· Define and track SLIs/SLOs to measure service health and reliability
· Lead root cause analysis (RCA) and implement preventive fixes
· Automation & Tooling
· Build and maintain automation for repetitive operational tasks
· Improve deployment pipelines and operational workflows
· Develop scripts/tools to enhance productivity and reduce human intervention
· Service Design & Transition
Job ID: 151732043
Skills:
proactive monitoring , Enhance productivity, SLIs, Develop scripts, Preventive fixes, Incident prevention, Automation Tooling, Operational workflows, Operational Stability SRE Practices, Improve deployment pipelines, Build and maintain automation, SLOs, Reduce human intervention