Role Title: Cloud Observability SRE – Senior Engineer
Location: Offshore – L3 Lead
Note: This role requires working one shift aligned to US business hours, preferably 10 AM CST – 6 PM CST (+/- 2 hrs. before or after).
Time : 8:30 PM - 4:30 AM (IST)
PRIMARY JOB FUNCTION
The Cloud Observability SRE – Senior Engineer is responsible for implementing, maintaining, and evolving cloud observability services for Abbott's Enterprise Cloud environments. This role operates within a DevSecOps model, enabling continuous integration, deployment, and scaling of observability practices. The engineer will deliver end‑to‑end observability solutions including requirements gathering, design, integration using Infrastructure-as-Code, and operational support. The role also involves evaluating vendor technologies, participating in proofs of concept, and ensuring adherence to observability best practices with a focus on resiliency, scalability, security, and automation. The ideal candidate is self-driven, highly collaborative, an effective communicator, and passionate about innovation, cost-efficient solutions, and bringing new ideas to the team.
CORE JOB RESPONSIBILITY
- IESDeploy and configure observability platforms using best practices, reusable IaC templates, SOPs, and standardized configuratio
- ns.Build dashboards for performance, availability, service health, synthetic monitoring, alerting, and SLO/SLI management across applications, platforms, and infrastructu
- re.Apply AI/ML‑enabled observability tools and data-driven approaches to detect anomalies, generate predictive insights, and perform deep analysis on latency, reliability, error rates, MTTD, MTBF, MTTR, and other key metri
- cs.Provide SME-level support to cross-functional teams for incident triage and troubleshooting across the full cloud application/infrastructure sta
- ck.Assist DevOps teams with troubleshooting cloud-native applications and infrastructure issues using observability insights, including end‑user experience metri
- cs.Enhance overall cloud service reliability, resiliency, deployment quality, CI/CD pipeline health, and visibility into IaaS/PaaS service observability across cloud providers
MINIMUM EXPERIENCE / TRAINING / SKI
- Bachelor's degree in technology, information systems, or related discipline (preferred).
- 6+ years of Observability/Monitoring SRE experience.
- AWS and/or Azure cloud certifications (required).
- Certified Observability Expert (required).
MUST HAVE SKILLS
- Experience with AI and Non-AI observability/monitoring tools such as New Relic (preferred), Datadog, Dynatrace, LogicMonitor.
- Hands-on experience integrating with ServiceNow ITSM for ticketing and event correlation.
- Strong knowledge of AWS/Azure IaaS/PaaS services including compute, storage, and networking.
- Proficiency with Terraform (IaC), Ansible (Configuration Management), Python, JSON.
- Experience with Git/GitHub Copilot, CI/CD platforms (Jenkins, Azure DevOps), and Big Panda (Event Management).
- Familiarity with cloud-native backup/recovery and site-recovery solutions for AWS/Azure.
- Basic understanding of database platforms including AWS RDS, Azure SQL, PostgreSQL, MongoDB Atlas, Cosmos DB.
- Exposure to RAG applications and LLM model evaluation platforms such as Azure OpenAI and AWS Bedrock.
PREFERRED (NICE-TO-HAVE) SKILLS
- Experience with Jira (Scrum/Sprint Management) and Confluence (Documentation).
- Understanding of Kubernetes, Docker, and CaaS platforms on Azure/AWS.