Job description
The Role
We are seeking a highly skilled Cloud Platform Engineer who is responsible for architecture , deployment and operations of our cloud platform. You will sit at the intersection of development and operations, building the abstractions and automation that allow our engineering teams to self-serve infrastructure while maintaining strict security and operational standards.
Key Responsibilities
- Infrastructure Automation (IaC): Design and maintain scalable cloud ( GCP/AWS) infrastructure using Infrastructure as Code principles, primarily Terraform.
- Platform Management: Provision and manage core GCP services, including Shared VPCs, IAM, Compute Engine, Cloud SQL, GCS, Pub/Sub, and Cloud Run.
- Security & Governance: Implement and enforce security controls such as IAM least-privilege models, VPC Service Controls, and Organization policies.
- CI/CD Pipeline Development: Standardize deployment pipelines using GitOps workflows to minimize manual intervention and ensure consistent environment provisioning.
- Observability & Monitoring: Design and maintain monitoring, alerting, and logging solutions using Google Cloud Monitoring and Logging to interpret service-level metrics and diagnose bottlenecks.
- Developer Enablement: Create developer-friendly abstractions and documentation (runbooks, templates ) to enable efficient self-service.
Requirements
Required Skills & Qualifications
- GCP Expertise: Deep understanding of Google Cloud Platform architecture, including networking (subnets, firewalls, PSC), compute, and managed services.
- Kubernetes & microservices deployment : Strong understanding of GKE setup & deployment with enterprise architectures for service mesh.
- Automation: Proven experience with Terraform for infrastructure provisioning and Python or Bash for scripting operational tasks.
- Networking & Security: Strong grasp of networking protocols (TCP/IP, DNS, SSL) and cloud security principles (IAM, encryption at rest/transit, Zero Trust).
- DevOps Philosophy: Experience with modern DevOps cultures, CI/CD tools (e.g., Google Cloud Build, Jenkins), and containerization (Kubernetes).
- Problem-Solving: Ability to debug issues across distributed systems and tune platform configurations for optimal performance and cost.
Preferred Experience
- Experience managing Shared VPC and multi-tenant environments with strict vertical isolation.
- Background in building Internal Developer Platforms (IDPs) that treat developers as customers.
- Experience in building Agents using ADK, MCP and working with LLMs ( Gemini, Claude) for agent led operations.
Benefits
Why Join TANUH
You will be managing the infrastructure that directly powers life-saving AI models. Your work will enable faster, more accurate detection of diseases like oral cancer, breast cancer, and diabetes, transforming frontline healthcare across India.