Search by job, company or skills

Site Reliability Engineer (SRE) Multiple Opportunities, Tech-driven and Innovative Environment

2-4 Years
Early Applicant
  • Posted 13 days ago
  • Be among the first 10 applicants

Job Description

My Clients:

I am currently supporting multiple technology companies in Singapore, including leading global internet platforms, large-scale consumer technology companies, and fast-growing AI infrastructure teams, who are actively expanding their Site Reliability Engineering (SRE) functions.

We are hiring across different SRE tracks, including large-scale production reliability, cloud-native infrastructure, platform engineering, and AI / GPU infrastructure.

Depending on your background, you may work on high-traffic consumer platforms, global distributed systems, Kubernetes-based cloud infrastructure, or next-generation AI computing platforms.

Job Responsibilities:

  • Design, operate, and improve the reliability of large-scale distributed systems and production infrastructure.
  • Build monitoring, observability, alerting, and automated recovery mechanisms to improve system availability.
  • Define and improve SLA / SLO / SLI frameworks, capacity planning, disaster recovery, and incident response processes.
  • Develop automation tools and internal platforms to improve infrastructure efficiency and reduce operational workload.
  • Deploy and operate cloud-native infrastructure using Kubernetes, containers, and public / hybrid cloud environments.
  • Troubleshoot complex production issues across applications, systems, networks, and infrastructure.

Job Requirements:

  • Bachelor's degree or above in Computer Science, Engineering, or a related technical discipline.
  • 2+ years of experience in SRE, Infrastructure Engineering, DevOps, Production Engineering, or related roles.
  • Strong Linux and production troubleshooting experience.
  • Hands-on experience with Kubernetes, Docker, and cloud platforms such as AWS, GCP, Azure, or OCI.
  • Familiarity with monitoring and observability systems, distributed systems, and high-availability architecture.
  • Programming or scripting experience in Python, Go, Bash, or similar languages.
  • Experience supporting large-scale internet platforms, high-traffic systems, or AI / GPU infrastructure is highly preferred.

What They Offer:

  • Exposure to large-scale production systems, global platforms, and cutting-edge AI computing environments.
  • Opportunities to specialize in Platform SRE, Infrastructure SRE, Cloud-Native Engineering, or AI Infra Reliability.
  • Competitive compensation and strong technical career growth opportunities.

About Us

Dada Consultants was established in 2017, with the commitment of providing the best recruitment services in Singapore. We are comprised of a dynamic head-hunting team dedicated to sourcing for highly competent professionals in IT industry. We provide enterprises with customized talent solutions, and bring talents to career advancement.

If it sounds like your next move, please don't hesitate to apply. Kindly note that only shortlisted candidates will be contacted. Appreciate your understanding. Data provided is for recruitment purposes only.

www.dadaconsultants.com

EA Registration Number: R26160730

Business Registration Number: 201735941W. Licence Number: 18S9037

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 151209489