We are looking for a Site Reliability Engineer who can design, build, and operate resilient infrastructure and CI/CD pipelines across multiple cloud platforms (AWS, Azure, GCP). This role is central to enabling reliable, scalable, and secure deployment of applications and AI/data platforms, with a strong focus on automation, observability, and infrastructure-as-code across heterogeneous cloud environments.
Site Reliability Engineer (Multi-Cloud Infrastructure & Pipelines)
Role Title: Site Reliability Engineer (SRE) - Multi-Cloud Infrastructure & CI/CD
Define and implement SRE practices: SLIs/SLOs/SLAs, error budgets, incident response, and on-call processes
Set up monitoring, logging, and observability stacks (Prometheus, Grafana, Datadog, Cloud Monitoring/Stackdriver, ELK/EFK) across environments
Architect and manage container orchestration platforms (Kubernetes — EKS/AKS/GKE) and containerization (Docker) for portable, cloud-agnostic workloads
Automate provisioning, scaling, patching, and configuration management (Ansible, Chef, or Puppet as needed)
Implement disaster recovery, backup, high-availability, and multi-region/multi-cloud failover strategies
Drive cost optimization and capacity planning across cloud providers
Build self-healing systems and automate incident remediation to reduce toil
Conduct root cause analysis (RCA) and post-incident reviews; drive reliability improvements
Implement security best practices: IAM policies, secrets management (Vault, Cloud KMS), network security, and compliance guardrails across clouds
Collaborate with development teams to embed reliability, observability, and deployment best practices into the software delivery lifecycle (DevSecOps/shift-left)
Support migration or workload portability initiatives between cloud providers or hybrid/on-prem environments
Requirements
Required Skills & Experience
5+ years of experience in SRE, DevOps, or Infrastructure Engineering roles
Proven hands-on experience across at least two of the three major cloud providers (AWS, Azure, GCP) — true multi-cloud experience strongly preferred
Strong expertise in Infrastructure-as-Code (Terraform mandatory; CloudFormation/Bicep/Pulumi a plus)
Deep knowledge of Kubernetes and container orchestration in production environments
Strong scripting/programming skills (Python, Go, or Bash)
Experience building and maintaining CI/CD pipelines end-to-end (build, test, deploy, rollback)
Solid understanding of networking fundamentals (VPCs, load balancers, DNS, service mesh — Istio/Linkerd a plus)
Experience with observability and monitoring tooling, and setting up alerting that ties to meaningful SLOs
Familiarity with GitOps practices (ArgoCD, FluxCD)
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for SRE Engineer
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!