Role Title : AVP, Site Reliability Engineer - CloudOps (L10)
Company Overview: Synchrony (NYSE: SYF) is a leading consumer financing company that has been at the heart of American commerce and opportunity for nearly a century. Synchrony delivers credit and banking products that empower tens of millions of consumers to improve their financial lives and access what matters most. Leveraging innovative solutions that are shaping the future of retail commerce, Synchrony supports the growth and success of some of the nation’s most respected brands, alongside hundreds of thousands of small and midsize businesses, including health and wellness providers. Committed to excellence in service and culture, Synchrony is proud to be named as #3 as a Great Place to Work® in India and is honored to be ranked the #1 Best Company to Work For® in the U.S. by Fortune magazine and Great Place to Work®. For more information, visit www.synchrony.com.
Organizational Overview
The CloudOps SRE team ensures reliability, stability, and performance of cloud and hybrid platforms through monitoring, incident response, and automation. Working closely with engineering and infrastructure teams, it enables secure, scalable operations and supports ongoing technology modernization.
Role Summary/Purpose
We are seeking an experienced Site Reliability Engineer (SRE) – CloudOps at the AVP level to join our Technology organization. The incumbent will be responsible for ensuring the reliability, scalability, performance, and operational excellence of mission-critical fintech platforms hosted across AWS, Pivotal Cloud Foundry (PCF/Tanzu), and Hybrid On-Premise environments. The role demands a strong engineering mindset with a passion for automation, observability, and continuous improvement, while partnering closely with Development, Architecture, and Infrastructure teams to embed reliability into every layer of the product lifecycle.
Key Responsibilities
Reliability Engineering: Define, measure, and govern Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Error Budgets across business-critical applications; drive engineering practices that uphold them.
Automation & Toil Reduction: Identify operational toil and engineer scalable automation solutions using
Cloud Migration & Modernization: Lead and contribute to migration of workloads from on-premise / PCF estates to AWS-native and containerized (Kubernetes/Docker) platforms, ensuring resilience and zero-business-impact cutovers.
Observability: Build and enhance the observability stack across Prometheus, New Relic, Grafana, Splunk, CloudWatch and create actionable dashboards, golden signals, and intelligent alerting that reduce noise and accelerate triage.
Incident & Change Management: Own major incident response, root cause analysis, and post-incident reviews; uphold ITIL-aligned change, problem, and release management disciplines.
Required Skills/Knowledge
Minimum 5+ years of hands-on experience in SRE, CloudOps, DevOps, or Production Engineering roles, preferably within the Fintech / Banking / Financial Services domain with overall 7+ years of industry experience.
Minimum 5+ years of expertise across Cloud Platforms: AWS (EC2, EKS, S3, RDS, IAM, CloudWatch, VPC), PCF / Tanzu, and Hybrid On-Premise environments.
Containers & Orchestration: Kubernetes and Docker at production scale.
Infrastructure as Code: Terraform (modules, state management, governance).
CI/CD Tooling: Jenkins, GitHub Actions, and ArgoCD (GitOps).
Observability: Prometheus, Grafana, Splunk, Dynatrace, and the ELK stack.
Databases: Oracle, PostgreSQL, MongoDB, Cassandra (operations and performance).
Practical knowledge of ITIL-aligned Incident, Problem, and Change Management.
Proven experience driving SLO/SLI frameworks, error budgets, and reliability KPIs in a regulated enterprise.
Demonstrated success in cloud migration, modernization, and FinOps-led cost optimization programs.
Strong analytical, problem-solving, and stakeholder communication skills with the ability to engage technology and business leadership.
Desired Skills/Knowledge
Hands-on experience in AWS/PCF/Hybrid environments with strong understanding of production support and cloud operations.
Working knowledge of Kubernetes, Docker, and Terraform (or equivalent IaC) for deployment and automation.
Experience with CI/CD tools (Jenkins/GitHub Actions/ArgoCD) and observability platforms (Prometheus, Grafana, Splunk, Dynatrace, or ELK).
Good incident/change management practices (ITIL-aligned), basic scripting (Shell/Python), and strong cross-team communication skills.
Eligibility Criteria
B.E. / B.Tech in Computer Science, Information Technology, or a related engineering discipline.
AWS Certification (e.g., AWS Certified Solutions Architect / SysOps Administrator / DevOps Engineer) will be considered an added advantage.
Work Timings: 2:00 PM to 11:00 PM IST
This role qualifies for Enhanced Flexibility offered in Synchrony India and will require the incumbent to be available between 06:00 AM Eastern Time – 11:30 AM Eastern Time (timings are anchored to US Eastern hours and will adjust twice a year locally). This window is for meetings with India and US teams. The remaining hours will be flexible for the employee to choose. Exceptions may apply periodically due to business needs) We are proud to offer flexibility at Synchrony. Our way of working allows you the option to work from home or workspaces in our Regional Engagement Hubs—Hyderabad, Bengaluru, Pune, Kolkata, or Delhi/NCR. Occasionally you may be required to commute or travel to Hyderabad or one of the Regional Engagement Hubs for in person engagement activities such as business or team meetings, trainings, and culture events.
For Internal Applicants
Understand the criteria or mandatory skills required for the role, before applying
Inform your manager and HRM before applying for any role on Workday
Ensure that your professional profile is updated (fields such as education, prior experience, other skills) and it is mandatory to upload your updated resume (Word or PDF format)
Must not be any corrective action plan (First Formal/Final Formal)
L8+ Employees who have completed 18 months in the organization and 12 months in current role and level are only eligible.
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for AVP Site Reliability Engineer - CloudOps L10
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!