hackajob is collaborating with J.P. Morgan to connect them with exceptional professionals for this role.
Job Description
Join a dynamic team at the forefront of technology, where your skills drive innovation and modernize mission-critical systems. Experience career growth and make a meaningful impact.
As a Site Reliability Engineer II at JPMorgan Chase within the Enterprise Technology, Engineering Services and Platform team, you will solve complex business problems with straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to iteratively improve existing solutions. You are a significant contributor by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. We foster a collaborative environment where your expertise shapes the future of technology.
Job Responsibilities
Guide and assist others in building appropriate level designs and gaining consensus from peers
Collaborate with software engineers and teams to design and implement deployment approaches using automated CI/CD pipelines
Design, develop, test, and implement availability, reliability, and scalability solutions in applications
Implement infrastructure, configuration, and network as code for applications and platforms
Collaborate with technical experts, stakeholders, and team members to resolve complex problems
Understand service level indicators and utilize service level objectives to proactively resolve issues
Support the adoption of site reliability engineering best practices within your team
Accelerate delivery and improve operational rigor through AI-assisted engineering adoption
Provide 24/7 production support for business-critical applications
Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements.
Applies enterprise-authorized AI capabilities within the work environment to identify recurring toil and reliability risks from operational signals, prioritizing reuse-first improvements and measurable SLO outcomes.
Required Qualifications, Capabilities And Skills
Formal training or certification on software engineering concepts and 2+ years applied experience
Proficient in site reliability culture and principles, and familiarity with implementation within applications or platforms
Proficient in at least one programming language such as Python, Java/Spring Boot, or .Net
Experience in observability including monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Elastic, Splunk
Experience with CI/CD tools like Jenkins, GitLab, or Terraform
Experience with event streaming platforms like Kafka
Experience with AI-assisted tools like GitHub Copilot, Claude
Working knowledge of using enterprise-authorized AI capabilities within the work environment to support SRE workflows (e.g., troubleshooting support and runbook drafting) with strong validation habits and awareness of data sensitivity.
Ability to assess AI-assisted operational recommendations for correctness and risk, and apply appropriate controls to maintain resiliency, security, and auditability.
Preferred Qualifications, Capabilities And Skills
Deep understanding of TCP/IP, DNS, load balancing, firewalls, and VPN technologies
Experience tuning Linux performance and troubleshooting system-level issues
Certifications: AWS Certified SysOps Administrator or Professional, Certified Kubernetes Administrator (CKA), or equivalent
Familiarity with container and orchestration technologies such as ECS, Kubernetes, and Docker
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for Site Reliability Engineer II
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!