As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications.
As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business.
What you’ll do
Design observability capabilities that give engineering teams clear insight into application health, platform performance, and issues affecting users
Build scalable collection pipelines for metrics, logs, and traces across cloud-based and on-premises environments
Develop actionable alerting standards that reduce noise, shorten incident response, and highlight the most important signals
Partner with application and infrastructure teams to define service health indicators and improve operational readiness before production launches
Automate monitoring configuration, dashboard deployment, and reliability checks to support consistent observability across the technology environment
Analyze production incidents to identify telemetry gaps and improve detection, diagnosis, and recovery
Create dashboards and reporting views that help teams understand trends, capacity risks, and reliability outcomes
Establish practical observability patterns, documentation, and self-service guidance for engineers across the organization
Contribute to a global technology team that improves the resilience of systems supporting the firm’s investment business
What’s Required
5-17 years of experience in observability, site reliability, systems engineering, production engineering, or infrastructure engineering roles
Experience using an administrator such as Datadog and Python for programing
Practical experience designing and operating telemetry solutions for metrics, logs, traces, alerting, and dashboards
Strong understanding of distributed systems, cloud-based environments, containerized workloads, and modern application architectures
Proficiency with at least one scripting or programming language for automation, tooling, and operational workflows
Experience deploying observability patterns across production environments with high availability and performance expectations
Experience troubleshooting production issues across applications, networks, operating systems, and infrastructure services
Working knowledge of incident management practices, root cause analysis, and continuous improvement of operational processes
Strong communication skills, with the ability to explain technical issues clearly and collaborate across global engineering teams
Commitment to the highest ethical standards
We take care of our people
We invest in our people, their careers, their health, and their well-being. When you work here, we provide:
Health care benefits
Maternity, Adoption & related leave policies
Generous paternity and family care leave policies
Employee Assistance Program & Mental wellness programs
Transportation support
Tuition assistance
About Point72
Point72 is a leading global alternative investment firm led by Steven A. Cohen. Building on more than 30 years of investing experience, Point72 seeks to deliver superior returns for its investors through fundamental and systematic investing strategies across asset classes and geographies. We aim to attract and retain the industry's brightest talent by cultivating an investor-led culture and committing to our people's long-term growth. For more information, visit https://point72.com/.
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for Observability Engineer
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!