Beacon by Clearwater is the AI-powered risk analytics and modeling arm of the Clearwater platform, giving institutional investors the tools to test scenarios and evaluate portfolio exposures in real time.
As Clearwater brings Beacon to more clients, the number of client environments we provision, monitor, and support grows with it and the only way that works is through standardization and automation. This team builds the tooling that keeps a growing fleet of client deployments consistent, observable, and supportable: automating away repetitive operational work, turning incident learnings into permanent platform fixes, and giving client-facing teams the self-service tools they need to onboard and support clients without engineering escalations.
What You’ll Do
Build internal tools and automation primarily in Python to monitor, diagnose, and support a fleet of client deployments across AWS and Azure.
Drive standardization across client environments: detect and remediate configuration and infrastructure drift, converge legacy deployments onto golden paths, and make “the standard way” the easy way.
Improve fleet-wide observability: build monitoring, alerting, and dashboards that surface problems across all client deployments before clients notice them.
Turn runbooks into code; converting the manual diagnostic and remediation steps support engineers perform today into automated checks, self-healing jobs, and one-click tools.
Extend the client provisioning and deployment pipeline (Terraform, configuration generation) to make onboarding new clients faster and more repeatable.
Work directly with client-facing teams (onboarding, support, client success) to find where operational toil lives.
What We’re Looking For
3-5 years of experience in software engineering, site reliability engineering, DevOps, or platform engineering.
Strong programming skills in Python (our platform core and tooling language); comfort writing production-quality code with tests, not just scripts.
Hands-on experience with at least one major cloud provider (AWS or Azure): networking (VPCs/VNets, subnets, security groups, load balancers, VPN), IAM/RBAC, storage, and compute.
Working knowledge of infrastructure-as-code, ideally Terraform, and what it means to manage many environments from shared modules and per-environment configuration.
Solid Linux fundamentals: you can read logs, trace a process, debug a service that won’t start, and automate what you did, so no one must do it by hand again.
An automation reflex: when you solve a problem twice, your instinct is to build a tool.
A collaborative, service-oriented mindset: your customers are internal teams, and your success is measured by how much easier you make their jobs.
Nice to Have
Experience operating multi-tenant or fleet-style environments (many similar deployments managed as one).
Exposure to financial services, fintech, or other regulated environments.
Why This Role
Direct, visible impact: every tool you ship makes onboarding the next client faster and supporting every existing client cheaper. This team is a force multiplier for the entire Beacon business.
Breadth: you’ll touch cloud infrastructure, a large Python platform codebase, deployment pipelines, and the human workflows of support and onboarding teams.
Growth: you’ll work across nearly every layer of a sophisticated financial-engineering platform, alongside experts in cloud infrastructure, quantitative finance, and large-scale SaaS operations.
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for Site Reliability Engineer
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!