For over 25 years, NVIDIA has been revolutionizing computer graphics, PC gaming, and accelerated computing. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people.
At NVIDIA, we are seeking a highly skilled Senior Engineer Operations Manager to join our world-class NGC Cloud team. In this role, you will help drive the efficiency, reliability, and scalability of the systems that power our global business operations. This is an exceptional opportunity to shape how we automate, streamline, and support critical operational workflows across the organization. You will define how we implement innovative automation and support solutions, enabling teams to operate seamlessly and deliver impact at global scale—all within an encouraging and inclusive environment.
What You'll Be Doing
Lead strategy, execution and operations for cloud services that provide container, artifact, and ML model registry capabilities for NVIDIA engineering teams.
Architect, Design, plan, implement and Operate complex PaaS for the GPU cloud services.
Partner with AI infrastructure, security, product, and engineering teams to define standards for model packaging, image build pipelines, artifact management, and deployment workflows.
Define and track KPIs and SLAs for registry and related services, including availability, latency, storage efficiency, reliability, and developer experience.
Drive operational excellence for secure, scalable cloud services
Mentor and coach engineering managers and senior individual contributors; build a strong leadership bench and a healthy, inclusive engineering culture.
Influence architecture and technical direction while empowering teams to own detailed design and implementation decisions.
Communicate clearly with senior leadership on strategy, risks, execution progress, and outcomes; represent the platform in multi-functional planning discussions.
What We Need To See
10+ overall years of software engineering experience with significant ownership of cloud platforms, distributed systems, developer platforms, artifact registry systems, storage systems, or infrastructure services.
5+ years of engineering leadership experience, including experience managing managers or leading multiple senior technical workstreams through other leaders.
Bachelors degree or equivalent experience.
Proven success operating customer-facing or company-critical services with demanding availability, latency, throughput, data integrity, and security expectations.
Strong technical judgment in registry, artifact management, or developer platform systems.
Deep cloud-native systems background across Kubernetes, object storage, relational or NoSQL databases, event streaming, caching, API design, service-to-service authentication, and observability.
Experience with containerization and registries such as Docker, Kubernetes, Docker Hub, Harbor, ECR, GCR, GAR, or similar technologies at enterprise scale.
Excellent people leadership skills, including hiring, performance management, career development, and building diverse, high-performing teams.
Strong communication skills, with experience presenting to senior executives and influencing cross-organization priorities.
Ways To Stand Out From The Crowd
Directly led teams building and operating distributed cloud platforms at enterprise scale.
You have led reliability transformations for large services, including SLO adoption, forecasting resource needs, load testing, and stress testing.
Worked in AI infrastructure, accelerated computing, model distribution, or regulated enterprise software delivery.
Track record of growing managers and senior technical leaders who can own ambiguous, high-impact platform areas without constant blocking issue.
Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/
Searching, interviewing and hiring are all part of the professional life. The TALENTMATE Portal idea is to fill and help professionals doing one of them by bringing together the requisites under One Roof. Whether you're hunting for your Next Job Opportunity or Looking for Potential Employers, we're here to lend you a Helping Hand.
Disclaimer: talentmate.com is only a platform to bring jobseekers & employers together.
Applicants
are
advised to research the bonafides of the prospective employer independently. We do NOT
endorse any
requests for money payments and strictly advice against sharing personal or bank related
information. We
also recommend you visit Security Advice for more information. If you suspect any fraud
or
malpractice,
email us at abuse@talentmate.com.
You have successfully saved for this job. Please check
saved
jobs
list
Applied
You have successfully applied for this job. Please check
applied
jobs list
Do you want to share the
link?
Please click any of the below options to share the job
details.
Report this job
Success
Successfully updated
Success
Successfully updated
Thank you
Reported Successfully.
Copied
This job link has been copied to clipboard!
Apply Job
Upload your Profile Picture
Accepted Formats: jpg, png
Upto 2MB in size
Your application for Senior Manager Cloud Services Platform
has been successfully submitted!
To increase your chances of getting shortlisted, we recommend completing your profile.
Employers prioritize candidates with full profiles, and a completed profile could set you apart in the
selection process.
Why complete your profile?
Higher Visibility: Complete profiles are more likely to be viewed by employers.
Better Match: Showcase your skills and experience to improve your fit.
Stand Out: Highlight your full potential to make a stronger impression.
Complete your profile now to give your application the best chance!