We are looking for a Senior DevOps Engineer to join our Data and Application Services team to improve its growing services infrastructure. At the core of our application services platform is our multi-tenant Kubernetes platform that is designed to run a variety of inhouse application services. You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure. Our team is a mix of varying levels of experience and CS backgrounds. We need a motivated, hardworking and focused individual who has a real passion for operational excellence, data systems, and automation.
What you'll be doing:
Own the services you build working with cross functional teams
Comfortable with frequent code testing and deployment
Continuously improve infrastructure provisioning and management using automation
Identify areas to improve service resiliency through industry standard practices
Support a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-prem
Determine root-cause for production level incidents and write corresponding high-quality RCA reports
Ensure the highest level of up-time and Quality of Service (QoS) to internal customers through operational excellence
Define service level objectives (SLOs) and service level indicators (SLIs) to represent and measure service quality
Participate in team's on-call rotation
What we need to see:
7+ years in operating services including web servers, load balancers, relational/non-relational databases, messaging systems and storage solutions
3+ years coding/scripting in at least two high level programming languages - Python, Go, Ruby, Groovy etc.
Deep understanding of linux operation system and TCP/IP fundamentals
Expertise with at least one major cloud service provider- AWS, GCP, Azure
Proficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)
Hands on experience running production quality observability stacks
Creative problem solver with excellent debugging skills
B.S. degree in Computer Science or related technical field (or equivalent experience)
Detail oriented with great communication and documentation skills
Ways to stand out from the crowd:
Linux certification from a well known vendor - RedHat, Oracle etc.
Prior experience managing large scale Kubernetes deployment in production
Strong skills in modern container networking and storage architecture
You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.
Other Jobs from NVIDIA
Senior Mixed Signal Design Engineer
Senior Data Processing Platform Engineer
System Validation Engineer
Decision Making and Planning Software Intern - 2025
Similar Jobs
Senior Lead Software Engineer, DevOps/SRE
Senior DevOps Engineer
Senior Platform Engineer
Platform Technical Architect (AWS)
Platform Solution Architect (AWS)
There are more than 50,000 engineering jobs:
Subscribe to membership and unlock all jobs
Engineering Jobs
60,000+ jobs from 4,500+ well-funded companies
Updated Daily
New jobs are added every day as companies post them
Refined Search
Use filters like skill, location, etc to narrow results
Become a member
π₯³π₯³π₯³ 401 happy customers and counting...
Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.
To try it out
For active job seekers
For those who are passive looking
Cancel anytime
Frequently Asked Questions
- We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
- We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
- We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
- We're the only job board *for* software engineers, *by* software engineersβ¦ in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. π οΈ
- Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. π
- Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. π―
- Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. π
What Fellow Engineers Say