NVIDIA

Senior DevOps Engineer

Tel Aviv, Israel
Python Ruby GCP Perl Go Groovy API AWS Azure Oracle Kubernetes
Description

We are seeking a Senior DevOps Engineer to join our Farm team to improve its growing services infrastructure. You will be working with a team of passionate and skilled engineers who are continuously working to provide better tools to build and manage our infrastructure. Our team is a mix of varying levels of experience. We need a motivated, hardworking and focused individual who has a real passion for operational excellence, data systems, and automation.

What you'll be doing:

  • Own the services you build working with cross functional teams
  • Comfortable with frequent code testing and deployment
  • Continuously improve infrastructure provisioning and management using automation
  • Identify areas to improve service resiliency through industry standard practices
  • Support a globally distributed, On-Prem environment (LSF)
  • Determine root-cause for production level incidents and write corresponding high-quality RCA reports
  • Ensure the highest level of up-time and Quality of Service (QoS) to internal customers through operational excellence
  • Participate in team's on-call rotation

What we need to see:

  • B.S. degree in Computer Science or related technical field or equivalent experience
  • 8+ years coding/scripting in at least two high level programming languages - Python, Perl, Go, Ruby, Groovy etc.
  • Build and maintain scalable web applications using modern front-end frameworks, back-end technologies, databases, APIs, and cloud platforms.
  • Good Knowledge in operating services including web servers, load balancers, relational/non-relational databases, messaging systems and storage solutions
  • Deep understanding of linux operation system and TCP/IP fundamental.
  • Knowledge in high-performance computing environments, including job schedulers (e.g., Slurm, PBS, or Grid Engine), parallel computing, and performance tuning.
  • Expertise with at least one major cloud service provider- AWS, GCP, Azure
  • Proficient in implementing and managing monitoring tools like Grafana and Prometheus, ensuring system performance, reliability, and real-time data visualization.
  • Proficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)
  • Detail oriented with great communication and documentation skills

Ways to stand out from the crowd:

  • Develop, fine-tune, and deploy advanced LLM-based solutions for [specific applications, e.g., NLP, chatbots, content generation, or data analysis
  • Linux certification from a well known vendor - RedHat, Oracle etc.
  • Prior experience managing large scale Kubernetes deployment in production
  • Strong skills in modern container networking and storage architecture

NVIDIA
NVIDIA
Artificial Intelligence (AI) GPU Hardware Software Virtual Reality

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

πŸ₯³πŸ₯³πŸ₯³ 401 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. πŸ› οΈ
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. πŸš€
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. πŸ“…

What Fellow Engineers Say