NVIDIA logo

Site Reliability Engineer

NVIDIA

On-site
Bengaluru, India
Full-time
Entry
Salary not listedPosted 1h ago

Real job — pulled straight from NVIDIA’s careers page · Verified August 22, 2026 · No reposts.

Job description

NVIDIA is hiring a Site Reliability Engineer — a full-time, based in Bengaluru, India role. Apply directly on NVIDIA's careers page below.

Site Reliability Engineer

Location: India, Bengaluru

Time Type: Full time

Job Description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people.

Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

What you'll be doing:

  • Support and contribute to SRE initiatives that improve reliability, scalability, and developer efficiency across enterprise systems.

  • Assist in building and maintaining distributed systems that power NVIDIA’s AI-powered enterprise products and services, learning modern architectural patterns along the way.

  • Help automate database operations — including provisioning, scaling, backup, and failover — for relational and vector database services.

  • Contribute to observability and monitoring efforts by building dashboards, alerts, and automation scripts to improve system performance and reliability.

  • Participate in incident response processes, learning to triage issues, reduce mean time to resolution (MTTR), and contribute to post-incident reviews.

  • Collaborate with Cloud, Platform, Security, and AI/ML teams to support platform reliability and help implement SRE best practices.

  • Learn to operate and troubleshoot complex systems — including Kubernetes-based and cloud-native infrastructure — following established standards in system design and incident management.

  • Explore and adopt AI-assisted engineering practices, including coding agents and LLM-powered tooling, to accelerate day-to-day development workflows.

What we need to see:

  • BS degree in Computer Science or a related technical field (e.g., physics, mathematics), or equivalent practical experience.

  • Foundational proficiency in at least one programming language such as Python, TypeScript, JavaScript, or Go.

  • Basic understanding of cloud platforms (AWS, Azure, or GCP) and containerisation technologies like Docker and Kubernetes.

  • Exposure to or coursework in infrastructure-as-code tools (e.g., Terraform, AWS CDK, CloudFormation) or willingness to learn.

  • Familiarity with Linux/Unix systems, networking fundamentals, and version control (Git).

  • Interest in observability concepts (logging, metrics, tracing) and tools such as OpenTelemetry, Prometheus, or Grafana.

  • Basic knowledge of relational databases (e.g., PostgreSQL, MySQL) — understanding of SQL, indexing, and simple query optimisation.

  • Strong problem-solving skills, curiosity, and a willingness to learn in a fast-paced, collaborative environment.

  • Good communication and teamwork skills, with the ability to ask the right questions and learn from senior engineers.

Ways to stand out from the crowd:

  • Personal projects, internships, or coursework involving cloud infrastructure, automation, or DevOps/SRE practices.

  • Contributions to open-source projects or active participation in hackathons, coding competitions, or technical communities.

  • Exposure to AI/ML concepts — e.g., building or deploying a simple ML model, experimenting with LLM APIs, or using AI-powered developer tools (Copilot, Cursor, etc.).

  • Hands-on experience with CI/CD pipelines, scripting for automation, or container orchestration (even in personal or academic projects).

  • A strong sense of ownership, curiosity, and initiative — you turn challenges into learning opportunities and aren’t afraid to dive into unfamiliar systems.

NVIDIA leads the charge in innovative breakthroughs in Artificial Intelligence, High-Performance Computing, and Visualization. The GPU, our invention, functions as the visual cortex of today’s computers and forms the core of our products and services. Our work opens new realms to explore, encourages outstanding creativity and discovery, and powers inventions once thought of as science fiction — from artificial intelligence to autonomous systems. NVIDIA is searching for outstanding talent like you to help us advance the next wave of artificial intelligence!

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/


 

Get Site Reliability Engineer jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Mastercard US logo

Senior Cloud Engineer

$115k–$184kAtlanta, GA
✓ From careers page· 10m ago
NVIDIA logo

Software Engineer, Infrastructure

$124k–$242kSanta Clara, CA
✓ From careers page· 10m ago
NVIDIA logo

Senior Security Engineer, Detection Engineering

$168k–$311kRemote · US-eligible
✓ From careers page· 11m ago
NVIDIA logo

Senior Software Engineer, AI Inference Systems

$170k–$275kToronto, ON
✓ From careers page· 14m ago

Frequently asked questions

What skills are required for Site Reliability Engineer at NVIDIA?

The required skills for Site Reliability Engineer at NVIDIA include: Python, TypeScript, JavaScript, Go, AWS, Azure, GCP, Docker, Kubernetes, Terraform, Git, OpenTelemetry, Prometheus, Grafana, PostgreSQL, MySQL, SQL, CI/CD.

What is the seniority level for Site Reliability Engineer at NVIDIA?

Site Reliability Engineer at NVIDIA is a Entry level position.

How do I apply for Site Reliability Engineer at NVIDIA?

You can view the full description and apply for Site Reliability Engineer at NVIDIA on EchoJobs: https://echojobs.io/job/nvidia-site-reliability-engineer-wzov4.