Genesis AI logo

Inference Engineer

Genesis AI

On-site
Bay Area
Full-time
Staff
8+ yrs
Salary not listedPosted 2mo ago

Real job — pulled straight from Genesis AI’s careers page · Verified August 11, 2026 · No reposts.

Job description

Genesis AI is hiring a Inference Engineer — a full-time, based in Bay Area role. Apply directly on Genesis AI's careers page below.

Inference

Department: Engineering & Research

Location: Bay Area

Employment Type: FullTime

What You’ll Do

  • Build low-latency inference pipelines for on-device deployment, enabling real-time next-token and diffusion-based control loops in robotics

  • Design and optimize distributed inference systems on GPU clusters, pushing throughput with large-batch serving and efficient resource utilization

  • Implement efficient low-level code (CUDA, Triton, custom kernels) and integrate it seamlessly into high-level frameworks

  • Optimize workloads for both throughput (batching, scheduling, quantization) and latency (caching, memory management, graph compilation)

  • Develop monitoring and debugging tools to guarantee reliability, determinism, and rapid diagnosis of regressions across both stacks

What You’ll Bring

  • Deep experience in distributed systems, ML infrastructure, or high-performance serving (8+ years)

  • Production-grade expertise in Python, with strong background in systems languages (C++/Rust/Go)

  • Low-level performance mastery: CUDA, Triton, kernel optimization, quantization, memory and compute scheduling

  • Proven track record scaling inference workloads in both throughput-oriented cluster environments and latency-critical on-device deployments

  • System-level mindset with a history of tuning hardware–software interactions for maximum efficiency, throughput, and responsiveness

Get Inference Engineer jobs like this→

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
GSK: Home logo

Portfolio Data Science Director

London, UK
✓ From careers page· 2d ago
GSK: Home logo

Data Analytics and Insight Manager

Ho Chi Minh City, Vietnam
✓ From careers page· 2d ago
GSK: Home logo

Platform & Engineering Senior Manager

London, UK
✓ From careers page· 2d ago
SteerBridge logo

Junior NLP Data Scientist

Vienna, VA
✓ From careers page· 2d ago

Frequently asked questions

What skills are required for Inference Engineer at Genesis AI?

The required skills for Inference Engineer at Genesis AI include: Python, C++, Rust, Go, Machine Learning.

What is the seniority level for Inference Engineer at Genesis AI?

Inference Engineer at Genesis AI is a Staff level position.

How do I apply for Inference Engineer at Genesis AI?

You can view the full description and apply for Inference Engineer at Genesis AI on EchoJobs: https://echojobs.io/job/genesis-ai-inference-ud1oh.