Deep Infra

Software Engineer

Palo Alto, CA
USD 150k - 195k
Python C++ CUDA PyTorch TensorFlow Git AI Machine Learning Transformers Diffusers NCCL
Description

Software Engineer

Location: Palo Alto, CA, USA

Department: Engineering

Location Type: IN_OFFICE

Employment Type: FULL_TIME

DeepInfra is looking for early-career Software Engineers to join our team. You’ll work on designing, building, and scaling infrastructure for serving top open-source AI models in production. This role is ideal for engineers who are already comfortable owning problems end-to-end and want to deepen their experience working on high-impact AI systems.

If you’re excited about AI/ML, have built and shipped projects, and are looking to work on real systems at scale — we’d love to meet you.

What You’ll Do


  • Design, develop, and test inference solutions for state-of-the-art AI models
  • Implement, optimize, and evaluate AI models using Python, C++, CUDA, and NCCL
  • Own and operate production model-serving systems, including monitoring and debugging
  • Build new features, improve system performance, and contribute to overall system design
  • Participate in code reviews and technical discussions to maintain high engineering standards
  • Explore and apply new AI/ML techniques to improve model performance and efficiency
  • Take ideas from concept to production

What You Bring


  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field
  • 1–4 years of relevant experience, including early full-time roles or research
  • Strong fundamentals in data structures, algorithms, and software design
  • Proficiency in Python and experience working with AI/ML frameworks (e.g., PyTorch, TensorFlow)
  • Hands-on experience building, shipping, and maintaining software systems
  • Familiarity with AI models, Transformers, and Diffusers
  • Experience working with version control (Git) and collaborative development workflows
  • Ability to debug, optimize, and improve existing systems
  • Strong communication skills and ability to work independently in a fast-paced environment

Bonus


  • Experience with C++, CUDA, or AI inference
  • Contributions to open-source ML projects

Why DeepInfra


  • Work on cutting-edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high-performance inference at scale.
  • Fast-paced environment with ownership, autonomy, and end-to-end responsibility.

Annual base salary range
$150,000 - $195,000
Deep Infra
Deep Infra

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say