Baseten

Engineering Manager - Model Performance

San Francisco, CA
USD 175k - 275k
Kubernetes Machine Learning Spark PyTorch Python C++ Go Docker
Description

ABOUT BASETEN

Join our dynamic team at Baseten, where we’re revolutionizing AI deployment with cutting-edge inference infrastructure. Backed by premier investors such as IVP, Spark Capital, Greylock, and Conviction, we’re trusted by leading enterprises and AI-driven innovators—including Descript, Bland.ai, Patreon, Writer, and Robust Intelligence—to deliver top-tier performance, security, and reliability for their production workloads. With our recent $75 million Series C funding, we’re poised to accelerate our mission to make AI accessible across all products. If you’re passionate about tackling impactful challenges and building transformative solutions from the ground up, we invite you to join us on this exciting journey!

THE ROLE

Are you passionate about advancing the frontiers of artificial intelligence while leading a team of exceptional engineers? We are looking for a Tech Lead Manager focused on ML performance and inference. This role is ideal for someone with a strong engineering background who is eager to lead and mentor a team while remaining hands-on with technology. If you thrive in a fast-paced startup environment and are excited about both leadership and technical challenges, we want to hear from you.

RESPONSIBILITIES:

  • Lead, mentor, and manage a team of engineers focused on developing and optimizing ML model inference and performance.

  • Oversee technical strategy and architecture decisions, driving improvements across our engineering organization.

  • Collaborate with cross-functional teams to ensure seamless integration and scalability of ML models in production environments.

  • Dive into the codebase of frameworks like TensorRT, PyTorch, CUDA, and others to identify and solve complex performance bottlenecks.

  • Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).

  • Own the full lifecycle of projects from inception through delivery, including planning, execution, and resource management.

  • Foster a collaborative, inclusive team environment that encourages continuous learning and growth.

REQUIREMENTS:

  • Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, or a related field.

  • 5+ years of professional experience in software engineering, with at least 2 years in a technical leadership role.

  • Proven experience managing and mentoring teams of engineers.

  • Expertise in one or more programming languages, such as Python, C++, or Go.

  • In-depth understanding of ML model performance optimization, especially using libraries such as PyTorch, TensorRT, and CUDA.

  • Strong knowledge of containerization (Docker) and orchestration systems (Kubernetes).

  • Experience with production-level AI/ML solutions, including scaling and deploying large models.

  • Ability to balance hands-on technical work with team leadership and project management.

BONUS POINTS:

  • Experience enhancing the performance of large language models (LLMs) or similar AI systems.

  • Familiarity with LLM optimization techniques such as quantization, speculative decoding, or continuous batching.

  • Deep knowledge of GPU architecture and performance tuning.

  • Previous experience in a high-growth startup environment.

BENEFITS:

  • Competitive compensation package (Unlimited PTO, 401k, covered healthcare premiums).

  • An opportunity to lead a talented engineering team at a rapidly growing startup in the machine learning space.

  • Inclusive and supportive work culture with ample opportunities for professional development.

  • Exposure to a wide range of ML use cases, offering unmatched learning and networking potential.

Baseten
Baseten
Artificial Intelligence Developer Tools Machine Learning Software Software Engineering

0 applies

22 views

Other Jobs from Baseten

Machine Learning Engineer - Fine Tuning

San Francisco, CA New York, NY

Software Engineer - Internal Platform

Montreal, Canada San Francisco, CA

Infrastructure Software Engineer

Remote San Francisco, CA

AI Support Engineer

San Francisco, CA New York, NY

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say

Sid avatar
Sid
Very nice portal for searching jobs in this rough market.
Mar 6, 2025
Michael Duran avatar
Michael Duran
Software Engineer
I've been using this job search site for a while now, and it’s honestly one of the best out there! The clean and easy-to-navigate UI makes the whole job-hunting process so much smoother. Plus, the job postings are always up-to-date, so I never feel like I’m wasting time. The cherry on top is the owner—super kind and always quick to respond. Definitely recommend checking it out if you're on the job hunt!
Aug 21, 2024
Sai avatar
Sai
It’s really great website for finding jobs based on skills it’s really helpful give a go
Aug 21, 2024
Adinadh avatar
Adinadh
What I like most about Echo Jobs is how easy it is to use. The platform helps me quickly find jobs that match my skills and interests, thanks to its great recommendations and filters. Yes, I would definitely recommend Echo Jobs to a friend. It makes job searching simple and efficient, making it a great tool for anyone looking for a new job.
Jul 23, 2024
As a student navigating the job market, I've found LinkedIn increasingly frustrating due to numerous fake postings by consultancies. In contrast, this job posting website has been a game-changer for me. It offers genuine opportunities and a straightforward application process, making it much easier to find and apply for real jobs. Highly recommend it to fellow students seeking reliable job listings!
Jul 16, 2024
Cliff Gor avatar
Echo Jobs has been exceptional in my job hunt where it provides one platform to job hunt and I don't have to open 10 websites just to look for a job. It has also helped me focus much on the job skill and the location filtering out the onsite jobs and remote ones. The only feature that I would request is to display fully remote jobs that are not restricted to a country since the one available shows ie, Remote, US yet. But if it could show remote only, that would be helpful not only to me but to other people applying for full remote and not tied to only US candidates
Apr 22, 2024
I found EchoJobs in 2022, and I love it. It has a lot of remote jobs. It's exclusive to software and technology jobs (helpful for devs like me). What I like the most are its filters and its API. If you're a tech professional seeking remote work, I highly recommend giving it a try to EchoJobs.
Mar 4, 2024
Would definitely recommend it! Excellent product, dedicated founder, Jobs are easier to find. Congrats 🎉 to the entire team!
Mar 3, 2024
Brandon Banks avatar
Brandon Banks
Echo Jobs is really impressive. It provides a great user experience with an ability to quickly search through the many job postings. There is an impressive amount of jobs here and it is quickly updated. The details in the each job posting is helpful when determining if it is worth pursuing. I would highly recommend using Echo Jobs to find the next step in your career.
Mar 2, 2024
Tyler Young avatar
Tyler Young
tylerayoung.com
Best wishes with EchoJobs—it's become my favorite job board overnight!
Dec 16, 2023
Simply put, it's the most up to date tech jobs aggregator I’ve found. I'm like... "I don't have to check 10+ jobs boards daily just to see if there's a new job listing? sign me up!" The filters are also quite helpful! The UI is very clean and straightforward. Love it!
Oct 5, 2023