Baseten

Tech Lead Manager - Model Performance

San Francisco, CA
USD 175k - 275k
C++ Go Docker Kubernetes Machine Learning Spark PyTorch Python
Search for More Jobs Talk to a recruiter now 💪
Description

ABOUT BASETEN

We’re a growing team of builders backed by top-tier investors, including IVP, Spark Capital, Greylock, and Sarah Guo at Conviction. ML teams at enterprises and category-defining AI-native companies like Descript, Bland.ai, Patreon, Writer, and Robust Intelligence use Baseten to power their core production workloads with best-in-class performance, security, and reliability. While we’ve unlocked PMF and secured Series B funding, the ML infrastructure market is massive, and we’re just getting started. If you’re excited to work on engaging and relevant problems while building something new from the ground up, come join us!

THE ROLE

Are you passionate about advancing the frontiers of artificial intelligence while leading a team of exceptional engineers? We are looking for a Tech Lead Manager focused on ML model performance and inference. This role is ideal for someone with a strong engineering background who is eager to lead and mentor a team while remaining hands-on with technology. If you thrive in a fast-paced startup environment and are excited about both leadership and technical challenges, we want to hear from you.

RESPONSIBILITIES:

  • Lead, mentor, and manage a team of engineers focused on developing and optimizing ML model inference and performance.

  • Oversee technical strategy and architecture decisions, driving improvements across our engineering organization.

  • Collaborate with cross-functional teams to ensure seamless integration and scalability of ML models in production environments.

  • Dive into the codebase of frameworks like TensorRT, PyTorch, CUDA, and others to identify and solve complex performance bottlenecks.

  • Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).

  • Own the full lifecycle of projects from inception through delivery, including planning, execution, and resource management.

  • Foster a collaborative, inclusive team environment that encourages continuous learning and growth.

REQUIREMENTS:

  • Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, or a related field.

  • 5+ years of professional experience in software engineering, with at least 2 years in a technical leadership role.

  • Proven experience managing and mentoring teams of engineers.

  • Expertise in one or more programming languages, such as Python, C++, or Go.

  • In-depth understanding of ML model performance optimization, especially using libraries such as PyTorch, TensorRT, and CUDA.

  • Strong knowledge of containerization (Docker) and orchestration systems (Kubernetes).

  • Experience with production-level AI/ML solutions, including scaling and deploying large models.

  • Ability to balance hands-on technical work with team leadership and project management.

BONUS POINTS:

  • Experience enhancing the performance of large language models (LLMs) or similar AI systems.

  • Familiarity with LLM optimization techniques such as quantization, speculative decoding, or continuous batching.

  • Deep knowledge of GPU architecture and performance tuning.

  • Previous experience in a high-growth startup environment.

BENEFITS:

  • Competitive compensation package (Unlimited PTO, 401k, covered healthcare premiums).

  • An opportunity to lead a talented engineering team at a rapidly growing startup in the machine learning space.

  • Inclusive and supportive work culture with ample opportunities for professional development.

  • Exposure to a wide range of ML use cases, offering unmatched learning and networking potential.

Baseten
Baseten
Artificial Intelligence Developer Tools Machine Learning Software Software Engineering

0 applies

3 views

Other Jobs from Baseten

Site Reliability Engineer

Remote San Francisco, CA

AI Support Engineer

Remote San Francisco, CA

Forward Deployed ML Engineer

Remote San Francisco, CA

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 389 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • Salaries for the engineering jobs on our site range from $100K-$200K. On average, senior engineer positions on our EchoJobs are about $160K.
  • The EchoJobs positions have been sourced and vetted from the top companies to work for in the US as a software engineer, including LinkedIn and other reputable job sites. We also have syndicated jobs from companies that have just raised funding, as well as those that have great unique products and culture. From all of these sources, our founder, Morgan, has also resourced the company's authenticity in terms of their website, public appearance, and more.
  • Yes, our users asked us for just this, so now our search filters allow you to search for your top jobs via location, as well as by onsite, remote, or both. Approximately 30% of our jobs are remote, so you’ve got the best options for you!
  • We have not yet implemented this option, but are considering doing so in the future. For the moment, you would need to cancel your subscription, and resubscribe when you wanted to come back.
  • We add new jobs to EchoJobs every day! We scan our sources for the newest jobs, verify them, and post them to EchoJobs within minutes. We add about 2,000-3,000 new jobs for you each day!
  • From starting your job search to getting hired, the entire job search process can take us software engineers anywhere between 3-6 months. However, at EchoJobs, we’re striving to shorten this duration by finding the best, newest jobs for you, so you can do less job searching, and more applying.
  • We’d recommend checking EchoJobs daily, as we add new jobs to the site each day. Additionally, if you got a chance to read our previous email on “what makes EchoJobs different from any other job search tools,” we also recommended that you set a job alert based on your job filters, so if you get emails on those new jobs, you could be checking more than once per day.
  • If you decide to continue with us after the 1-month trial, we definitely recommend this, as we all know it usually takes 3-6 months to find a quality job as a software engineer these days. So to best support you, we just adjusted our membership options at EchoJobs to monthly, 3 months, or 12 months (this option is more for passive job seekers looking a little bit for the future if they want to come back to work or make a job switch potentially. This lets you see what’s out there in case an even better fit job becomes available.)
  • EchoJobs is truly the only job site of its kind. We want to be THE spot for you to find the best job for you, and haven’t encountered any other company doing this. Other job sites are in niches besides software engineering or focus on a small portion of engineering jobs (like a specific coding language). In the words of Morgan, our founder, “I think what makes EchoJobs different is the amount of jobs, frequency that we add new jobs (we add 2,000-3,000 new jobs daily!), and the powerful search engines to find exactly the job you want more easily and efficiently. We can provide you with the most jobs that are vetted by us, we’ll continually find more new jobs for you, and we make it easier for you to apply and get hired.

What Fellow Engineers Say