ABOUT BASETEN
We’re a growing team of builders backed by top-tier investors, including IVP, Spark Capital, Greylock, and Sarah Guo at Conviction. ML teams at enterprises and category-defining AI-native companies like Descript, Bland.ai, Patreon, Writer, and Robust Intelligence use Baseten to power their core production workloads with best-in-class performance, security, and reliability. While we’ve unlocked PMF and secured Series B funding, the ML infrastructure market is massive, and we’re just getting started. If you’re excited to work on engaging and relevant problems while building something new from the ground up, come join us!
THE ROLE
Are you passionate about advancing the frontiers of artificial intelligence while leading a team of exceptional engineers? We are looking for a Tech Lead Manager focused on ML performance and inference. This role is ideal for someone with a strong engineering background who is eager to lead and mentor a team while remaining hands-on with technology. If you thrive in a fast-paced startup environment and are excited about both leadership and technical challenges, we want to hear from you.
RESPONSIBILITIES:
Lead, mentor, and manage a team of engineers focused on developing and optimizing ML model inference and performance.
Oversee technical strategy and architecture decisions, driving improvements across our engineering organization.
Collaborate with cross-functional teams to ensure seamless integration and scalability of ML models in production environments.
Dive into the codebase of frameworks like TensorRT, PyTorch, CUDA, and others to identify and solve complex performance bottlenecks.
Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
Own the full lifecycle of projects from inception through delivery, including planning, execution, and resource management.
Foster a collaborative, inclusive team environment that encourages continuous learning and growth.
REQUIREMENTS:
Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, or a related field.
5+ years of professional experience in software engineering, with at least 2 years in a technical leadership role.
Proven experience managing and mentoring teams of engineers.
Expertise in one or more programming languages, such as Python, C++, or Go.
In-depth understanding of ML model performance optimization, especially using libraries such as PyTorch, TensorRT, and CUDA.
Strong knowledge of containerization (Docker) and orchestration systems (Kubernetes).
Experience with production-level AI/ML solutions, including scaling and deploying large models.
Ability to balance hands-on technical work with team leadership and project management.
BONUS POINTS:
Experience enhancing the performance of large language models (LLMs) or similar AI systems.
Familiarity with LLM optimization techniques such as quantization, speculative decoding, or continuous batching.
Deep knowledge of GPU architecture and performance tuning.
Previous experience in a high-growth startup environment.
BENEFITS:
Competitive compensation package (Unlimited PTO, 401k, covered healthcare premiums).
An opportunity to lead a talented engineering team at a rapidly growing startup in the machine learning space.
Inclusive and supportive work culture with ample opportunities for professional development.
Exposure to a wide range of ML use cases, offering unmatched learning and networking potential.
0 applies
14 views
Other Jobs from Baseten
Developer Success Engineer
AI Support Engineer
Forward Deployed ML Engineer
Tech Lead Manager - Infrastructure Engineering
Site Reliability Engineer
There are more than 50,000 engineering jobs:
Subscribe to membership and unlock all jobs
Engineering Jobs
60,000+ jobs from 4,500+ well-funded companies
Updated Daily
New jobs are added every day as companies post them
Refined Search
Use filters like skill, location, etc to narrow results
Become a member
🥳🥳🥳 452 happy customers and counting...
Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.
To try it out
For active job seekers
For those who are passive looking
Cancel anytime
Frequently Asked Questions
- We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
- We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
- We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
- We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
- Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
- Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
- Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅
What Fellow Engineers Say