
Real job — pulled straight from Modular’s careers page · Verified August 21, 2026 · No reposts.
Job description
Modular is hiring a Developer Advocate, MAX Inference & Serving (Remote) — a full-time, remote role ($150k–$225k). Apply directly on Modular's careers page below.
Developer Advocate, MAX Inference & Serving
Location: United States / Canada
Department: Developer Relations
Location Type: REMOTE
Employment Type: FULL_TIME
About the role:
What you will do:
- Provide support and respond to questions from customers evaluating MAX for inference and serving, including some of the world's largest corporations.
- Establish, run and publish benchmarks comparing MAX to serving frameworks like vLLM, Triton Inference Server, and TensorRT-LLM.
- Foster an inclusive and welcoming environment for ML engineers and practitioners deploying models with MAX.
- Collaborate with engineering and product teams to create tutorials, video guides, and examples demonstrating how the MAX Platform handles inference and serving workloads efficiently on both CPUs and GPUs.
- Write blog posts and other educational content that inform prospective and current customers about MAX's inference and serving performance and functionality.
- Engage with the community across GitHub, Discord, Twitter/X, and LinkedIn by facilitating discussions, answering questions, and providing support.
- Act as a voice for the developer community internally, feeding inference and serving feedback back to engineering and product teams, shaping the future of the MAX Platform.
- Represent Modular at conferences, summits, and industry events through presentations, panel discussions, and networking with industry professionals.
- Contribute to Modular's broader developer relations, product, and marketing strategies.
What success looks like after 6 months
- You've published a steady cadence of inference and serving content that developers reference when deploying models with MAX.
- Your benchmarks and comparisons against tools like vLLM, Triton Inference Server, and TensorRT-LLM get cited in community discussions.
- Code examples and cookbooks you've built get linked in forum and Discord answers by other community members.
- You've identified and closed content gaps that were blocking adoption of MAX for production inference.
What you bring to the table:
- Demonstrable experience creating technical content for developer audiences. Send us a portfolio: blog posts, tutorials, videos, docs, or courses.
- You understand the ML inference stack, including model serving architectures, GPU acceleration, and how MAX compares to vLLM, Triton Inference Server, and TensorRT-LLM.
- Strong Python skills; systems programming experience (C++, Rust, or similar) is an advantage.
- You learn new tools fast and produce accurate content quickly. A feature ships Tuesday, your tutorial goes out Thursday.
- You can record, edit, and publish a technical video without a production team. Clear audio, good pacing, technically correct, not necessarily polished.
- You write well. You explain complex ideas without losing precision, and you cut the filler.
- You plan your own content calendar because you're in the community and know where developers get stuck.
- A growth and leadership mindset, with a collaborative attitude that seeks to learn more from our customers, team members, and the broader market.
Minimum Qualifications:
- Bachelor's degree in Engineering, Information Systems, Computer Science, or technical related field.
- 2+ years of Product Management or related work experience.
Helpful, but not required
- Experience programming GPUs using CUDA or ROCm.
- Familiarity with the Mojo 🔥 programming language and MAX AI framework.
- Familiarity with open source software development practices and communities.
- Experience producing high-quality videos covering technical topics.
What Modular brings to the table:
- Amazing Team. We are a progressive and agile team with some of the industry’s best engineering and product leaders.
- World-class Benefits. In order to attract the best, we need to offer the best. Premier insurance plans, up to 5% 401k matching, flexible paid time off, and more are available to you! Please note that specific benefit packages may vary based on your location.
- Competitive Compensation. We offer very strong compensation packages, including stock options. We want people to be focused on their best work and believe in tailoring compensation plans to meet the needs of our workforce.
- Team Building Events. We organize regular team onsites and local meetups in Los Altos, CA as well as different cities. Traveling 2-4 times a year is expected for all roles.
Get Developer Advocate, MAX Inference & Serving (Remote) jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What is the salary for Developer Advocate, MAX Inference & Serving (Remote) at Modular?
The estimated salary range for Developer Advocate, MAX Inference & Serving (Remote) at Modular is $150,000 - $225,000 USD per year.
Is Developer Advocate, MAX Inference & Serving (Remote) at Modular a remote job?
Yes, Developer Advocate, MAX Inference & Serving (Remote) at Modular is a remote position. Candidates in Los Altos, CA may be preferred.
What skills are required for Developer Advocate, MAX Inference & Serving (Remote) at Modular?
The required skills for Developer Advocate, MAX Inference & Serving (Remote) at Modular include: Python, C++, Rust, Machine Learning, GitHub.
What is the seniority level for Developer Advocate, MAX Inference & Serving (Remote) at Modular?
Developer Advocate, MAX Inference & Serving (Remote) at Modular is a Mid Level / Senior level position.
How do I apply for Developer Advocate, MAX Inference & Serving (Remote) at Modular?
You can view the full description and apply for Developer Advocate, MAX Inference & Serving (Remote) at Modular on EchoJobs: https://echojobs.io/job/modular-developer-advocate-max-inference-serving-76pnq.