
Real job — pulled straight from Jobgether’s careers page · Verified July 27, 2026 · No reposts.
Job description
Jobgether is hiring a AI Systems Performance Specialist — a full-time, remote role ($100k–$150k). Apply directly on Jobgether's careers page below.
AI Systems Performance Specialist
Team: IT
Location: US
Commitment: Full-time
Workplace Type: remote
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Systems Performance Specialist based in United States.
The AI Systems Performance Specialist will optimize large-scale artificial intelligence systems by improving performance, efficiency, and scalability across training and inference workloads.
This role focuses on maximizing throughput, reducing latency, and lowering infrastructure costs through advanced optimization techniques.
You will work across the AI technology stack, from GPU-level optimization and distributed computing to model efficiency and production deployment.
The ideal candidate combines deep machine learning systems expertise with strong engineering discipline and a passion for measurable performance improvements.
This position offers the opportunity to solve complex challenges in AI infrastructure while collaborating with engineering teams building next-generation intelligent systems.
You will contribute to performance standards, optimization strategies, and technical innovations that directly impact production AI capabilities.
Accountabilities:
- Profile and optimize end-to-end AI training and inference pipelines to improve throughput, latency, and cost efficiency.
- Identify performance bottlenecks across data pipelines, model execution, memory usage, communication layers, and infrastructure components.
- Implement optimization strategies including quantization, sparsity, pruning, and other model efficiency techniques.
- Optimize distributed training systems using approaches such as tensor parallelism, pipeline parallelism, FSDP, and ZeRO-style sharding.
- Improve large language model serving performance through techniques such as KV cache optimization, continuous batching, and speculative decoding.
- Develop and apply compiler-level optimizations using technologies such as Triton, XLA, TorchInductor, or TVM.
- Optimize data loading, storage access patterns, and dataset sharding strategies for high-performance AI workloads.
- Build and maintain benchmarking frameworks, regression testing systems, and performance measurement tools.
- Collaborate with machine learning and platform engineering teams to integrate optimization best practices into production workflows.
- Drive cost optimization initiatives through improvements in model architecture, hardware utilization, and workload scheduling.
- Evaluate emerging AI hardware and software technologies and recommend adoption strategies.
- Create technical documentation, optimization playbooks, and knowledge-sharing materials for engineering teams.
- Stay current with AI systems research and translate new developments into practical production improvements.
- Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field.
- 6+ years of experience in performance engineering, machine learning systems, distributed computing, or high-performance computing environments.
- Strong programming skills in Python and C++.
- Hands-on experience optimizing deep learning workloads on modern GPU architectures.
- Deep understanding of distributed training and inference architectures.
- Experience using profiling and performance analysis tools across CPU, GPU, and distributed systems.
- Strong knowledge of memory hierarchies, communication primitives, and parallel computing strategies.
- Familiarity with model compression techniques and understanding their impact on accuracy and performance.
- Excellent measurement, debugging, and analytical reasoning abilities.
- Strong communication and collaboration skills with the ability to work effectively across engineering teams.
- Experience optimizing large language model inference systems at production scale.
- Contributions to AI infrastructure projects such as vLLM, TensorRT-LLM, DeepSpeed, or similar technologies.
- Experience developing custom GPU kernels using Triton, CUTLASS, or related frameworks.
- Familiarity with FinOps practices for managing AI infrastructure costs.
- Technical publications, conference presentations, or community contributions related to AI systems performance.
- Fully remote work opportunity within the Continental United States.
- Competitive annual salary range of approximately $100,000 - $150,000, depending on experience and qualifications.
- Full-time direct employment opportunity.
- Opportunity to work on advanced AI systems and large-scale machine learning infrastructure.
- Career growth opportunities within an innovative technology environment.
- Exposure to cutting-edge AI optimization techniques and emerging technologies.
- Collaborative culture focused on engineering excellence, learning, and continuous improvement.
- Opportunity to contribute to impactful cloud, AI, and enterprise technology solutions.
- Inclusive workplace committed to equal opportunity and professional development.
The AI Systems Performance Specialist will lead efforts to improve the efficiency and reliability of advanced AI workloads through profiling, optimization, and engineering best practices. This role requires strong technical ownership, analytical thinking, and the ability to collaborate across machine learning and infrastructure teams.
Requirements:
The successful candidate will bring extensive experience in AI systems, performance engineering, or high-performance computing, with a strong ability to analyze and optimize complex machine learning workloads. The ideal profile combines software engineering expertise, deep understanding of modern AI infrastructure, and strong problem-solving skills.
Preferred qualifications include:
Benefits:
Get AI Systems Performance Specialist jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What is the salary for AI Systems Performance Specialist at Jobgether?
The estimated salary range for AI Systems Performance Specialist at Jobgether is $100,000 - $150,000 USD per year.
Is AI Systems Performance Specialist at Jobgether a remote job?
Yes, AI Systems Performance Specialist at Jobgether is a remote position. This role is open to remote candidates.
What skills are required for AI Systems Performance Specialist at Jobgether?
The required skills for AI Systems Performance Specialist at Jobgether include: Python, C++, Machine Learning, Deep Learning, AI, LLM.
What is the seniority level for AI Systems Performance Specialist at Jobgether?
AI Systems Performance Specialist at Jobgether is a Senior level position.
How do I apply for AI Systems Performance Specialist at Jobgether?
You can view the full description and apply for AI Systems Performance Specialist at Jobgether on EchoJobs: https://echojobs.io/job/jobgether-ai-systems-performance-specialist-bzljs.