
Real job — pulled straight from Jobgether’s careers page · Verified August 1, 2026 · No reposts.
Job description
Jobgether is hiring a Machine Learning Operations Engineer — a full-time, remote role ($151k–$173k). Apply directly on Jobgether's careers page below.
Senior Machine Learning Ops Engineer
Team: IT
Location: US
Commitment: Full-time
Workplace Type: remote
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Machine Learning Ops Engineer based in the United States.
This role offers the opportunity to shape and scale enterprise machine learning infrastructure within a growing data platform organization.
You will work at the intersection of machine learning, software engineering, and cloud operations to bring advanced models into reliable production environments.
The position focuses on building scalable platforms, improving deployment workflows, and enabling data teams to deliver impactful AI solutions faster.
You will collaborate closely with Data Science, Data Engineering, and Applied AI teams to establish best practices across the ML lifecycle.
This is an ideal opportunity for an experienced engineer who enjoys solving complex infrastructure challenges and driving technical innovation.
You will play a key role in improving reliability, automation, observability, and governance for critical machine learning systems.
Accountabilities:
- Design, deploy, and maintain scalable ML infrastructure supporting model training, batch processing, and real-time inference workloads.
- Build and manage cloud-based infrastructure and services using AWS, Snowflake, and related platforms through Infrastructure-as-Code practices.
- Develop and maintain containerized ML deployment solutions using Docker, FastAPI, and modern software delivery patterns.
- Create and improve CI/CD pipelines, automation frameworks, testing processes, and deployment standards for machine learning systems.
- Partner with Data Science and Data Engineering teams to productionize models and accelerate machine learning delivery.
- Establish monitoring and observability frameworks, including model performance tracking, drift detection, data quality monitoring, and automated alerting.
- Improve platform reliability, scalability, security, governance, and operational efficiency across ML workflows.
- Support architecture decisions, engineering standards, and best practices for enterprise ML platforms.
- Document technical architecture, deployment processes, and operational procedures to ensure maintainability and knowledge sharing.
- Contribute to the development of reusable ML infrastructure components and data products.
- Support both batch and low-latency inference workflows while optimizing system performance.
- Help define the future direction of machine learning operations and platform capabilities.
- Bachelor’s degree in Computer Science, Data Engineering, or a related technical field; advanced degree preferred.
- 6+ years of experience in MLOps, platform engineering, DevOps, data engineering, or related infrastructure roles.
- 3+ years of hands-on experience working with AWS cloud infrastructure.
- Strong Python engineering skills, including API development, automation, and backend service development.
- Experience building and operating production machine learning systems.
- Strong knowledge of Docker, containerized application deployment, and modern deployment practices.
- Experience with Kubernetes, ECS, EKS, or similar container orchestration platforms.
- Experience managing Infrastructure-as-Code projects using tools such as Terraform, OpenTofu, or CloudFormation.
- Strong SQL skills and experience with modern data warehouse platforms such as Snowflake, Databricks, or BigQuery.
- Experience implementing CI/CD workflows and software engineering best practices.
- Experience with workflow orchestration tools such as Airflow, Dagster, or Prefect.
- Experience with testing frameworks such as pytest, including unit, integration, and end-to-end testing approaches.
- Strong understanding of Bash and Unix-based environments.
- Experience with backend frameworks such as FastAPI, Flask, or Django.
- Knowledge of ML observability and experiment tracking tools such as MLflow, Arize, Evidently, WhyLabs, or Monte Carlo is a plus.
- Experience designing feature stores or reusable ML data products is preferred.
- Experience supporting Generative AI, LLM deployment workflows, financial services, fintech, or regulated industries is a plus.
- Strong communication skills with the ability to collaborate effectively across Data Science, Data Engineering, Product, and technical teams.
- Ability to manage multiple priorities, work independently, and thrive in a fast-paced environment.
- Competitive salary range of approximately $150,500 to $173,000 annually, depending on location, skills, experience, and qualifications.
- Medical, dental, and vision insurance coverage.
- 401(k) retirement plan with company match.
- Paid holidays, vacation time, sick days, and volunteer time off.
- 12 weeks of paid parental leave.
- Pre-tax transit benefits.
- Company-paid life insurance.
- Voluntary benefits options.
- Discounted pet health insurance.
- Wellness incentive programs.
- Employee mentorship and leadership development opportunities.
- Team-oriented culture with opportunities for professional growth.
- Remote work flexibility.
The Senior Machine Learning Ops Engineer will be responsible for designing, implementing, and maintaining scalable machine learning infrastructure that supports production AI initiatives. This role requires strong engineering expertise, operational ownership, and the ability to collaborate across technical teams to create reliable and efficient ML platforms.
Requirements:
The ideal candidate is a technically strong engineer with experience building production machine learning systems, cloud infrastructure, and scalable software platforms. They should be comfortable operating independently while collaborating with cross-functional teams to solve complex technical challenges.
Benefits:
Get Machine Learning Operations Engineer jobs like this→
New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.
Email me new jobsSimilar jobs




Frequently asked questions
What is the salary for Machine Learning Operations Engineer at Jobgether?
The estimated salary range for Machine Learning Operations Engineer at Jobgether is $151,000 - $173,000 USD per year.
Is Machine Learning Operations Engineer at Jobgether a remote job?
Yes, Machine Learning Operations Engineer at Jobgether is a remote position. This role is open to remote candidates.
What skills are required for Machine Learning Operations Engineer at Jobgether?
The required skills for Machine Learning Operations Engineer at Jobgether include: Python, API, AWS, Snowflake, Docker, FastAPI, Kubernetes, Terraform, SQL, CI/CD, Airflow, Bash, Flask, MLflow, Generative AI, LLM.
What is the seniority level for Machine Learning Operations Engineer at Jobgether?
Machine Learning Operations Engineer at Jobgether is a Senior / Staff level position.
How do I apply for Machine Learning Operations Engineer at Jobgether?
You can view the full description and apply for Machine Learning Operations Engineer at Jobgether on EchoJobs: https://echojobs.io/job/jobgether-senior-machine-learning-ops-engineer-aczst.