FriendliAI

Software Engineer, Platform Security

San Francisco, CA
Kubernetes Docker Terraform Helm AWS GCP Oracle Cloud Python Bash Shell API Machine Learning Deep Learning Prometheus Grafana ELK OTEL
Description

Software Engineer – Platform Security

Department: Engineering

Location: San Francisco

Employment Type: FullTime

About the job

FriendliAI is seeking a Forward Deployed Engineer (FDE) to assist enterprises in deploying, scaling, and operating generative and agentic AI workloads on FriendliAI infrastructure. You will work directly with customers to solve and implement production-grade applications using our products, such as Serverless Endpoints, Dedicated Endpoints, or Container.

Friendli Container is our service that allows customers to download our inference engine as Docker images and deploy it in their chosen environment, such as private clouds or on-premises. Our Friendli Container can be adopted directly to AWS EKS clusters using our EKS add-on product.

You will work directly on our customers’ projects, collaborating with their engineering teams to solve AI inference challenges like scaling, orchestration, and monitoring. This is a hands-on, customer-embedded role. If you have worked in DevOps, platform engineering, or SRE for AI applications, this is your ideal position.

Key Responsibilities

  • Design and implement large-scale deployment architectures for LLM and multimodal inference

  • Deploy and manage containerized workloads across Kubernetes clusters

  • Diagnose production issues, such as performance bottlenecks, and implement temporary fixes as needed

  • Collaborate with customers’ DevOps teams to integrate FriendliAI’s infrastructure into their CI/CD workflows

  • Develop scripts, Helm charts, and Terraform modules that simplify repeated deployments

  • Contribute field insights to shape our platform reliability, observability, and scaling strategies

  • Lead workshops, technical sessions, or webinars to help customers master infrastructure best practices

Qualifications

  • 3+ years of experience in cloud infrastructure, DevOps, or reliability engineering

  • Bachelor’s or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent

  • Proficiency with Kubernetes, Docker, Terraform, and Helm

  • Strong foundation in distributed systems, networking, and performance tuning

  • Experience with GPU-based computing and generative AI model serving workloads

  • Strong technical background in backend systems or AI tooling

  • Experience operating workloads on AWS, GCP, or OCI

  • Excellent problem-solving and debugging skills in real-world environments

Preferred Experience

  • Experience deploying large models (LLMs, diffusion models) on GPUs or clusters

  • Familiarity with inference frameworks (Triton, vLLM, TensorRT, DeepSpeed-Inference)

  • Familiarity with observability stacks (Prometheus, Grafana, Loki, ELK, OTEL)

  • Understanding of networking security and compliance frameworks (e.g., SOC 2)

  • Experience supporting on-prem or hybrid-cloud deployments

Benefits

  • A front-row seat to the generative AI infrastructure revolution

  • Competitive compensation and benefits package

  • Daily lunch and dinner provided; unlimited snacks and beverages

  • Health check-up and top-tier hardware support

  • Flexible working hours and a highly collaborative environment

About us

FriendliAI is building the next-generation AI inference platform that accelerates the deployment of large language and multimodal models with unmatched performance and efficiency. Our infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 510,000 open-source models. We are on a mission to deliver the world’s best platform for AI inference.

FriendliAI
FriendliAI

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say