Level AI

Research Intern, Reinforcement Learning

Bay Area, CA
Reinforcement Learning AI Machine Learning Python LLM SLM
Description

Research Intern – Reinforcement Learning (RL) - Onsite

Team: Engineering

Location: Bay Area, California

Workplace Type: onsite

🚀 Build the next generation of Agentic AI with us

Our platform combines conversation intelligence, multimodal understanding, and agentic AI systems to power both human agents and autonomous AI agents across the entire customer experience lifecycle.

A core part of this vision is our investment in custom Small Language Models (SLMs)—purpose-built for CX workflows—paired with reinforcement learning systems that continuously improve decision-making in real-world environments.

We’re looking for a Research Intern (Reinforcement Learning) to join us in shaping this future.


What you’ll do

 
  • Design and build reinforcement learning environments that model real-world customer interaction workflows.

  • Design RL agents that learn from these environments using real-world interaction data, rewards, and feedback loops

  • Define reward models and feedback loops using real-world signals (outcomes and human feedback)

  • Enable learning from production data by structuring interaction traces into training-ready datasets for offline and online learning

  • Experiment with multi-agent systems and simulation frameworks for complex coordination and decision-making

  • Collaborate with engineering and product teams to deploy, evaluate, and iterate on learning systems in production at scale.

 


What we’re looking for

  • Currently pursuing (or recently completed) a degree in Computer Science, AI, Machine Learning, or related field

  • Strong understanding of reinforcement learning fundamentals

  • Familiarity with RL environments and training libraries such as Verl and Tinker

  • Strong foundation in probability, math, and optimization

  • Passion for building real-world AI systems


Nice to have

  • Experience with RLHF, LLM/SLM fine-tuning, or model alignment

  • Exposure to agent-based systems or multi-agent RL

  • Prior research, projects, or publications in RL or applied ML

  • Experience working with large-scale or production datasets

 


Why Level AI

  • Work on production-grade Agentic AI systems used by leading enterprises

  • Build alongside a team with deep expertise from Amazon, Google, and Meta

  • Be part of a fast-growing Series C AI company.

  • Direct exposure to 0→1 AI innovation in CX and decisioning systems

Level AI
Level AI

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say