InstaDeep

Research Engineer

Cape Town, South Africa Remote Hybrid
C++ Python Docker TensorFlow PyTorch GCP Machine Learning Deep Learning
Search for More Jobs Talk to a recruiter now 💪
Description
InstaDeep, founded in 2014, is a pioneering AI company at the forefront of innovation. With strategic offices in major cities worldwide, including London, Paris, Berlin, Tunis, Kigali, Cape Town, Boston, and San Francisco, InstaDeep collaborates with giants like Google DeepMind and prestigious educational institutions like MIT, Stanford, Oxford, UCL, and Imperial College London. We are a Google Cloud Partner and a select NVIDIA Elite Service Delivery Partner. We have been listed among notable players in AI, fast-growing companies, and Europe's 1000 fastest-growing companies in 2022 by Statista and the Financial Times. Our recent acquisition by BioNTech has further solidified our commitment to leading the industry.

Join us to be a part of the AI revolution!

The Role:
We seek a research engineer who can collaborate closely with our research and machine learning engineering teams. You will work at the intersection of cutting-edge research and engineering with an emphasis on scaling machine learning models to leverage large compute clusters and push the boundaries of what is possible. If you enjoy writing highly performant code, find joy in utilising all the FLOPS that the underlying hardware provides, and love diving deep into algorithm optimisation, this is the position for you.

TLDR:
Optimise models and pipelines to run as fast as possible in a distributed setup, taking advantage of the underlying hardware and applying them to some of the most exciting research problems in the industry.

Responsibilities

  • Algorithmic Optimisation: Research and understand the latest deep learning literature to implement and optimise state-of-the-art algorithms and architectures, ensuring compute efficiency and performance.
  • Scaling Expertise: Design and implement strategies to efficiently scale machine learning models across different accelerator platforms (GPU/TPU).
  • Performance Optimisation: Analyse and profile ML systems under heavy load, pinpointing bottlenecks and implementing targeted optimisations.
  • Distributed Systems Architecture: Create robust distributed training and inference solutions for maximum computational efficiency.
  • Low-Level Mastery: Understand how to take advantage of the underlying hardware. Although not a prerequisite, be comfortable working with technologies like C/C++, XLA, Pallas, Triton, and/or CUDA code to achieve performance breakthroughs.

Required Skills

  • A Masters or equivalent in Computer Science, Engineering, Mathematics, or related field.
  • 2+ years of work experience.
  • Expertise in Python.
  • Experience working with Linux systems.
  • Experience with Docker and container orchestration.
  • Experience with at least one modern machine learning framework (JAX, Tensorflow, PyTorch, etc.)
  • Experience with software profiling, identifying bottlenecks, and delivering efficient solutions.

Highly Desirable

  • Experience working with JAX and packages within the JAX ecosystem.
  • Track record of successfully building and scaling ML models.
  • Experience with distributed training frameworks (Ray, Dask, PyTorch Lightning, etc.)
  • Experience working with HPC clusters and distributing programs over multiple hosts.
  • Understanding of GPU/TPU architectures and their implications for efficient ML systems.
  • Experience using statically typed languages like C, C++, etc.
  • Fundamentals of modern Deep Learning and experience with Reinforcement Learning.
  • Actively following ML trends.

What we offer:

  • Real-World Impact: Directly contribute to the performance and reach of our AI solutions.
  • Cutting-Edge Challenges: Tackle complex problems at the forefront of machine learning and large-scale system design.
  • Growth-Oriented Environment: Expand your expertise with a team of talented engineers dedicated to advancing ML scalability.
Our commitment to our people
We empower individuals to celebrate their uniqueness here at InstaDeep. Our team comes from all walks of life, and we’re proud to continue encouraging and supporting applicants from underrepresented groups across the globe. Our commitment to creating an authentic environment comes from our ability to learn and grow from our diversity, and how better to experience this than by joining our team? We operate on a hybrid work model with guidance to work at the office at least 2 to 3 days per week to encourage close collaboration and innovation. We are continuing to review the situation with the well-being of InstaDeepers at the forefront of our minds.

Right to work: Please note that you will require the legal right to work in the location you are applying for.
InstaDeep
InstaDeep
Artificial Intelligence (AI) Information Technology

0 applies

3 views

Other Jobs from InstaDeep

Senior DevOps Engineer

London, UK Paris, France

Machine Learning Engineer

Cape Town, South Africa Remote Hybrid

Software Engineer intern

Tunis, Tunisia Remote Hybrid

Dev Ops / ML Ops Intern

Tunis, Tunisia Remote Hybrid

AI Research Intern

Tunis, Tunisia Remote Hybrid

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 401 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say