Near AI logo

LLM Inference Engineer (Remote)

Near AI

Hybrid
Full-time
Senior
Salary not listedPosted 8h ago

Real job — pulled straight from Near AI’s careers page · Verified September 9, 2026 · No reposts.

Job description

Near AI is hiring a LLM Inference Engineer (Remote) — a full-time, remote role. Apply directly on Near AI's careers page below.

LLM Inference Engineer

Location: San Francisco or Remote

Department: Near AI

Locations: San Francisco or Remote

About The Role

The NEAR AI team is building decentralized and confidential machine learning infrastructure to enable user-owned AI. Our mission is to build highly scalable and efficient infrastructure for open-source AI at a global scale.

We are specifically seeking an expert in high-performance LLM serving systems and inference optimization. In this role, you will push the boundaries of how large language models are served.

What You'll Be Doing

  • Architect and maintain production high-traffic LLM serving systems.
  • Optimize throughput, latency, and cost for leading open-source LLMs.

What We're Looking For

  • Strong hands-on experience in LLM inference, with expertise debugging and optimizing major inference engines such as SGLang, vLLM, or TensorRT.
  • Deep knowledge of state-of-the-art GPU architectures, and effectively exploit them using PyTorch, Triton, CuTe, CUDA, etc.
  • Proven track record in designing and maintaining end-to-end high-traffic LLM serving systems.
  • Strong problem-solving skills and ability to communicate technical ideas clearly.

We'd Love If You Have

  • Experience with Trusted Execution Environments (TEE).
  • Active contributor to open-source LLM inference engines.

Please let us know if you require any special requirements for your interview and we'll do our best to accommodate.

Get LLM Inference Engineer (Remote) jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Zeta Global logo

Senior Data Scientist

Berlin, Germany
✓ From careers page· 20m ago
SIXT logo

SIXT

New

Applied AI Engineer

Munich, BY
✓ From careers page· 23m ago
Stability AI logo

Forward Deployed Engineer (Remote)

Remote · US-eligible
✓ From careers page· 2h ago
Slack logo

Software Engineer, Machine Learning

$130k–$179kToronto, ON
✓ From careers page· 2h ago

Frequently asked questions

Is LLM Inference Engineer (Remote) at Near AI a remote job?

Yes, LLM Inference Engineer (Remote) at Near AI is a remote position. Candidates in San Francisco, CA may be preferred.

What skills are required for LLM Inference Engineer (Remote) at Near AI?

The required skills for LLM Inference Engineer (Remote) at Near AI include: PyTorch.

What is the seniority level for LLM Inference Engineer (Remote) at Near AI?

LLM Inference Engineer (Remote) at Near AI is a Senior level position.

How do I apply for LLM Inference Engineer (Remote) at Near AI?

You can view the full description and apply for LLM Inference Engineer (Remote) at Near AI on EchoJobs: https://echojobs.io/job/near-ai-llm-inference-engineer-7wodd.