Hiring.Camp

LLM Inference Engineer (SR)

Job Board

·

Mar 30, 2026

Workplace
Remote
Type
Full-time
Department
Marketing
Experience
5+ years
Closing date
Today
Source
Vincere

Description

Career Opportunity for a LLM Inference Engineer in Japan!

 

■ LLM Inference Engineer

 

■ Company Overview

Japan-based AI technology company developing advanced large-scale language models and next-generation AI platforms, focused on delivering innovative AI-driven products and competing with leading global AI players.

 

■ Your Role and Responsibilities 

● Build systems that translate the value of large-scale language models into real user-facing applications

● Design and manage development processes to improve engineering productivity

● Develop and operate large-scale distributed systems requiring high scalability and high availability

● Build and maintain infrastructure for machine learning model inference and online serving

● Optimize system performance for large-scale AI workloads

● Collaborate with engineering teams to design reliable and scalable AI platforms

● Contribute to continuous improvement of system architecture, monitoring, and operational processes

 

■ Experience and Qualifications

● 5+ years of experience in software engineering, machine learning infrastructure, or related fields

● Experience developing or operating large-scale distributed systems with high scalability and availability

● Strong system design and problem-solving skills

● Experience building or operating production systems

● Strong commitment to building high-quality and scalable systems

 

■ Additional Preferred Qualifications

● Experience designing systems running on on-premises or cloud GPU clusters

● Experience building high-availability systems across multiple data centers or regions

● Experience with distributed databases or large-scale search engines

● Experience designing online serving infrastructure for machine learning models

● Knowledge of inference optimization and performance acceleration for ML models

● Experience using LLM inference frameworks such as vLLM, SGLang, or TensorRT-LLM

● Experience designing monitoring and observability systems for distributed infrastructure

● Experience contributing to open-source software, publishing technical papers, or participating in technical communities

 

■ Good Reasons to Join

● Opportunity to work on advanced AI technologies and large-scale machine learning systems

 

■ Work Location

Tokyo, Japan
 

Details will be provided during the meeting.

Skills

Machine Learning

Similar Jobs

16

Principal LLM Inference Engineer

D Matrix · Santa Clara · Hybrid

1 month ago

AI Computing Software Development Engineer, LLM Inference

Nvidia · China, Shanghai +1

1 month ago

AI Computing Software Development Engineer, LLM Inference

Nvidia · Shanghai, Shanghai,CN, CN +1

1 month ago

Senior Deep Learning Research Engineer, LLM Inference

Nvidia · Israel, Tel Aviv

2 months ago

Senior Deep Learning Research Engineer, LLM Inference

Nvidia · Tel Aviv-Yafo, Tel Aviv District,IL, IL

2 months ago

Distributed LLM Inference Engineer

Anyscale · San Francisco +1 · Hybrid

2 months ago

Distributed Training & Inference Optimization Engineer (LLM) - GPU Optimization Department (GPUOD)

Rakuten · Rakuten Crimson House, Japan · Remote, Hybrid

9 months ago

Sr. Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

1 week ago

Senior Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

4 weeks ago

Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

2 months ago

Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

2 months ago

Machine Learning Engineer, Inference & Serving (Speech LLM) - San Francisco

Plaud · San Francisco, CA · Hybrid

2 months ago

Senior Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

4 months ago

Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

5 months ago

Lead AI Engineer (FM Hosting, LLM Inference)

Capitalone · New York, NY, United States of America +3

6 months ago

LLM Inference Frameworks and Optimization Engineer

Together · San Francisco, Singapore, Amsterdam +1 · Remote

1+ year ago