Archived on

Preferred Networks is hiring a
LLM Inference Optimization Engineer
About Preferred Networks
Preferred Networks is an AI company based in Tokyo working across the stack, from AI chips and computing infrastructure to LLMs and products. You may already know us indirectly if you\'ve used software we\'ve built, such as Optuna or CuPy (or Chainer, back in the day). We are designing in-house chips (MN-Core series) and training LLMs (PLaMo series). Our team is actively hiring for two roles related to these endeavors.
Job Description
Improve the inference engine powering our API service and maintain PLaMo implementations in open source projects such as vLLM.Remote
Remote Conditions
Remote within Japan; relocation to Japan required; visa and relocation support provided.
Salary
Not Specified
Benefits
Visa and relocation support
Tech Tags
ApisC##LinuxPerformance OptimizationPythonVllm
Date Listed
02 May, 2026 (3 months ago)
Loading...
Share this job
This job is archived, but you can still apply.
Hiring engineers?
Reach thousands of tech candidates from the Hacker News community.
Post a Job β $99Similar Jobs
No tags
π Berlin, Germany; Freiburg, Germany; New York City, USA
84% match
No tags
π Berlin, Germany; Freiburg, Germany; New York City, USA
83% match
No tags
π Berlin, Germany; Freiburg, Germany; New York City, USA
82% match
No tags
π Berlin, Germany; Freiburg, Germany; New York City, USA
82% match









