Jobs › US jobs › Software Engineer, Inference

Software Engineer, Inference

Lumaai · Redwood City, CA · All 22 Lumaai jobs

Work pattern
Hybrid
Employment
Permanent
Posted
24 Jul 2026

Email me new Lumaai jobs

Every new job Lumaai posts in the US. One email a week, and only in a week that has new ones.

We email you once to confirm; nothing else is sent until you click the link. We keep your email address and this alert, nothing more, and delete both when you unsubscribe, which takes one click from any email. Privacy policy

Required skills, as the advert states them

  • Kubernetes
  • Python
  • Pytorch
  • LLMs
  • Scheduling
  • Fleet Management
  • Linux
  • Docker

Nice to have

  • CI/CD
  • Redis
  • Networking

About this opportunity

You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs. This is large-scale inference systems work: scheduling, fleet management, deployment pipelines, and reliability across clusters and hardware providers. It fits a strong systems engineer comfortable with model serving and Kubernetes at scale.

Read the full advert and apply on Lumaai's site →

Collected from Lumaai's own careers site (Ashby). Posted 24 Jul 2026. TUNAI shows an excerpt and the facts it read from the advert; the employer's page has the full description and the application.