Staff AI Infrastructure Engineer
You'll own the reliability of Luma's 10k+ GPU fleet: the scheduling, efficiency, and resilience that research and products depend on.
2,004 live vacancies in San Mateo, California.
You'll own the reliability of Luma's 10k+ GPU fleet: the scheduling, efficiency, and resilience that research and products depend on.
You'll ship real products to real users at Luma, working close to experienced technical leaders and using AI coding agents to build at a pace that wasn't possible a few years ago.
You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs.
You'll help define the simulation substrate Luma uses to train general-purpose robot policies — a faithful, controllable simulation of the world built on our generative video and 3D models.
Team: Infra Reliability · SF Bay Area / Remote (US) You'll own the GPU infrastructure Luma's research and product run on — thousands of NVIDIA and AMD GPUs across on-prem and multi-cloud (AWS and OCI).
You'll lead Luma's regional sales — owning strategic accounts yourself while hiring and coaching a team of Account Executives — and place our creative AI technology with the world's leading brands and agencies.
As Luma's founding Robotics Engineer, you'll bring up commercial robot platforms — humanoids, arms, mobile bases — wire up the sensors and data pipelines our world models need, and run the experiments that tell us whether the model…
You'll turn Luma's industry-leading generative video models into world models: interactive, controllable, physically faithful, and useful as a substrate for embodied reasoning.
You'll build the distributed systems that train Luma's large-scale multimodal models across thousands of GPUs, so researchers can focus on innovation on top of reliable, efficient, scalable infrastructure.
You'll build the systems that make reinforcement learning work at frontier scale — coupling policy optimization with large fleets of inference workers, agentic environments, and the reward and verification systems that turn model behavior…
You'll make Luma's multimodal models fast — profiling and optimizing GPU, CPU, and accelerator code so they train efficiently and deploy at scale without sacrificing quality.
You'll build and train large-scale multimodal agentic models — systems that reason, plan, code, and call tools to do complex, multi-step work over pixels. This is core research shaping how users interact with what Luma's models can do.
You'll own growth end to end at Luma — acquisition, activation, retention, and revenue — scaling the product from strong early traction toward mass adoption, and defining how growth works here from first principles.
You'll build Luma's enterprise products from the model up, turning frontier multimodal capabilities into solutions marketing, advertising, and entertainment customers will pay for and scale.
You'll own the roadmap for Luma's core creative product, turning frontier AI research into pro-level photo and video tools that creators love, driving adoption and setting PM best practices from day one.
You'll own the bridge between Luma's frontier research and its products — Canvas, Agents, and the model platform — making sure what researchers build is shaped by what customers need, and what ships takes full advantage of what research…
Forward Deployed Engineers turn what Luma's models can do into systems customers actually rely on.
You'll be Luma's creative lead on the ground with major brands and agencies, turning what our models can do into production workflows they actually use.
Why Sony Interactive Entertainment? Sony Interactive Entertainment isn’t just the Best Place to Play — it’s also the Best Place to Work.
Why Sony Interactive Entertainment? Sony Interactive Entertainment isn’t just the Best Place to Play — it’s also the Best Place to Work.