Base Career helps you apply smarter for this job.
Key skills for this role
We're looking for a Runtime Engineer to design and build the multi-target runtime that sits at the heart of our AI compiler stack. This is a systems-level role where you'll take the output of our optimizing compiler and make it execute — efficiently, correctly, and at scale — across a diverse landscape of hardware targets.
You'll work on low-level parallelization, kernel scheduling, and performance analysis, and collaborate closely with our compiler and product teams to push the boundaries of what's possible on modern AI hardware.
At Lemurian Labs, we're reimagining the foundations of computing to make AI accessible to everyone. Our mission is to remove the limits of scale, hardware, and cost that hold back innovation, so the people solving humanity's hardest problems can move faster.
We're building a new kind of software stack: a hardware-agnostic platform that makes every system — from a laptop to a supercomputer — feel like one seamless engine. Developers can write once, run anywhere, and get state-of-the-art performance across any chip, any cloud, at any scale. It's a complete rethink of how software and hardware interact — designed for the era beyond Moore's Law.
We're not looking for the comfortable or the conventional; we're looking for the bold. The engineers who crave frontier problems, who want to bend the limits of what's possible, who see infrastructure not as a constraint but as a canvas. If you want to build the foundation for the next era of AI and change what humanity can achieve in the process, join us.
We're looking for a Runtime Engineer to design and build the multi-target runtime that sits at the heart of our AI compiler stack. This is a systems-level role where you'll take the output of our optimizing compiler and make it execute — efficiently, correctly, and at scale — across a diverse landscape of hardware targets.
You'll work on low-level parallelization, kernel scheduling, and performance analysis, and collaborate closely with our compiler and product teams to push the boundaries of what's possible on modern AI hardware.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
Santa Clara, USA
Santa Clara, USA
Santa Clara, USA
, USA
, USA
Design, develop, maintain, and improve our multi-target runtime.
Apply the latest techniques in parallelization and partitioning to automate kernel generation and exploit highly optimized execution paths.
Rapidly prototype and data-drive exploration of new runtime ideas.
Benchmark and analyze the outputs produced by our optimizing compiler on target hardware.
Build tools to collect and analyze performance bottlenecks.
Work closely with our product team to understand the evolving needs of ML engineers and drive improvements in runtime architecture.
Build the runtime that makes next-generation AI infrastructure actually go fast.
Work across the full stack — from hardware intrinsics to compiler output to distributed execution.
Join a team that approaches infrastructure as a canvas, not a constraint.
Competitive compensation including equity, medical/dental/vision, retirement savings, and wellness benefits.
Lemurian Labs is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees, regardless of gender identity, race, ethnicity, sexual orientation, disability status, age, or background.
Compensation depends on experience and geographic location and will be narrowed during the interview process. Additional benefits include equity, company bonus opportunities, medical, dental, and vision coverage, a retirement savings plan, and supplemental wellness benefits.
Full-time
Mid · 4+ years experience
Remote
Apply faster on company sites with our extension.