Base Career helps you apply smarter for this job.
Key skills for this role
Build the model runtime within an inference engine for complex frontier models running at scale on custom silicon.
Develop a production-grade runtime that translates demanding inference workloads into efficient execution while optimizing throughput, latency, utilization, and reliability.
Work across model architecture, distributed systems, compilers, kernels, and silicon.
Build the model runtime within an inference engine for complex frontier models running at scale on custom silicon.
Develop a production-grade runtime that translates demanding inference workloads into efficient execution while optimizing throughput, latency, utilization, and reliability.
Work across model architecture, distributed systems, compilers, kernels, and silicon.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
San Francisco, USA
San Francisco, USA
San Francisco, USA
, USA
, USA
San Francisco, USA
San Francisco, USA
San Francisco, USA
Candidates may need to meet certain legal status requirements under U.S. export control laws and regulations.
OpenAI is an AI research and deployment company developing AI systems and products with safety and human needs at their core.
$266K – $445K • Offers Equity
Full-time
Remote
Apply faster on company sites with our extension.