{bc}
Live Jobs · Updated Daily

"Software Engineer, Inference Systems" Jobs in United States

Get matched with opportunities that fit your skills, experience, and career goals.

Jobs found — scroll to load more

Clear filters

Software Engineer, Inference Systems

River AI Inc. · Palo Alto

MidOnsite

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu

Discovered Today

Apply

Principal System Software Engineer, AI Inference Execution

D Matrix · Santa Clara

SeniorFull-timeHybrid

At d-Matrix, we are focused on unleashing the potential of generative AI to power the transformation of technology. We are at the forefront of software and hardware innovation, pushing the boundaries of what is possible.

Skills

cpluspluskuberneteslinuxpython

Discovered 1 weeks ago

Apply

Software Engineer -Gen AI Inferencing | Onsite - (Dallas, Charlotte, New York)

Photon Group ·

MidOnsite

"Required qualifications: 5+ years OOP in Python/Scala/Java programming experience with expert level development skills Experience with AI/ML/GenAI Lifecycle Management and Development and its Ecosystem. -Hands on experi

Discovered Today

Apply

Staff+ Software Engineer, ML Inference Path

Anthropic · San Francisco

SeniorOnsite

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co

Discovered Today

Apply

Staff + Sr. Software Engineer, Scaling

Anthropic ·

SeniorOnsite

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co

Discovered Today

Apply

AI Systems Engineer (OCI/AI Infrastructure)

CLBPTS - All Lines of Business · Nashville

MidOnsite

Oracle Hardware Platform Development Engineering is seeking a highly driven AI Systems Engineer to evaluate and characterize next-generation GPU and AI accelerator platforms for Oracle Cloud Infrastructure (OCI). This is

Discovered Today

Apply

Software Engineer -Gen AI Inferencing | Onsite - (Dallas, Charlotte, New York)

Photon Career Site ·

MidOnsite

"Required qualifications: 5+ years OOP in Python/Scala/Java programming experience with expert level development skills Experience with AI/ML/GenAI Lifecycle Management and Development and its Ecosystem. -Hands on experi

Discovered Today

Apply

Software Engineer, Distributed Training

River AI Inc. · Palo Alto

MidOnsite

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu

Discovered Today

Apply

Software Engineer, GPU Kernels

River AI Inc. · Palo Alto

MidOnsite

At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructu

Discovered Today

Apply

AI Systems Engineer (OCI/AI Infrastructure)

Oracle · Nashville

MidOnsite

Oracle Hardware Platform Development Engineering is seeking a highly driven AI Systems Engineer to evaluate and characterize next-generation GPU and AI accelerator platforms for Oracle Cloud Infrastructure (OCI). This is

Discovered Today

Apply

Software Engineer III, Copilot Models Inference

GitHub, Inc. ·

USD 107700-285900 yearly / yearSeniorFull-timeRemote

As a software engineer at GitHub, you will enhance the collaboration experience at GitHub by working closely with a community of engineers and designers with a distributed, diverse and passionate team delivering the serv

Skills

CC++C#JavaScript

Discovered Yesterday

Apply

Software Engineer, Inference

Luma · Redwood City

MidFull-timeHybrid

You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs. Thi

Skills

PythonPyTorchHugging FacevLLM

Discovered Yesterday

Apply

Software Engineer, AI Infrastructure - LVM Inference & Evaluation

Ambient.ai · Redwood City

USD 168000-205000 yearly / yearEntryFull-timeHybrid

Build a safer world with us, one incident at a time. Ambient.ai is the category creator and leader in Agentic Physical Security. Powered by Ambient Pulsar, the first reasoning Vision-Language Model purpose-built for phy

Skills

PythonvLLMTriton Inference ServerCUDA

Discovered 2 days ago

Apply

Staff Software Engineer, Inference Cloud

Cerebras Systems · Sunnyvale

SeniorFull-timeOnsite

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale

Skills

GoC++Python

Discovered 1 weeks ago

Apply

Staff Software Engineer, ML Training and Inference Infrastructure

Rivian · Palo Alto

USD 228000-285000 yearly / yearSeniorFull-timeOnsite

About Rivian Rivian is on a mission to keep the world adventurous forever. This goes for the emissions-free Electric Adventure Vehicles we build, and the curious, courageous souls we seek to attract. As a company, we con

Skills

PyTorchPyTorch LightningRayCUDA

Discovered 1 weeks ago

Apply

Showing 15 jobs

Your Career Platform

Launch Your Career Today

Browse thousands of jobs, tailor your resume in 60 seconds, and start applying with confidence.

Start Free Now

Free plan available · No credit card required