Base Career helps you apply smarter for this job.
Key skills for this role
We’re looking for a systems-minded AI Software Engineer to join our core inference platform team. You’ll design and extend the low-level serving stack — hacking open-source frameworks like vLLM, SGLang, and TensorRT-LLM , building new model sharding and scheduling logic, and integrating deeply with our proprietary AI accelerator . This role sits at the intersection of ML systems, compiler/runtime engineering, and hardware-software co-design .
We are building the next-gen AI inference platform.
Job Title: Software Engineer, AI Inference Platform
Company: ElastixAI, Inc.
ElastixAI is an early-stage startup building the next-generation AI inference infrastructure — co-designed across ML software and custom accelerator hardware . Our platform dynamically optimizes inference efficiency and scalability across diverse deployments, enabling adaptive, high-performance AI serving.
We’re looking for a systems-minded AI Software Engineer to join our core inference platform team. You’ll design and extend the low-level serving stack — hacking open-source frameworks like vLLM, SGLang, and TensorRT-LLM , building new model sharding and scheduling logic, and integrating deeply with our proprietary AI accelerator . This role sits at the intersection of ML systems, compiler/runtime engineering, and hardware-software co-design .
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
Experience with hardware-aware ML optimization , compiler/runtime integration, or accelerator SDKs.
Hands-on experience profiling GPU/accelerator workloads .
Familiarity with containerized deployments (Docker/Kubernetes) .
Exposure to distributed systems or large-scale inference clusters.
Contributions to open-source ML or serving frameworks.
Full-time
Mid · 3+ years experience
Hybrid
Apply faster on company sites with our extension.