Senior Software Engineering Manager – KV Cache Platform
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Key Skills for This Role
Full Job Posting
Responsibilities
- Lead, mentor, and grow a geographically distributed team of software engineers and technical leaders, fostering a culture of technical excellence, innovation, ownership, and collaboration.
- Define and execute the technical strategy and roadmap for the KV Cache Platform, ensuring scalability, reliability, security, and operational excellence.
- Drive the architecture, development, and delivery of distributed systems supporting AI inference, GPU memory optimization, distributed caching, RDMA networking, GPUDirect Storage, NVIDIA BlueField DPUs, and emerging AI infrastructure technologies.
- Partner closely with Product Management, Sales, Customer Engineering, NVIDIA, and strategic technology partners to prioritize customer requirements, drive proof-of-concepts (POCs), influence product direction, and successfully deliver customer deployments.
- Own day-to-day engineering execution, including feature development, release planning, bug triage, production issues, customer escalations, and cross-functional execution to ensure timely, high-quality software delivery.
- Establish engineering best practices for software quality, observability, automation, performance, testing, and production readiness.
- Collaborate across engineering, infrastructure, and hardware teams to deliver scalable, production-ready AI infrastructure while developing future engineering leaders and driving continuous improvement.
Required
15+ years of experience building distributed systems, cloud infrastructure, storage platforms, or AI infrastructure software.
7+ years leading high-performing software engineering organizations, including geographically distributed teams.
Proven experience delivering large-scale distributed infrastructure products from architecture through production deployment.
Strong background in distributed systems, Linux, networking, performance engineering, and cloud-native architectures.
Hands-on programming experience with Go and Python; experience with C/C++ is a plus.
Demonstrated ability to lead cross-functional initiatives and influence technical direction across multiple organizations.
Experience building AI infrastructure, LLM serving platforms, distributed caching systems, or high-performance storage solutions.
Experience with technologies such as NVIDIA Dynamo, TensorRT-LLM, Triton, RDMA, GPUDirect Storage, BlueField DPUs, Kubernetes, or related AI infrastructure.
Background in HPC, distributed storage, networking, or enterprise infrastructure software.
Experience working directly with strategic customers, technology partners, OEMs, or hyperscalers to deliver enterprise AI solutions.
About DDN
Private American data storage and AI infrastructure company serving enterprises, cloud providers, governments, and research institutions.
Visit company websiteJobs and hiring trendsApply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at DDN
Senior Sales Engineer
, USA
About DDN DDN powers the world's most demanding data workloads—from AI research to drug discovery, autonomous vehicles to financial modeling. For over two decades, our massively parallel storage solutions have enabled br
Sr. Technical Product Manager, AI Data Platforms
Santa Clara, USA
Senior Sales Engineer
, USA
QA Engineer III
Pune, IND
Staff Security Engineer
, USA
Staff Engineer - Customer Facing
Santa Clara, USA
Platform Support Architect
, USA
Senior/Staff Fuse Developer
, USA
Sr/Staff Lustre Engineer
Sacramento, USA
Sr. Technical Product Manager, AI Data Platforms
Santa Clara, USA