Base Career helps you apply smarter for this job.
Key skills for this role
We are seeking a Lead Machine Learning Engineer to own and advance Cephable’s core ML systems. This role is highly hands-on and technical, with responsibilities spanning model development, fine-tuning, optimization, quantization, testing, and deployment across on-device environments.
You will lead the design and implementation of ML models for speech recognition, natural language understanding, generative and reasoning tasks, and multimodal inference—ensuring they run efficiently, reliably, and privately on end-user devices.
Cephable is building the future of privacy-first, on-device AI that helps people control, create, and automate across software on-device agents, speech, computer use, and more. Our technology runs locally—offline, secure, and fast—across consumer and enterprise environments, including productivity software and games.
We work at the intersection of speech recognition, multimodal ML, generative and reasoning models, and real-time systems, optimized for modern CPUs, GPUs, and NPUs.
We are seeking a Lead Machine Learning Engineer to own and advance Cephable’s core ML systems. This role is highly hands-on and technical, with responsibilities spanning model development, fine-tuning, optimization, quantization, testing, and deployment across on-device environments.
You will lead the design and implementation of ML models for speech recognition, natural language understanding, generative and reasoning tasks, and multimodal inference—ensuring they run efficiently, reliably, and privately on end-user devices.
Design, train, fine-tune, and evaluate ML models for speech recognition, generative and reasoning models, and multimodal inference
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
Adapt open-source and foundation models using Hugging Face and related tooling
Translate research ideas into production-ready systems
Optimize models for low-latency, low-power, offline execution
Perform quantization, pruning, and distillation
Deploy models via ONNX Runtime and OpenVINO targeting CPU, GPU, and NPU backends
Build pipelines for training, evaluation, benchmarking, and regression testing
Define and improve accuracy, latency, and resource metrics
Partner with application and platform engineers to ensure seamless ML integration
Communicate model performance, architectural decisions, and technical tradeoffs clearly to both technical and non-technical stakeholders
Own Cephable’s ML architecture
Set best practices and mentor team members
Evaluate new tools, frameworks, and hardware
Mentor engineers across the team on ML concepts and practices as the org grows
You'll be joining a small, highly collaborative engineering team of engineers. You will be the dedicated ML lead — there is significant greenfield opportunity here to shape systems, practices, and architecture from the ground up. Close partnership with application and platform engineers is a core part of the role.
Mission-driven impact: Your models run on real devices for real users — many of whom depend on Cephable as a primary way to interact with technology
Greenfield ML ownership: Shape Cephable's ML architecture from the ground up with the trust and autonomy to do it right
Cutting-edge stack: On-device inference, multimodal input, NPU targeting, and privacy-first AI
Small team, high trust: Work directly with senior leadership in a low-bureaucracy environment
Seed-stage momentum: Backed by top investors with enterprise partnerships at scale
What we offer:
Meaningful equity grants (options) in a company with existing revenue and clear growth trajectory
Standard 4-year vesting with 1-year cliff
Health insurance
medical
dental
vision
120 hours per year accrued per pay period
Up to 40 hours carry over year to year (we want you taking vacation)
8 additional hours per year of tenure
15 paid Holidays
11 Federal Holidays, plus,
Fridays before Labor and Memorial Day
Extra day for July 4th
Wed before and Friday after Thanksgiving
Private on-device AI productivity software helping individuals, teams, and developers control, create, and automate across applications.
Visit company websiteJobs and hiring trendsUSD 150000-180000 yearly / year
Full-time
Senior · 4+ years experience
Remote
Apply faster on company sites with our extension.