{bc}
ashby

Senior Software Engineer - Computer Vision Deployment

Claryo
San Francisco, USA
Full-time
Senior · 7+ years experience
Onsite
USD 170000-190000 yearly / year
Discovered 1 weeks ago
PythonPyTorchTensorFlowDeepSpeedTorchServeTensorFlow Serving
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

PythonPyTorchTensorFlow
Smart Apply

Full Job Posting

Responsibilities

  • Develop and maintain distributed cloud GPU infrastructure for large-scale world model training and low-latency inference.
  • Build end-to-end computer vision pipelines — from data ingestion and preprocessing through model training, evaluation, and deployment — and integrate them into core product workflows.
  • Deploy and optimize state-of-the-art machine learning models in the cloud using model serving platforms and inference optimization techniques, including VLMs and VLAs.
  • Design and operate orchestration systems that enable both engineers and non-engineers to build and manage data and ML pipelines.
  • Establish monitoring, benchmarking, and evaluation frameworks to ensure model performance and reliability in production environments.

Required Experience

B.S. / M.S. in Computer Science, Robotics, or similar technical field, or equivalent practical experience.

7+ years of professional software engineering experience, with at least 3 years in machine learning infrastructure — developing, scaling, training, deploying, and optimizing large-scale ML systems from data to model.

Track record of deploying computer vision models in production environments with real-world constraints.

Experience with distributed messaging and compute systems (Kafka, gRPC, ROS2, or similar).

Strong programming skills in Python with solid software engineering practices.

Preferred Experience

Experience developing, running, and managing orchestration systems (Flyte, Temporal, Airflow, or similar) for ML and data pipelines.

Proficiency with ML frameworks (PyTorch, TensorFlow, DeepSpeed) and model serving platforms (TorchServe, TensorFlow Serving, NVIDIA Triton Inference Server, or similar).

Deep understanding of state-of-the-art machine learning models such as auto-regressive transformers and familiarity with inference optimization techniques (TensorRT, quantization, custom kernels).

Experience with C++ or CUDA programming for GPU acceleration.

Prior experience working at autonomous vehicles or robotics companies.

Equal Opportunity Statement

We’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Benefits

  • At Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, parental leave, and unlimited vacation.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Claryo