{bc}
breezy

Senior Development Operations Engineer (DevOps)

Rapta, Inc
Remote, USA
Full-time
Senior · 10+ years experience
Remote
USD 140000-180000 bi-weekly / bi weekly
Discovered 1 weeks ago
BazelPythonAnsibleGolangLLMDocker
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

BazelPythonAnsible
Smart Apply

Full Job Posting

Full-Time Position | Remote (US)

About Us

Rapta is revolutionizing American manufacturing with our AI-powered vision systems. Our cutting-edge software platform expands manufacturing capacity 30% by reducing errors 90%+ and automating quality control and inspection processes. We're looking for a talented senior development operations engineer to join our mission and make a real impact.

Position Overview

We're seeking an experienced senior development operations (DevOps) engineer to own the build, test, and deployment pipeline for our distributed computer vision platform running at edge sites across customer manufacturing floors. This is a full-time remote position working directly with our engineering team to harden our release process, drive deployment automation, and pioneer LLM-driven test generation and validation at Rapta.

What You'll Do

Own and evolve the end-to-end release pipeline — branching strategy, build orchestration, artifact promotion, and rollback — across our Bazel monorepo and Python deployable units

Design and maintain Ansible-driven fleet automation for heterogeneous Linux edge nodes (Ubuntu LTS, NVIDIA driver stacks, Docker with NVIDIA runtime)

Manage all update tooling, currently written in Golang

Build LLM-powered automated testing systems: test generation from specs, flake triage, log/failure analysis, regression diffing, and release-note synthesis from commit and ticket history

Harden CI/CD for offline and bandwidth-constrained deployment targets (airgap wheel distribution, signed artifacts, deterministic builds)

Drive observability for releases — deployment telemetry, version drift detection, and post-deploy health validation across the fleet

Mentor engineers on release hygiene, reproducible builds, and infrastructure-as-code practices

What We're Looking For

10+ years of professional experience in release engineering, DevOps, or SRE roles shipping production Linux systems

Deep curiosity for software, infrastructure, and applied AI — particularly using LLMs as production engineering tools, not just chat assistants

Expert-level Python (3.8+) with a strong grasp of packaging, dependency resolution, and PEP 440 versioning discipline

Demonstrated ownership of Linux fleets at scale — kernel, systemd, networking, package management

Excellence in technical communication, runbook authorship, and post-incident documentation

Strong systems thinking — comfortable reasoning about failure modes across hardware, OS, container, and application layers

Required Technical Skills

Expert proficiency with Ansible (roles, dynamic inventory, idempotent design); working knowledge of Terraform

Expert proficiency with Docker, including creation and lifecycle management of containers, image hardening, registry management and installing & configuring the NVIDIA container runtime

Production experience with Linux administration: systemd, networking (VLANs, DHCP, DNS), kernel/driver management (especially NVIDIA/DKMS), package and APT internals

Strong Python skills focused on tooling, automation, packaging (wheels, pip, private indexes), and subprocess/CI integration

Proficiency with Git workflows, branching strategies, and modern CI/CD systems (GitHub Actions, GitLab CI, or equivalent)

Experience designing and operating automated test infrastructure — unit, integration, hardware-in-the-loop, and end-to-end

Practical experience using LLMs (Anthropic, OpenAI, or local) as part of engineering workflows — test generation, code review augmentation, log analysis, or agentic tooling

Nice to Have

Bazel or similar monorepo build systems

Edge or embedded deployment experience

Tailscale, WireGuard, or zero-trust networking in production

gRPC/protobuf service ecosystems

Vault, PKI, or secrets management at fleet scale

Background in regulated or compliance-driven environments (CMMC, ISO 27001, SOC 2)

Why Join Us?

Work on cutting-edge AI infrastructure with real-world impact on American manufacturing

Build the release engineering foundation for a late-seed company actively scaling

Direct collaboration with the CTO and engineering leadership

Remote-first culture with flexible hours

Meaningful equity in a company solving hard problems

Location Requirements

Remote (US)

Equal Opportunity

Rapta is committed to hiring and retaining a diverse workforce. We are proud to be an Equal Opportunity/Affirmative Action Employer, making decisions without regard to race, color, religion, creed, sex, sexual orientation, gender identity, marital status, national origin, age, veteran status, disability, or any other protected class.

How to Apply

  • No recruiters or agencies — we only accept applications directly from applicants.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Rapta, Inc