Senior Lead Engineer – Cloud Platform, DevOps & Site Reliability
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Role Overview
Lead architecture, reliability, scalability, and operational excellence for mission-critical cloud services in the IoT organization.
Design and operate large-scale distributed systems, drive modernization, influence architecture decisions, and establish engineering best practices.
Partner with development, quality engineering, product management, operations, architects, and customers on secure and resilient cloud-native platforms.
Key Skills for This Role
Full Job Posting
Role Summary
Lead architecture, reliability, scalability, and operational excellence for mission-critical cloud services in the IoT organization.
Design and operate large-scale distributed systems, drive modernization, influence architecture decisions, and establish engineering best practices.
Partner with development, quality engineering, product management, operations, architects, and customers on secure and resilient cloud-native platforms.
Technical Leadership
- Provide technical leadership and architectural guidance for highly scalable cloud-native platforms and services.
- Drive reliability, performance, security, automation, and operational excellence initiatives.
- Lead design reviews, architecture discussions, and technology evaluations.
- Mentor engineers and define long-term platform and infrastructure roadmaps.
Cloud and Platform Engineering
- Architect, develop, and operate cloud infrastructure across Azure and/or AWS environments.
- Define Infrastructure as Code standards and governance.
- Lead cloud migration, modernization, platform optimization, and cost governance initiatives.
- Ensure compliance with security, regulatory, and operational requirements.
DevOps and Automation
- Drive CI/CD strategy and platform engineering across multiple products.
- Build secure, scalable, automated deployment pipelines.
- Promote automation for provisioning, remediation, testing, and operational workflows.
- Champion GitOps, configuration management, and infrastructure automation.
Reliability and Performance
- Lead capacity planning, scalability assessments, performance benchmarking, and load testing.
- Drive root cause analysis and preventive actions for production issues.
- Conduct chaos engineering exercises to improve system resilience.
- Establish performance and reliability standards across engineering teams.
Required Qualifications
- Bachelor’s or Master’s degree in Computer Science, Engineering, Information Systems, or a related field.
- 8+ years of software engineering, cloud infrastructure, platform engineering, DevOps, or SRE experience.
- 5+ years leading large-scale cloud platform or infrastructure initiatives.
- Experience leading cross-functional engineering teams and driving technical strategy.
Technical Requirements
- Expert-level experience with Azure and/or AWS and enterprise-scale cloud platforms.
- Deep understanding of compute, networking, storage, identity, security, high availability, and disaster recovery.
- Strong expertise with Kubernetes, Docker, Helm, and cloud-native ecosystems.
- Experience with service mesh technologies such as Istio or Linkerd.
- Experience with multi-cluster and multi-region deployment architectures.
- Hands-on experience with Terraform, Pulumi, ARM/Bicep, or CloudFormation.
- Strong automation and scripting skills using Python, Bash, Go, or similar languages.
- Experience implementing GitOps with tools such as ArgoCD or Flux.
- Experience with Jenkins, GitHub Actions, Azure DevOps, GitLab CI/CD, or similar platforms.
- Expertise in deployment automation, release management, software delivery pipelines, SDLC, DevSecOps, and shift-left practices.
- Expertise in monitoring, logging, and observability platforms such as Prometheus and Grafana.
Apply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at Qualcomm
FY27 Intern - Low-Power AI, Audio, and Sensors Subsystem Engineering Internship - ASIC Firmware / Embedded - Canada (12 or 16 months)
Markham, CAN
Company: ------------ Qualcomm Canada ULC Job Area: ------------- Interns Group, Interns Group > Interim Engineering Intern - HW Qualcomm Overview: ---------------------- Qualcomm is a company of inventors that unlocked
FY27 Intern - Compute DSP/AI Processor Engineering Internship - Canada (16 months)
Markham, CAN
Qualcomm is seeking interns for its Compute DSP/AI Processor hardware engineering program in Markham. Interns may design, verify, model, synthesize, optimize, and debug ASIC hardware IP and subsystems using hardware desc
FY27 Intern - Machine Learning Compiler & Performance Engineering Intern - Canada (16 months)
Markham, CAN
Qualcomm is seeking a Machine Learning Compiler and Performance Engineering Intern to optimize AI workloads for Qualcomm Neural Processing Units and develop performance models and simulation tools. Applicants must be enr
FY27 Intern - Display IP Engineering Internship - Canada (16 months)
Markham, CAN
Qualcomm is building a 2027 Canada internship class for display IP engineering projects spanning ASIC design, verification, silicon validation, emulation, and systems work. Candidates should be enrolled in a relevant bac
Firmware Engineer, Staff
Toronto, CAN
Qualcomm is seeking a Staff Firmware Engineer to lead architecture, development, integration, bring-up, validation, and debugging for SERDES PHY products used in AI data centers. Candidates need strong C/C++ and Python s
RISCV CPU Design Verification (Multiple Levels)
, IND
Qualcomm is seeking design verification engineers to verify high-performance, low-power RISC-V CPUs. The role covers verification planning, simulation and formal methods, testbench development, system validation, emulati
Engineer - Camera Systems (3A)
Hyderabad, IND
Qualcomm is seeking an entry-level systems engineer for multimedia, vision, and image-processing work across embedded product platforms. The role involves designing and optimizing algorithms, supporting hardware accelera
FY27 Intern - Low-Power AI, Audio, and Sensors Subsystem Engineering Internship - ASIC Firmware / Embedded - Canada (12 or 16 months)
Markham, CAN
FY27 Intern - Compute DSP/AI Processor Engineering Internship - Canada (16 months)
Markham, CAN
FY27 Intern - Machine Learning Compiler & Performance Engineering Intern - Canada (16 months)
Markham, CAN
FY27 Intern - Display IP Engineering Internship - Canada (16 months)
Markham, CAN
Firmware Engineer, Staff
Toronto, CAN
RISCV CPU Design Verification (Multiple Levels)
, IND
Engineer - Camera Systems (3A)
Hyderabad, IND
System Performance Modeling Engineer
San Diego, USA