Collaborate with AppSec and compliance teams to meet standards such as ISO 27001, NIST 800-53, SOC 2, and CIS Benchmarks.
Define and own the organization's SRE strategy, roadmap, and operating model, including SLI/SLO/SLA frameworks and error budget policies.
Design and implement observability architectures using APM tools, distributed tracing, log aggregation, and real-time alerting (e.g., Prometheus, Grafana, Datadog, Azure Monitor).
Lead incident management, RCA/post-mortem processes, and drive systemic improvements to reduce MTTR and recurrence.
Champion chaos engineering and resilience testing practices to proactively identify failure modes before they impact production.
Establish toil reduction programs through automation, enabling SRE teams to focus on reliability engineering over manual operations.
Design high-availability, fault-tolerant architectures for distributed systems and microservices on cloud platforms.
Drive capacity planning, performance engineering, and scalability reviews across critical services.
Build Internal Developer Platforms (IDPs) and self-service tooling to accelerate developer productivity while maintaining security and compliance guardrails.
Apply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
Honeywell Connected Industrial (HCI) Line of business is a leading provider of software offerings in the Process Industry which includes Oil and Gas, Refining, Petrochemicals, Paper, MMM industry. These offerings drive m
Manager, Product Management & Business Development – Industrial Sensing
Richardson, USA
ManagerFull-time
As a Lead Business Development Analyst here at Honeywell, you will play a key role in driving business growth and identifying new opportunities for Honeywell
We are seeking an experienced Electrical Estimator / Proposal Engineer with 2–8 years of experience in estimation and proposal preparation for electrical systems including switchgear, transformers, MV systems, RMU, and P
You will report directly to our Director Materials Management, and work from one of the following locations: Houston, TX, Rosemont, IL, Mobile, AL, Charlotte, NC or Shreveport, LA on a Hybrid work schedule.
The Materials
As a Lead Account Manager here at Honeywell, you will play a critical role in managing key customer accounts and driving business growth within your territory. You will be responsible for developing strong customer relat
As a Lead Account Manager here at Honeywell, you will play a critical role in managing key customer accounts and driving business growth within your territory. You will be responsible for developing strong customer relat
As Lead Field Service Technician at Honeywell, you will play a crucial role in ensuring the optimal performance of our building automation systems. The main function of this position is to respond and resolve complex ser
Establish GitOps workflows for continuous delivery and configuration management at scale.
Oversee container orchestration strategy using Kubernetes (EKS, AKS, Rancher) and define standards for workload security, networking, and observability.
Guide, mentor, and upskill cross-functional DevSecOps and SRE teams, fostering a culture of reliability, security, and continuous improvement.
Facilitate development operations and identify gaps, setbacks, and shortcomings across processes.
Collaborate with product, engineering, and business stakeholders to align platform strategy with organizational goals.
Define and track KPIs and engineering metrics (deployment frequency, change failure rate, DORA metrics, mean time to detect).
Deliver internal training, conduct architecture reviews, and represent the platform function in senior leadership forums.
Requirements
Bachelor's or Master's degree in Computer Science, Information Systems, or a related field.
5-10 years of total industry experience, with 3+ years in a DevSecOps or SRE Lead role.
Proven track record of leading large-scale cloud-native transformation programs.
Certification of Cloud Platform and Infrastructure (AWS/Azure)
Certification of DevOps tools.
Strong analytical, problem-solving, and communication skills
Ability to lead geographically distributed and cross-functional teams
Qualifications & Skills
Technical Skills
Cloud, Virtualization, and Infrastructure Microsoft Azure, AWS, GCP, VMware, and cloud services Windows Server and Linux Server administration RedHat, CentOS, and Ubuntu environments System sizing, software installation, maintenance, and upgrades Azure / VMware backup, restore, and disaster recovery tools
Microsoft Azure, AWS, GCP, VMware, and cloud services
Windows Server and Linux Server administration
RedHat, CentOS, and Ubuntu environments
System sizing, software installation, maintenance, and upgrades
Azure / VMware backup, restore, and disaster recovery tools
DevOps, CI/CD, and Automation DevOps pipelines and CI/CD practices Bamboo, Octopus, Jenkins, Flux, and SonarQube Terraform, PowerShell, Azure CLI, Shell, and Perl scripting Software development and deployment automation Power BI
DevOps pipelines and CI/CD practices
Bamboo, Octopus, Jenkins, Flux, and SonarQube
Terraform, PowerShell, Azure CLI, Shell, and Perl scripting
Software development and deployment automation
Power BI
Containers and Kubernetes Kubernetes cluster installation, configuration, and management Docker containerization Containerized application deployment and operations KEDA, Scale Set Operations, ISTIO
Kubernetes cluster installation, configuration, and management
Docker containerization
Containerized application deployment and operations
KEDA, Scale Set Operations, ISTIO
Networking and Directory Services TCP/IP, DNS, and load balancing NGINX, HAProxy Active Directory, DNS configuration, Group Policy Objects, and OU structure
TCP/IP, DNS, and load balancing
NGINX, HAProxy
Active Directory, DNS configuration, Group Policy Objects, and OU structure
API, Architecture, and Reliability Swagger and Postman Mission-critical system architecture High availability design Performance monitoring Reliable infrastructure operations
Swagger and Postman
Mission-critical system architecture
High availability design
Performance monitoring
Reliable infrastructure operations
Core SRE Competencies
Deep expertise in secure SDLC and threat modelling
Experience designing SLO frameworks and error budget management
Strong knowledge of compliance frameworks: NIST, ISO 27001, SOC 2, CIS, FIPS
Experience with Agile and SAFe methodologies
Preferred / Good to Have
Certifications: CISSP, CISM, CKA, AWS/Azure Security Specialist, Google SRE
Experience with AI/ML-driven DevOps automation and AIOps platforms
Knowledge of FinOps and cloud cost optimization strategies
Familiarity with Platform Engineering and Internal Developer Portals (Backstage)
Six Sigma or ITIL certification for process excellence