Base Career helps you apply smarter for this job.
Key skills for this role
We are looking for an Observability Expert to be part of our Nestlé Nespresso Digital and Tech Team. At Nespresso, our Digital & Tech teams are at the heart of our innovation journey, a space where we continue to invest, evolve, and grow.
Location: Bengaluru, Karnataka, India
Type of Contract: Permanent
Grade : Band 2
Type of work: Hybrid
Work Language: Fluent Business English
As a Observability Platform Expert at Nespresso, you will be responsible for the hands-on implementation, configuration, operation, and continuous improvement of enterprise observability capabilities across our critical digital services, applications, infrastructure, networks, and customer journeys.
You will use your deep expertise in monitoring, logging, distributed tracing, Real User Monitoring, and telemetry correlation to improve End-to-End visibility across complex hybrid and multi-cloud environments. You will help teams detect service degradation earlier, investigate incidents faster, identify root causes, and reduce customer and business impact.
Working closely with application owners, operations, platform teams, Enterprise Architecture, security, and external partners, you will onboard services into the observability platform, validate instrumentation coverage, develop dashboards and actionable alerts, troubleshoot telemetry pipelines, and support production incident investigations.
The role requires a strong hands-on mindset. You will implement solutions directly, lead technical evaluations and Proofs of Concept, and translate business and operational requirements into measurable observability outcomes.
Knowledge of observability architecture is a strong advantage. You may contribute to target architecture, platform integration patterns, telemetry standards, application onboarding models, security and data-governance requirements, scalability, resilience, and cost optimization. You will also support architecture reviews and help ensure solutions are aligned with OpenTelemetry and other relevant industry standards.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
Rhodes, AUS
Laval, CAN
Bengaluru, IND
Crawley, GBR
Calgary, CAN
Burlington, CAN
Montréal, CAN
Brampton, CAN
You will stay current with emerging observability practices and technologies, including AIOps, AI-assisted investigation, automation, and Agentic AI, evaluating them based on practical value, operational readiness, governance, and cost.
Maintain, run, and continuously improve our monitoring infrastructure.
Setup new monitoring capabilities for the existing/new infrastructure and applications (dashboards, alerts, metrics…)
Ensure the stability of the mission-critical systems by setting up alerts, notification, and auto-remediation.
Setup event-driven automation based on events from monitoring systems.
Collaborate with an international, self-responsible and agile team who is working with the newest toolset.
Bachelor’s degree in Computer Science, Systems Engineering or a related discipline, or equivalent professional experience.
Minimum five years of experience in observability, monitoring, Site Reliability Engineering or production operations.
Strong hands-on experience implementing and operating Grafana in complex enterprise environments.
Advanced experience with:
Grafana dashboards, variables, transformations and data-source configuration.
Prometheus and PromQL.
Grafana Mimir or managed Prometheus environments.
Grafana Loki and LogQL.
Grafana Tempo and distributed tracing.
Grafana Alloy for telemetry collection and forwarding.
Grafana Alerting, contact points, notification policies and alert routing.
Experience designing and troubleshooting dashboards and alerts for applications, infrastructure, databases, networks and business services.
Strong knowledge of OpenTelemetry, including:
OpenTelemetry Collector architecture.
OTLP ingestion and export.
Metrics, logs and traces.
Context propagation and trace correlation.
Sampling and telemetry-processing strategies.
Experience onboarding applications and infrastructure into the Grafana observability ecosystem.
Experience correlating metrics, logs and traces to support End-to-End troubleshooting and root-cause analysis.
Experience troubleshooting production incidents across hybrid and multicloud environments.
Working knowledge of Linux, containers and Kubernetes.
Experience with at least one major cloud platform such as Azure, AWS, OCI or GCP.
Experience integrating Grafana with ITSM, incident-management and DevOps tools such as ServiceNow, Jira, PagerDuty or CI/CD platforms.
Proven experience leading technical evaluations, proofs of concept and production implementations.
Strong English communication skills and experience working with global, distributed teams.
We offer more than just a job. We put people first and inspire you to become the best version of yourself.
Flexible work policies including core hours and options for working from home. Discuss with us during the recruitment process to understand what flexibility could look like for you!
Genuine opportunities for career and personal development through ongoing training and constant career opportunities reflecting our conviction that people are our most important asset.
Modern "smart office" locations providing agile workspaces. Our state-of-the-art campus is equipped with areas to co-create, network, and chill!
International, dynamic & inclusive working environment with attractive additional benefits.
The pride to work for a B Corp certified company and one of the world’s most trusted brands.
Your Application: Submit your application, and we'll review it carefully (make sure your CV is in English as the hiring team is international).
Initial Screening: Relevant candidates will be contacted by our Talent Acquisition team for an initial interview.
Hiring Manager Interview: Selected candidates will then meet with the hiring manager to discuss the role and their experience in more detail.
Stakeholder Interview: Candidates will engage with potential team members to assess fit and collaboration.
Leadership & HRBP Interaction: Candidates will have a discussion with our leadership team & HRBP.
Feedback: After interviews, we provide feedback to all candidates.
Job Offer: Successful candidates will receive a formal offer.
First Working Day: Once the offer is accepted, we’ll welcome you on your first day!
The Nespresso story began with a simple but revolutionary idea: enable anyone to create the perfect cup of espresso coffee.
Since 1986, Nespresso has redefined and revolutionized the way millions of people enjoy their coffee.
We are a Company committed with the Climate change and we aim to achieve carbon neutrality as soon as possible and net-zero GHG emissions by 2050 at the latest.
In 2019 we created the digital hub in Barcelona to offer the best customer experience and innovation to B2C and B2B channels.
We encourage the diversity of applicants across gender, age, ethnicity, nationality, sexual orientation, social background, religion or belief and disability.
People are at the heart of our success – all 14,000 of them. We actively cultivate diversity, inclusion and belonging in the workplace. We celebrate individuality, believing that your authenticity and uniqueness can help us to grow and thrive together
Step outside your comfort zone; share your ideas, way of thinking and working to make a difference to the world, every single day. You own a piece of the action – make it count.
Bangalore, IN, 560103
Verified company details for this employer are not available yet.
Full-time
Senior · 5+ years experience
Hybrid
Apply faster on company sites with our extension.