Observability Expert (Grafana)
Observability Expert (Grafana)
OBSERVABILITY EXPERT (GRAFANA)
We are looking for an Observability Expert to be part of our Nestlé Nespresso Digital and Tech Team. At Nespresso, our Digital & Tech teams are at the heart of our innovation journey, a space where we continue to invest, evolve, and grow.
Position Snapshot:
- Location: Bengaluru, Karnataka, India
- Type of Contract: Permanent
- Grade: Band 2
- Type of work: Hybrid
- Work Language: Fluent Business English
The Role:
As a Observability Platform Expert at Nespresso, you will be responsible for the hands-on implementation, configuration, operation, and continuous improvement of enterprise observability capabilities across our critical digital services, applications, infrastructure, networks, and customer journeys.
You will use your deep expertise in monitoring, logging, distributed tracing, Real User Monitoring, and telemetry correlation to improve End-to-End visibility across complex hybrid and multi-cloud environments. You will help teams detect service degradation earlier, investigate incidents faster, identify root causes, and reduce customer and business impact.
Working closely with application owners, operations, platform teams, Enterprise Architecture, security, and external partners, you will onboard services into the observability platform, validate instrumentation coverage, develop dashboards and actionable alerts, troubleshoot telemetry pipelines, and support production incident investigations.
The role requires a strong hands-on mindset. You will implement solutions directly, lead technical evaluations and Proofs of Concept, and translate business and operational requirements into measurable observability outcomes.
Knowledge of observability architecture is a strong advantage. You may contribute to target architecture, platform integration patterns, telemetry standards, application onboarding models, security and data-governance requirements, scalability, resilience, and cost optimization. You will also support architecture reviews and help ensure solutions are aligned with OpenTelemetry and other relevant industry standards.
You will stay current with emerging observability practices and technologies, including AIOps, AI-assisted investigation, automation, and Agentic AI, evaluating them based on practical value, operational readiness, governance, and cost.
In This Role, You Will:
- Maintain, run, and continuously improve our monitoring infrastructure.
- Setup new monitoring capabilities for the existing/new infrastructure and applications (dashboards, alerts, metrics…)
- Ensure the stability of the mission-critical systems by setting up alerts, notification, and auto-remediation.
- Setup event-driven automation based on events from monitoring systems.
- Collaborate with an international, self-responsible and agile team who is working with the newest toolset.
What We’re Looking For:
- Bachelor’s degree in Computer Science, Systems Engineering or a related discipline, or equivalent professional experience.
- Minimum five years of experience in observability, monitoring, Site Reliability Engineering or production operations.
- Strong hands-on experience implementing and operating Grafana in complex enterprise environments.
Advanced experience with:
- Grafana dashboards, variables, transformations and data-source configuration.
- Prometheus and PromQL.
- Grafana Mimir or managed Prometheus environments.
- Grafana Loki and LogQL.
- Grafana Tempo and distributed tracing.
- Grafana Alloy for telemetry collection and forwarding.
- Grafana Alerting, contact points, notification policies and alert routing.
- Experience designing and troubleshooting dashboards and alerts for applications, infrastructure, databases, networks and business services.
- Strong knowledge of OpenTelemetry, including:
- OpenTelemetry Collector architecture.
- OTLP ingestion and export.
- Metrics, logs and traces.
- Context propagation and trace correlation.
- Sampling and telemetry-processing strategies.
- Experience onboarding applications and infrastructure into the Grafana observability ecosystem.
- Experience correlating metrics, logs and traces to support End-to-End troubleshooting and root-cause analysis.
- Experience troubleshooting production incidents across hybrid and multicloud environments.
- Working knowledge of Linux, containers and Kubernetes.
- Experience with at least one major cloud platform such as Azure, AWS, OCI or GCP.
- Experience integrating Grafana with ITSM, incident-management and DevOps tools such as ServiceNow, Jira, PagerDuty or CI/CD platforms.
- Proven experience leading technical evaluations, proofs of concept and production implementations.
- Strong English communication skills and experience working with global, distributed teams.
Extra Skills That Set You Apart:
- Experience having worked in a global environment and with virtual teams.
- Solid understanding and hands on experience using Ansible and Kubernetes.
- Experience working with Ecommerce platforms.
- Good understanding of MS Azure.
- Experience with UNIX/Linux and UNIX shell scripting knowledge (Bash and Python).
- Architecture experience — strong advantage
- Experience designing enterprise observability architectures for hybrid, multicloud and on-premises environments.
- Ability to define telemetry collection, processing, routing and storage architectures.
- Experience designing OpenTelemetry-based architectures that minimize vendor lock-in.
We Offer You:
We offer more than just a job. We put people first and inspire you to become the best version of yourself.
- Flexible work policies including core hours and options for working from home. Discuss with us during the recruitment process to understand what flexibility could look like for you!
- Genuine opportunities for career and personal development through ongoing training and constant career opportunities reflecting our conviction that people are our most important asset.
- Modern "smart office" locations providing agile workspaces. Our state-of-the-art campus is equipped with areas to co-create, network, and chill!
- International, dynamic & inclusive working environment with attractive additional benefits.
- The pride to work for a B Corp certified company and one of the world’s most trusted brands.
The Hiring Process:
- Your Application: Submit your application, and we'll review it carefully (make sure your CV is in English as the hiring team is international).
- Initial Screening: Relevant candidates will be contacted by our Talent Acquisition team for an initial interview.
- Hiring Manager Interview: Selected candidates will then meet with the hiring manager to discuss the role and their experience in more detail.
- Stakeholder Interview: Candidates will engage with potential team members to assess fit and collaboration.
- Leadership & HRBP Interaction: Candidates will have a discussion with our leadership team & HRBP.
- Feedback: After interviews, we provide feedback to all candidates.
- Job Offer: Successful candidates will receive a formal offer.
- First Working Day: Once the offer is accepted, we’ll welcome you on your first day!
About Nespresso:
The Nespresso story began with a simple but revolutionary idea: enable anyone to create the perfect cup of espresso coffee.
Since 1986, Nespresso has redefined and revolutionized the way millions of people enjoy their coffee.
We are a Company committed with the Climate change and we aim to achieve carbon neutrality as soon as possible and net-zero GHG emissions by 2050 at the latest.
In 2019 we created the digital hub in Barcelona to offer the best customer experience and innovation to B2C and B2B channels.
We encourage the diversity of applicants across gender, age, ethnicity, nationality, sexual orientation, social background, religion or belief and disability.
People are at the heart of our success – all 14,000 of them. We actively cultivate diversity, inclusion and belonging in the workplace. We celebrate individuality, believing that your authenticity and uniqueness can help us to grow and thrive together
Step outside your comfort zone; share your ideas, way of thinking and working to make a difference to the world, every single day. You own a piece of the action – make it count.
Join Nestlé #beaforceforgood
OBSERVABILITY EXPERT (GRAFANA)
We are looking for an Observability Expert to be part of our Nestlé Nespresso Digital and Tech Team. At Nespresso, our Digital & Tech teams are at the heart of our innovation journey, a space where we continue to invest, evolve, and grow.
Position Snapshot:
- Location: Bengaluru, Karnataka, India
- Type of Contract: Permanent
- Grade: Band 2
- Type of work: Hybrid
- Work Language: Fluent Business English
The Role:
As a Observability Platform Expert at Nespresso, you will be responsible for the hands-on implementation, configuration, operation, and continuous improvement of enterprise observability capabilities across our critical digital services, applications, infrastructure, networks, and customer journeys.
You will use your deep expertise in monitoring, logging, distributed tracing, Real User Monitoring, and telemetry correlation to improve End-to-End visibility across complex hybrid and multi-cloud environments. You will help teams detect service degradation earlier, investigate incidents faster, identify root causes, and reduce customer and business impact.
Working closely with application owners, operations, platform teams, Enterprise Architecture, security, and external partners, you will onboard services into the observability platform, validate instrumentation coverage, develop dashboards and actionable alerts, troubleshoot telemetry pipelines, and support production incident investigations.
The role requires a strong hands-on mindset. You will implement solutions directly, lead technical evaluations and Proofs of Concept, and translate business and operational requirements into measurable observability outcomes.
Knowledge of observability architecture is a strong advantage. You may contribute to target architecture, platform integration patterns, telemetry standards, application onboarding models, security and data-governance requirements, scalability, resilience, and cost optimization. You will also support architecture reviews and help ensure solutions are aligned with OpenTelemetry and other relevant industry standards.
You will stay current with emerging observability practices and technologies, including AIOps, AI-assisted investigation, automation, and Agentic AI, evaluating them based on practical value, operational readiness, governance, and cost.
In This Role, You Will:
- Maintain, run, and continuously improve our monitoring infrastructure.
- Setup new monitoring capabilities for the existing/new infrastructure and applications (dashboards, alerts, metrics…)
- Ensure the stability of the mission-critical systems by setting up alerts, notification, and auto-remediation.
- Setup event-driven automation based on events from monitoring systems.
- Collaborate with an international, self-responsible and agile team who is working with the newest toolset.
What We’re Looking For:
- Bachelor’s degree in Computer Science, Systems Engineering or a related discipline, or equivalent professional experience.
- Minimum five years of experience in observability, monitoring, Site Reliability Engineering or production operations.
- Strong hands-on experience implementing and operating Grafana in complex enterprise environments.
Advanced experience with:
- Grafana dashboards, variables, transformations and data-source configuration.
- Prometheus and PromQL.
- Grafana Mimir or managed Prometheus environments.
- Grafana Loki and LogQL.
- Grafana Tempo and distributed tracing.
- Grafana Alloy for telemetry collection and forwarding.
- Grafana Alerting, contact points, notification policies and alert routing.
- Experience designing and troubleshooting dashboards and alerts for applications, infrastructure, databases, networks and business services.
- Strong knowledge of OpenTelemetry, including:
- OpenTelemetry Collector architecture.
- OTLP ingestion and export.
- Metrics, logs and traces.
- Context propagation and trace correlation.
- Sampling and telemetry-processing strategies.
- Experience onboarding applications and infrastructure into the Grafana observability ecosystem.
- Experience correlating metrics, logs and traces to support End-to-End troubleshooting and root-cause analysis.
- Experience troubleshooting production incidents across hybrid and multicloud environments.
- Working knowledge of Linux, containers and Kubernetes.
- Experience with at least one major cloud platform such as Azure, AWS, OCI or GCP.
- Experience integrating Grafana with ITSM, incident-management and DevOps tools such as ServiceNow, Jira, PagerDuty or CI/CD platforms.
- Proven experience leading technical evaluations, proofs of concept and production implementations.
- Strong English communication skills and experience working with global, distributed teams.
Extra Skills That Set You Apart:
- Experience having worked in a global environment and with virtual teams.
- Solid understanding and hands on experience using Ansible and Kubernetes.
- Experience working with Ecommerce platforms.
- Good understanding of MS Azure.
- Experience with UNIX/Linux and UNIX shell scripting knowledge (Bash and Python).
- Architecture experience — strong advantage
- Experience designing enterprise observability architectures for hybrid, multicloud and on-premises environments.
- Ability to define telemetry collection, processing, routing and storage architectures.
- Experience designing OpenTelemetry-based architectures that minimize vendor lock-in.
We Offer You:
We offer more than just a job. We put people first and inspire you to become the best version of yourself.
- Flexible work policies including core hours and options for working from home. Discuss with us during the recruitment process to understand what flexibility could look like for you!
- Genuine opportunities for career and personal development through ongoing training and constant career opportunities reflecting our conviction that people are our most important asset.
- Modern "smart office" locations providing agile workspaces. Our state-of-the-art campus is equipped with areas to co-create, network, and chill!
- International, dynamic & inclusive working environment with attractive additional benefits.
- The pride to work for a B Corp certified company and one of the world’s most trusted brands.
The Hiring Process:
- Your Application: Submit your application, and we'll review it carefully (make sure your CV is in English as the hiring team is international).
- Initial Screening: Relevant candidates will be contacted by our Talent Acquisition team for an initial interview.
- Hiring Manager Interview: Selected candidates will then meet with the hiring manager to discuss the role and their experience in more detail.
- Stakeholder Interview: Candidates will engage with potential team members to assess fit and collaboration.
- Leadership & HRBP Interaction: Candidates will have a discussion with our leadership team & HRBP.
- Feedback: After interviews, we provide feedback to all candidates.
- Job Offer: Successful candidates will receive a formal offer.
- First Working Day: Once the offer is accepted, we’ll welcome you on your first day!
About Nespresso:
The Nespresso story began with a simple but revolutionary idea: enable anyone to create the perfect cup of espresso coffee.
Since 1986, Nespresso has redefined and revolutionized the way millions of people enjoy their coffee.
We are a Company committed with the Climate change and we aim to achieve carbon neutrality as soon as possible and net-zero GHG emissions by 2050 at the latest.
In 2019 we created the digital hub in Barcelona to offer the best customer experience and innovation to B2C and B2B channels.
We encourage the diversity of applicants across gender, age, ethnicity, nationality, sexual orientation, social background, religion or belief and disability.
People are at the heart of our success – all 14,000 of them. We actively cultivate diversity, inclusion and belonging in the workplace. We celebrate individuality, believing that your authenticity and uniqueness can help us to grow and thrive together
Step outside your comfort zone; share your ideas, way of thinking and working to make a difference to the world, every single day. You own a piece of the action – make it count.
Join Nestlé #beaforceforgood
Bangalore, IN, 560103
Bangalore, IN, 560103