Skip to content
Site Reliability Engineer

Site Reliability Engineer

Technical Product Owner — Observability Event Management

Procter & Gamble

Warsaw

Apply on the employer's site

Role description

Job Location
WARSAW DOWNTOWN OFFICE

Job Description
Mission
Own the functional design and capability roadmap for an
enterprise observability event management platform
— the intelligence layer that turns raw telemetry, alerts, events, and tickets into actionable incidents, faster diagnosis, and automated remediation. Event management sits at the center of the detect-to-correct value chain, converting high-volume operational signals into fewer, higher-quality, context-rich incidents and closed-loop recovery.

You will help advance a modern,
vendor-agnostic, composable event management platform
built on open industry standards, competing with the best commercial tools in a rapidly evolving market. Working at the intersection of
SRE, incident intelligence, AI/ML, and engineering
, you will translate product strategy into coherent capabilities — event correlation, noise reduction, anomaly detection, AI-assisted root cause analysis, and human-governed self-healing — that strengthen enterprise reliability across highly complex, interdependent application and platform landscapes. This is a high-impact role for a technically strong product leader who wants to shape state-of-the-art incident intelligence and automation at enterprise scale.

What You Will Enable

  • Fewer, higher-quality signals — Give SRE and operations teams correlated, enriched, de-noised incidents that surface real problems fast, cutting alert fatigue and accelerating triage.
  • Faster, AI-assisted diagnosis — Deliver contextual root cause analysis and incident intelligence that maps symptoms to probable causes across shared dependencies and cascading failures.
  • Closed-loop, human-governed remediation — Enable automated and agentic recovery for recurring incidents, improving MTTD, MTTR, and self-healing rate while keeping humans in control.
  • Reliability expressed in business terms — Turn operational evidence into quantified risk, SLO compliance, and readiness confidence for launches and peak demand.
  • Continuous reliability improvement — Feed operational learning back into correlation logic, detection models, and automation so recurring weaknesses are systematically eliminated over time.

Responsibilities

  • Own end-to-end functional design — Define how event management capabilities work individually and together across event ingestion, correlation, enrichment, anomaly detection, incident intelligence, and remediation orchestration; eliminate functional gaps, duplication, and inconsistent experiences.
  • Own the capability roadmap — Translate product strategy, SRE needs, and the competitive event management landscape into a clear capability-level roadmap, defining what each capability must deliver to strengthen the detect-to-correct value chain and enable agentic self-healing.
  • Lead deep technical discovery — Work directly with SREs, operations teams, platform engineers, application owners, and architects to understand operational pain points, failure modes, event patterns, and the complex technical setups of observed applications and platforms.
  • Define rigorous product requirements — Turn discovery into functional specifications, user journeys, correlation and enrichment logic, integration requirements, non-functional expectations, user stories, and precise acceptance criteria that engineering teams can implement with confidence.
  • Guide engineering design — Partner closely with engineers and architects on APIs, event and data models, integration patterns, correlation and ML approaches, scalability, security, and extensibility; challenge design choices and make functional trade-offs grounded in technical reality.
  • Shape event management and agentic capabilities — Translate needs for event correlation, anomaly detection, incident intelligence, remediation orchestration, and human-governed agentic workflows into a coherent, forward-looking capability set informed by leading platforms and emerging patterns.
  • Create a coherent, state-of-the-art product experience — Ensure consistent concepts, workflows, defaults, and automation patterns across capabilities and personas, benchmarked against leading event management and observability platforms and the evolving market.
  • Assure functional quality and measurable value — Validate that delivered capabilities match design intent, work coherently end to end, and contribute to MTTD, MTTR, incident volume reduction, automation coverage, self-healing rate, SLO compliance, and user satisfaction.

Required
Job Qualifications

  • Engineering or deeply technical product background strongly preferred — for example, prior experience as a Software Engineer, Platform Engineer, SRE, Observability Engineer, Solution Architect, Data/ML Engineer, or Technical Product Owner working closely with complex platform engineering teams.
  • Experience shaping enterprise-grade, data-intensive, AI-powered, infrastructure, observability, or automation products, with evidence of taking ambiguous operational problems through discovery, functional design, roadmap ownership, and engineering delivery.
  • Strong technical depth across cloud and distributed systems, event-driven architectures, APIs, data pipelines, telemetry, and AI/ML fundamentals; able to reason about architecture and implementation trade-offs without needing to be the primary coder.
  • Proven ability to define functional product architecture, detailed specifications, user journeys, non-functional expectations, and precise acceptance criteria for complex technical capabilities.
  • Excellent written and verbal English, with the ability to guide engineers and architects, explain design decisions clearly, and connect technical choices to SRE, reliability, user, and business outcomes.

Strong plus

  • Hands-on or design-level experience with observability and event management technologies such as Datadog, Dynatrace, Splunk ITSI, ServiceNow ITOM, PagerDuty, BigPanda, Moogsoft, New Relic, Grafana, Prometheus, or OpenTelemetry.
  • Deep familiarity with SRE practices and the operations value chain across observe and detect, correlate and triage, investigate and diagnose, resolve, and learn and prevent.
  • Experience designing event correlation logic, anomaly detection, enrichment and topology-aware integrations, remediation orchestration, runbook automation, multi-tenant systems, or reusable and configuration-driven product models.
  • Experience with GenAI or agentic-AI products, retrieval-augmented generation, evaluation metrics, MCP, or human-in-the-loop automation.
  • Familiarity with ITIL 4 event, incident, problem, and change management and the detect-to-correct value stream.
  • SAFe PO/PM, PSPO, SRE, cloud, or relevant engineering certification.

We offer

  • P&G-sized projects and access to world leading IT partners and technologies from Day 1.
  • Wide range of self-development possibilities (training and certifications paths).
  • Competitive starting salary and benefits program (private health care, P&G stock, saving plans, sport cards).
  • Regular salary increases and possible promotions - in line with your results and performance.
  • Opportunity to change role every few years to be in the best place for you and best for P&G.

At Procter & Gamble we embrace a
hybrid work model
that combines the flexibility of remote work with the collaborative benefits of in-office engagement. Employees can enjoy the option to work from home two days a week while also spending time in the office to foster teamwork and enhance communication.

Watch this video to learn more about our full recruiting process: https://www.youtube.com/watch?v\=0bicvbpy0gI

Kindly be advised that at P&G, employment is exclusively extended on the basis of an "Umowa o Pracę" (Full-time Employment Contract). Apply only if you agree to these conditions.

About Us
We produce globally recognized brands and we grow the best business leaders in the industry. With a portfolio of trusted brands as diverse as ours, it is paramount our leaders can lead with courage the vast array of brands, categories and functions. We serve consumers around the world with one of the strongest portfolios of trusted, quality, leadership brands, including Always®, Ariel®, Gillette®, Head & Shoulders®, Herbal Essences®, Oral-B®, Pampers®, Pantene®, Tampax® and more. Our community includes operations in approximately 70 countries worldwide.

Visit http://www.pg.com to know more.

We are an equal opportunity employer and value diversity at our company. We do not discriminate against individuals on the basis of race, color, gender, age, national origin, religion, sexual orientation, gender identity or expression, marital status, citizenship, disability, HIV/AIDS status, or any other legally protected factor.

Job Schedule
Full time

Job Number
R000158047

Job Segmentation
Experienced Professionals

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Ingénieur Système, Sauvegarde et réseaux à Paris H/F

Free-Work

Remotefulltimesenior

Apply on the employer's site

Role description

Dans le cadre du développement de notre équipe IT chez l'un de nos clients grands comptes, nous recherchons un(e) Ingénieur(e) Système Sauvegarde réseaux H/F afin d'assurer l'administration, la maintenance et l'évolution des infrastructures systèmes de nos environnements clients.

Vos missions seront les suivantes :

  • Participer aux projets d'évolutions de la plateforme technique de la Video Factory ( tête de réseau OTT et IPTV
  • Conception / participation au POC avec le N3 (les experts) / intégration / ingénierie / recette unitaire / recette des workflow et documentation des briques techniques et des procédures.
  • Assurer la maintenance en condition opérationnelle de la plateforme
  • Organisation des sauvegardes et des montées de version des équipements IT (VM, NAS, OS (linux et windows)...) et Vidéos (encodeurs, DCM, serveurs d'origine, sondes ) de la plateforme.
  • Analyse des risques et des impacts potentiels et planification en HNO le cas échéant.
  • Assurer le « run » de la plateforme en tant que support niveau 2 en soutien des équipes support de niveau 0 et 1
  • Analyse d'incidents et résolutions, escalade au N3 et aux fournisseurs (ouverture et suivi des tickets) le cas échéant, communication sur les avancées les plus significatives et les impacts majeurs,
  • Organiser les activités des fournisseurs (mises à jour et évolution technique),
  • Assurer le suivi des déploiements et mettre en place les contrôles (recette).
  • Participer à l'amélioration de l'organisation du support global de la plate-forme technique :
  • Formation des équipes de maintenance, documentation des procédures d'exploitation.

     Des missions complémentaires peuvent être confiées.

Référence de l'offre : cmk6d3cswx

Profil candidat:

Profil recherché :

Issu(e) d'une formation supérieure en informatique (BAC+5 / Diplôme d'ingénieur),

Vous justifiez d'au moins 5 ans d'expérience sur un poste similaire.

Compétences techniques souhaitées :

  • Maîtrise des environnements Windows Server, Linux et Vmware (des connaissances sur Nutanix est un plus).
  • Bonne connaissance des environnements vidéo et audio sur IP.
  • Connaissance des réseaux IP et de l'adressage multicast.
  • Maitrise des outils d'analyse de qualité vidéo ainsi que des outils de supervision.
  • Bonnes capacités d'analyse et de résolution de problèmes.
  • Autonomie, rigueur et bon relationnel.
  • Capacité à travailler en équipe et à intervenir dans des environnements de production

Environnement technique :

DCM, Anevia, Imagine, Nevion, Harmonic xOS, OpenHeadEnd, Elemental, USP.

Produits systèmes, virtualisation : VMware, Wallix, Nutanix, NAS, Debian, Windows Server, FTP

Ce que nous vous proposons :

Valeurs : en plus de nos 3 fondamentaux que sont l'audace, la bonne foi et la réactivité, nous garantissons un management à l'écoute et de proximité, ainsi qu'une ambiance familiale.

Contrat : CDI ou Freelance

Localisation : Paris

Package rémunération & avantages :

  • Le salaire : rémunération annuelle brute selon profil et compétences
  • Les basiques : mutuelle familiale et prévoyance, titres restaurant, remboursement transport en commun à 50%, avantages du CSE (culture, voyage, chèque vacances, et cadeaux), RTT (jusqu'à 12 par an), plan d'épargne, prime de participation
  • Nos plus : forfait mobilité douce & Green (vélo/trottinette et covoiturage), prime de cooptation de 1000 € brut, e-shop de matériel informatique à des prix préférentiels (smartphone, tablette, etc.)
  • Votre carrière : plan de carrière, dispositifs de formation techniques & fonctionnels, passage de certifications, accès illimité à Microsoft Learn
  • La qualité de vie au travail : télétravail avec indemnité, évènements festifs et collaboratifs, accompagnement handicap et santé au travail, engagements RSE

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

SAP DevOps Engineer (f/m/d)

E.ON Digital Technology

Remote

Apply on the employer's site

Role description

You have a passion for technology and want to make the world a greener place?
Then become a playmaker (f/m/d) and join our team as SAP DevOps Engineer (f/m/d) at E.ON Digital Technology.

We play a key role in shaping the energy transition by leading E.ON's digital transformation across Europe. We explore new paths by developing ideas, breaking new ground, making visions reality, and bringing new technologies to life. We deliver sustainable technology solutions because…

… it’s on us to make new energy work!
The Team
– your impact

At E.ON, the SAP Engineering Chapter is a collective of innovative minds dedicated to delivering world-class SAP architecture, SAP software development and SAP engineering capabilities across our segments and product teams. By joining us, you will play a critical role in keeping E.ON's SAP architecture and capabilities modern, secure and excellent.

Your Role –
meaningful & rewarding

As a SAP DevOps Engineer (f/m/d) at E.ON, you will be responsible for designing, developing, training and maintaining SAP DevOps solutions for our SAP system landscape and platforms. You will work closely with the SAP DevOps teams and developers, and other stakeholders to ensure the provisioning of a modern, user-friendly, state-of-the-art SAP DevOps pipeline.

  • Design, implement and continuously improve DevOps concepts, CI/CD templates, pipelines and automation for SAP landscapes (e.g. SAP BTP, S/4HANA)
  • Enable reliable build, test and deployment processes across SAP Cloud and on-premise environments
  • Establish observability concepts for SAP BTP and on-premise applications for monitoring, health checks, alerts and APM by using tools such as SAP Cloud ALM, New Relic and Uptrends
  • Collaborate closely with SAP development, architecture, security and operations teams
  • Conduct DevOps maturity assessments and guide teams on their DevOps Journey
  • Ensure high availability, performance, security and compliance of SAP systems
  • Troubleshoot complex problems and support root cause analysis
  • Continuously evaluate new SAP DevOps tools, technologies and best practices

Your Profile
– authentic & open-minded

  • Strong experience in DevOps, platform engineering or operations within SAP environments
  • Hands-on experience with CI/CD and security, as well as quality tooling (e.g. GitLab CI/CD, SonarQube, Renovate or similar)
  • Strong knowledge in development, automated testing and testing strategies
  • Experience with scripting and automation (e.g. Bash, Groovy)
  • Solid knowledge of SAP S/4HANA, SAP BTP and related deployment and transport mechanisms
  • Knowledge of Kubernetes and container technologies and IaaC principles is a plus
  • Strong understanding of security, monitoring and reliability concepts
  • Structured, proactive and solution-oriented working style
  • Fluent in English

Our Benefits
– smart & useful

  • Advance your development: We grow and we want you to grow with us. Learning on the job, exchanging with others, or taking part in an individualtraining – our learning culture enables you to bring your personal and professional development to the next level.
  • Recharge your battery: You have 30 days of paid vacation per year plus Christmas and New Year's Eve off. Your battery still needs charging? You canexchange parts of your salary for more paid vacation or you can take a sabbatical.
  • Enjoy hybrid work: We combine office collaboration with focused work from home. It’s also possible to go on workation for up to 20 days per year withinEurope.
  • Stay active & healthy: Benefit from a company-sponsored health membership.
  • Elevate your mobility: From car and bike leasing offers to a subsidised Deutschland-Ticket – your way is our way.
  • Think ahead: With our company pension scheme and a great insurance package we take care of your future.
  • This is by far not all… We are looking forward to speaking with you about further benefits during the hiring process.

Do you have questions?
For further information please contact Agneta Lierl, EDT_Talent_Acquisition@eon.com.

What you need to know:
Contract type: Permanent

Working time: Full time

Company: E.ON Digital Technology GmbH

Location: Essen, Hannover, München, Berlin, Würzburg, Hamburg, Frankfurt am Main

Function area: IT/Digital; Engineering

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Security Engineer (m/w/d)

Rocken®

Remotesenior

Apply on the employer's site

Role description

Steuerverwaltungen brauchen clevere Software – und dahinter stehen clevere Köpfe. Unser Rocken Partner entwickelt die führende Business Lösung für kantonale und kommunale Steuerverwaltungen. Die Anwendung deckt den kompletten Verwaltungsprozess ab: vom Steuerregister über Veranlagungen und Fakturierung bis zum Inkasso und zur Verlustscheinbewirtschaftung. Derzeit entsteht eine neue Software-Generation – ein spannendes Projekt, das technisches Know-how und Innovationskraft vereint. Die Arbeitsweise ist geprägt von überlegter Fokussierung, lebendigem Austausch zwischen Teams und intellektueller Courage. Elegante Lösungen entstehen durch Wissen, Geist und Teamenergie. Bereit, an einem Produkt zu arbeiten, das echten Impact hat? Unser Rocken Partner sucht Menschen, die sich für eine gemeinsame Idee begeistern und die digitale Zukunft der Steuerverwaltung mitgestalten möchten.

Verantwortung

  • Du analysierst Schwachstellen und setzt passende Sicherheitslösungen um.
  • Du arbeitest an Themen wie Zugriff, Netzwerk, Pentesting und Monitoring.
  • Du entwickelst die Sicherheitsarchitektur weiter und beachtest ISO 27001.
  • Du automatisierst in Windows/Linux und unterstützt Kubernetes-Setups.

Qualifikationen

  • Du hast eine IT-Ausbildung und Erfahrung als System Engineer mit Security-Fokus.
  • Du denkst vernetzt, erkennst Risiken und arbeitest agil mit modernen Tools.
  • Du kennst Azure, VMware und gängige Security-Lösungen.
  • Du brennst für IT-Security und entwickelst im Team nachhaltige Lösungen.

Benefits

  • Interessante und abwechslungsreiche Tätigkeiten/Projekte
  • Attraktive Weiterbildungs- und Entwicklungsmöglichkeiten
  • Flexible Arbeitszeitgestaltung
  • Homeoffice
  • Offene Unternehmenskultur
  • Beteiligung oder Übernahme ÖV-Abonnements
  • Beteiligung oder Übernahme Parkplatz
  • Attraktive Mitarbeiterrabatte
  • Kostenlose Früchte und Getränke

ROCKEN Jobs
https://rocken.jobs

Profil Erstellen
https://rocken.jobs/application/profil\-erstellen/

  • new

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Microsoft Cloud & System Engineer (m/w/d)

Rocken®

Remote

Apply on the employer's site

Role description

Unser Rocken® Partner ist spezialisiert auf IT-Security, Cloud Lösungen, sowie IT-Outsourcing und Support. Sie setzen auf innovative Services und Dienstleistungen, bewegen und orientieren sich am Puls der Technik. Daraus ergeben sich für unsere Kunden viele Vorteile gegenüber den traditionellen IT-Lösungen, die weniger flexibel und selten skalierbar sind.

Verantwortung

  • Sicherstellung einer stabilen IT-Infrastruktur und hohen Systemverfügbarkeit
  • Betreuung und Weiterentwicklung von Microsoft-Umgebungen im Kundenumfeld
  • Planung und Umsetzung von Infrastruktur-, Rollout- und Migrationsprojekten
  • Analyse und Behebung technischer Störungen im 1st- und 2nd-Level-Support
  • Direkter Kundensupport, technische Beratung und Betreuung vor Ort

Qualifikationen

  • Microsoft 365, Azure, Entra ID und Microsoft Intune
  • Kenntnisse in Client-/Server-Systemen, Virtualisierung, Monitoring und Backup
  • Erfahrung im 1st-/2nd-Level-Support und technischen Troubleshooting
  • Kenntnisse in Infrastrukturplanung, System-Rollouts und Migrationen
  • Sehr gute Deutschkenntnisse sowie gute Englischkenntnisse; weitere Sprachen von Vorteil

Benefits

  • Flexible Arbeitszeitgestaltung
  • Homeoffice
  • Zahlreiche Mitarbeiterevents
  • Beteiligung oder Übernahme Parkplatz
  • Kostenlose Früchte und Getränke
  • Attraktive Weiterbildungs- und Entwicklungsmöglichkeiten

ROCKEN Jobs
https://rocken.jobs

Profil Erstellen
https://rocken.jobs/application/profil\-erstellen/

  • new

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

241 more openings in this category and country

Technical Product Owner — Observability Event ManagementProcter & Gamble

Apply on the employer's site
Technical Product Owner — Observability Event Management — Procter & Gamble | mentors.coach