Skip to content
Site Reliability Engineer

Site Reliability Engineer

Electrical Infrastructure Reliability Manager, EU AMZL RME

Amazon Web Services

Milanfulltime

Apply on the employer's site

Role description

DESCRIPTION

Operations is at the heart of everything Amazon does, and reliability is one of the ways we optimise our processes. The Amazon Reliability Maintenance and Engineering team keeps all our equipment in excellent working order, from printers to fully automated sortation systems.

As a Electrical Infrastructure Reliability Manager, you'll have a proven track record of improving the reliability of electrical systems, via maintenance commissioning and inspection programs. You’ll make sure that our automated systems are able to work at their best so that our customers get their orders on time. By exploring data on everything from control systems to sensors, you’ll minimise EU wide downtime due to power outages. Prior experience in any RME capacity at an Amazon FC/SC/DS site is valuable. German, Italian, Spanish or French language proficiency is considered a preferred qualification.

Key job responsibilities

  • Lead Power Resilence improvement workstreams involving multiple internal teams: central engineering teams and site teams including operational leadership, maintenance and safety
  • Manage vendor relationships and conduct technical review of vendor deliverables for accuracy, completeness, and compliance
  • Deliver regular status reports and executive-level briefings to senior stakeholders across engineering, operations, and safety functions
  • Be capable of being a subject matter expert in electrical distribution (LV, MV and HV) for site auditing, root cause analysis and problem resolution.
  • Lead end-to-end Power System Studies, including data collection, modelling, technical review of vendor deliverables, on-site implementation of findings, lessons learned.
  • Evaluate commissioning, maintenance, and testing documentation for key electrical assets (e.g., transformers, switchgear) to identify risks, verify compliance, and support data-driven maintenance decisions.
  • Ensure full compliance with EU electrical safety standards and regulations (BS 7671, local EU statutory requirements).
  • Apply deep knowledge of electrical operational risks, including arc flash hazards, to develop and implement effective mitigation strategies, safe systems of work, and engineering controls
  • Leverage data and analytics to support evidence-based decisions on asset reliability and program prioritization
  • Support annual budgeting cycles and planning

A day in the life

You will identify status trends, and assess audit findings related to electrical and associated risks & drive corrective actions EU wide, through structured problem-solving and root cause analysis and serve as the connective tissue between engineering teams, site-based RME teams, vendors, and senior stakeholders. You will lead end-to-end execution of electrical programs from design through commissioning and ensure electrical safety, reliability, resilience, and sustainability standards are met across Amazon's European operations. Key internal stakeholders are D&C, RME, Operations, and Global Building Design.

About the team

The goal of Amazon Logistics is to build a world class last mile operation. Amazon Logistics aims to exceed the expectations of our customers by ensuring that their orders, no matter how large or small, are delivered as quickly, accurately, and cost effectively as possible.

Our Reliability Maintenance Engineering or RME team keep our equipment performing at its best. We're a technically minded team, made up of excellent team players and guided by experienced leaders. We work together improve, troubleshoot and standardise equipment across our network of delivery stations. Everything we do focuses on reducing downtime in Amazon's crucial operations sites, so customers get their orders on time.

BASIC QUALIFICATIONS

  • Bachelor's degree
  • Experience and strong technical background in relevant fields of automated or non-automated material handling equipment
  • Work flexible schedule including weekends, nights, and holidays
  • Bachelor's degree in electrical engineering or equivalent
  • Knowledge of electrical engineering principles including switchgear, UPS, transformers, and circuit breakers
  • Experience leading large-scale, technical or engineering programs with a proven record of thought leadership, business case development, realizing customer benefits, and successful program completion
  • Knowledge of network analysis fundamentals and robust troubleshooting
  • Experience in vendor management, or experience in electrical or mechanical engineering
  • 5+ years of progressive electrical discipline specific experience in electrical engineering and/or technical program management on electrical subjects
  • Experience conducting Power System Studies, using electrical engineering software including short circuit analysis (IEC 60909 / ANSI/IEEE methods), protection coordination studies (relay grading, IDMT curves), arc flash hazard analysis (IEEE 1584 / NFPA 70E) and load flow and power quality assessments for LV/MV/HV systems
  • Experience in asset reliability including LV/MV/HV switchgear inspection, maintenance, and life-cycle management, Transformer condition monitoring including DGA (Dissolved Gas Analysis), Thermal imaging, partial discharge testing, and insulation resistance assessments and developing and implementing asset reliability frameworks and PMPs

PREFERRED QUALIFICATIONS

  • Master's degree in engineering, mechanical, operations, supply chain, business administration, or equivalent STEM field
  • Experience in Lean Management, Six Sigma and other operations engineer tools
  • Experience leading complex global or country-wide initiatives
  • Professional Engineer License
  • Speak, write, and read fluently in French, Spanish, German or Italian

Amazon is an equal opportunities employer. We believe passionately that employing a diverse workforce is central to our success. We make recruiting decisions based on your experience and skills. We value your passion to discover, invent, simplify and build. Protecting your privacy and the security of your data is a longstanding top priority for Amazon. Please consult our Privacy Notice (https://www.amazon.jobs/en/privacy\_page) to know more about how we collect, use and transfer the personal data of our candidates.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how\-we\-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The minimum gross base salary for this position is listed below. The base salary listed corresponds to working on a full-time basis. For part-time hours, the salary will be pro-rated.

Amazon reserves the right to offer a higher salary and/or level, depending on the candidate's skills, competencies, and experience.Amazon's package may include a sign on payment. In addition, the candidate may be eligible to participate in a restricted stock unit scheme operated independently by Amazon.com Inc. in USA.

Your recruiting team will share final salary and any restricted stock unit scheme if applicable, depending on skills and requirements.

In addition to statutory benefits, and those applicable to the relevant CBA, company supplementary benefits may apply subject to further terms.

Milan, ITA - 49,900.00 EUR Annually

Rome, ITA - 49,900.00 EUR Annually

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Senior AI Tech Architect - Milan | BCG Platinion

BCG Platinion

Milan

Apply on the employer's site

Role description

Who We Are

At BCG Platinion, we close the gap between strategy and technology. We design and build AI tech strategies, AI solutions, and next-generation platform architectures that turn ambition into execution, turning buzzwords into a true business engine and real competitive advantage.

Working at the intersection of boardroom strategy and hands-on tech delivery, we help leading organizations tackle their toughest challenges, with access to C-suite leaders and the unmatched resources of BCG.

What You'll Do

As a Senior AI Tech Architect, you will hold end-to-end ownership of the overall solution architecture — spanning API strategy, data platform design, cloud infrastructure, and AI engineering.

The role requires the ability to operate with authority across all dimensions: from deep technical governance and architecture reviews to C-suite-level communication and strategic advisory with client leadership.

  • You lead the design and evolution of enterprise-grade AI engineering platforms, setting the architectural standards and delivery patterns to be implemented across engagements
  • You own the end-to-end governance framework for AI-powered systems — defining trust boundaries, identity models, validation hierarchies, and audit mechanisms — and ensure these standards are adopted consistently across multi-team delivery environments
  • You architect and oversee evaluation and security programs at scale: from red-teaming strategies and adversarial testing frameworks to compliance-by-design approaches embedded in CI/CD pipelines
  • You translate executive-level business intent into multi-horizon AI architectures, balancing immediate delivery velocity with long-term technical sustainability, and define the integration path within existing system landscapes and operating models
  • You act as a technical authority in client-facing settings, shaping AI architecture strategy at C-suite and CTO level, and codifying reusable patterns and governance playbooks

What You'll Bring

  • You hold a degree in Computer Science, Information Systems, AI Engineering, or a related field
  • You have 6-9 years of experience in software development or technical roles, with a strong track record in cloud, backend, or system integration at enterprise scale
  • You lead the analysis, design, and optimization of complex IT architectures across multi-team engagements, with deep expertise in legacy modernization, enterprise solutions, and advanced market trends including agentic AI, LLM orchestration, and AI-native architectures
  • You have extensive hands-on experience with AI systems — including LLM-based applications, multi-agent architectures, and ML platforms — and have delivered these in production enterprise environments
  • You define and govern modern software delivery standards, including DevSecOps practices, CI/CD pipelines, automated testing strategies, and platform engineering approaches across delivery teams
  • You have deep command of concepts like agent assembly lines, execution harnesses, and context engineering, and own the architectural decisions that embed them securely and scalably into enterprise landscapes
  • You translate executive-level business intent into multi-horizon technical architectures and communicate with authority across C-suite, CTO-level, and technical audiences
  • You thrive in interdisciplinary teams and global environments, and are flexible and willing to travel as needed
  • You are willing to travel to client sites in Italy and abroad (frequency varies by project)
  • You have strong written and verbal communication skills in both Italian and English (minimum B2, C1 preferred) — senior client-facing interaction and executive communication are core components of the role

Who You'll Work With

We make sure you never stop growing. You'll tackle exciting, high-impact challenges every day, supported by a passionate team of talented colleagues who share your drive for excellence and keep connect long after your case wraps up. Our collaborative, open culture gives you the freedom to explore and perfect your strengths, backed by individual learning opportunities and recognized certifications.

Additional info

At BCG Platinion, our people and relationships are at the heart of everything we do. We believe that in-person collaboration plays an essential role in our culture, mentorship, and professional development. That's why we generally operate on a hybrid model, with office presence shaped by business needs, team norms, and local policies and practices. We aim to provide a dynamic and collaborative environment that fosters connection and teamwork.

Compensation Information

Total compensation for this role includes base salary, annual discretionary performance bonus, and a market-leading benefits package described below. Our Reward Philosophy is to take a holistic approach to offer a comprehensive package which includes market-competitive base salary, discretionary performance bonus and strong benefits that are locally relevant and support both physical and mental health. Our philosophy drives our actions and aims at being competitive, fair and easy for employees to understand. We are committed to ensuring that BCG Platinion is the best place for you to thrive and grow.

The base salary range for this role is
EUR 62.000,00 - EUR 66.000,00
.

This is an estimated range. Any final offer would be based on objective, role-related criteria such as your relevant experience, qualifications, and demonstrated skills. Employees may also be eligible for an annual discretionary performance bonus.

BCG Platinion offers a best-in-class benefits package designed to provide comprehensive care and meaningful support to you and your family. Benefits may vary based on role, contract type, and eligibility criteria.

Our Benefits Include

  • Health insurance: comprehensive medical coverage for employees with the option to extend coverage to eligible family members
  • Meal vouchers and corporate discounts: daily meal allowance and access to a range of partner discounts
  • Life/Accident /Disability Insurance for employees
  • Employee Assistance Program (EAP): confidential psychological, legal, and financial support services
  • Well-being initiatives: programs supporting physical health, mental well-being, and family needs
  • Flexible benefits/Welfare: an annual flexible benefits allowance that can be used for eligible services such as education, health, transport, and leisure activities
  • Company car plan: available for eligible roles

Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws.

BCG is an E - Verify Employer. Click here for more information on E-Verify.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Data Center Network Specialist – Presidio

Gruppo Informet

Milan

Apply on the employer's site

Role description

Sede di lavoro: Roma, Napoli, Milano oppure Smart Working

Contratto: Tempo indeterminato, P.Iva

Stipendio: Da 25.000 - 40.000 € lordi annui in base alle proprie competenze

Ricerchiamo
Sistemisti in ambito Networking e/o Security.
L’ inserimento iniziale è previsto nel team di
presidio
per clienti della Pubblica Amministrazione. Saranno possibili, ove richiesto, interazioni con i team di Ingegneria su tematiche di progetto o di pre-vendita. La risorsa sarà inserita nella Technology Unit che eroga servizi professionali afferenti le area tecnologiche di Networking, Hybrid Cloud e Cyber Security.

Sede e Orario di lavoro

  • L’attività viene svolta in modalità full remote, ma con possibilità di trasferta sul territorio nazionale.
  • Turni di 8h all’interno della fascia oraria 7.30-19.30
  • Reperibilità fuori dall’orario di presidio stimata in 1 settimana al mese (inclusi weekend e festivi)
  • Attività programmate fuori orario di presidio

Inviare il proprio curriculum a
curriculum@informet.it
Requisiti richiesti

Le Competenze Tecniche Richieste Sono

  • Buona conoscenza di apparati di rete router e switch in tecnologia Cisco e preferibilmente anche Juniper.
  • Buona conoscenza di reti IP.
  • Buona conoscenza dei protocolli di switching quali STP, FabricPath, vPC.
  • Buona conoscenza di protocolli di routing statico e dinamico come OSPF e BGP
  • Buona conoscenza di architetture SDN quali Cisco ACI, VMware NSX.
  • Buona conoscenza delle architetture di Load Balancing su vendor F5
  • Conoscenza di architetture DNS basate su Infoblox.
  • Conoscenza delle architetture di rete in Cloud Azure e/o AWS.
  • Buona conoscenza della lingua inglese scritto e parlato.
  • Aver maturato esperienza in ambito di conduzione (operation)

Principali responsabilità

La Risorsa Deve Possedere Buone Competenze In Ambito Data Center Networking Su Architetture Legacy e Innovative (SDN), Incluso Tematiche Di Load Balancing, Con Focus Sui Vendor Cisco Ed F5. L’attività è Finalizzata Alla Conduzione Della Infrastruttura (operation) Ed Interventi Di Analisi e Risoluzione Di Problematiche. A Titolo Indicativo, La Risorsa Deve Essere In Grado Di Erogare Con Efficacia Attività Di

  • Supporto Specialistico di II livello per problematiche di maggiore complessità tecnica rilevate dal I livello o rilevato autonomamente in proattività.
  • Supporto sistemistico ai tecnici on-site ed ai referenti tecnici di I livello.
  • Elaborazione di eventuali proposte di miglioramento e di tuning delle piattaforme da sottoporre ai referenti del Cliente.
  • Supporto nella implementazione di processi di controllo del corretto funzionamento delle applicazioni e piattaforme tramite l’integrazione con i tool di monitoring
  • Implementazione di major e minor changes e/o configurazioni

Caratteristiche Personali

La Risorsa Deve

  • Avere comprovata esperienza relativamente alle competenze richieste
  • Parlare correntemente la lingua italiana e avere buona conoscenza della lingua inglese
  • possedere ottime doti comunicative e relazionali;
  • Avere un forte orientamento al cliente e alla soluzione di problemi;
  • Possedere caratteristiche di dinamicità, efficienza, flessibilità ed autonomia;

Costituiscono Titolo Preferenziale Una Delle Seguenti Certificazioni

  • CISCO CCNP Data Center
  • F5 CERTIFIED TECHNICAL SPECIALIST (BIG-IP LTM)

Benefit

  • Assicurazione sulla vita
  • Buoni pasto
  • Cellulare aziendale
  • Computer aziendale
  • Formazione e Certificazione professionale
  • Lavoro da casa
  • Orario flessibile
  • Supporto allo sviluppo professionale

Tipi Di Retribuzione Supplementare

  • Bonus annuale
  • Premio di produzione

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Software Engineer

Worldline

Milansenior

Apply on the employer's site

Role description

Who We Are
Worldline helps businesses of all shapes and sizes to accelerate their growth journey - quickly, simply, and securely. We are the innovators at the heart of the payments technology industry, shaping how the world pays and gets paid. Our technology powers the growth of millions of businesses across 5 continents. And just as we help our customers accelerate their business, we are committed to helping our people accelerate their careers. Together, we shape the evolution.

Objective

  • Design and develop software solutions with regard to full technical stack in line with business requirements, applying architectural best practices for efficient code creation.
  • Ensure software quality throughout the entire software development lifecycle to deliver reliable products.

Activity

Concept and Design

  • Collaborate with stakeholders, including product managers and designers, to understand project requirements and objectives.
  • Contribute to the conceptualization and design of software solutions, ensuring alignment with business goals.
  • Translate high-level requirements into detailed technical specifications and system designs.
  • Apply software architecture principles to create scalable and modular software structures.

Planning

  • Participate in project planning and estimation, providing insights into technical feasibility and effort required.
  • Break down software development tasks into actionable items and prioritize them based on project goals and timelines.
  • Identify potential risks and challenges early in the planning phase and propose mitigation strategies.

Updates and Maintenance

  • Perform regular updates, enhancements, and optimizations to existing software systems.
  • Debug and troubleshoot issues reported by users or identified during maintenance cycles.
  • Collaborate with the operation team to ensure smooth operation, stability, and reliability of software products.
  • Implement backward-compatible changes and updates to maintain software integrity.

Coding and Testing

  • Write clean, efficient, and maintainable code according to coding standards and best practices.
  • Develop software components and features using appropriate programming languages and frameworks.
  • Implement automated unit tests, integration tests, and regression tests to ensure software quality.
  • Debug and resolve issues identified during testing phases, maintaining a focus on code quality and performance.

Analysis

  • Analyze complex technical problems and propose innovative solutions to improve software functionality and performance.
  • Conduct thorough code reviews, providing constructive feedback to peers and fostering a culture of code quality.
  • Perform performance analysis to identify bottlenecks and areas for optimization in software systems.
  • Use data-driven insights to make informed decisions about software design, architecture, and improvements.

Footer

Perks & Benefits
At Worldline you’ll get the chance to be at the heart of the global payments technology industry and shape how the world pays and gets paid. On top of that, you will also:

  • Be part of a company guided by a strong purpose to do good and recognized as top 1% of the most sustainable companies in all sectors worldwide.
  • Work with inspiring colleagues and be empowered to learn, grow and accelerate your career.
  • Work in an international environment with cutting edge technologies.
  • Enjoy a wide range of benefits: medical insurance, pension found, tickets restaurant, company bonus, 50% of remote working.

Learn more about life at Worldline at
jobs.worldline.com
We are proud to be an Equal Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as an individual with a disability, or any applicable legally protected characteristics.

Salary Information

In accordance with the EU Pay Transparency Directive, the salary range for this position is between:

Italy: 48,000.00 euro and 65,000.00 euro gross annual salary.

The final offer will be determined based on the candidate's professional experience, specific skills, and qualifications.

Shape the evolution
We are on an exciting journey towards the next frontiers of payments technology, and we look for big thinkers, people with passion, can-do attitude and a hunger to learn and grow. Here you'll work with ambitious colleagues from around the world, take on unique challenges as a team, and make a real impact on the society. With an empowering culture, strong technology and extensive training opportunities, we help you accelerate your career - wherever you decide to go. Join our global team of 18,000 innovators and shape a tomorrow that is yours to own.

Learn more about life at Worldline at jobs.worldline.com

We are proud to be an Equal Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex, sexual orientation, gender identity, gender expression, age, status as an individual with a disability, or any applicable legally protected characteristics.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Infrastructure Engineer

ToolsGroup

Milansenior

Apply on the employer's site

Role description

About Us
We are a dynamic, rapidly growing global company and the innovators of service-driven supply chain planning software. We help companies make better, faster supply chain decisions that reduce inventory, improve customer satisfaction, and deliver powerful financial results amid increasing complexity, product proliferation, and uncertainty.

Our solutions have been recognized by customers globally and analyst firms, such as Gartner, for our ability to support service and inventory trade-offs, while dramatically improving planner productivity. ToolsGroup has been successfully deployed worldwide in more than 44 countries, and we have one of the highest customer retention rates in our industry.

About The Role
We are looking for an experienced Site Reliability Engineer who can also lead IT service operations. You will lead a small IT/Ops team, remain the senior technical escalation point for production services, and act as the operational interface between IT/Ops, Engineering, Product, Security, Support, and business teams.

This is not a coordination-only service management role or a generalist infrastructure position. You will diagnose distributed-system failures using logs, metrics, traces, commands, and platform tooling; make safe recovery decisions; automate recurring work; and engineer lasting reliability improvements.

Main Responsibilities

  • Lead major incidents from impact assessment and containment through recovery, stakeholder communication, root-cause analysis, and corrective actions.
  • Troubleshoot complex issues across Windows and Linux systems, Kubernetes and container workloads, hybrid networking and DNS, cloud infrastructure, identity, authentication, databases, storage, APIs, and service dependencies.
  • Operate and improve Azure, OCI, or comparable cloud environments, including monitoring, access controls, backup and recovery, reliability, and cost-aware scaling.
  • Define and improve service-level indicators and objectives, observability, alert quality, capacity, resilience, dependency mapping, and recovery readiness for critical services.
  • Automate operational tasks and controls using PowerShell, Python, infrastructure as code, or CI/CD pipelines, with validation, logging, secure credential handling, and rollback.
  • Apply incident, change, and problem management pragmatically, protecting service availability without introducing unnecessary process.
  • Connect technical and business teams: clarify service ownership and dependencies, translate business needs into reliability and infrastructure requirements, frame risk and tradeoffs, align priorities, and ensure decisions have accountable owners and realistic commitments.

What We Are Looking For

  • We care more about demonstrated production engineering experience than a checklist of certifications. Strong candidates will bring all of the following:
  • A strong SRE or Production Engineering background, typically 5+ years operating business-critical, customer-facing, or high-availability services. Recent work must include direct technical ownership, not only coordination or people management.
  • Recent ownership of high-severity incidents, including technical triage, recovery decisions, clear communications, and measurable follow-through.
  • Strong systems and network troubleshooting fundamentals: Windows and Linux, TCP/IP, DNS, routing, firewalls, proxies or load balancers, and the ability to isolate faults across service layers.
  • Hands-on cloud operations experience in Azure, OCI, or a similar platform, including compute, storage, networking, IAM, observability, backup, and recovery.
  • Production experience with containers and Kubernetes, including workload health, scheduling, networking, persistent storage, secrets, deployment and rollback, scaling, and backup or recovery considerations.
  • Deep observability and reliability engineering practice: metrics, logs, distributed tracing, actionable alerting, SLI/SLO design, capacity and saturation analysis, failure-mode thinking, and post-incident engineering.
  • Ability to diagnose database-backed and API-driven services across application, query, connection-pool, storage, certificate, network, and downstream dependency layers.
  • Practical identity and access management experience with Active Directory and Microsoft Entra ID or equivalent, including hybrid identity, privileged access, MFA, service identities or gMSAs, lifecycle controls, dependency mapping, and controlled recovery from identity failures.
  • Evidence of safe automation and infrastructure-as-code work using PowerShell, Python, Terraform or comparable tooling and CI/CD. You should be able to explain testing, idempotency, error handling, credential security, rollout, rollback, and measurable impact.
  • Experience leading, mentoring, or acting as the senior escalation point for other technical professionals.
  • Strong business-facing and cross-functional leadership. You can translate technical complexity into business impact and options, challenge unsafe or unrealistic requests constructively, negotiate priorities, and communicate decisions clearly to engineers, executives, customers, and non-technical stakeholders.

Additional Relevant Experience:

  • Microsoft 365, endpoint management, EDR, device compliance, and hybrid workplace operations.
  • Formal ITIL, cloud, security, or infrastructure certifications.

What Success Looks Like

  • Incidents are diagnosed and resolved with greater speed, structure, and confidence.
  • Monitoring, runbooks, automation, recovery controls, and change practices reduce repeat failures and operational toil.
  • The team becomes more capable and accountable without depending on a single hero.
  • Technical and business teams share clear service ownership, priorities, risk decisions, and delivery expectations.

Our hiring process
The process includes a scenario-based SRE technical discussion. We will ask you to think aloud through realistic production incidents involving cloud, Kubernetes, identity, networking, databases, storage, APIs, and service dependencies. You will be expected to describe the logs, metrics, traces, commands, tools, tradeoffs, and recovery criteria you would use. We will also assess how you align technical and business stakeholders when priorities, risk, and customer commitments conflict.

Our Vision, Purpose, and Values
Our VISION: Unparalleled control over demand and supply to deliver certainty.

Our PURPOSE: Problem Solvers Welcome.

Our VALUES: Deliver the Goods – Have Deep Care – Find the Right Answer, Not the First Answer – Creativity That Endures – Brilliant But Not Loud.

Salary range:
55-68k/year, plus 10% bonus based on personal and company objectives.

Equal Opportunity Employer
ToolsGroup provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

ToolsGroup is an E-Verify employer, to learn more please visit E-Verify.gov

Notice to third parties
The ToolsGroup Human Resources team manages the recruitment and employment process. Note that ToolsGroup will only accept resumes from a third-party recruiter or placement agency if a fully executed, written search agreement is in place at the start of the recruitment effort for a specific position and only if ToolsGroup HR has engaged with that firm for support.

Unsolicited resumes sent to ToolsGroup from third-party recruiters and placement agencies do not (a) constitute any type of relationship with ToolsGroup or (b) obligate ToolsGroup to pay fees should we elect to hire from those resumes. Please do not contact or present candidates directly to any ToolsGroup personnel.

Applying to this job the candidate consents that his/her data are treated by ToolsGroup in compliance with the GDPR n. 2016/679 GDPR and Transparency Document

U.S. applicant notice: This employer participates in E-Verify and will provide the federal government with your Form I-9 information to confirm that you are authorized to work in the U.S. ToolsGroup is CCPA/CPRA compliant.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

23 more openings in this category and country

Electrical Infrastructure Reliability Manager, EU AMZL RMEAmazon Web Services · Italy

Apply on the employer's site