Skip to content
Site Reliability Engineer

Site Reliability Engineer

DCEO Engineer, Data Center Engineering Operations

Amazon Web Services (AWS)

Pariscontract

Apply on the employer's site

Role description

Description
Amazon Web Services is looking for motivated, rigorous, and technically passionate individuals to join our teams as Data Center Engineering Operations (DCEO) Engineers. This position offers an exciting opportunity to join a team dedicated to the reliability and availability of AWS critical infrastructure, serving millions of customers worldwide.

AWS Infrastructure Services (AIS)

AWS Infrastructure Services is responsible for the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we are the people who keep the cloud running. We support all AWS data centers as well as all the servers, storage equipment, networks, power systems, and cooling systems that ensure our customers have continuous access to the innovations they rely on. We work on the most complex problems, with thousands of variables impacting our operations, and we are looking for talent who want to help us tackle these challenges.

You will join a diverse team of software, hardware, and network engineers, critical systems specialists, security experts, operations managers, and other essential roles. You will collaborate with colleagues across AWS to maintain the highest safety standards while delivering seemingly infinite capacity at the lowest possible cost for our customers. You will thrive in an inclusive culture that welcomes bold ideas and empowers you to see them through.

We are currently looking for a DCEO Engineer to provide technical support for critical infrastructure within one of our data centers. This position contributes to ensuring the availability and reliability of all electrical and mechanical infrastructure in a data center environment. This equipment supports critical servers and must maintain an uptime rate greater than 99.999%.

The role requires a self-directed individual capable of taking initiative and proposing effective solutions to complex technical problems.

Successful candidates will have a direct and immediate impact on the resilience, efficiency, and capacity of our facilities, particularly by ensuring the maintenance, operation, and troubleshooting of critical infrastructure: diesel backup generators and associated fuel systems, three-phase electrical systems (switchboards, UPS, PDUs, batteries, etc.), cooling systems (CRAC units, centrifugal chillers, cooling towers, chemical water treatment, air handling units), pumps, and motors.

Key job responsibilities

  • Operate and maintain all mechanical, electrical, and HVAC (heating, ventilation, air conditioning) equipment in the data center.
  • Supervise contractors performing maintenance or preventive maintenance operations.
  • Develop intervention plans for emergency repairs on critical assets.
  • Work independently with minimal direct supervision.
  • Provide on-call support and participate in a rotation schedule based on operational needs.
  • Apply basic support concepts: ticketing systems, root cause analysis, task prioritization.
  • Perform on-site shifts according to the defined schedule.
  • Strictly comply with all physical security procedures and policies.
  • Ensure strict adherence to safety procedures during the execution of work.
  • Perform rack power installation, PDU replacement, and rack ATS replacement.
  • Verify electrical and cooling capacities before any new rack installation.
  • Create and close work orders with appropriate data (labor hours, equipment maintenance, parts used) in the asset management system.
  • Record state changes in critical infrastructure as part of corrective and preventive maintenance operations.
  • Perform quality, performance, safety, and reliability testing of products, equipment, and processes.
  • Serve as first-line responder for troubleshooting electrical and mechanical equipment: air handling units (AHUs), chillers, cooling towers, chemical treatment systems, pumps, motors, variable frequency drives (VFDs), and building management systems (BMS).
  • Define and maintain the highest safety standards, and actively promote a world-class safety culture in all aspects of operational procedures.
  • Drive continuous improvement efforts for infrastructure through standardization of procedures and policies, while ensuring performance against defined metrics.
  • Enable the operations organization to achieve 100% uptime across all customer-supporting infrastructure.
  • Collaborate effectively with internal and external stakeholders to ensure operational excellence for all AWS customers.
  • Ensure the planning and execution of preventive maintenance on critical infrastructure according to AWS procedures.
  • Ensure facility supervision and monitoring (via BMS, PMS, surveillance rounds, etc.).
  • Actively participate in the delivery of construction and renovation projects for data center infrastructure.
  • Ensure the organization's ability to respond appropriately to any event that may impact customers, on any electrical or mechanical component.
  • Analyze incident reports, document periodic trends, and formulate recommendations to management.
  • Write, update, and maintain operating procedures, standard operating procedures (SOPs), emergency procedures, preventive maintenance programs, and all technical documentation related to DCEO.

About The Team
Why AWS?

Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and have never stopped innovating — which is why customers, from the most successful startups to Global 500 enterprises, trust our robust suite of products and services to power their businesses.

Diverse Experiences

Amazon values diverse experiences. Even if you do not meet all the preferred qualifications and skills listed in the job description, we encourage you to apply. If your career is just beginning, has not followed a traditional path, or includes alternative experiences, do not let that stop you from applying.

Inclusive Team Culture

At AWS, it is in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that allows us to be proud of our differences. Continuous events and learning experiences, including our CORE (Conversations on Race and Ethnicity) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness.

Mentorship and Career Growth

We continuously raise our performance bar in our quest to become Earth's Best Employer. That is why you will find endless knowledge sharing, mentorship, and other professional development resources here to help you become a more well-rounded professional.

Work-Life Balance

We value work-life harmony. Succeeding at work should never come at the expense of sacrifices at home, which is why we strive to offer flexibility in our work culture. When we feel supported in the workplace and at home, there is nothing we cannot achieve in the cloud.

Basic Qualifications

  • Intermediate level in English and/or French (spoken and written)
  • Bachelor's degree in Electrical Engineering, Mechanical Engineering, HVAC, Industrial Technology, or Electrotechnical Engineering
  • Experience with critical infrastructure systems (UPS, generators, switchboards, HVAC, chilled water/cooling systems, pumps)
  • Familiarity with 24/7/365 critical operational environments

Preferred Qualifications

  • Fluency in English (spoken, written, and reading)
  • Knowledge of BMS (Building Management Systems) control systems and electrical power monitoring systems (EPMS)
  • Experience in data centers, manufacturing, or the Oil & Gas sector

Amazon est un employeur engagé pour l'égalité des chances. Nous sommes convaincus qu'une main d'oeuvre diversifée est essentielle à notre réussite. Nous prenons nos décisions de recrutement en fonction de votre expérience et de vos compétences. Nous apprécions votre envie de découvrir, d'inventer, de simplifier et de construire. La protection de votre vie privée et la sécurité de vos données constituent depuis longtemps une priorité absolue pour Amazon. Veuillez consulter notre Politique de Confidentialité pour en savoir plus sur la façon dont nous collectons, utilisons et traitons les données personnelles de nos candidats.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how\-we\-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Company
- Amazon Data Services France SAS

Job ID: A10468632

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Site Reliability Engineer - SRE

Scaleway

Paris

Apply on the employer's site

Role description

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
WHY WE NEED YOU?
Our growth is driving us to strengthen our SRE team to support and scale our production environments.

Your mission will be to build and maintain reliable, observable, and secure infrastructure in order to ensure optimal service availability for our customers around the world.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together.

You will be part of a team of experienced Site Reliability Engineers. The team is responsible for maintaining and evolving core infrastructure and observability tools, supporting product teams, and improving the reliability of Scaleway’s services.

YOUR DAILY ROUTINE

  • Build and optimize tooling to automate monitoring, diagnosis, and remediation of production incidents
  • Troubleshoot high-impact production issues in collaboration with other engineering teams
  • Participate in an on-call rotation to handle incidents and ensure service continuity
  • Implement and maintain observability solutions to monitor infrastructure and application health
  • Contribute to infrastructure lifecycle management across different environments
  • Promote and apply best practices in terms of stability, resiliency, scalability, and security
  • Maintain clear technical documentation for tools and procedures
  • Contribute to system and tool evolution based on production feedback
  • Collaborate closely with development teams to ensure infrastructure readiness
  • Participate in team rituals and knowledge-sharing initiatives

About You
SOFTSKILLS :

  • Proactive and solution-oriented mindset
  • Passion for automation and continuous improvement
  • Strong collaboration and communication skills
  • Ability to work independently and in a team
  • Willingness to mentor and share knowledge

💻 HARDSKILLS :

  • Experience with Go, Python or Rust
  • Strong scripting skills (Bash, Python)
  • Hands-on experience with Linux systems (Ubuntu/Debian)
  • Knowledge of networking (TCP/IP, DNS, BGP, load-balancing, IPv6, etc.)
  • Experience in cloud environments and infrastructure (bare metal, VMs, containers, orchestrators)
  • Familiarity with monitoring and logging tools (Prometheus, Grafana, Elastic, etc.)
  • Comfortable with Infrastructure-as-Code (Ansible, Salt, AWX, etc.)
  • Experience managing relational databases (PostgreSQL)
  • Understanding of CI/CD pipelines (GitLab)
  • Comfortable with English (written and spoken)

WHAT YOU WILL FIND AT SCALEWAY ++++

  • Hybrid work: We offer up to 3 days of remote work per week
  • Offices: Our offices are spacious, dynamic workspaces with bold design, conveniently located near public transport. Most of our offices feature outdoor spaces (terraces) and bike parking facilities
  • Dining: Our chef provides a healthy meal service at the headquarters, and breakfast is available across all our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches
  • Well-being commitments: Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life

International environment: With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.

  • Career & Mobility: Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers

🚀
Why join the Scaleway adventure ?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜 THE NEXT STEPS …

  • Discovery call with a recruiter (30 min)
  • Interview with the manager to understand your technical skills and approach to the role (45 min)
  • Technical interview to validate your expertise (1h)
  • Interview with the Head of the Tribe to deepen your discussions and assess your fit with the team (45 min)
  • HR interview to tour our offices and meet your future colleagues

Version française disponible ici

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Site Reliability Engineer (SRE) - Network Products

Scaleway

Paris

Apply on the employer's site

Role description

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
WHY WE NEED YOU ?
Our growth is driving us to strengthen our
Network SRE Products
team to ensure the high reliability, performance, and scalability of our storage platforms.

Your mission will be to automate, monitor, and improve the reliability, performance, and scalability of our infrastructure. You will maximize availability, optimize fault tolerance, and reduce operational overhead—ensuring robust and efficient systems for our products and services.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together. You will be part of a team of Site Reliability Engineers reporting to a Lead SRE and integrated into the SRE Guild, a collective focused on fostering best practices across engineering.

The team collaborates daily with Dev, Product, and Ops teams to improve resiliency, support service scalability, and ensure a seamless customer experience across our network solutions.

YOUR DAILY ROUTINE

  • Develop automation tools and frameworks to streamline infrastructure management
  • Build and maintain CI/CD pipelines using Infrastructure as Code best practices
  • Implement and refine monitoring and alerting systems (OpenMetrics, OpenTelemetry)
  • Ensure system reliability through incident response and root cause analysis
  • Collaborate with developers and product teams to bake resilience into network systems
  • Participate in architecture reviews and provide SRE perspective early in the design
  • Apply principles of fault-tolerance, load balancing, and energy efficiency optimization
  • Share knowledge within the team and broader engineering org via the SRE Guild
  • Contribute to the reliability and performance of services

About You
Hardskills :

  • Strong experience with Infrastructure as Code (IaC) and CI/CD pipelines
  • Solid expertise in Linux systems and production troubleshooting
  • Proficiency with monitoring/logging tools (OpenMetrics, OpenTelemetry)
  • Programming skills in Python, Go, or Rust
  • Good understanding of network systems is a great bonus: BGP, BGP EVPN, VXLAN

Softskills:

  • Collaborative mindset and team-first approach
  • Curious, continuous learner with a drive for operational excellence
  • Clear and effective communicator (written & verbal)
  • Comfortable working in English and French and across multidisciplinary teams

WHAT YOU WILL FIND AT SCALEWAY ++++
Hybrid work:
We offer up to 3 days of remote work per week.

Offices:
Our offices are spacious, dynamic workspaces with bold design, conveniently located near public transport. Most of our offices feature outdoor spaces (terraces) and bike parking facilities.

Dining:
Our chef provides a healthy meal service at the headquarters, and breakfast is available across all our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches.

Well-being commitments:
Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life.

International environment:
With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.

Career & Mobility:
Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers.

🚀
WHY JOIN THE SCALEWAY ADVENTURE?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜
THE NEXT STEPS …
Initial call
with a recruiter to get to know each other (30 min)

Technical interview
with the Head of SRE to assess your skills and approach (1h)

Manager x team interview
to explore your background and team fit (1h)

Final interview
with HR and office visit to meet your future teammates and discover our workspaces

Version française disponible ici

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Software Engineer - Kubernetes Specialist

Scaleway

Paris

Apply on the employer's site

Role description

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
OUR STORY
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow!

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is currently investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux, and Lyon.
WHY WE NEED YOU ?
Our growth is driving us to strengthen our
Non-Relational Database
team as a Software Engineer to replace a departing member and maintain our momentum on scaling our high-demand products.

Your mission will be to
ensure a robust, scalable, and high-performance experience for our customers across new sovereign regions
.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together.

You will be part of a team of
5 people
, including a Manager, a Product Manager, and 3 Developers. This young team was born from the strategic split of the Database department to focus exclusively on non-relational databases. You will work on customer-facing products with modern stacks and no heavy legacy, participating in a strong dynamic of camaraderie and technical excellence.

YOUR DAILY ROUTINE

  • Contribute to the consolidation and development of major database products in Scaleway's product ecosystem : Redis and MongoDB.
  • Maintain the production to high standards of quality for performance, stability and resilience
  • Deploy products to new geographical regions as part of Scaleway's global expansion
  • Perform code reviews and provide constructive feedback to teammates
  • Contribute to internal documentation and the development of technical best practices
  • Participate in Agile rituals (two-week sprints) and team meetings
  • Collaborate with cross-functional teams across Engineering, Product, and Sales

About You
HARD SKILLS:

  • Strong expertise in Kubernetes administration and cluster management
  • Experience with deployment tools like Helm, ArgoCD, and Flux
  • Proficiency in GoLang development preferred, but otherwise demonstrated experience in a programming language
  • Familiarity with Terraform for Infrastructure as Code
  • Familiarity with Redis, MongoDB, or other NoSQL technologies

SOFT SKILLS:

  • Strong team player with a focus on cohesion and benevolence
  • Excellent communication skills and openness to feedback
  • High level of professional maturity and ability to mentor junior colleagues
  • Proactive mindset in problem-solving and documentation
  • Comfortable working in a hybrid environment with remote teammates

WHAT YOU WILL FIND AT SCALEWAY ++++

  • Hybrid work: We offer up to 3 days of remote work per week.
  • Offices: Our offices are spacious, dynamic workspaces with bold design. Most of our offices feature outdoor spaces (terraces) and bike parking facilities.
  • Dining: Our chef provides a healthy meal service at the headquarters, and breakfast is available across almost all of our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches.
  • Well-being commitments: Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life.
  • International environment: With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.
  • Career & Mobility: Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers

🚀
Why join the Scaleway adventure?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜 THE NEXT STEPS …

  • Discovery call with a recruiter to discuss your background and expectations (30 min)
  • Interview with the Manager to understand your approach to the role (45 min)
  • Technical interview focused on Go and Kubernetes to validate your expertise (1h)
  • Interview with the Head of Tribe, Loïc Martinez Varizat, to deepen discussions and assess your fit (45 min)
  • Final on-site visit to tour our offices and meet your future colleagues

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Ingénieur Windows et Virtualisation (H/F/X)

ORNESS

Boulogne-Billancourt

Apply on the employer's site

Role description

  • Localisation : Boulogne-Billancourt
  • Type de contrat : CDI
  • Mode de travail : Hybride – 2 jours de télétravail par semaine
  • Rémunération : 58K - 68K

Contexte de l’organisation
Notre client est une organisation de grande taille opérant dans un environnement de production critique, avec de fortes exigences de disponibilité, de sécurité et de rigueur opérationnelle. Son infrastructure IT, fortement mutualisée, supporte des activités nécessitant une continuité de service permanente.

Dans un contexte de croissance et de transformation de son système d’information, l’organisation renforce son pôle
Windows Infrastructure Platform (WIP)
et recherche un·e
Ingénieur·e Infrastructure et Virtualisation expérimenté·e (N3)
(H/F/X).

L’environnement est exigeant, avec des infrastructures critiques, de fortes contraintes de disponibilité et de sécurité, et une dynamique continue de standardisation, d’industrialisation et d’automatisation.

Rôle
L’Ingénieur·e Infrastructure et Virtualisation (H/F/X) est responsable de l’exploitation, de l’évolution et du maintien en conditions opérationnelles des plateformes d’infrastructure.

Il/elle intervient en niveau expert (N3) sur des sujets complexes et contribue activement à la stabilité, l’évolution et la sécurisation des environnements.

Responsabilités

  • Exploitation & production (MCO)
  • Assurer la stabilité, la performance et la disponibilité des infrastructures
  • Prendre en charge les incidents complexes de niveau 3
  • Piloter les escalades avec les éditeurs et fournisseurs
  • Mettre en place des correctifs durables et industrialisés
  • Assurer la supervision et le monitoring des environnements critiques
  • Réaliser les analyses de cause racine (RCA)
  • Ingénierie & architecture
  • Concevoir et faire évoluer les architectures d’infrastructure et de virtualisation
  • Définir et optimiser les standards, référentiels et configurations
  • Mettre en œuvre les solutions d’infrastructure et d’exploitation
  • Définir des architectures cibles à partir des besoins métiers et assurer leur déploiement
  • Participer à la modernisation des plateformes d’infrastructure
  • Automatisation & industrialisation
  • Contribuer à l’automatisation des tâches d’exploitation (Terraform, Puppet, Chocolatey, PowerShell, API)
  • Industrialiser les processus de provisioning et de configuration
  • Améliorer la standardisation des environnements
  • Sécurité & conformité
  • Assurer le hardening et le patch management des systèmes
  • Gérer les vulnérabilités et la conformité des infrastructures
  • Garantir la traçabilité via logs et audits
  • Participer à la sécurisation des accès et authentifications (AD / Azure AD / Entra ID)
  • Respecter les normes internes et réglementaires
  • Support & amélioration continue
  • Intervenir dans un environnement structuré N1/N2/N3
  • Contribuer à la résolution des incidents et à la montée en compétence des équipes N1/N2
  • Capitaliser et documenter les solutions (procédures, bonnes pratiques)
  • Participer à l’amélioration continue de la qualité de service
  • Assurer une veille technologique active
  • Projets & transverse
  • Participer aux projets IT transverses (infrastructure, sécurité, transformation)
  • Collaborer avec les équipes infrastructure, sécurité, support et métiers
  • Contribuer aux projets de modernisation des plateformes

Environnements techniques

  • Expertise
  • VMware VCF (virtualisation)
  • Stockage SAN (PureStorage)
  • Stockage NAS (NetApp)
  • Automatisation : Terraform, Puppet, Chocolatey
  • Scripting : PowerShell, API
  • Bon niveau
  • Windows Server 2019 / 2022
  • Active Directory / Azure AD / Entra ID (hybride)
  • DNS et services Microsoft
  • Rubrik (sauvegarde)
  • Sécurité infrastructure (chiffrement, authentification, hardening)
  • Compléments
  • Linux

Profil recherché

  • Ingénieur ou équivalent avec expérience significative en infrastructure et virtualisation
  • Forte autonomie dans la gestion des environnements complexes
  • Excellent sens de l’analyse et de la résolution de problèmes
  • Capacité à documenter, transmettre et accompagner les équipes
  • Esprit d’initiative et force de proposition
  • Bon niveau d’anglais professionnel requis (minimum B2/C1)
  • Connaissance des environnements Citrix (appréciée)
  • Expérience en industrialisation et automatisation (appréciée)
  • Expérience en environnements critiques type banque, e-commerce ou production 24/7 (appréciée)

Qui sommes-nous ?
Orness, c’est une ESN à taille humaine (200 consultant·es), spécialisée en infrastructure, réseaux et ingénierie de systèmes.

Depuis plus de 20 ans, nous accompagnons les plus grands noms de la finance et de l’énergie sur des projets où fiabilité, performance et sécurité sont essentiels.

Processus de recrutement
Nous privilégions un processus simple et rapide :

  • Entretien RH approfondi (environ 1 heure)
  • Entretien technique (environ 1 heure)
  • Rencontre avec le client (30 minutes à 1 heure)

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

91 more openings in this category and country

DCEO Engineer, Data Center Engineering OperationsAmazon Web Services (AWS) · France

Apply on the employer's site