Skip to content
Site Reliability Engineer

Site Reliability Engineer

Senior Software Engineer - Incident Insights & Readiness

Datadog

Parissenior

Откликнуться на сайте работодателя

Описание вакансии

The Incident Insights & Readiness SRE team at Datadog fosters a resilient culture by using incidents as learning opportunities and catalysts for growth. Our users are Datadog engineers, and we build the software, tooling, and operational frameworks that help them prepare for, respond to, and learn from incidents. We work closely with engineering teams across Datadog to analyze incidents and turn those insights into better tools, stronger incident response, and organizational learning. Our efforts empower Datadog to navigate unexpected failures confidently, efficiently, and with a commitment to continuous learning and systems improvement.

At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them.
The Incident Insights & Readiness SRE team at Datadog fosters a resilient culture by using incidents as learning opportunities and catalysts for growth. Our users are Datadog engineers, and we build the software, tooling, and operational frameworks that help them prepare for, respond to, and learn from incidents. We work closely with engineering teams across Datadog to analyze incidents and turn those insights into better tools, stronger incident response, and organizational learning. Our efforts empower Datadog to navigate unexpected failures confidently, efficiently, and with a commitment to continuous learning and systems improvement.

At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them.
What You’ll Do:

  • Own and improve the on-call experience for the company by establishing best practices and building platforms to support on-call rotations and compensation.
  • Define how we respond to incidents, lead the design and implementation of software to streamline the process, and collaborate with product teams to improve incident response across Datadog. Our aim is to fully support our incident responders in dealing with complexity.
  • Contribute to the post-mortem process for the company, collaborating with teams on writing them, and identifying opportunities to reduce friction and enhance learning value for the organization. Our team also runs a weekly postmortem reading group.
  • Support various teams in facilitating incident reviews that emphasize learning and blamelessness. Help them share their learnings across the organization to improve the resilience of our people.
  • Provide technical leadership and day-to-day coaching to team members, accelerating their growth through design reviews, collaborative problem-solving and operational excellence best practices.
  • Train our on-callers in incident and post-mortem processes, sharing expertise in incident management best practices. This involves both introducing newcomers to on-call responsibilities and refreshing the knowledge of existing engineers.
  • Lead cross-functional initiatives in engineering organizations across Datadog, embedding with teams to understand their challenges and drive lasting improvements to reliability and operational excellence.

Who You Are:

  • At least 5 years of experience building software that solves real user problems. Experience designing new features and collaborating on code and technical design reviews. We primarily develop in Go and Python, with a bit of TypeScript.
  • Experience building or operating distributed systems, with familiarity with Kubernetes and an understanding of complex failure modes.
  • Demonstrated ability to independently own ambiguous technical problems from design through delivery while balancing long-term engineering quality with pragmatic execution.
  • Experience analyzing incidents, identifying systemic risks, and driving engineering improvements informed by operational learnings.
  • Experience participating in on-call rotations and improving incident response processes. Experience serving as an incident commander or incident coordinator is a plus.
  • Empathy, collaboration, and communication skills in English to cultivate strong relationships across various teams in the organization
  • Experience mentoring engineers, driving cross-functional initiatives, and influencing technical direction without relying on organizational authority.
  • We welcome candidates from a variety of backgrounds, including software engineering, site reliability engineering, production engineering, infrastructure, and other roles focused on building reliable systems or improving incident response.

Datadog values people from all walks of life. We understand not everyone will meet all the above qualifications on day one. That's okay. If you’re passionate about technology and want to grow your skills, we encourage you to apply.
Benefits and Growth:

  • New hire stock equity (RSUs) and employee stock purchase plan (ESPP)
  • Continuous professional development, product training, and career pathing
  • Intradepartmental mentor and buddy program for in-house networking
  • An inclusive company culture, ability to join our Community Guilds (Datadog employee resource groups)
  • Access to Inclusion Talks, our internal panel discussions
  • Free, global mental health benefits for employees and dependents age 6+
  • Competitive global benefits

Benefits and Growth listed above may vary based on the country of your employment and the nature of your employment with Datadog.
About Datadog:
Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure, data, models, and security into one place, using AI to detect and resolve issues before they impact customers. Trusted globally by Fortune 500 companies and high-growth AI leaders, Datadog enables businesses to move faster with clarity and confidence. Learn more about #DatadogLife on Instagram, LinkedIn, and Datadog Learning Center.

Equal Opportunity at Datadog:
Datadog is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and other characteristics protected by law. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. Here are our Candidate Legal Notices for your reference.

Datadog endeavors to make our Careers Page accessible to all users. If you would like to contact us regarding the accessibility of our website or need assistance completing the application process, please complete this form. This form is for accommodation requests only and cannot be used to inquire about the status of applications.

Privacy and AI Guidelines:
Any information you submit to Datadog as part of your application will be processed in accordance with Datadog’s Applicant and Candidate Privacy Notice. For information on our AI policy, please visit Interviewing at Datadog AI Guidelines.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Скоро на этой странице

Резюме под эту вакансию — и билет в розыгрыш

Мы разбираем объявление до настоящих требований и переписываем ваше резюме под него — вопросами, а не выдумкой: ни одна строка не появится без вашего подтверждения. Войдите, чтобы получить это первым, — и попасть в розыгрыш.

  • Резюме под конкретную вакансию, а не «универсальное»
  • Ответы хранятся: правится любой, а не весь разговор заново
  • Всё в аккаунте — открывается с любого устройства

Разыгрываем

Скидка на сопровождение

Победителей выбираем случайно среди заявок с подтверждённой почтой. Дата розыгрыша и полные правила — на странице розыгрыша.

Правила розыгрыша

Site Reliability Engineer

Site Reliability Engineer - SRE

Scaleway

Paris

Откликнуться на сайте работодателя

Описание вакансии

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
WHY WE NEED YOU?
Our growth is driving us to strengthen our SRE team to support and scale our production environments.

Your mission will be to build and maintain reliable, observable, and secure infrastructure in order to ensure optimal service availability for our customers around the world.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together.

You will be part of a team of experienced Site Reliability Engineers. The team is responsible for maintaining and evolving core infrastructure and observability tools, supporting product teams, and improving the reliability of Scaleway’s services.

YOUR DAILY ROUTINE

  • Build and optimize tooling to automate monitoring, diagnosis, and remediation of production incidents
  • Troubleshoot high-impact production issues in collaboration with other engineering teams
  • Participate in an on-call rotation to handle incidents and ensure service continuity
  • Implement and maintain observability solutions to monitor infrastructure and application health
  • Contribute to infrastructure lifecycle management across different environments
  • Promote and apply best practices in terms of stability, resiliency, scalability, and security
  • Maintain clear technical documentation for tools and procedures
  • Contribute to system and tool evolution based on production feedback
  • Collaborate closely with development teams to ensure infrastructure readiness
  • Participate in team rituals and knowledge-sharing initiatives

About You
SOFTSKILLS :

  • Proactive and solution-oriented mindset
  • Passion for automation and continuous improvement
  • Strong collaboration and communication skills
  • Ability to work independently and in a team
  • Willingness to mentor and share knowledge

💻 HARDSKILLS :

  • Experience with Go, Python or Rust
  • Strong scripting skills (Bash, Python)
  • Hands-on experience with Linux systems (Ubuntu/Debian)
  • Knowledge of networking (TCP/IP, DNS, BGP, load-balancing, IPv6, etc.)
  • Experience in cloud environments and infrastructure (bare metal, VMs, containers, orchestrators)
  • Familiarity with monitoring and logging tools (Prometheus, Grafana, Elastic, etc.)
  • Comfortable with Infrastructure-as-Code (Ansible, Salt, AWX, etc.)
  • Experience managing relational databases (PostgreSQL)
  • Understanding of CI/CD pipelines (GitLab)
  • Comfortable with English (written and spoken)

WHAT YOU WILL FIND AT SCALEWAY ++++

  • Hybrid work: We offer up to 3 days of remote work per week
  • Offices: Our offices are spacious, dynamic workspaces with bold design, conveniently located near public transport. Most of our offices feature outdoor spaces (terraces) and bike parking facilities
  • Dining: Our chef provides a healthy meal service at the headquarters, and breakfast is available across all our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches
  • Well-being commitments: Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life

International environment: With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.

  • Career & Mobility: Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers

🚀
Why join the Scaleway adventure ?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜 THE NEXT STEPS …

  • Discovery call with a recruiter (30 min)
  • Interview with the manager to understand your technical skills and approach to the role (45 min)
  • Technical interview to validate your expertise (1h)
  • Interview with the Head of the Tribe to deepen your discussions and assess your fit with the team (45 min)
  • HR interview to tour our offices and meet your future colleagues

Version française disponible ici

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Site Reliability Engineer (SRE) - Network Products

Scaleway

Paris

Откликнуться на сайте работодателя

Описание вакансии

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
WHY WE NEED YOU ?
Our growth is driving us to strengthen our
Network SRE Products
team to ensure the high reliability, performance, and scalability of our storage platforms.

Your mission will be to automate, monitor, and improve the reliability, performance, and scalability of our infrastructure. You will maximize availability, optimize fault tolerance, and reduce operational overhead—ensuring robust and efficient systems for our products and services.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together. You will be part of a team of Site Reliability Engineers reporting to a Lead SRE and integrated into the SRE Guild, a collective focused on fostering best practices across engineering.

The team collaborates daily with Dev, Product, and Ops teams to improve resiliency, support service scalability, and ensure a seamless customer experience across our network solutions.

YOUR DAILY ROUTINE

  • Develop automation tools and frameworks to streamline infrastructure management
  • Build and maintain CI/CD pipelines using Infrastructure as Code best practices
  • Implement and refine monitoring and alerting systems (OpenMetrics, OpenTelemetry)
  • Ensure system reliability through incident response and root cause analysis
  • Collaborate with developers and product teams to bake resilience into network systems
  • Participate in architecture reviews and provide SRE perspective early in the design
  • Apply principles of fault-tolerance, load balancing, and energy efficiency optimization
  • Share knowledge within the team and broader engineering org via the SRE Guild
  • Contribute to the reliability and performance of services

About You
Hardskills :

  • Strong experience with Infrastructure as Code (IaC) and CI/CD pipelines
  • Solid expertise in Linux systems and production troubleshooting
  • Proficiency with monitoring/logging tools (OpenMetrics, OpenTelemetry)
  • Programming skills in Python, Go, or Rust
  • Good understanding of network systems is a great bonus: BGP, BGP EVPN, VXLAN

Softskills:

  • Collaborative mindset and team-first approach
  • Curious, continuous learner with a drive for operational excellence
  • Clear and effective communicator (written & verbal)
  • Comfortable working in English and French and across multidisciplinary teams

WHAT YOU WILL FIND AT SCALEWAY ++++
Hybrid work:
We offer up to 3 days of remote work per week.

Offices:
Our offices are spacious, dynamic workspaces with bold design, conveniently located near public transport. Most of our offices feature outdoor spaces (terraces) and bike parking facilities.

Dining:
Our chef provides a healthy meal service at the headquarters, and breakfast is available across all our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches.

Well-being commitments:
Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life.

International environment:
With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.

Career & Mobility:
Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers.

🚀
WHY JOIN THE SCALEWAY ADVENTURE?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜
THE NEXT STEPS …
Initial call
with a recruiter to get to know each other (30 min)

Technical interview
with the Head of SRE to assess your skills and approach (1h)

Manager x team interview
to explore your background and team fit (1h)

Final interview
with HR and office visit to meet your future teammates and discover our workspaces

Version française disponible ici

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Software Engineer - Kubernetes Specialist

Scaleway

Paris

Откликнуться на сайте работодателя

Описание вакансии

OUR STORY:
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow !

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍 Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux and Lyon.
OUR STORY
🇪🇺 Join Scaleway and shape the sovereign cloud of tomorrow!

Since 1999, we have been designing secure, sustainable infrastructures aimed at supporting the most ambitious companies.

Historically known for our dedicated servers (Dedibox), we made a strategic shift to cloud computing in 2015. Staying true to our principles of simplicity, flexibility, and technical excellence, we have become one of the leading players in Europe in the sector.

With the rise of artificial intelligence, we have strengthened our commitment, supported by the Iliad Group, which is currently investing €3 billion to develop a serious, sovereign AI alternative to American and Asian giants.

Every day, thanks to our fast-growing portfolio of cloud and AI products (bare metal, containerization, serverless, AI, etc.), Scaleway proudly serves thousands of customer across the private and public sector, from corporations like France Télévisions or Hachette Livre, to fast-growing startups like Photoroom and Biolevate, to institutions like the City of Copenhagen.

📍Our offices are located in
Paris, Lille, Toulouse, Rennes, Rouen, Bordeaux, and Lyon.
WHY WE NEED YOU ?
Our growth is driving us to strengthen our
Non-Relational Database
team as a Software Engineer to replace a departing member and maintain our momentum on scaling our high-demand products.

Your mission will be to
ensure a robust, scalable, and high-performance experience for our customers across new sovereign regions
.

YOUR FUTURE TEAM
We work in a collaborative and international environment where the diversity of Scalers, combined with a spirit of sharing, helps bring new projects to life every day, advancing our ambitions together.

You will be part of a team of
5 people
, including a Manager, a Product Manager, and 3 Developers. This young team was born from the strategic split of the Database department to focus exclusively on non-relational databases. You will work on customer-facing products with modern stacks and no heavy legacy, participating in a strong dynamic of camaraderie and technical excellence.

YOUR DAILY ROUTINE

  • Contribute to the consolidation and development of major database products in Scaleway's product ecosystem : Redis and MongoDB.
  • Maintain the production to high standards of quality for performance, stability and resilience
  • Deploy products to new geographical regions as part of Scaleway's global expansion
  • Perform code reviews and provide constructive feedback to teammates
  • Contribute to internal documentation and the development of technical best practices
  • Participate in Agile rituals (two-week sprints) and team meetings
  • Collaborate with cross-functional teams across Engineering, Product, and Sales

About You
HARD SKILLS:

  • Strong expertise in Kubernetes administration and cluster management
  • Experience with deployment tools like Helm, ArgoCD, and Flux
  • Proficiency in GoLang development preferred, but otherwise demonstrated experience in a programming language
  • Familiarity with Terraform for Infrastructure as Code
  • Familiarity with Redis, MongoDB, or other NoSQL technologies

SOFT SKILLS:

  • Strong team player with a focus on cohesion and benevolence
  • Excellent communication skills and openness to feedback
  • High level of professional maturity and ability to mentor junior colleagues
  • Proactive mindset in problem-solving and documentation
  • Comfortable working in a hybrid environment with remote teammates

WHAT YOU WILL FIND AT SCALEWAY ++++

  • Hybrid work: We offer up to 3 days of remote work per week.
  • Offices: Our offices are spacious, dynamic workspaces with bold design. Most of our offices feature outdoor spaces (terraces) and bike parking facilities.
  • Dining: Our chef provides a healthy meal service at the headquarters, and breakfast is available across almost all of our sites year-round. Scalers working from regional sites enjoy a Swile card for lunches.
  • Well-being commitments: Whether it’s access to a gym, daycare places, or discounted services for caring services, Scaleway is committed to supporting Scalers in maintaining a balanced life.
  • International environment: With dozens of nationalities, Scaleway offers a stimulating environment where English is as widely spoken as French.
  • Career & Mobility: Our managers value internal mobility, and opportunities to transition to other entities within the Iliad Group are accessible to all Scalers

🚀
Why join the Scaleway adventure?
✔
A rich and diverse product offering:
Scaleway offers over 100 public cloud products in IaaS, PaaS, and AI.

✔
A cutting-edge technical environment:
Scaleway provides modern infrastructures, including high-performance bare metal servers, to tackle exciting technical challenges.

✔
Commitment to responsible cloud:
Scaleway is dedicated to a more responsible cloud, with data centers powered solely by renewable energy since 2017, minimizing our ecological footprint and holding top-level certification.

🔜 THE NEXT STEPS …

  • Discovery call with a recruiter to discuss your background and expectations (30 min)
  • Interview with the Manager to understand your approach to the role (45 min)
  • Technical interview focused on Go and Kubernetes to validate your expertise (1h)
  • Interview with the Head of Tribe, Loïc Martinez Varizat, to deepen discussions and assess your fit (45 min)
  • Final on-site visit to tour our offices and meet your future colleagues

At Scaleway, we are committed to building an inclusive and respectful workplace where everyone has a fair opportunity to thrive.
All applications are considered with care, regardless of age, gender, sexual orientation, ethnic or social background, religion, disability, or any other characteristic.
We believe great ideas come from everywhere, and everyone which is why you should definitely apply.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Ingénieur Windows et Virtualisation (H/F/X)

ORNESS

Boulogne-Billancourt

Откликнуться на сайте работодателя

Описание вакансии

  • Localisation : Boulogne-Billancourt
  • Type de contrat : CDI
  • Mode de travail : Hybride – 2 jours de télétravail par semaine
  • Rémunération : 58K - 68K

Contexte de l’organisation
Notre client est une organisation de grande taille opérant dans un environnement de production critique, avec de fortes exigences de disponibilité, de sécurité et de rigueur opérationnelle. Son infrastructure IT, fortement mutualisée, supporte des activités nécessitant une continuité de service permanente.

Dans un contexte de croissance et de transformation de son système d’information, l’organisation renforce son pôle
Windows Infrastructure Platform (WIP)
et recherche un·e
Ingénieur·e Infrastructure et Virtualisation expérimenté·e (N3)
(H/F/X).

L’environnement est exigeant, avec des infrastructures critiques, de fortes contraintes de disponibilité et de sécurité, et une dynamique continue de standardisation, d’industrialisation et d’automatisation.

Rôle
L’Ingénieur·e Infrastructure et Virtualisation (H/F/X) est responsable de l’exploitation, de l’évolution et du maintien en conditions opérationnelles des plateformes d’infrastructure.

Il/elle intervient en niveau expert (N3) sur des sujets complexes et contribue activement à la stabilité, l’évolution et la sécurisation des environnements.

Responsabilités

  • Exploitation & production (MCO)
  • Assurer la stabilité, la performance et la disponibilité des infrastructures
  • Prendre en charge les incidents complexes de niveau 3
  • Piloter les escalades avec les éditeurs et fournisseurs
  • Mettre en place des correctifs durables et industrialisés
  • Assurer la supervision et le monitoring des environnements critiques
  • Réaliser les analyses de cause racine (RCA)
  • Ingénierie & architecture
  • Concevoir et faire évoluer les architectures d’infrastructure et de virtualisation
  • Définir et optimiser les standards, référentiels et configurations
  • Mettre en œuvre les solutions d’infrastructure et d’exploitation
  • Définir des architectures cibles à partir des besoins métiers et assurer leur déploiement
  • Participer à la modernisation des plateformes d’infrastructure
  • Automatisation & industrialisation
  • Contribuer à l’automatisation des tâches d’exploitation (Terraform, Puppet, Chocolatey, PowerShell, API)
  • Industrialiser les processus de provisioning et de configuration
  • Améliorer la standardisation des environnements
  • Sécurité & conformité
  • Assurer le hardening et le patch management des systèmes
  • Gérer les vulnérabilités et la conformité des infrastructures
  • Garantir la traçabilité via logs et audits
  • Participer à la sécurisation des accès et authentifications (AD / Azure AD / Entra ID)
  • Respecter les normes internes et réglementaires
  • Support & amélioration continue
  • Intervenir dans un environnement structuré N1/N2/N3
  • Contribuer à la résolution des incidents et à la montée en compétence des équipes N1/N2
  • Capitaliser et documenter les solutions (procédures, bonnes pratiques)
  • Participer à l’amélioration continue de la qualité de service
  • Assurer une veille technologique active
  • Projets & transverse
  • Participer aux projets IT transverses (infrastructure, sécurité, transformation)
  • Collaborer avec les équipes infrastructure, sécurité, support et métiers
  • Contribuer aux projets de modernisation des plateformes

Environnements techniques

  • Expertise
  • VMware VCF (virtualisation)
  • Stockage SAN (PureStorage)
  • Stockage NAS (NetApp)
  • Automatisation : Terraform, Puppet, Chocolatey
  • Scripting : PowerShell, API
  • Bon niveau
  • Windows Server 2019 / 2022
  • Active Directory / Azure AD / Entra ID (hybride)
  • DNS et services Microsoft
  • Rubrik (sauvegarde)
  • Sécurité infrastructure (chiffrement, authentification, hardening)
  • Compléments
  • Linux

Profil recherché

  • Ingénieur ou équivalent avec expérience significative en infrastructure et virtualisation
  • Forte autonomie dans la gestion des environnements complexes
  • Excellent sens de l’analyse et de la résolution de problèmes
  • Capacité à documenter, transmettre et accompagner les équipes
  • Esprit d’initiative et force de proposition
  • Bon niveau d’anglais professionnel requis (minimum B2/C1)
  • Connaissance des environnements Citrix (appréciée)
  • Expérience en industrialisation et automatisation (appréciée)
  • Expérience en environnements critiques type banque, e-commerce ou production 24/7 (appréciée)

Qui sommes-nous ?
Orness, c’est une ESN à taille humaine (200 consultant·es), spécialisée en infrastructure, réseaux et ingénierie de systèmes.

Depuis plus de 20 ans, nous accompagnons les plus grands noms de la finance et de l’énergie sur des projets où fiabilité, performance et sécurité sont essentiels.

Processus de recrutement
Nous privilégions un processus simple et rapide :

  • Entretien RH approfondi (environ 1 heure)
  • Entretien technique (environ 1 heure)
  • Rencontre avec le client (30 minutes à 1 heure)

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Ещё 91 вакансия по этой категории в этой стране

Senior Software Engineer - Incident Insights & ReadinessDatadog · France

Откликнуться на сайте работодателя
Senior Software Engineer - Incident Insights & Readiness — Datadog | mentors.coach