Skip to content
SRE

SRE

Ingénieur Infrastructure IA Cloud & HPC (F/H)

Valeo

Créteil

Apply on the employer's site

Role description

Rejoignez les équipes de Valeo engagées pour réinventer la mobilité !

Valeo est un équipementier automobile, partenaire de tous les constructeurs dans le monde. Nous proposons des systèmes et équipements innovants au service de la voiture du futur – un véhicule intuitif et connecté, autonome, économe en énergie, accessible à tous et respectueux de l’environnement.

Nos équipes développent des solutions IA de pointe, en particulier dans la vision par ordinateur pour véhicule autonome. Pour soutenir nos projets ambitieux, nous recrutons un/une
Ingénieur Infrastructure Cloud, IA et GPC.
📍
Poste basé à Créteil (94)
Vos Missions
→ Gestion de l'infrastructure interne :

  • Administrer et maintenir nos clusters de calcul GPU (NVIDIA).
  • Gérer le stockage et les réseaux associés aux activités d'IA/ML.
  • Assurer la disponibilité et la performance des ressources locales.

→ Gestion De L'infrastructure Cloud (Google Cloud Platform)

  • Concevoir, déployer et gérer des architectures cloud sur Google Cloud Platform (GCP) pour les projets d'Intelligence Artificielle / Machine Learning.
  • Optimiser l'utilisation des services GCP pertinents (Compute Engine, Kubernetes Engine (GKE), Cloud Storage, Networking, etc.)

→ Environnement De Simulation HPC (High Performance Computing)

  • Gérer et optimiser les infrastructures liées au HPC telles que Rescale / CCRT / AtNorth HPCs.
  • Administrer et exploiter les plateformes HPC.

De plus, vous assurez une veille technologique de solutions d'infrastructures et d'outils MLOps ; vous développez et améliorez les outils d'orchestration sur Python, et vous travaillez en étroite collaboration avec les équipes Data Science et Machine Learning pour leur fournir des solutions d'infrastructure adaptées.

À Propos De Vous

  • Diplôme d'ingénieur ou Bac+5 en sciences informatiques ou équivalent.
  • Expérience confirmée (5 ans minimum) dans la gestion d'infrastructures informatiques, y compris une partie dédiée à l'IA/ML ou aux environnements de calcul à haute performance (HPC).
  • Compétences avancées en développement Python pour l'automatisation, les scripts et l'orchestration.
  • Proactivité et curiosité technique.
  • Niveau d'anglais professionnel.

Nos Avantages

  • Plan d’actionnariat, Action Logement
  • Présence d’un comité social & économique (CSE)
  • Télétravail partiel possible (2 jours par semaine)
  • Restaurant d’entreprise

Pourquoi Valeo ?

  • Pour rejoindre un leader technologique et industriel, pionnier français dans l’innovation automobile
  • Pour une carrière dynamique avec des possibilités de mobilité nationale ou internationale, adaptée à vos aspirations
  • Pour contribuer au développement d’une mobilité plus propre, plus sûre et plus intelligente

Valeo accorde une grande importance à la diversité, qu’elle soit culturelle, intergénérationnelle, de genre ou qu’elle concerne les personnes en situation de handicap.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

Senior Security Engineer, Exploits, Google Threat Intelligence Group

Google

Remote

Apply on the employer's site

Role description

Note: By applying to this position you will have an opportunity to share your preferred working location from the following:

In-office locations: Zürich, Switzerland.
Remote location(s): Switzerland.
Minimum qualifications:

  • Bachelor's degree in Computer Science, Cybersecurity, a related field, or equivalent practical experience.
  • 5 years of experience in threat intelligence, intrusion analysis, vulnerability researcher, or a similar security role.
  • Experience with threat intelligence platforms and tools (e.g., VirusTotal, SIEMs).

Preferred qualifications:

  • Knowledge of Android and Chrome security and internals.
  • Deep understanding of attacker Tactics, Techniques, and Procedures (TTPs).
  • Proven ability to lead complex threat research projects independently.
  • Strong analytical, problem-solving, and communication skills.
  • Skills in malware analysis, reverse engineering, or vulnerability analysis.
  • Proficiency in scripting or querying languages (e.g., Python, GoogleSQL).

About The Job
Our Security team works to create and maintain the safest operating environment for Google's users and developers. Security Engineers work with network equipment and actively monitor our systems for attacks and intrusions. In this role, you will also work with software engineers to proactively identify and fix security flaws and vulnerabilities.

Join the Google Threat Intelligence Group's (GTIG) Exploits Mission. The Exploits Mission focuses on protecting users from targeted exploitation, primarily from government-backed attackers and Commercial Surveillance Vendors (CSVs), through the detection, analysis, and ultimate prevention of vulnerabilities and exploits, with a special focus on 0-day attacks.

We provide timely, actionable intelligence and coordinate with internal and external partners to fix critical vulnerabilities and secure user devices.

As a Security Engineer on our team, you will conduct in-depth research on threat groups, their Tactics, Techniques, and Procedures (TTPs), and the malware they employ. You'll utilize Google's powerful internal intelligence platforms, Nirvana and mGraph, to model threat activity and generate actionable insights. This role involves close collaboration with various teams across GTIG and Google to develop and implement effective countermeasures, contributing directly to threat disruption and enhancing our collective security posture. We are looking for engineers passionate about threat research who can lead projects and mentor others.

Responsibilities

  • Lead complex technical analyses, modeling threat activity, TTPs, and Indicators of Compromise (IOCs) across internal platforms (mGraph, Nirvana).
  • Create and deploy detection signatures (autoqueries, Watchtower rules) to maintain visibility over threat actors and assist in closing security gaps.
  • Produce polished, high-quality technical intelligence documentation and actor profiles to deliver actionable insights to internal and external stakeholders.
  • Influence technical direction within your scope, mentor junior engineers, and collaborate with cross-functional Google teams to support threat disruption efforts. Collaborate with security engineers and product teams in designing innovative exploit mitigations.
  • Identify and execute opportunities for continuous improvement: analytic collection, process optimization, and automation.

Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form .

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

Senior Database Reliability Engineer (DBRE) (worldwide remote)

CloudLinux

Remotefulltimesenior

Apply on the employer's site

Role description

CloudLinux / TuxCare is a remote-first infrastructure and security company. More than 300 engineers build and operate products used by hosting providers, enterprises, and internal service teams worldwide. Our Infrastructure Department runs the platforms behind CloudLinux OS, Imunify, KernelCare, TuxCare ELS, and our engineering systems.

We are hiring a
Senior Database Reliability Engineer
to join the Infrastructure DBA cell. This is a hands-on production ownership role, not a narrow ticket-processing DBA position. You will keep critical database services reliable, automate repeated work, support engineering teams, and reduce single-person dependency in our PostgreSQL, ClickHouse, MongoDB, and Redis operations.

PostgreSQL is the main requirement. ClickHouse experience is a strong plus, but it is not a day-one blocker. We need a senior engineer with enough database, Linux, automation, and incident-response depth to learn our ClickHouse environment quickly and operate it safely.

Your Responsibilities:

  • Own production PostgreSQL reliability: HA design, Patroni, PgBouncer, replication, failover, upgrades, vacuum/bloat control, query tuning, locks, indexes, capacity, backups, PITR, and restore validation
  • Improve disaster recovery and operational evidence: tested restores, documented recovery paths, measurable RTO/RPO targets, runbooks, and safe maintenance plans
  • Support the wider database estate: ClickHouse, MongoDB, and Redis. You will troubleshoot incidents, review access and data-safety changes, improve monitoring, and learn the production ClickHouse patterns already in use
  • Automate DBA workflows with Ansible, Terraform/OpenTofu, GitLab CI/CD, scripts, and reproducible runbooks for provisioning, grants, backups, restores, health checks, and ownership metadata
  • Help build DBaaS-style self-service capabilities so engineering teams can request databases, access, credentials, and operational checks with less manual DBA intervention
  • Improve observability and incident response through Grafana, metrics, logs, SLOs, alert rules, Opsgenie routing, and clear communication during production issues

What Success Looks Like:

  • PostgreSQL clusters have tested backup and restore paths, useful dashboards, clear ownership, and documented failover procedures
  • Repeated DBA tickets become automation or self-service workflows
  • ClickHouse operational knowledge is no longer a single-person dependency
  • Database incidents have owners, runbooks, evidence, and measurable recovery paths
  • Product and engineering teams get database help faster without sacrificing safety, auditability, or reliability

Why CloudLinux?

  • You will work on real production infrastructure used across CloudLinux and TuxCare products.
  • You will have a direct impact on reliability, incident response, developer experience, and operational resilience.
  • You will also work in an AI-assisted engineering culture where automation, documentation, Claude, Codex, and careful human verification are part of the daily operating model

Requirements
What We Expect From You:

  • Deep hands-on PostgreSQL experience in business-critical production environments, typically 5+ years or equivalent depth
  • Strong understanding of PostgreSQL internals and operations: MVCC, WAL, transactions, locks, indexes, query planning, replication, autovacuum, bloat, major upgrades, backups, PITR, and restore testing
  • Proven experience with highly available databases and the ability to reason about quorum, split-brain risk, failover, rollback, and recovery
  • Strong Linux and infrastructure fundamentals: systemd, networking, storage, filesystems, CPU/memory/disk bottlenecks, TLS, DNS, firewalls, and root-cause troubleshooting
  • Automation skills with Ansible and scripting. Terraform/OpenTofu, GitLab CI/CD, and merge-request based delivery are strong advantages
  • Ability to support more than one database engine. You do not need to be a ClickHouse expert on day one, but you must be ready to learn it quickly and take responsibility for it
  • Practical use of AI engineering assistants such as Claude and Codex. We expect you to use them to improve speed and quality, while personally verifying generated SQL, commands, scripts, and operational conclusions
  • English - upper-intermediate or higher - to ensure clear communication of progress within the teams

Nice to Have:

  • ClickHouse operations: replication, Keeper/ZooKeeper, MergeTree engines, distributed DDL, grants, row policies, backups, query troubleshooting, and cluster recovery
  • MongoDB replica sets and Percona Backup for MongoDB
  • Redis/Sentinel and broker/cache failure modes
  • Database observability, SLOs, golden signals, alert tuning, and executable incident runbooks
  • Building internal platforms, self-service portals, or DBaaS workflows for engineering teams

Benefits
What's in it for you?

  • A focus on professional development
  • Interesting and challenging projects
  • Fully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide
  • Paid 24 days of vacation per year, 10 days of national holidays, and unlimited sick leaves
  • Compensation for private medical insurance
  • Co-working and gym/sports reimbursement
  • Budget for education
  • The opportunity to receive a reward for the most innovative idea that the company can patent

By applying for this position, you agree with
CloudLinux Privacy Policy
and give us your consent to maintain and process your personal data with this respect. Please read our Privacy Policy for more information.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

Enterprise Services Engineer

Tanium

Remotefulltimesenior

Apply on the employer's site

Role description

The Basics
At Tanium, our Enterprise Services Engineers (ESEs) fulfill a vital role in our organization by serving Tanium's customers as hands-on-keyboard experts for both remote and on-site engagements. You will operate in a highly team-based environment with Technical Support Engineers, fellow Engineers, Customer Success Managers, and Sales Account Managers. ESEs contribute to each customer's success by operating and maintaining the Tanium platform and its modules to achieve customer-specific Operations, Risk, Compliance, Asset, and Security-focused outcomes, either remotely or on on-site with the customer. You’ll have continuous opportunities and challenges which will require you to apply your best technical chops in large enterprise environments, all leveraging the power of Tanium.

What You’ll Do

  • Work closely with our customers to;
  • Operationalize, administer and maintain the Tanium Platform to solve complex technical issues independently or with the help of teammates
  • Identify opportunities for our customers to get greater value from the Tanium platform
  • Consistently and cogently address our customers’ needs through astute verbal and written communication skills
  • Conduct daily health-checks on assigned accounts & Work with Technical Support Engineers on strategic customer activities
  • Contribute to and track activity, after action, root cause and daily status reports
  • Document best practices and Play Book entries
  • Work closely with CSMs on improving Tanium operational status within key accounts
  • Provide technical direction to customer IT support staff

We’re Looking For
Security Clearance: minimum Secret but Top Secret or DV preferred

Education:
BA/BS or equivalent experience required

Experience

  • 5+ years of experience in IT/Consulting
  • 5+ years of experience in Systems Administration/Engineering
  • Experience working with public sector customers
  • Ability to speak German and English
  • Strong troubleshooting skills
  • Ability to articulate and communicate
  • Broad knowledge across several technical domains, coupled with deep knowledge in one or more of the following:
  • Endpoint Security
  • Endpoint Support/Troubleshooting
  • Incident response
  • Systems Management
  • Systems Administration
  • Software Engineering
  • Utility Scripting (e.g. Bash, PowerShell, VBScript, Python, etc.)
  • Demonstrated critical thinking skills
  • Ability to break a problem down into manageable, ordered piece parts
  • Ability to convey problem statement and plan of attack to others
  • Hands-on Tanium experience will be a major plus

About Tanium
Tanium is the Autonomous IT company. Driven by AI and real-time endpoint intelligence, Tanium Autonomous IT empowers IT and security teams to make their organizations unstoppable.

Many of the world’s leading organizations trust Tanium’s single, unified platform for endpoint management and security to innovate faster, stay resilient and move business forward with confidence, at scale. To learn how Tanium delivers Autonomous IT for unstoppable business – visit www.tanium.com and follow us on LinkedIn and X.

On a mission. Together.
At Tanium, we are stewards of a culture that emphasizes the importance of collaboration, respect, and diversity. In our pursuit of revolutionizing the way some of the largest enterprises and governments in the world solve their most difficult IT challenges, we are strengthened by our unique perspectives and by our collective actions.

As a global organization with stakeholders around the world, it’s imperative that the diversity of our customers and communities is reflected internally in our team members. We strive to create a diverse and inclusive environment where everyone feels they have opportunities to succeed and grow because we know that only together can we do great things.

Our commitment to excellence and innovation has earned us a place on the Forbes Cloud 100 list for ten consecutive years, and we continue to be recognized worldwide as a great place to work.

Taking care of our team members
Each of our team members has 5 days set aside as volunteer time off (VTO) to contribute to the communities they live in and give back to the causes they care about most.

Tanium is an Equal Opportunity and Affirmative Action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, gender identity, sexual orientation, disability, protected Veteran status, or other legally protected categories. If you require a reasonable accommodation in searching for a job opening, completing an application, interviewing, or completing any pre-employment testing or requirements, please contact accommodations@tanium.com. For more information refer to the “Know Your Rights” poster which is available here - https://www.eeoc.gov/poster.

Please be aware of job offers coming from people claiming to be Tanium employees. Tanium employees will only use @tanium.com email addresses to communicate with you, will have video interviews with you, and will never ask you for money.

This link leads to the machine readable files that are made available in response to the federal Transparency in Coverage Rule and includes negotiated service rates and out-of-network allowed amounts between health plans and healthcare providers. The machine-readable files are formatted to allow researchers, regulators, and application developers to more easily access and analyze data.

For more information on how Tanium processes your personal data, please see our Privacy Policy.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

That is every opening in this category and country

Ingénieur Infrastructure IA Cloud & HPC (F/H)Valeo

Apply on the employer's site