Skip to content
SRE

The week's list

Every role like this one, in one letter

You are reading one posting. There are hundreds like it on the board, and new ones every week. Pick what you want, leave an email, and the list comes to you — no searching, no coming back here.

Counting what came out this past week…

The first letter arrives right away, then one a week. Unsubscribe in one click from any letter — the address goes nowhere else.

SRE

Senior System Engineer Cloud Storage (m/f/d)

Schwarz Digits

Bucharestsenior

Apply on the employer's site

Role description

Schwarz Digits creates the technological foundation for digital sovereignty in Europe. As the IT and digital division of the Schwarz Group, we develop and manage the IT infrastructures for the retail divisions Lidl and Kaufland, as well as Schwarz Production and PreZero. At the same time, we operate as an independent provider in the external market to support companies across Europe in their digital transformation. We bundle our core services in the areas of Cloud, Cyber Security, Data & AI, Communication, and Workspace.

Senior System Engineer Cloud Storage Project metadata
The Impact You Will Create

  • Architecture: Together with your team, you are responsible for designing and maintaining a robust, efficient, and forward-looking storage architecture.
  • Stability & Reliability: You are responsible for maintaining and optimizing the stability and availability of our highly resilient storage infrastructure (Block, Object, Backup, and File Storage). You ensure this through proactive monitoring, independently resolving incidents, and implementing measures to prevent future occurrences.
  • Incident & Post-Mortem Analysis: You will take charge of processing major incidents involving storage as part of our incident and problem management process. Your goal is to derive and successfully implement mitigating measures for the future.
  • Automation: You automate provisioning and operating processes in the storage environment with a continuous drive to improve our products and get a little better every day.
  • Performance & Capacity Planning: You will analyze and optimize the performance of existing systems to support the future scaling of our landscape, including proactive, forward-looking capacity planning.

Experience And Skills You Will Need

  • Storage Expertise: Proven experience in the architecture, operations, and troubleshooting of storage systems. You must have deep hands-on experience with NetApp ONTAP or NetApp StorageGRID,
  • Protocols & Architectures: Strong overall understanding of storage architectures and deep experience with associated protocols, specifically NFSv3, NFSv4.x, NVMe-oF, and S3.
  • Operating Systems: Solid, practical experience working with Linux operating systems.
  • Automation: Strong experience in automating configurations and operational tasks.
  • API Management: Experience in working with REST APIs.
  • Curiosity & Interface Topics: A strong motivation to dig into and learn new technologies. You have a keen interest in topics surrounding the storage ecosystem, such as virtualization, backup, monitoring, and logging (e.g., Prometheus, Grafana, Elasticsearch).
  • Mindset & Language: You are a motivated team player who enjoys the challenges of operating complex storage systems (lifecycle, high availability, performance analysis).
  • Excellent communication skills in English are required for our international, agile teams (German is an optional bonus).

Nice-to-Have Skills

  • Containerization: Experience using Kubernetes (K8s) and working within containerized system landscapes in a storage context.
  • You are proficient with Infrastructure as Code (IaC) frameworks and tools such as Ansible, Terraform, and OpenTofu, as well as scripting/programming (e.g., Golang, Python, Bash).

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

Production Administrator - L1 Application Support

SCOR

Bucharest

Apply on the employer's site

Role description

* Production Administrator supports the company's strategic and operating objectives by maintaining operational excellence of SCOR Products and Information System components.

* Production Administrators contributes to SCOR Information System through monthly patching

As a leading global reinsurer, SCOR offers its clients a diversified and innovative range of reinsurance and insurance solutions and services to control and manage risk. Applying "The Art & Science of Risk," SCOR uses its industry-recognized expertise and cutting-edge financial solutions to serve its clients and contribute to the welfare and resilience of society in around 160 countries worldwide.

Working at SCOR means engaging with some of the best minds in the industry - actuaries, data scientists, underwriters, risk modelers, engineers, and many others - as we work together to find solutions to pressing challenges facing societies.

As an international company, our common culture is defined by "The SCOR Way." Serving both to build momentum that drives the Group forward and as a compass to guide our actions and choices, The SCOR Way is anchored by five core values, reflecting the input of employees at all levels of the Group. We care about clients, people, and societies. We perform with integrity. We act with courage. We encourage open minds. And we thrive through collaboration.

SCOR supports inclusion and the diversity of talents, and all positions are open to people with disabilities.

At our brand-new Scor Bucharest, we offer a dynamic environment where career growth is actively supported through internal mobility, globally recognized certifications, and continuous professional development. We value work-life balance, offering flexible work arrangements, and wellbeing initiatives that help you thrive both personally and professionally.

Now, let's explore this exciting opportunity so that you can be part of our mission.

Under the responsibility of Manager IT Operations, Production Administrator supports the company's strategic and operating objectives through maintaining high level of excellence of Products delivered to End Users. Within SCOR IT Technical / Operations Chapter, your key function will consist in:

- Monitoring SCOR group's information systems to maintain it in optimal operating condition

- Maintaining daily processes such as backups while keeping an eye on issues impacting Disaster Recovery capacity

- Managing monthly patching

- Contributing to continuous improvement on monitoring, scheduling or other tools and Operations Chapter documentation

- Sharing and training peers on your domain of expertise

Your role requires you to be detail oriented, thorough, open-minded, information-sharing and problem-solving posture, while working in a demanding environment with various stakeholders, in Technical department but also in close collaboration with Squads and Solution Owners.

Candidate Success Factors

- Accountability, reliability and thoroughness are key values required from an Operation Administrator

- Curiosity and willingness to learn from peers and past experiences

- Continuous improvement and service driven while keeping it simple

Objectives for the first year (or 2 years)

- Deliver Product(s) on time and in budget, in collaboration with Solution Owner

- Maintain Product(s) high level of availability and reliability

Abide to and enforce SCOR Operation and Security practices

Required experience & competencies

  • 2+ years of experience maintaining IT Production systems
  • ITSM tools such as ServiceNow
  • Backup, patching and disaster recovery procedures and tools
  • Hands on experience with Monitoring, Patching, Backup, Process Automation, Disaster Recovery,… is mandatory
  • Observability experience such as Elastic is desirable
  • Fluent in English

Hard skills

  • Experience in technical IT support.
  • Knowledge of monitoring tools (Nagios XI) and scheduling tools (Absyss Visual TOM).
  • Familiarity with ticket management systems (ServiceNow).
  • Availability to work in 24/7 rotating shifts.

Required Education

<li data-aria-level="1" data-aria-posinset="1" data-leveltext="·" data-font="Symbol" data-list-defn-props="{"335552541":1,"335559685":720,"335559991":360,"469769226":"Symbol","469769242":[8226],"469777803":"left","469777804":"·","469777815":"hybridMultilevel"}" data-listid="1">Degree in Computer Science, Management of Information Systems, or related analytical field; or equivalent experience

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

AI Specialist

BMC Helix

Bucharest

Apply on the employer's site

Role description

Looking for details about our benefits? You can learn more about them by clicking
HERE
This Is Helix. Powered by You.
At
BMC Helix
, we don’t do ordinary. We’re the
AI-native engine
behind the world’s most forward-thinking IT organizations, helping them focus on what matters most. What are we passionate about? We're here to reset the economics of enterprise IT and help others realize the ROI of AI.

We are a mix of
curious minds, creative thinkers, and courageous builders
who believe tech should change the game—not just play it. We celebrate wins, support each other, and laugh a lot.

We are the change makers. With decades of leadership and established trust in IT service and operations management, we’re scaling with purpose—through organic innovation, strategic acquisitions, and relentless R&D. Our open-first
Agentic AI platform
empowers autonomous agents to drive real outcomes with speed, accountability, and precision.

We are laser-focused on delivering real value to our customers by accelerating innovation and the application of applying agentic AI in digital service and operations management for IT organizations around the world.

Role Overview
BMC Helix is seeking an AI Specialist who brings both hands-on technical depth and the strategic discipline to deploy, govern, and continuously measure AI systems in live production environments. This is not a build-only role. The successful candidate will be the prime Customer Success organisation’s authority on responsible AI deployment—designing governance frameworks, defining performance baselines before go-live, monitoring model behaviour at scale, and embedding best practice across every team that touches AI.

Key Responsibilities
AI Governance & Risk Management

  • Design and own the AI governance framework covering model risk, bias detection, fairness criteria, and explainability requirements aligned to applicable regulations (EU AI Act, ISO/IEC 42001, SOC 2 AI controls).
  • Establish and maintain an AI model registry tracking all models in production—lineage, version, training data provenance, risk classification, and owner accountability.
  • Define acceptable use policies and escalation paths for high-risk AI decisions, in partnership with Legal, Compliance, and Security.
  • Lead AI ethics reviews for new model deployments; document decisions and maintain an audit trail sufficient for regulatory inspection.
  • Conduct periodic governance health checks and produce a governance scorecard for senior leadership on a quarterly cadence.

Production Deployment & Operations

  • Architect and implement production-grade AI deployment pipelines in alignment with IT and Operations — CI/CD for models, canary releases, shadow mode testing, and staged rollouts with defined success gates.
  • Define pre-production readiness criteria: latency SLAs, throughput requirements, fallback behaviour, and failure mode documentation for every model before it enters production.
  • Own the model versioning and rollback strategy; ensure any model can be reverted to a known-good state within a defined RTO window.
  • Partner with IT, Platform, DevOps, and Security teams to harden AI workloads—container security, secrets management, data-in-transit encryption, and adversarial input handling.
  • Manage inference infrastructure optimisation: batch vs. real-time trade-offs, cost-per-inference tracking, and resource right-sizing.

Performance Measurement & Monitoring

  • Build and maintain a production AI observability stack: model drift detection, data quality monitoring, prediction confidence tracking, and business-outcome correlation.
  • Define the metrics hierarchy for each deployed model—distinguishing technical metrics (accuracy, F1, latency, p99) from business outcome metrics (deflection rate, resolution time, cost per ticket) and lagging indicators in alignment with the team.
  • Set performance baselines at deployment and alert thresholds for degradation; own the on-call escalation path when models breach SLAs.
  • Produce a monthly AI Performance Report covering all production models: drift signals, retraining triggers, cost trend, and business impact vs. baseline.
  • Drive model retraining and fine-tuning cycles informed by production data; define the feedback loop from human-in-the-loop review back into training pipelines.

Best Practice & Centre of Excellence

  • Author and maintain the organisation’s AI Engineering Standards—the canonical reference for how AI is built, tested, deployed, governed, and retired at BMC Helix.
  • Run a regular AI Practice community of practice (bi-weekly); present case studies, post-mortems, and emerging patterns.
  • Evaluate and recommend tooling across the MLOps lifecycle: experiment tracking, feature stores, model serving, monitoring platforms, and vector databases.
  • Mentor engineers across teams on production AI patterns, responsible AI principles, and governance obligations.
  • Represent BMC Helix externally in AI governance forums, standards bodies, or industry working groups as appropriate.

Requirements
Essential

  • 5+ years in AI/ML engineering with at least 3 years focused on production systems (not research or prototyping)
  • Demonstrable experience designing and implementing an AI governance or model risk framework in a regulated or enterprise environment
  • Hands-on proficiency with MLOps tooling: MLflow, Kubeflow, Seldon, BentoML, or equivalent
  • Production monitoring experience: model drift detection, data quality pipelines, alerting (Prometheus/Grafana or equivalent)
  • Strong command of Python and familiarity with LLMOps patterns (prompt versioning, retrieval-augmented generation, evaluation harnesses)
  • Track record of defining and measuring AI business impact metrics—not just technical accuracy scores
  • Experience with cloud-native deployment on AWS, Azure, or GCP including containerised model serving

Highly Desirable

  • Familiarity with EU AI Act risk tiers, ISO/IEC 42001, or NIST AI RMF
  • Experience with agentic AI systems, multi-model orchestration, or AI safety evaluation
  • Background in ITSM, ServiceOps, or enterprise IT operations domains
  • Contributions to open-source AI tooling or published writing on AI governance / production ML
  • Experience in a scale-up or product company where AI is a core revenue driver, not a side initiative
  • Relevant certifications: AWS ML Specialty, Google Professional ML Engineer, or equivalent

Why Work Here? Because You’ll Matter.
We’re not hiring for roles—we’re hiring for
impact
. At Helix, you’ll solve hard problems, build smart solutions, and work with people who challenge and champion you. You’ll see your ideas come to life—and your work make a difference.

We believe in
trust, transparency, and grit
. Our culture is inclusive, flexible, and built for people who want to stretch themselves - and support others doing the same. Whether you’re remote or in-office, you’ll find space to show up fully and contribute meaningfully. You won’t be boxed in—you’ll be backed up.

Make Your Mark At Helix
If Helix excites you but you're unsure if you meet every qualification,
apply anyway
. We value diverse perspectives and believe the best ideas come from everywhere.

EEOC Statement
Helix is committed to equal opportunity employment regardless of race, age, sex, creed, color, religion, citizenship status, sexual orientation, gender, gender expression, gender identity, national origin, disability, marital status, pregnancy, disabled veteran or status asa protected veteran. If you need a reasonable accommodation for any part of the application and hiring process, visit the accommodation request page.

BMC Helix maintains a strict policy of not requesting any form of payment in exchange for employment opportunities, upholding a fair and ethical hiring process.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

Sr. Expert Cloud & Compute Design

OMV Petrom

Bucharestsenior

Apply on the employer's site

Role description

Your tasks

  • Responsible for overall architectural coordination for the engineering and evolution of Cloud and Compute IT Infrastructure services, ensuring consistency, scalability, and alignment with enterprise technology standards and strategy.
  • Acts as a senior expert in complex business and Digital Infrastructure initiatives, delivering advanced technical consultancy and architectural guidance, and shaping solution implementation from a senior Cloud & Compute Design and Architecture perspective.
  • Coordination of the implementation and customization by our strategic partners of Cloud & Compute technical solutions and processes, ensuring proper technical implementations along with services delivery based on best practices.
  • Continuously assesses emerging technologies, industry developments, and innovation trends, evaluating their relevance and impact on strategic IT developments and future business capabilities.
  • Advanced technical contribution to tender and vendor evaluations, assessing architectural quality, technical feasibility, and alignment of vendor solutions with Cloud & Compute standards and enterprise requirements.
  • Advanced technical support role in tenders evaluation, including technical challenges of the vendors
  • Acts as a senior technical advisor across the IT organization, supporting project managers, service managers, and business consultants on complex Cloud & Compute topics, and reporting on KPIs, risks, and root‑cause analyses related to critical incidents or systemic issues.
  • Evaluates vendor strategies, market trends, and technology developments, determining their enterprise value and strategic fit, while functionally coordinating external suppliers and contributing to detailed requirements, specifications, and Cloud & Compute implementation proposals.
  • Ensures continuous improvement of Cloud & Compute services with a focus on innovation, quality, and sustainability, maintaining high‑quality technical documentation and ensuring controlled enhancement of existing solutions in line with business requirements and data integrity standards.

Your profile

  • Master degree in IT&C
  • Relevant professional experience: > 9 years
  • Has advanced understanding of the business context, all interfaces and interdependencies and manages complex workflows with other disciplines. Has knowledge of the capabilities and limitations of systems, peripheral devices, and the most sophisticated information technologies.
  • ITIL Service Management certification
  • Excellent knowledge of business‑specific IT&C areas, including Agile and Project Management methodologies; design and governance of hybrid and multi‑cloud infrastructure; Storage, File, Backup, Hyper‑Converged and Software‑Defined Storage technologies; Cloud storage and backup services (e.g. Azure Files, Blob); security protections (antivirus, ransomware); archiving; and virtualization technologies.
  • Excellent experience in automation of Cloud & Compute service provisioning, including Infrastructure‑as‑Code, standardized automation frameworks, and repeatable design patterns.
  • Excellent knowledge of DevOps practices and automation, with focus on Microsoft Azure and AWS DevOps toolchains (e.g. Azure DevOps, AWS DevOps, Ansible).
  • English – advanced level, written and verbal.
  • Relevant professional certifications (e.g. ITIL, Cloud platform certifications, infrastructure or security certifications) are considered an advantage.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

SRE

DATA Ops

Inetum

Bucharest

Apply on the employer's site

Role description

Company Description
Inetum is a European leader in digital services. For businesses, public sector organizations, and society as a whole, the group’s 26,000 consultants and specialists work every day to create tangible digital impact: solutions that contribute to performance, innovation, and the common good.

With a presence in 19 countries, working closely with local communities, and alongside its major software developer partners, Inetum supports organizations in their digital transformation challenges with proximity, flexibility, and responsibility. Driven by its purpose, Inetum champions a vision of technology that is useful and well-managed, capable of unlocking the full potential of organizations and society: “Let’s make tech right.”

In 2025, the group generated revenue of 2.2 billion euros.

More information at: www.inetum.com

Job Description
Are you a tech-savvy Data Ops, with a passion for managing and optimizing databases? We're looking for a new colleague who is proactive, self-motivated, and has a talent for identifying opportunities for improvement.

As a Data Ops You Will

  • Administer and monitor the IT service solutions in the area of responsibility, supervise and continuously track their compliance with the agreed performance parameters in the SLA
  • Ensure the use and compliance with legal regulations and internal rules regarding IT equipment
  • Proactively initiate measures to optimize functional parameters or corrections when necessary, propose corresponding activities within the Service Improvement Plan (SIP)
  • Periodically evaluate the IT system in the area of responsibility, in terms of technical requirements, audit and security, work norms and procedures, availability, and continuity
  • Develop, test, implement, and continuously update operational work procedures for monitoring and service failures at the enterprise level
  • Define alerts for system monitoring and proactive intervention in the event of defined alert thresholds being exceeded
  • Analyze incidents and problems that occur in the IT Data Flows to identify the causes of errors and solutions
  • Resolve requests and incidents assigned within the terms provided in the SLA, ensuring the minimization of impact and service downtime
  • Apply internal procedures established with process owners in the event of incidents with high impact and potential financial losses
  • Collaborate with other IT Data Squads to maintain inter-cooperation to resolve incidents and improve service levels
  • Participate in testing business continuity and ensure the functioning of IT systems in the conditions provided in the BIA (Business Impact Analysis) in the event of planned or unplanned switchovers of IT systems
  • Provide expertise, recommend technical solutions, participate in testing (technical, failure, availability, performance, volume) in IT projects for IT systems and platforms in the area of responsibility
  • Install and configure platforms and solutions in the area of responsibility, promote new objects on test, pre-production, or production environments (according to projects and change requests)
  • Maintain the knowledge database (Knowledge Base)

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or a related field
  • At least 2 years experience in a similar role (Banking or Telecommunications fields represent an advantage)
  • Understanding and experience in ETL methodologies and Data Warehousing principles
  • Solid knowledge of PL/SQL code and scripting
  • Knowledge of working with databases Knowledge of IT security
  • English language proficiency
  • Excellent troubleshooting and problem-solving skills
  • Strong communication and interpersonal skills
  • Ability to work independently and as part of a team
  • Ability to prioritize and manage multiple tasks simultaneously

Nice To Have

  • ODI (Oracle Data Integrator) or other ETL tool
  • Understanding Data Mart environment/ concepts, Windows, UNIX and Linux Systems

Benefits
Additional Information

  • Full access to foreign language learning platform
  • Personalized access to tech learning platforms
  • Tailored workshops and trainings to sustain your growth
  • Medical insurance
  • Meal tickets
  • Monthly budget to allocate on flexible benefit platform
  • Access to 7 Card services
  • Wellbeing activities and gatherings

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

20 more openings in this category and country

Senior System Engineer Cloud Storage (m/f/d)Schwarz Digits · Romania

Apply on the employer's site
Senior System Engineer Cloud Storage (m/f/d) — Schwarz Digits | mentors.coach