Skip to content
DevOps Engineer

DevOps Engineer

Senior Cloud Engineer (d/f/m)

flexa

Munichsenior

Apply on the employer's site

Role description

Short Facts
flexa runs a virtual power plant. We control the home batteries, heat pumps and EVs across the Enpal customer fleet and trade their flexibility on European energy markets in real time. Every dispatch decision is money - for the customer and for us - and it has to land within seconds, on hardware we don't own, under regulation that is paying attention. You would own the AWS platform underneath that.

  • Location: Munich, Germany, Office-first work setup
  • Employment Type: Full-Time, indefinite term
  • Salary Range: € 95.000 - 115.000 per year gross depending on the seniority level
  • Language Requirement: C1 Level English

Your Responsibilities

  • Own the AWS platform. Architecture and evolution of the AWS-native platform that ingests asset telemetry, runs our dispatch and optimisation workloads, and serves the customer-facing applications. You set the direction, and you build.
  • Make shipping boring. Infrastructure-as-Code, CI/CD, environments and guardrails good enough that engineers deploy without asking you first.
  • Reliability where it counts. SLOs, observability and incident practice for systems where minutes of downtime are lost revenue and missed market commitments.
  • Security and compliance as design constraints. We touch critical energy infrastructure. IAM, secrets, network boundaries and audit trails get designed in, not retrofitted before an audit.
  • Raise the bar around you. Standards, reviews and mentoring across the tech team — at the Staff end of this role, that's a large part of the job.

Your Profile
What we need

  • 5+ years building and operating production infrastructure, at least 4 of them deep in AWS - and you've been on call for what you built.
  • Real Infrastructure-as-Code ownership (Terraform, Pulumi or CDK), including the unglamorous parts: state, module design, drift, provider upgrades.
  • Strong judgement about AWS services - Lambda, ECS/EKS, API Gateway, DynamoDB, IAM - including when the right answer is a plain EC2 instance or RDS rather than something clever.
  • You've run AWS at the account level, not just the resource level - multi-account setups with Control Tower or Organizations, SCPs, centralised logging, and the boring discipline of keeping new accounts consistent with the ones that came before.
  • Security is part of how you build, not a review gate at the end: least-privilege IAM you actually enforce, findings from Security Hub / GuardDuty / GitHub Advanced Security triaged rather than accumulated, secrets and supply chain handled properly. Bonus if you've worked under a formal regime (ISO 27001, KRITIS or similar).
  • You write code, not just config. Python or Go to a standard where other engineers want to use your tooling.
  • Containers and CI/CD in production, including rollbacks and how they fail.
  • You've run monitoring and alerting that people actually trust (DataDog, Grafana, Prometheus or equivalent) and can tell a useful alert from noise.

What would set you apart

  • Energy, grid, trading, or any domain where being late costs money.
  • Time-series data at scale.
  • Fleets of devices you don't physically control - IoT, edge, intermittent connectivity.
  • Streaming and data platforms (Kafka, Kinesis, Flink).
  • Working in an audited or regulated environment.

This role may not be for you if

  • You want to join a finished platform. Parts of this are greenfield, parts need replacing, and you'll have real influence over which is which - but nothing is handed to you tidy.
  • On-call is not a formality here. These systems move money, and you'd share the rotation.
  • Compliance is genuine engineering work in this role, not a checkbox someone else ticks.
  • The team is small enough that you own your own operations, which some weeks means unglamorous work.
  • You need remote or hybrid-heavy. This is an office-first role in Munich.

A
t flexa,
we are committed to diversity - of backgrounds and experiences. You don’t need 100% of the preferred qualifications to add incredible value to our team. If you’re passionate about what you could accomplish here, we’d love to hear from you.
Your Benefits

  • Competitive Compensation Package: Including salary, benefits, and potential for growth.
  • Professional Development: Annual development budget of €3,000 for coaching, training, books, etc.
  • Health & Sport Subsidy: Company-subsidized sports facility memberships.
  • Public Transportation Subsidy: Monthly subsidy for your public transport ticket.
  • Lunch/Dinner Allowance Vouchers: Digital meal vouchers for workdays.
  • Work Equipment: Your choice of MacBook or Windows Laptop and ergonomic workplace setup.
  • Regular Team Events: Knowledge sessions, afterwork hangouts, sports events, and company offsites.

A Short Note from Your Future Lead
Willi Richert, VP of Technology — flexa

Hi there!

I'm Willi, VP of Technology at flexa. I've spent the last 15 years building engineering teams around systems that have to make good decisions fast and at scale. Most recently at Lyft, where I led the mapping organization — 30+ engineers across five countries — and we moved more than 96% of rides off Google Maps onto our own mapping product. Before that I worked on machine learning and conversational AI at Microsoft Bing, and I did a PhD on learning in heterogeneous robot groups, which is a long way of saying that distributed decision-making has held my attention for a while.

What pulled me to flexa is that it's the same class of problem with something physical at the other end. We dispatch energy of the Enpal customer fleet in real time against energy markets. If our infrastructure is a few seconds late or a few percent off, customers lose money and the grid gets less flexibility than it could have had. That makes cloud engineering a first-order product concern here rather than a supporting function — which is why this role sits close to the decisions that actually matter.

How I work: I'd rather hand you the whole problem, context and constraints included, than a ticket. I care that engineers can see the consequences of what they build, and I'll be direct with feedback and expect the same back — the fastest way to lose a year is for everyone to stay polite about an architecture that isn't working.

You won't find everything already built. Some of it is greenfield, some of it needs replacing, and you'll have real influence over which is which. If that sounds like your kind of problem, I'd like to hear from you — and if you're not sure your profile is a perfect match, apply anyway and let's talk.

Looking forward to meeting you,

Willi

About Us
Flexa, a Joint Venture between Enpal and Entrix, is chartered with delivering a Virtual Power Plant (VPP) delivering exceptional energy cost savings while supporting the transition to a 100% renewable electricity future.

The combination of delivering complete residential energy systems at great cost with savvy market participation in several revenue streams sets us up to deliver real world customer savings while improving customer satisfaction enjoying the advantages of a fully electrified and energy producing home.

Through a deep integration into the installed hardware and a direct connection to the customer interfaces, Flexa’s solution controls the energy assets (such as EV, heat pump, home storage and others) of the entire Enpal energy community with an exceptional level of accuracy, speed, transparency, and thus customer satisfaction. With intelligent real-time dispatching algorithms, Flexa maximizes Enpal customers’ usable flexibility and its returns on energy markets including costs, such as grid fees, asset degradation.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

DevOps Engineer

Senior HPC Cluster Administrator - Deep Learning Frameworks Infrastructure

NVIDIA

Berlinfulltimesenior

Apply on the employer's site

Role description

NVIDIA's Deep Learning Frameworks (DLFW) Infrastructure team is looking for a deeply technical Senior HPC Cluster Administrator to lead the design, deployment, and reliability of our large-scale GPU compute clusters. These systems run the most demanding deep learning training, inference, and high-performance computing workloads in the industry — from DGX/HGX platforms to ground-breaking Grace Blackwell systems. You will drive architectural decisions across compute, networking, and storage, and partner closely with software, research, and product teams to keep our infrastructure ahead of the workloads it supports.

What You'll Be Doing

  • Own the full lifecycle of GPU compute clusters — procurement, provisioning, configuration management, monitoring, and deprecation — across heterogeneous Linux environments (DGX, HGX, embedded systems)
  • Design and scale storage solutions (NFS, Lustre, WekaFS, or equivalent) with a clear roadmap for capacity and performance growth
  • Lead automation of infrastructure using modern IaC tools (Ansible, Terraform) and CI/CD pipelines (GitLab)
  • Manage and optimize job scheduling via Slurm, including fair-share policies, reservation management, and MIG/GPU partitioning strategies
  • Maintain and improve observability stacks (Prometheus, Grafana, DCGM) and drive proactive resolution of hardware and software incidents
  • Collaborate with ML engineers and software teams to tune cluster configuration for large-scale distributed training workloads
  • Evaluate and introduce new technologies — networking fabrics (InfiniBand, NVLink, EFA/RDMA), storage tiers, container runtimes — to improve performance and reliability
  • Mentor junior engineers and contribute to team-wide engineering standards

What We Need To See

  • BS/MS in CS, EE, CE, or equivalent hands-on experience
  • 5+ years of experience deploying and administering large-scale HPC or ML training clusters
  • Deep expertise in Linux systems administration at scale
  • Strong scripting and automation skills in Python and/or bash
  • Hands-on experience with Slurm (scheduling, accounting, cgroup configuration)
  • Proficiency with configuration management and IaC (Ansible required; Terraform a plus)
  • Experience with container technologies (Docker, Apptainer/Singularity, Kubernetes)
  • Solid understanding of high-speed networking (InfiniBand, RoCE, RDMA, EFA)
  • Experience with distributed/parallel filesystems and storage architecture
  • Ability to own problems end-to-end and communicate clearly with engineering and management stakeholders

Ways To Stand Out From The Crowd

  • Experience with NVIDIA GPU infrastructure tools (DCGM, nvidia-smi, MIG, NVSwitch diagnostics)
  • Familiarity with cluster management platforms (Colossus, Bright Cluster Manager, xCAT, or similar)
  • Experience supporting large-scale distributed deep learning workloads (PyTorch, JAX, Megatron)
  • Knowledge of BMC/IPMI/Redfish for out-of-band management and hardware lifecycle
  • Background in MLOps tooling or ML platform engineering

Join our team of world-class engineers and be part of the groundbreaking work we do at NVIDIA. We are committed to encouraging a collaborative and inclusive environment, where every team member has the opportunity to thrive and make a significant impact!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 221,250 PLN - 383,500 PLN for Level 3, and 292,500 PLN - 507,000 PLN for Level 4. , , JR2015529

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

DevOps Engineer

DevOps Engineer *German required (GCP) (m/w/d)

ventx - we make IT!

Munichfulltime

Apply on the employer's site

Role description

Zur Verstärkung unseres Teams im Großraum München suchen wir, zum nächstmöglichen Termin, engagierte Leute die Lust haben, deine bisher erworbenen IT-Kenntnisse zu erweitern und sich firmenintern zu einem DevOps Cloud Engineer (m/w/d) mit Schwerpunkt AWS Cloud ausbilden zu lassen.

Aufgaben

  • Lust sich in neue Themengebiete einzuarbeiten

  • Mut sich zum Spezialisten ausbilden zu lassen

  • CI/CD Implementationen

  • Erstellung von Infrastrucure as Code

  • Consulting von Projekten im Bereich IT-Infrastruktur

  • Gute Deutsch- und Englischkenntnisse

Qualifikation

Dein Profil

eine erfolgreich abgeschlossene IT-Ausbildung bzw. ein abgeschlossenes Studium der Informatik oder entsprechende Berufserfahrung

  • sicherer Umgang mit Linux

  • beherrschen mindestens einer Script-Sprache wie Bash oder Python

  • Kenntnisse in der Administration von Netzwerken

  • gute Infrastrukturkenntnisse und Coding-Skills

Wünschenswert, aber nicht zwingend erforderlich

Erfahrungen in dem Bereich von Cloud Providern – insbesondere GCP

  • Know-How über Infrastruktur als Code – Cloudformation, Terraform

  • Kenntnisse der Tools CI/CD, GIT, Ansible, Chef, Puppet, Jenkins

  • Kenntnisse in der Administration von Netzwerken

  • praktische Erfahrung mit Monitoring-Tools wie Grafana, Graphite, Prometheus sowie Logging-Tools wie ElasticSearch, lostack, Kibana

Benefits

  • neben attraktiver Vergütung und abwechslungsreicher Tätigkeit, Spaß bei der Arbeit

  • flache Hierarchien mit kurzen Entscheidungswegen

  • ein Team dass sich gerne bei Problemen und Fragen gegenseitig unterstützt

  • moderne Büroräume mit aktuellster Hard- und Software

  • Team-Events

  • Kostenlose Getränke im Office

BEWIRB DICH JETZT! Wir antworten in kürzester Zeit.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

DevOps Engineer

Senior Systems Engineer – Business Applications (m/w/d)

BWI GmbH

Strausbergfulltimesenior

Apply on the employer's site

Role description

Sorge gemeinsam mit uns für die digitale Zukunftsfähigkeit der Bundeswehr.

Deine Aufgaben:

  • 2nd Level Support für die LAMP (Linux/Apache/MySQL/PHP) - Plattformen und Applikationen in unserer eigenen Cloud
  • Mitwirkung bei der automatisierten Bereitstellung von Plattformkomponenten durch CI/CD
  • Erstellung, Test und Implementierung von Containerbuilds
  • Bearbeitung von Störungen gemäß Spezifikation der BWI (Störungsannahme, Fehleranalyse, Fehlerbehebung)
  • Identifikation von Risiken, sowie Absicherung und Behebung von Security Risiken
  • Mitwirkung bei der Entwicklung von Schulungskonzepten
  • Pflege und Weiterentwicklung der technischen Dokumentation

Dein Profil:

  • Erfahrungen mit einem oder mehreren der folgenden Technologien: Webservern (Apache, NGINX), Linux Betriebssysteme (z.B. CentOS, SuSE, RHEL, Debian), Datenbanken (z.B.: MySQL, MariaDB, PostgreSQL), LAMP basierte Anwendungen
  • Gute Kenntnisse in Kubernetes und Cloud-nativen Konzepten
  • Kenntnisse von Anwendungsdeployments as Code via Helm-Charts, kustomize, k8s-Operators
  • Erfahrungen mit Automatisierungswerkzeugen (z.B. GitLab-CI, Terraform, Crossplane, Ansible, ArgoCD und FluxCD)
  • Erfahrungen mit Container-Builds (z.B. Docker, Buildah) und Artefaktverwaltung (Nexus, Harbor)
  • Vertrautheit mit Monitoring- , Logging- und Alerting-Stacks (z.B. Prometheus, Grafana, Instana)
  • Erfahrungen mit Skalierung und Ressourcenoptimierung in Container basierten Umgebungen sowie Git/Git-Ops
  • Grundverständnis TCP Netze wie Firewalling, Netzstruktur, Sicherheit
  • Erfahrungen mit CMDB/Asset-Managementsystemen
  • ITIL Grundkenntnisse (Change/Incident/Problem-Management)
  • Fließende Deutsch- und gute Englischkenntnisse

Wir bieten:

  • Durch abwechslungsreiche und gesellschaftlich relevante Aufgaben gewährleisten wir den reibungslosen IT-Betrieb und die Digitalisierung der Bundeswehr
  • Das Ziel eint uns. Dabei sind für uns ein wertschätzender Umgang miteinander sowie ein großer Teamgeist elementar
  • Die Vergütung liegt zwischen 57.800 € und 84.680 €. Die tatsächliche Höhe wird basierend auf deinem Verantwortungsbereich sowie Erfahrungen und Kompetenzen festgelegt
  • Wir bieten 30 Tage Jahresurlaub, 1 Brauchtumstag plus Optionen auf individuelle Anpassungen
  • Über unsere Benefit-App erhält man ein monatliches Guthaben und kann sich zusätzlich Steuervergünstigungen auf Tickets für den ÖPNV sichern
  • Wir ermöglichen Flexibilität, um Beruf und Privatleben in Einklang zu bringen, etwa durch mobiles Arbeiten oder Vertrauensarbeitszeit
  • Wir unterstützen die berufliche und persönliche Weiterbildung durch individuelle Maßnahmen sowie einen kostenfreien Zugriff auf LinkedIn Learning
  • Unser Jobradangebot ermöglicht das Leasing von bis zu 2 Fahrrädern

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

DevOps Engineer

IT Expert DevOps / Platform Engineering (m/w/d)

BWI GmbH

Berlinfulltime

Apply on the employer's site

Role description

Sorge gemeinsam mit uns für die digitale Zukunftsfähigkeit der Bundeswehr.

Ihre Aufgaben:

  • Betreuung und (Weiter-)Entwicklung einer Cloud-Native-Plattform für Data Analytics und KI Tools für unseren Kunden Bundeswehr
  • Realisierung und Sicherstellung der Betreibbarkeit eines PaaS/SaaS auf Basis von Kubernetes mit hohem Automatisierungsgrad in der Provisionierung und Deployment in einem agil aufgestellten Team
  • Automatisierte Bereitstellung von containerisierten Plattformkomponenten durch CI/CD sowie Überwachung und Optimierung der Plattform durch den gezielten Einsatz von Logging, Monitoring und Tracing
  • Kontinuierliche und systematische Härtung durch die Realisierung von Anforderungen der IT-Sicherheit
  • Mitwirkung an der Erstellung von code-naher technischen Dokumentationen

Ihr Profil:

  • Abgeschlossenes fachbezogenes Hochschulstudium oder eine vergleichbare Ausbildung
  • Etwa 2 Jahre Berufserfahrung in einem oder mehreren der folgenden Themenfelder Cloud-native Plattformen, Deployment von Microservices und/oder Automatisierung containerisierter Anwendungen
  • Gute Kenntnisse in Kubernetes und den Cloud-native Konzepten, API-Gateways, Service-Mesh (z.B. Istio), Authentifizierung und Observability
  • Erfahrungen im Umgang mit Automatisierungswerkzeugen und Infrastructure as Code und GitOps, z.B. mit GitLab-CI, Terraform, Crossplane oder ArgoCD
  • Erfahrung im Deployment von Machine Learning Systemen (MLOps) wünschenswert
  • Gute Kommunikations- und Teamfähigkeiten, ein analytisches Denkvermögen Lernbereitschaft sowie eine sorgfältige und gewissenhafte Arbeitsweise
  • Verhandlungssichere Deutschkenntnisse und ein sicheres Verständnis der englischen Sprache

Wir bieten:

  • Durch abwechslungsreiche und gesellschaftlich relevante Aufgaben gewährleisten wir den reibungslosen IT-Betrieb und die Digitalisierung der Bundeswehr
  • Das Ziel eint uns. Dabei sind für uns ein wertschätzender Umgang miteinander sowie ein großer Teamgeist elementar
  • Die Vergütung liegt zwischen 57.800 € und 84.680 €. Die tatsächliche Höhe wird basierend auf deinem Verantwortungsbereich sowie Erfahrungen und Kompetenzen festgelegt
  • Wir bieten 30 Tage Jahresurlaub, 1 Brauchtumstag plus Optionen auf individuelle Anpassungen
  • Über unsere Benefit-App erhält man ein monatliches Guthaben und kann sich zusätzlich Steuervergünstigungen auf Tickets für den ÖPNV sichern
  • Wir ermöglichen Flexibilität, um Beruf und Privatleben in Einklang zu bringen, etwa durch mobiles Arbeiten oder Vertrauensarbeitszeit
  • Wir unterstützen die berufliche und persönliche Weiterbildung durch individuelle Maßnahmen sowie einen kostenfreien Zugriff auf LinkedIn Learning
  • Unser Jobradangebot ermöglicht das Leasing von bis zu 2 Fahrrädern

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

206 more openings in this category and country

Senior Cloud Engineer (d/f/m)flexa · Germany

Apply on the employer's site