Skip to content
Site Reliability Engineer

Site Reliability Engineer

Principal AI-Native Engineer, @Ionic Partners

Sparkrock

Remotelead

Apply on the employer's site

Role description

Most companies are running AI pilots. We are rebuilding how our companies actually operate.

Ionic Partners is a holding company with a portfolio of operating businesses. We are not adding AI features to products or rolling assistants out to employees. We are rebuilding the operating systems of entire business functions — engineering, support, finance, marketing, revenue operations — as agentic systems that do the work, in production, every day.

We are looking for Principal AI-Native Engineers to build them.

This is not a plan we are pitching. Agents already write and review code inside our engineering pipeline. Our own orchestration framework is what those systems run on, and it is on its second generation because the first one taught us what breaks. The knowledge our agents work from is maintained as structured, machine-readable artifacts rather than scraped from wikis, because we learned early that context quality sets the ceiling on everything else. Every hard-won lesson in that sentence cost us something. You will inherit all of it, and the next set of lessons is yours to learn.

This is a central team, deliberately flat, that builds agentic systems for every company in the portfolio. There are no internal layers, no leads, and no junior tier. Everyone on it is a principal-level builder reporting directly to the executive who owns the mandate. Scope is concentrated rather than distributed: each engineer carries an unusually large share of the outcome and an unusually direct line to the decisions that matter.

You will not be starting from scratch. We already run a production agentic platform in-house — our own orchestration framework, machine-readable context artifacts that give agents grounded knowledge of our products and customers, and automated development and QA agents working inside our engineering pipeline. Your first job is to ramp into a live system that real businesses depend on. Your ongoing job is to extend it, harden it, and use it to make each new function agentic faster than the last.

The team is domain-agnostic by design. Subject-matter experts inside each business define what their function needs. You design and build the systems that deliver it. You do not need to be an expert in finance, support, or revenue operations. You need to be exceptional at turning a well-specified problem into a reliable agentic system that runs without you standing over it.

And you own what you build. Not a prototype handed to another team. Not a proof of concept that dies in a demo. You ship it, you run it, you watch its evals, you carry its cost line, you improve it. If it breaks, it is yours. If it transforms how a company works, that is yours too.

This is not an applied research role, and it is not an AI enablement or transformation role. Your work product is not a strategy, a recommendation, or a rollout plan — it is a working system in production, operating a real business function, measured on whether that function got materially better. The distance between an impressive agentic prototype and a system a business will actually depend on is where this entire job lives: the eval harnesses, the failure modes nobody anticipated, the guardrails, the traces, the fallback paths, the cost per run at volume, the second year of maintenance.

It is also not a management role, and it does not become one. There is no reporting line beneath you and no expectation that you build one. Influence here comes from what you ship and the standard you set for the engineers and operators around you. We built the team flat specifically so we would never have to ask an elite builder to stop building to advance.

The bar is high, and we intend to keep it there. Each hire is a large fraction of this team, and we would rather stay understaffed than lower it.

If you want to build agentic systems that run real businesses — and stay technical while doing it — we would like to hear from you.

Responsibilities

  • Partner with subject-matter experts across the portfolio companies to turn functional specs into agentic system designs, pressure-testing scope and surfacing failure modes before build
  • Design agent architectures: agent boundaries, orchestration and control flow, tool surfaces, state and memory, human-in-the-loop placement, and the split between model-driven and deterministic logic
  • Build, ship, and operate production agentic systems that run real business functions end to end
  • Extend and harden the in-house orchestration framework — execution control, integrations, permissioning, observability, failure, and recovery behavior — so each build raises the floor for the next
  • Improve the automated development and QA agents already running inside our engineering pipeline
  • Create and maintain machine-readable context artifacts covering products, customers, processes, and operational data, and keep them current as the businesses change
  • Design retrieval and context-assembly strategies, manage context budgets, and treat prompting as versioned, tested, measured engineering
  • Integrate agentic systems with live business platforms — product codebases, CRMs, ERPs, support desks, data warehouses, and internal services — including auth, permissioning, rate and cost limits, idempotency, and blast-radius controls
  • Define success criteria and build eval harnesses, baselines, regression suites, and human review sampling where automated scoring is insufficient
  • Instrument systems with tracing and observability that make agent behavior diagnosable in production
  • Build guardrails, fallbacks, retries, circuit breakers, approval gates, audit trails, and rollback paths proportional to what a system can affect
  • Track and control token and tool cost per unit of work, and design for unit economics that hold at full volume rather than pilot volume
  • Monitor eval trendlines, cost curves, and failure patterns for systems in production, and act on what they show
  • Investigate and resolve regressions, incidents, and quality drift in systems you own
  • Evolve deployed systems as the business processes they serve change
  • Ramp quickly into large, unfamiliar production codebases across our companies and extend them safely rather than working around them
  • Extract reusable components, patterns, and abstractions from solved problems so recurring classes of work get cheaper across functions and companies
  • Evaluate emerging models, frameworks, agentic techniques, and tooling against real workloads, and make adoption calls based on measured results rather than capability claims
  • Define and apply practices for responsible autonomous operation: data exposure, IP protection, security boundaries, and human review requirements
  • Review agentic work built elsewhere in the organization and raise the technical standard through direct engagement and demonstrated results
  • Work alongside engineers and operators whose day-to-day work these systems change, explaining behavior, limits, and intent
  • Retire systems and approaches that measurement shows are not delivering, and document why

Requirements

  • Bachelor's degree or higher in Computer Science, Computer Engineering, Software Engineering, or a related field, or equivalent practical experience. We weigh demonstrated engineering work far more heavily than credentials; a strong body of shipped systems fully substitutes for a degree
  • 8+ years of hands-on software engineering experience building and operating production systems
  • 3+ years building AI-Native or agentic systems that reached production and were used for real work — not prototypes, internal demos, hackathon projects, or evaluations that stopped at a pilot
  • Demonstrated end-to-end ownership: systems you designed, built, shipped, and then operated and improved over time, including responsibility for their failures
  • Experience integrating AI systems with the platforms a business actually runs on — production codebases, CRMs, ERPs, support and ticketing systems, data warehouses, internal services — including authentication, permissioning, and controls on what an autonomous system is allowed to do
  • Experience designing and running evaluations for non-deterministic systems: defining correctness criteria, building eval harnesses, establishing baselines, and detecting regressions
  • Experience operating LLM-based systems in production, including tracing, failure diagnosis, guardrails, and cost management at volume
  • Experience in becoming productive quickly inside large, complex codebases you did not write, and extending them safely
  • Experience working from specifications or requirements set by domain experts outside your own area of expertise
  • Experience with platform, infrastructure, or developer-tooling work — building the systems that other engineers or systems depend on
  • People management experience is not required and is not an advantage. This is a principal-level individual contributor role with no reporting line, and we are explicitly open to engineers who have deliberately stayed technical
  • Ownership. You treat what you ship as yours indefinitely — the outcomes, the failures, the cost, and the maintenance. You do not look for the point where responsibility transfers to someone else
  • Autonomy. You operate from intent rather than instruction, set your own sequence, and make progress in genuine ambiguity without waiting for the picture to resolve
  • Intellectual honesty. You report what the data shows, including when it undermines your own work. You retire your own systems when they stop earning their keep and say plainly what did not work
  • Judgment. You weigh speed against reliability, autonomy against safety, and elegance against maintainability, and you consistently choose well without a rule to follow
  • Adaptability. You expect your techniques to be obsolete within a year and treat that as normal rather than destabilizing
  • Collaboration across expertise boundaries. You work well with domain experts who know their function far better than you do, take their specs seriously, and push back with substance when a spec is wrong
  • Communication. You explain complex system behavior clearly to technical and non-technical audiences, and you write well enough that your reasoning survives without you in the room
  • Curiosity with discipline. You stay current on a fast-moving field and test new capabilities against real workloads before believing it
  • High standards, low ceremony. You hold a high bar for yourself and the people around you without needing a process to enforce it

Nice to have

  • Experience building agentic systems in enterprise environments, with the constraints that imply: legacy systems, compliance requirements, security review, real data sensitivity, and organizational change
  • Experience making a non-engineering business function agentic — support, finance, marketing, revenue operations, or similar
  • Experience building shared platform or framework layers that multiple downstream systems depend on
  • Experience with enterprise SaaS, ERP systems, or mission-critical business applications
  • Experience applying agentic development or QA workflows inside a real engineering pipeline
  • Experience modernizing or extending complex legacy systems using AI-assisted approaches
  • Experience working across multiple companies, business units, or product lines rather than a single product
  • Experience in a fully remote, globally distributed organization
  • Open-source contributions, technical writing, or public work on agentic systems

Benefits
We don't call them perks; they're part of what makes working here great.

  • Access to frontier models, agentic tooling, and infrastructure without procurement friction, plus the authority to evaluate, choose, and replace what the team builds on
  • A live production agentic platform to inherit and extend, rather than a blank page — and the mandate to make every company in the portfolio run on what you build
  • Unusual breadth. Because you work across a portfolio rather than a single product, you will touch more distinct problem domains in a year here than in almost any single-company role, and see your work compound across all of them
  • A principal-level individual contributor path with real scope, real autonomy, and no expectation that you move into management to advance
  • We are 100% remote and global. Live your best life wherever that may be, and never lose out on career opportunities because of it
  • Flexible work hours. We work asynchronously and don't care when you're online, just that you deliver great results and are there for our customers
  • We are dedicated to your growth with consistent and meaningful feedback, support in achieving your personal career goals, and access to leading-edge tools, playbooks, and technology to amplify your experience
  • Introductions to thought leaders in the space and webinars on cutting-edge tech hot topics
  • Stipend to help set up your ideal home office
  • Focus on culture: coffee chats, happy hours, cooking classes, book clubs, and more!

All open roles are for existing vacancies unless otherwise communicated to the candidate. We are committed to keeping candidates informed throughout the process and will notify all interviewed applicants of our hiring decision within 45 days of their interview. The company retains all job postings and related recruitment information for a minimum of three years.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Ingénieur Système, Sauvegarde et réseaux à Paris H/F

Free-Work

Remotefulltimesenior

Apply on the employer's site

Role description

Dans le cadre du développement de notre équipe IT chez l'un de nos clients grands comptes, nous recherchons un(e) Ingénieur(e) Système Sauvegarde réseaux H/F afin d'assurer l'administration, la maintenance et l'évolution des infrastructures systèmes de nos environnements clients.

Vos missions seront les suivantes :

  • Participer aux projets d'évolutions de la plateforme technique de la Video Factory ( tête de réseau OTT et IPTV
  • Conception / participation au POC avec le N3 (les experts) / intégration / ingénierie / recette unitaire / recette des workflow et documentation des briques techniques et des procédures.
  • Assurer la maintenance en condition opérationnelle de la plateforme
  • Organisation des sauvegardes et des montées de version des équipements IT (VM, NAS, OS (linux et windows)...) et Vidéos (encodeurs, DCM, serveurs d'origine, sondes ) de la plateforme.
  • Analyse des risques et des impacts potentiels et planification en HNO le cas échéant.
  • Assurer le « run » de la plateforme en tant que support niveau 2 en soutien des équipes support de niveau 0 et 1
  • Analyse d'incidents et résolutions, escalade au N3 et aux fournisseurs (ouverture et suivi des tickets) le cas échéant, communication sur les avancées les plus significatives et les impacts majeurs,
  • Organiser les activités des fournisseurs (mises à jour et évolution technique),
  • Assurer le suivi des déploiements et mettre en place les contrôles (recette).
  • Participer à l'amélioration de l'organisation du support global de la plate-forme technique :
  • Formation des équipes de maintenance, documentation des procédures d'exploitation.

     Des missions complémentaires peuvent être confiées.

Référence de l'offre : cmk6d3cswx

Profil candidat:

Profil recherché :

Issu(e) d'une formation supérieure en informatique (BAC+5 / Diplôme d'ingénieur),

Vous justifiez d'au moins 5 ans d'expérience sur un poste similaire.

Compétences techniques souhaitées :

  • Maîtrise des environnements Windows Server, Linux et Vmware (des connaissances sur Nutanix est un plus).
  • Bonne connaissance des environnements vidéo et audio sur IP.
  • Connaissance des réseaux IP et de l'adressage multicast.
  • Maitrise des outils d'analyse de qualité vidéo ainsi que des outils de supervision.
  • Bonnes capacités d'analyse et de résolution de problèmes.
  • Autonomie, rigueur et bon relationnel.
  • Capacité à travailler en équipe et à intervenir dans des environnements de production

Environnement technique :

DCM, Anevia, Imagine, Nevion, Harmonic xOS, OpenHeadEnd, Elemental, USP.

Produits systèmes, virtualisation : VMware, Wallix, Nutanix, NAS, Debian, Windows Server, FTP

Ce que nous vous proposons :

Valeurs : en plus de nos 3 fondamentaux que sont l'audace, la bonne foi et la réactivité, nous garantissons un management à l'écoute et de proximité, ainsi qu'une ambiance familiale.

Contrat : CDI ou Freelance

Localisation : Paris

Package rémunération & avantages :

  • Le salaire : rémunération annuelle brute selon profil et compétences
  • Les basiques : mutuelle familiale et prévoyance, titres restaurant, remboursement transport en commun à 50%, avantages du CSE (culture, voyage, chèque vacances, et cadeaux), RTT (jusqu'à 12 par an), plan d'épargne, prime de participation
  • Nos plus : forfait mobilité douce & Green (vélo/trottinette et covoiturage), prime de cooptation de 1000 € brut, e-shop de matériel informatique à des prix préférentiels (smartphone, tablette, etc.)
  • Votre carrière : plan de carrière, dispositifs de formation techniques & fonctionnels, passage de certifications, accès illimité à Microsoft Learn
  • La qualité de vie au travail : télétravail avec indemnité, évènements festifs et collaboratifs, accompagnement handicap et santé au travail, engagements RSE

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

SAP DevOps Engineer (f/m/d)

E.ON Digital Technology

Remote

Apply on the employer's site

Role description

You have a passion for technology and want to make the world a greener place?
Then become a playmaker (f/m/d) and join our team as SAP DevOps Engineer (f/m/d) at E.ON Digital Technology.

We play a key role in shaping the energy transition by leading E.ON's digital transformation across Europe. We explore new paths by developing ideas, breaking new ground, making visions reality, and bringing new technologies to life. We deliver sustainable technology solutions because…

… it’s on us to make new energy work!
The Team
– your impact

At E.ON, the SAP Engineering Chapter is a collective of innovative minds dedicated to delivering world-class SAP architecture, SAP software development and SAP engineering capabilities across our segments and product teams. By joining us, you will play a critical role in keeping E.ON's SAP architecture and capabilities modern, secure and excellent.

Your Role –
meaningful & rewarding

As a SAP DevOps Engineer (f/m/d) at E.ON, you will be responsible for designing, developing, training and maintaining SAP DevOps solutions for our SAP system landscape and platforms. You will work closely with the SAP DevOps teams and developers, and other stakeholders to ensure the provisioning of a modern, user-friendly, state-of-the-art SAP DevOps pipeline.

  • Design, implement and continuously improve DevOps concepts, CI/CD templates, pipelines and automation for SAP landscapes (e.g. SAP BTP, S/4HANA)
  • Enable reliable build, test and deployment processes across SAP Cloud and on-premise environments
  • Establish observability concepts for SAP BTP and on-premise applications for monitoring, health checks, alerts and APM by using tools such as SAP Cloud ALM, New Relic and Uptrends
  • Collaborate closely with SAP development, architecture, security and operations teams
  • Conduct DevOps maturity assessments and guide teams on their DevOps Journey
  • Ensure high availability, performance, security and compliance of SAP systems
  • Troubleshoot complex problems and support root cause analysis
  • Continuously evaluate new SAP DevOps tools, technologies and best practices

Your Profile
– authentic & open-minded

  • Strong experience in DevOps, platform engineering or operations within SAP environments
  • Hands-on experience with CI/CD and security, as well as quality tooling (e.g. GitLab CI/CD, SonarQube, Renovate or similar)
  • Strong knowledge in development, automated testing and testing strategies
  • Experience with scripting and automation (e.g. Bash, Groovy)
  • Solid knowledge of SAP S/4HANA, SAP BTP and related deployment and transport mechanisms
  • Knowledge of Kubernetes and container technologies and IaaC principles is a plus
  • Strong understanding of security, monitoring and reliability concepts
  • Structured, proactive and solution-oriented working style
  • Fluent in English

Our Benefits
– smart & useful

  • Advance your development: We grow and we want you to grow with us. Learning on the job, exchanging with others, or taking part in an individualtraining – our learning culture enables you to bring your personal and professional development to the next level.
  • Recharge your battery: You have 30 days of paid vacation per year plus Christmas and New Year's Eve off. Your battery still needs charging? You canexchange parts of your salary for more paid vacation or you can take a sabbatical.
  • Enjoy hybrid work: We combine office collaboration with focused work from home. It’s also possible to go on workation for up to 20 days per year withinEurope.
  • Stay active & healthy: Benefit from a company-sponsored health membership.
  • Elevate your mobility: From car and bike leasing offers to a subsidised Deutschland-Ticket – your way is our way.
  • Think ahead: With our company pension scheme and a great insurance package we take care of your future.
  • This is by far not all… We are looking forward to speaking with you about further benefits during the hiring process.

Do you have questions?
For further information please contact Agneta Lierl, EDT_Talent_Acquisition@eon.com.

What you need to know:
Contract type: Permanent

Working time: Full time

Company: E.ON Digital Technology GmbH

Location: Essen, Hannover, München, Berlin, Würzburg, Hamburg, Frankfurt am Main

Function area: IT/Digital; Engineering

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Security Engineer (m/w/d)

Rocken®

Remotesenior

Apply on the employer's site

Role description

Steuerverwaltungen brauchen clevere Software – und dahinter stehen clevere Köpfe. Unser Rocken Partner entwickelt die führende Business Lösung für kantonale und kommunale Steuerverwaltungen. Die Anwendung deckt den kompletten Verwaltungsprozess ab: vom Steuerregister über Veranlagungen und Fakturierung bis zum Inkasso und zur Verlustscheinbewirtschaftung. Derzeit entsteht eine neue Software-Generation – ein spannendes Projekt, das technisches Know-how und Innovationskraft vereint. Die Arbeitsweise ist geprägt von überlegter Fokussierung, lebendigem Austausch zwischen Teams und intellektueller Courage. Elegante Lösungen entstehen durch Wissen, Geist und Teamenergie. Bereit, an einem Produkt zu arbeiten, das echten Impact hat? Unser Rocken Partner sucht Menschen, die sich für eine gemeinsame Idee begeistern und die digitale Zukunft der Steuerverwaltung mitgestalten möchten.

Verantwortung

  • Du analysierst Schwachstellen und setzt passende Sicherheitslösungen um.
  • Du arbeitest an Themen wie Zugriff, Netzwerk, Pentesting und Monitoring.
  • Du entwickelst die Sicherheitsarchitektur weiter und beachtest ISO 27001.
  • Du automatisierst in Windows/Linux und unterstützt Kubernetes-Setups.

Qualifikationen

  • Du hast eine IT-Ausbildung und Erfahrung als System Engineer mit Security-Fokus.
  • Du denkst vernetzt, erkennst Risiken und arbeitest agil mit modernen Tools.
  • Du kennst Azure, VMware und gängige Security-Lösungen.
  • Du brennst für IT-Security und entwickelst im Team nachhaltige Lösungen.

Benefits

  • Interessante und abwechslungsreiche Tätigkeiten/Projekte
  • Attraktive Weiterbildungs- und Entwicklungsmöglichkeiten
  • Flexible Arbeitszeitgestaltung
  • Homeoffice
  • Offene Unternehmenskultur
  • Beteiligung oder Übernahme ÖV-Abonnements
  • Beteiligung oder Übernahme Parkplatz
  • Attraktive Mitarbeiterrabatte
  • Kostenlose Früchte und Getränke

ROCKEN Jobs
https://rocken.jobs

Profil Erstellen
https://rocken.jobs/application/profil\-erstellen/

  • new

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Microsoft Cloud & System Engineer (m/w/d)

Rocken®

Remote

Apply on the employer's site

Role description

Unser Rocken® Partner ist spezialisiert auf IT-Security, Cloud Lösungen, sowie IT-Outsourcing und Support. Sie setzen auf innovative Services und Dienstleistungen, bewegen und orientieren sich am Puls der Technik. Daraus ergeben sich für unsere Kunden viele Vorteile gegenüber den traditionellen IT-Lösungen, die weniger flexibel und selten skalierbar sind.

Verantwortung

  • Sicherstellung einer stabilen IT-Infrastruktur und hohen Systemverfügbarkeit
  • Betreuung und Weiterentwicklung von Microsoft-Umgebungen im Kundenumfeld
  • Planung und Umsetzung von Infrastruktur-, Rollout- und Migrationsprojekten
  • Analyse und Behebung technischer Störungen im 1st- und 2nd-Level-Support
  • Direkter Kundensupport, technische Beratung und Betreuung vor Ort

Qualifikationen

  • Microsoft 365, Azure, Entra ID und Microsoft Intune
  • Kenntnisse in Client-/Server-Systemen, Virtualisierung, Monitoring und Backup
  • Erfahrung im 1st-/2nd-Level-Support und technischen Troubleshooting
  • Kenntnisse in Infrastrukturplanung, System-Rollouts und Migrationen
  • Sehr gute Deutschkenntnisse sowie gute Englischkenntnisse; weitere Sprachen von Vorteil

Benefits

  • Flexible Arbeitszeitgestaltung
  • Homeoffice
  • Zahlreiche Mitarbeiterevents
  • Beteiligung oder Übernahme Parkplatz
  • Kostenlose Früchte und Getränke
  • Attraktive Weiterbildungs- und Entwicklungsmöglichkeiten

ROCKEN Jobs
https://rocken.jobs

Profil Erstellen
https://rocken.jobs/application/profil\-erstellen/

  • new

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

241 more openings in this category and country

Principal AI-Native Engineer, @Ionic PartnersSparkrock

Apply on the employer's site