Skip to content
Site Reliability Engineer

Site Reliability Engineer

Staff Software Engineer, Cluster Orch (SUNK)

CoreWeave

Warsawlead

Apply on the employer's site

Role description

CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.

We're proud to be a Living Wage accredited Employer.

What You'll Do
CoreWeave’s Cluster Orchestration team builds and operates the Kubernetes-native foundation that powers AI training and inference at scale. We eliminate infrastructure bottlenecks and create next-generation orchestration capabilities—including SUNK (Slurm on Kubernetes) and beyond—ensuring demanding workloads run seamlessly, reliably, and efficiently across massive GPU clusters.

About The Role
As a Staff Software Engineer (IC5), you will act as a principal technical leader shaping the long-term architecture and strategy for CoreWeave’s orchestration platform. You will define the technical direction, own critical components of our orchestration layers and managed services, and drive major cross-organizational initiatives in scheduling, multi-tenant quota enforcement, and hyperscale scaling. This high-impact role requires you to establish organization-wide best practices for platform reliability and observability, resolve complex distributed bottlenecks under intense demand, and mentor senior engineers across teams to elevate engineering standards.

Who You Are

  • 8–12 years of professional software engineering experience.
  • Advanced software development proficiency in Go and strong distributed systems design principles.
  • Proven track record designing, operating, and scaling large-scale distributed systems in production environments.
  • Deep technical expertise in Kubernetes internals, Slurm schedulers, or cloud-native development.
  • Demonstrated experience setting technical direction and successfully influencing cross-team architecture and platform goals.
  • Proven ability to mentor senior engineers, review complex technical designs, and elevate organizational operational standards.

Preferred

  • Familiarity with modern orchestration, workflow, and mesh technologies such as Ray, Kubeflow, Kueue, Istio, Knative, or Argo Workflows.
  • Deep experience with distributed workloads, GPU-based cloud applications, or machine learning pipelines.
  • Strong structural knowledge of advanced scheduling concepts, including quota enforcement, resource pre-emption, and intelligent scaling strategies.
  • Practical exposure to enterprise reliability practices, including defining SLOs, telemetry monitoring, and leading post-incident reviews.
  • Direct exposure to large-scale AI infrastructure workloads (ML training, inference, or high-performance computing).

Wondering if you're a good fit?
We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk.

  • You love to define the long-term architecture for cloud platform systems operating at global hyperscale.
  • You're curious about pioneering orchestration beyond SUNK and uncovering novel ways to evolve platforms for next-generation AI workloads.
  • You're an expert in mentorship, technical leadership, and solving structural issues that balance infrastructure cost, performance, and reliability.

Why CoreWeave?
About
At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:

  • Be Curious at Your Core
  • Act Like an Owner
  • Empower Employees
  • Deliver Best-in-Class Client Experiences
  • Achieve More Together

We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and enables the development of innovative solutions to complex problems. As we get set for takeoff, the organisation's growth opportunities are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!

The base salary range for this role is
369,000 PLN to 493,000 PLN
gross annually. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).

To fulfill our obligation to protect client data, successful applicants offered employment with CoreWeave will be required to complete a basic criminal record check, conducted in compliance with GDPR. Employment offers are conditional upon receiving satisfactory check results
What We Offer
In addition to a competitive salary, we offer a variety of benefits to support your needs, including:

  • Family-level Medical Insurance
  • Family-level Dental Insurance
  • Generous Pension Contribution
  • Life Assurance at 4x Salary
  • Critical Illness Cover
  • Employee Assistance Programme
  • Tuition Reimbursement
  • Work culture focused on innovative disruption

Benefits may vary by location.
Equal Opportunity
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
Recruitment Agencies
CoreWeave does not accept speculative CVs. Any unsolicited CVs received will be treated as the property of CoreWeave and your Terms & Conditions associated with the use of CVs will be considered null and void.
Any unsolicited CVs sent by your company to us – that is to say, in any situation where we have not directly engaged your company in writing to supply candidates for a specific vacancy – will be considered by us to be a “free gift”, leaving us liable for no fees whatsoever should we choose to contact the candidate directly and engage the candidate’s services, and will in no way establish any prior claim by your company to representation of that candidate should the candidate’s details also be submitted by any other party.
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C.

  • 1157, or (iv) asylee under 8 U.S.C.
  • 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.

Updated privacy notice - UK and EU Job Applications
When you apply to a job on this site, the personal data contained in your application will be collected by CoreWeave UK Ltd. (“Controller”), which is located at
Phosphor (6th Floor), 133 Park Street, London, SE1 9EA
and can be contacted by emailing
careers.eu@coreweave.com
. Controller’s data protection officer can be contacted at
privacy@coreweave.com
. Your personal data will be processed for the purposes of managing Controller’s recruitment related activities, which include setting up and conducting interviews and tests for applicants, evaluating and assessing the results thereto, and as is otherwise needed in the recruitment and hiring processes. Such processing is legally permissible under Art. 6(1)(f) of (i) Regulation (EU) 2016/679 (General Data Protection Regulation (“GDPR”) and (ii) the GDPR as it forms part of the laws of the UK (“UK GDPR”), as necessary for the purposes of the legitimate interests pursued by the Controller, which are the solicitation, evaluation, and selection of applicants for employment. Your personal data will be shared with Greenhouse Software, Inc., a cloud services provider located in the United States of America and engaged by Controller to help manage its recruitment and hiring process on Controller’s behalf. With respect to transfers originating from the UK or the European Economic Area ("EEA") to a country outside the UK or the EEA, we implement the appropriate transfer mechanism(s) and other appropriate solutions to address cross-border transfers as required by applicable law. You may request a copy of the suitable mechanisms we have in place by contacting us at
privacy@coreweave.com
Your personal data will be retained by Controller as long as Controller determines it is necessary to evaluate your application for employment. Where permitted by applicable law, we may also retain your personal data for a limited period after the recruitment process ends in order to consider you for future job opportunities, respond to legal claims, or comply with record-keeping obligations. Under the GDPR and the UK GDPR, you have the right to request access to your personal data, to request that your personal data be rectified or erased, and to request that processing of your personal data be restricted. You also have the right to data portability. In addition, you may lodge a complaint with the relevant supervisory authority: (i) A list of Europe’s data protection authorities can be found
here
; and (ii) for the UK, this is the
Information Commissioner's Office
.
For additional information, please see our
Privacy Policy
.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Senior Software Engineer, Platform

WP Engine

Krakow

Apply on the employer's site

Role description

We engage the most inspired minds to do their best work wherever they work best—powering the freedom to create worldwide.

WP Engine empowers companies and agencies of all sizes to build, power, manage, and optimize their WordPress websites and applications with confidence. Serving 1.5 million customers across 150+ countries, the global technology company provides premium, enterprise-grade solutions, tools, and services, including specialized platforms for WordPress, industry-tailored eCommerce and agency solution suites, and developer-centric tools like Local, Advanced Custom Fields, and more. WP Engine’s innovative technology and industry-leading expertise are why 8% of the web visits a WP Engine-powered site daily. Learn more at wpengine.com.
Senior Software Engineer, Platform (Poland)
In this role, you will architect and implement backend capabilities that drive our global, high-scale cloud infrastructure services.This position transcends standard backend development and pure DevOps; you will operate at the intersection of software engineering, distributed systems, and production reliability, maintaining full ownership of services from initial design through observability and technical support. Your work will focus on complex challenges in resiliency, asynchronous workflows, automation, and system scalability. We value engineers with deep technical fundamentals who are eager to adapt to evolving architectures, new languages, and emerging cloud-native technologies.

What's Cool About This Job

  • Design, build, test, and operate scalable backend services and platform capabilities.
  • Develop APIs and service integrations used across distributed and cloud-based systems.
  • Make technical design decisions around data flow, asynchronous workflows, scalability, resiliency, and performance.
  • Own services throughout their lifecycle, from design and implementation to deployment, observability, troubleshooting, and production support.
  • Design and operate services across different cloud and platform environments, understanding how application, infrastructure, networking, and runtime decisions affect production systems.
  • Work across modern and established platform technologies, making pragmatic decisions based on system requirements.
  • Diagnose complex problems across application, infrastructure, networking, and data layers.
  • Improve CI/CD, automation, testing, and engineering workflows to reduce manual work and increase delivery quality.
  • Contribute to technical designs through ADRs, RFCs, and architecture discussions.
  • Collaborate with engineers across teams and help raise engineering standards through technical feedback and knowledge sharing.
  • Use AI-assisted engineering tools to accelerate development, investigate problems, and improve engineering productivity.

Your expertise and passion
You have strong software engineering fundamentals and a passion for building scalable, resilient production systems. You value engineering principles over specific languages and are comfortable operating across the entire software lifecycle.

  • Backend & Cloud Architecture: Skilled in designing REST/gRPC APIs, relational databases (SQL), and cloud-native architectures across GCP and AWS environments.
  • Distributed Systems: Deep understanding of event-driven architecture, asynchronous processing, scalability, fault tolerance, and eventual consistency.
  • DevOps & Security: Committed to Infrastructure as Code (Terraform), observability (monitoring/tracing), and integrating security (IAM, OAuth) throughout the development process.
  • Engineering Discipline: Practitioner of SOLID principles, automated testing, and collaborative development through code reviews and technical documentation.
  • AI Integration: Proactive in leveraging modern AI tools to accelerate development and improve code quality while maintaining technical ownership.

Helpful experience to have

  • Languages & Runtimes: Go, PHP, Ruby, Python.
  • Cloud Services: GCP (GKE, Pub/Sub, IAM, BigQuery, Cloud SQL) and AWS (EC2, S3).
  • Orchestration & Infrastructure: Kubernetes, Docker, Terraform.
  • Observability Tools: OpenTelemetry, Google Cloud Operations, Prometheus, Datadog

Perks and Benefits

  • Company Stock Options (Every employee is an owner in the company)
  • Health Benefits (LuxMed Premium Rehab + Dentistry III)
  • Pension Scheme with a match
  • Life Insurance, 100% premium paid by WP Engine
  • Free Multisport Card - Light and Classic
  • Short-Term Disability & Sick Leave: 100% salary coverage for the first 33 days
  • Employee Assistance Program
  • Generous Vacation Time (Who doesn’t like time off)
  • 1 floating holiday
  • 4 Company Wellness Days
  • Home Office Stipend
  • On-going education through LinkedIn Learning, Workday Learning and our Career Growth Portal

At WP Engine, we strive to have the broadest possible view of diversity, going beyond visible differences to include the background, experiences, skills, and perspectives that make each person unique. WP Engine is proud to be an equal opportunity workplace and is committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, Veteran status, or any other basis protected by federal, state, or local law.

Base Salary Range
zł244 000,00 - zł335 500,00

We believe that compensation should be reflective of the impact you have within the organization relative to the market value of your role. The estimated base salary range for this position is as listed above. Some roles may also be eligible for overtime pay. Our salary ranges are determined by job role and responsibilities and level. The range displayed on each job posting reflects the minimum and maximum target for salaries for the position nationwide. The actual base pay will vary based on various factors including job-related skills and individual qualifications objectively assessed during the interview process. Your talent acquisition partner can share more about the total rewards package at WP Engine including any additional total rewards components such as equity, variable pay plans (if applicable), and benefits during the hiring process.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Cloud Engineer (Blockchain) - Digital Assets

EPAM Systems

Krakowcontractsenior

Apply on the employer's site

Role description

We are looking for a
Senior Cloud Engineer
to join our Digital Assets Crew, working on pioneering blockchain-based projects including tokenization of financial assets. Our team operates within two Scrum teams and builds cloud-hosted applications that leverage cutting-edge technologies to deliver value and innovation to the finance industry

Responsibilities

  • Develop cloud solutions that scale and ensure high availability
  • Nurture a DevOps culture across the team
  • Contribute to DLT-related projects, particularly blockchain
  • Design and maintain infrastructure using Terraform and Helm
  • Implement GitOps and CI/CD pipelines
  • Monitor distributed systems using Grafana and Prometheus
  • Manage secret management solutions across cloud environments
  • Collaborate with colleagues in English on cross-functional initiatives

Requirements

  • 5+ years of experience in DevOps/SRE roles
  • Expertise in Azure Cloud and Kubernetes
  • Proficiency in Terraform, Helm and GitOps/CI-CD
  • Experience with Grafana, Prometheus and distributed systems architecture at enterprise scale
  • Knowledge of secret management solutions
  • Good communication skills, comfortable interacting with colleagues in English (B2+)

Nice to have

  • Experience with Istio Service Mesh and zero-trust architecture
  • Background in banking or regulated environments
  • Familiarity with blockchain technology

We offer

  • We gather like-minded people:
  • Top tech minds driving innovation in AI, cloud and digital platform modernization
  • Supportive team and agile, startup-like culture
  • Hybrid by design mode and opportunity to work remotely within Poland
  • Chance to work abroad for up to 60 days annually
  • Business-driven relocation opportunities
  • We provide growth opportunities:
  • Career development programs
  • Thought leadership, mentoring, soft skills and well-being programs
  • Certification (Anthropic, Gemini, GCP, Azure, AWS)
  • English classes
  • We cover it all:
  • Stable pay
  • Participation in the Employee Stock Purchase Plan with a 15% discount
  • Benefits package (health insurance, multisport, shopping vouchers)
  • Referral bonuses up to $2,000
  • Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and more
  • Corporate, social and well-being events
  • Please, note:
  • Benefits listed above are available to employees only
  • We are open for working with Contractors. Terms of B2B cooperation agreements are agreed individually
  • We will reach out to selected candidates exclusively

EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

ITIL Service Management Expert (EU Institution) - hybrid in Poland

NRB

Warsaw

Apply on the employer's site

Role description

Job Description
Who are we?
KEYES
is a dynamic global organization that takes pride in being the trusted partner of
EU Institutions.
With strong commitment to excellence and a
30-years track record
of delivering high-quality solutions, we are dedicated to supporting the growth and success of our clients. Our Mission is to help our clients keep up with the challenges of
digital transformation
by providing the right talent at the right time for the right job. To this end, we are constantly looking for talented professionals who are interested in working on
challenging international projects
and able to deliver high-quality results within multicultural environments. Our services include (but are not limited to)
modernization of solutions, digital workspaces, cloud technologies and IT security
. Our Headquarters are in Brussels and we have active accounts and offices across Europe (i.e. Luxembourg, Amsterdam, Athens, Stockholm, Geneva).

Is this YOU?
For our customer based in
Warsaw, Poland -
an European Institution, we are looking for
ITIL Service Manager Expert based in Poland
to join a long-term mission in the area of
cybersecurity, public sector
and
law enforcement
. You will play a key role in managing business architecture, guiding solution designs, and fostering collaboration among business architects and project teams to ensure coherence with the customer’s long-term vision.

Please note that the role requires 80-100% hybrid presence at the client’s headquarters in Warsaw, meaning that being based 2-3 hours by ground transportation from Warsaw will be required.
More specifically, you will be responsible for…

  • Supporting execution of the existing IT service management processes
  • Analysing existing IT service management processes
  • Identifying risks and weak points and recommending improvements
  • Advising on the best IT service management practices
  • Preparing, documenting and implementing improvement plans
  • Supporting implementation and customisation of ITSM solution
  • Participation in projects related with implementation and customization of ITSM solution
  • Preparing required policies and Standard Operating Procedures (SOPs) supporting IT service management processes
  • Creating a positive customer experience
  • Other specific duties as assigned by the ICT SCSMT team leader.

Job Requirements
Are you the perfect match?

  • University degree (BSc/MSc)
  • At least 3 years (full time) experience at the similar position
  • At least 2 years (full time) experience in managing ICT service delivery and/or ICT operations
  • At least ITIL Intermediate certificate (ITIL Expert certificate preferred)
  • Excellent practical knowledge of IT Service Management
  • Practical experience with ITSM solutions
  • Well understanding of complex information systems and their interoperability
  • Excellent understanding and practical knowledge of IT technologies and information systems technical components
  • Familiarity with project management approaches, tools and phases of the project lifecycle Skills:
  • Excellent communication skills (in written and verbal communication)
  • Very strong sense of responsibility
  • Accuracy and attention to details
  • Very good organizational skills
  • Very good reporting skills
  • Supportive and helpful personality with co-operative and service oriented attitude
  • Forward looking with a holistic approach
  • High level of motivation and initiative.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

IT/OT Systems Engineer

ALGOTEQUE Innovation Hub

Warsawsenior

Apply on the employer's site

Role description

ALGOTEQUE is an IT consultancy firm that helps startups, mid-sized and large corporations to create and deliver innovative technologies.

Our team has a successful track record in designing, developing, implementing, and integrating software solutions (AI, ML, BI, Web, Automation) for Telecom, Energy, Bank, Insurance, Pharma, Automotive, Industry, e-commerce. We deliver our services both in fixed-price and time-and-materials models, helping our customers achieve their business and IT strategies.

Job Description
Poszukujemy Inżyniera Systemów IT/OT, który wesprze rozwój i utrzymanie infrastruktury technologicznej zakładu produkcyjnego. Konsultant będzie kluczowym ogniwem łączącym świat IT (infrastruktura, sieci, systemy) z OT (automatyka, linie produkcyjne, systemy sterowania), wspierając inicjatywy z obszaru Industry 4.0.

Zakres Projektu

  • Utrzymanie ciągłości działania systemów produkcyjnych (SCADA, HMI, MES)
  • Integracja systemów OT z infrastrukturą IT (serwery, sieci, środowiska chmurowe)
  • Wsparcie cyfryzacji procesów produkcyjnych i wdrożeń Industry 4.0
  • Diagnostyka i rozwiązywanie problemów na styku IT/OT (incydenty, wydajność)
  • Zarządzanie infrastrukturą sieciową w środowisku produkcyjnym (VLAN, segmentacja)
  • Współpraca z działami utrzymania ruchu, automatykami i zespołami IT
  • Wdrażanie i utrzymanie standardów bezpieczeństwa w środowisku OT
  • Tworzenie dokumentacji technicznej oraz procedur operacyjnych

Profile / Requirements

  • 3–5 lat doświadczenia w środowiskach IT/OT lub infrastrukturalnych
  • Praktyczna znajomość systemów przemysłowych (SCADA, MES, PLC – mile widziane)
  • Doświadczenie w pracy w środowisku produkcyjnym (fabryka / zakład przemysłowy)
  • Znajomość sieci przemysłowych i IT (TCP/IP, VLAN, routing, firewall)
  • Systemy operacyjne: Windows Server i/lub Linux
  • Umiejętność pracy w środowisku krytycznym (wysoka dostępność, uptime)
  • Język angielski min. B2

Mile Widziane

  • Znajomość standardów bezpieczeństwa OT (ISA/IEC 62443)
  • Doświadczenie z systemami MES / integracją danych produkcyjnych
  • Chmura (Azure / AWS) w kontekście zbierania i analizy danych z produkcji
  • Podstawy automatyki (PLC – Siemens, Rockwell)
  • Skrypty / automatyzacja (PowerShell, Python)

AO4344

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

126 more openings in this category and country

Staff Software Engineer, Cluster Orch (SUNK)CoreWeave · Poland

Apply on the employer's site