Skip to content
Site Reliability Engineer

Site Reliability Engineer

Senior Software Engineer, Fleet Monitoring Analysis

CoreWeave

Warsawfulltimesenior

Откликнуться на сайте работодателя

Описание вакансии

CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.

We're proud to be a Living Wage accredited Employer.

What You’ll Do
The Fleet Monitoring & Analysis (FMA) team within Fleet Engineering builds and operates the metrics, monitoring, and alerting systems that power CoreWeave’s automated provisioning and lifecycle management of our global hardware fleet. As the team behind many of the exporters, dashboards, Slack bots, and alerting rules used across Fleet Engineering and Operations, FMA plays a central role in enabling zero-touch, high-reliability operations for GPU servers and the environments they run in.

About The Role
As an Engineer on the Fleet Monitoring & Analysis team, you’ll help build, run, and refine the metrics, alerts, visualizations, and data-driven insights that keep CoreWeave’s ever-expanding fleet of hardware nodes healthy and observable.

You’ll join a mixed-skill engineering team focused on elevating the art of managing high-performance hardware at scale—partnering closely with Fleet Engineering and Operations, and our observability platform teams to turn telemetry into automation, operational efficiency, and a better experience for CoreWeave’s customers.

In This Role, You Will

  • Design and implement large-scale server observability solutions that improve the stability and reliability of CoreWeave’s global hardware fleet.
  • Adapt, extend, and implement open-source monitoring and alerting tooling (e.g., Prometheus-compatible exporters, AlertManager/Victoria Metrics) to deepen our visibility into fleet and environmental health.
  • Generate and maintain tailored reports, alarms, and visualizations used by Fleet Engineering and FROps to understand, respond to, and plan for fleet growth and change.
  • Create and evolve test plans, deployment automation, dashboards, alerts, and insights around fleet operations, and participate in the Fleet Engineering Developers’ on-call rotation.
  • Collaborate with teammates across FMA to invest in each other’s growth, share ideas, and continuously improve how we monitor and automate CoreWeave’s infrastructure.

Who You Are

  • 2+ years of experience in a software or infrastructure engineering role in industry.
  • Experience with automation and orchestration workflows, and familiarity with server hardware, components, and strategies for managing physical infrastructure at scale.
  • Experience implementing metrics collection and alerting on standard monitoring platforms (for example, Prometheus/Victoria Metrics, AlertManager, Grafana, or similar observability stacks).
  • Proficiency working in Linux-based environments and using at least one scripting or programming language (such as Python, Go, or similar) to build tooling and automation.

Preferred

  • Experience designing or operating time-series monitoring at scale using Prometheus-compatible systems, Victoria Metrics, and related observability tooling (e.g., Grafana, AlertManager).
  • Experience deploying and maintaining application services in Kubernetes.
  • Experience building or maintaining Slack bots, webhooks, or automation that consume alerts and drive lifecycle actions across a fleet (for example, via Kubernetes controllers, Ansible, or similar tools).
  • Experience with data warehousing, SQL, and building reporting pipelines or dashboards for operational analytics in environments like Grafana or similar BI tools.

Wondering if you’re a good fit?

We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk.

  • You enjoy digging into metrics, logs, and alerts to understand how large-scale systems behave over time.
  • You’re excited to collaborate with operations and engineering teams to turn manual runbooks into automation and improve fleet reliability.
  • You like working at the intersection of hardware, software, and data—using observability to make complex infrastructure easier to operate.

The base salary range for this role is zł321,000 to zł428,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).

To fulfill our obligation to protect client data, successful applicants offered employment with CoreWeave will be required to complete a basic criminal record check, conducted in compliance with GDPR. Employment offers are conditional upon receiving satisfactory check results
What We Offer
In addition to a competitive salary, we offer a variety of benefits to support your needs, including:

  • Family-level Medical Insurance
  • Family-level Dental Insurance
  • Generous Pension Contribution
  • Life Assurance at 4x Salary
  • Critical Illness Cover
  • Employee Assistance Programme
  • Tuition Reimbursement
  • Work culture focused on innovative disruption

Benefits may vary by location.
Equal Opportunity
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
Recruitment Agencies
CoreWeave does not accept speculative CVs. Any unsolicited CVs received will be treated as the property of CoreWeave and your Terms & Conditions associated with the use of CVs will be considered null and void.
Any unsolicited CVs sent by your company to us – that is to say, in any situation where we have not directly engaged your company in writing to supply candidates for a specific vacancy – will be considered by us to be a “free gift”, leaving us liable for no fees whatsoever should we choose to contact the candidate directly and engage the candidate’s services, and will in no way establish any prior claim by your company to representation of that candidate should the candidate’s details also be submitted by any other party.
Export Control Compliance
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C.

  • 1157, or (iv) asylee under 8 U.S.C.
  • 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.

Updated privacy notice - UK and EU Job Applications
When you apply to a job on this site, the personal data contained in your application will be collected by CoreWeave UK Ltd. (“Controller”), which is located at
Phosphor (6th Floor), 133 Park Street, London, SE1 9EA
and can be contacted by emailing
careers.eu@coreweave.com
. Controller’s data protection officer can be contacted at
privacy@coreweave.com
. Your personal data will be processed for the purposes of managing Controller’s recruitment related activities, which include setting up and conducting interviews and tests for applicants, evaluating and assessing the results thereto, and as is otherwise needed in the recruitment and hiring processes. Such processing is legally permissible under Art. 6(1)(f) of (i) Regulation (EU) 2016/679 (General Data Protection Regulation (“GDPR”) and (ii) the GDPR as it forms part of the laws of the UK (“UK GDPR”), as necessary for the purposes of the legitimate interests pursued by the Controller, which are the solicitation, evaluation, and selection of applicants for employment. Your personal data will be shared with Greenhouse Software, Inc., a cloud services provider located in the United States of America and engaged by Controller to help manage its recruitment and hiring process on Controller’s behalf. With respect to transfers originating from the UK or the European Economic Area ("EEA") to a country outside the UK or the EEA, we implement the appropriate transfer mechanism(s) and other appropriate solutions to address cross-border transfers as required by applicable law. You may request a copy of the suitable mechanisms we have in place by contacting us at
privacy@coreweave.com
Your personal data will be retained by Controller as long as Controller determines it is necessary to evaluate your application for employment. Where permitted by applicable law, we may also retain your personal data for a limited period after the recruitment process ends in order to consider you for future job opportunities, respond to legal claims, or comply with record-keeping obligations. Under the GDPR and the UK GDPR, you have the right to request access to your personal data, to request that your personal data be rectified or erased, and to request that processing of your personal data be restricted. You also have the right to data portability. In addition, you may lodge a complaint with the relevant supervisory authority: (i) A list of Europe’s data protection authorities can be found
here
; and (ii) for the UK, this is the
Information Commissioner's Office
.
For additional information, please see our
Privacy Policy
.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Скоро на этой странице

Резюме под эту вакансию — и билет в розыгрыш

Мы разбираем объявление до настоящих требований и переписываем ваше резюме под него — вопросами, а не выдумкой: ни одна строка не появится без вашего подтверждения. Войдите, чтобы получить это первым, — и попасть в розыгрыш.

  • Резюме под конкретную вакансию, а не «универсальное»
  • Ответы хранятся: правится любой, а не весь разговор заново
  • Всё в аккаунте — открывается с любого устройства

Разыгрываем

Скидка на сопровождение

Победителей выбираем случайно среди заявок с подтверждённой почтой. Дата розыгрыша и полные правила — на странице розыгрыша.

Правила розыгрыша

Site Reliability Engineer

Senior Cloud Engineer (Blockchain) - Digital Assets

EPAM Systems

Krakowcontractsenior

Откликнуться на сайте работодателя

Описание вакансии

We are looking for a
Senior Cloud Engineer
to join our Digital Assets Crew, working on pioneering blockchain-based projects including tokenization of financial assets. Our team operates within two Scrum teams and builds cloud-hosted applications that leverage cutting-edge technologies to deliver value and innovation to the finance industry

Responsibilities

  • Develop cloud solutions that scale and ensure high availability
  • Nurture a DevOps culture across the team
  • Contribute to DLT-related projects, particularly blockchain
  • Design and maintain infrastructure using Terraform and Helm
  • Implement GitOps and CI/CD pipelines
  • Monitor distributed systems using Grafana and Prometheus
  • Manage secret management solutions across cloud environments
  • Collaborate with colleagues in English on cross-functional initiatives

Requirements

  • 5+ years of experience in DevOps/SRE roles
  • Expertise in Azure Cloud and Kubernetes
  • Proficiency in Terraform, Helm and GitOps/CI-CD
  • Experience with Grafana, Prometheus and distributed systems architecture at enterprise scale
  • Knowledge of secret management solutions
  • Good communication skills, comfortable interacting with colleagues in English (B2+)

Nice to have

  • Experience with Istio Service Mesh and zero-trust architecture
  • Background in banking or regulated environments
  • Familiarity with blockchain technology

We offer

  • We gather like-minded people:
  • Top tech minds driving innovation in AI, cloud and digital platform modernization
  • Supportive team and agile, startup-like culture
  • Hybrid by design mode and opportunity to work remotely within Poland
  • Chance to work abroad for up to 60 days annually
  • Business-driven relocation opportunities
  • We provide growth opportunities:
  • Career development programs
  • Thought leadership, mentoring, soft skills and well-being programs
  • Certification (Anthropic, Gemini, GCP, Azure, AWS)
  • English classes
  • We cover it all:
  • Stable pay
  • Participation in the Employee Stock Purchase Plan with a 15% discount
  • Benefits package (health insurance, multisport, shopping vouchers)
  • Referral bonuses up to $2,000
  • Offices featuring entertainment and relaxation zones, table tennis and football, free snacks, coffee and more
  • Corporate, social and well-being events
  • Please, note:
  • Benefits listed above are available to employees only
  • We are open for working with Contractors. Terms of B2B cooperation agreements are agreed individually
  • We will reach out to selected candidates exclusively

EPAM is global leader in AI transformation engineering and integrated consulting, serving Forbes Global 2000 companies and ambitious startups. With over thirty years of expertise in custom software, product and platform engineering, we empower our clients to become AI-Native enterprises, driving measurable value from innovation and digital investments.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

ITIL Service Management Expert (EU Institution) - hybrid in Poland

NRB

Warsaw

Откликнуться на сайте работодателя

Описание вакансии

Job Description
Who are we?
KEYES
is a dynamic global organization that takes pride in being the trusted partner of
EU Institutions.
With strong commitment to excellence and a
30-years track record
of delivering high-quality solutions, we are dedicated to supporting the growth and success of our clients. Our Mission is to help our clients keep up with the challenges of
digital transformation
by providing the right talent at the right time for the right job. To this end, we are constantly looking for talented professionals who are interested in working on
challenging international projects
and able to deliver high-quality results within multicultural environments. Our services include (but are not limited to)
modernization of solutions, digital workspaces, cloud technologies and IT security
. Our Headquarters are in Brussels and we have active accounts and offices across Europe (i.e. Luxembourg, Amsterdam, Athens, Stockholm, Geneva).

Is this YOU?
For our customer based in
Warsaw, Poland -
an European Institution, we are looking for
ITIL Service Manager Expert based in Poland
to join a long-term mission in the area of
cybersecurity, public sector
and
law enforcement
. You will play a key role in managing business architecture, guiding solution designs, and fostering collaboration among business architects and project teams to ensure coherence with the customer’s long-term vision.

Please note that the role requires 80-100% hybrid presence at the client’s headquarters in Warsaw, meaning that being based 2-3 hours by ground transportation from Warsaw will be required.
More specifically, you will be responsible for…

  • Supporting execution of the existing IT service management processes
  • Analysing existing IT service management processes
  • Identifying risks and weak points and recommending improvements
  • Advising on the best IT service management practices
  • Preparing, documenting and implementing improvement plans
  • Supporting implementation and customisation of ITSM solution
  • Participation in projects related with implementation and customization of ITSM solution
  • Preparing required policies and Standard Operating Procedures (SOPs) supporting IT service management processes
  • Creating a positive customer experience
  • Other specific duties as assigned by the ICT SCSMT team leader.

Job Requirements
Are you the perfect match?

  • University degree (BSc/MSc)
  • At least 3 years (full time) experience at the similar position
  • At least 2 years (full time) experience in managing ICT service delivery and/or ICT operations
  • At least ITIL Intermediate certificate (ITIL Expert certificate preferred)
  • Excellent practical knowledge of IT Service Management
  • Practical experience with ITSM solutions
  • Well understanding of complex information systems and their interoperability
  • Excellent understanding and practical knowledge of IT technologies and information systems technical components
  • Familiarity with project management approaches, tools and phases of the project lifecycle Skills:
  • Excellent communication skills (in written and verbal communication)
  • Very strong sense of responsibility
  • Accuracy and attention to details
  • Very good organizational skills
  • Very good reporting skills
  • Supportive and helpful personality with co-operative and service oriented attitude
  • Forward looking with a holistic approach
  • High level of motivation and initiative.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

IT/OT Systems Engineer

ALGOTEQUE Innovation Hub

Warsawsenior

Откликнуться на сайте работодателя

Описание вакансии

ALGOTEQUE is an IT consultancy firm that helps startups, mid-sized and large corporations to create and deliver innovative technologies.

Our team has a successful track record in designing, developing, implementing, and integrating software solutions (AI, ML, BI, Web, Automation) for Telecom, Energy, Bank, Insurance, Pharma, Automotive, Industry, e-commerce. We deliver our services both in fixed-price and time-and-materials models, helping our customers achieve their business and IT strategies.

Job Description
Poszukujemy Inżyniera Systemów IT/OT, który wesprze rozwój i utrzymanie infrastruktury technologicznej zakładu produkcyjnego. Konsultant będzie kluczowym ogniwem łączącym świat IT (infrastruktura, sieci, systemy) z OT (automatyka, linie produkcyjne, systemy sterowania), wspierając inicjatywy z obszaru Industry 4.0.

Zakres Projektu

  • Utrzymanie ciągłości działania systemów produkcyjnych (SCADA, HMI, MES)
  • Integracja systemów OT z infrastrukturą IT (serwery, sieci, środowiska chmurowe)
  • Wsparcie cyfryzacji procesów produkcyjnych i wdrożeń Industry 4.0
  • Diagnostyka i rozwiązywanie problemów na styku IT/OT (incydenty, wydajność)
  • Zarządzanie infrastrukturą sieciową w środowisku produkcyjnym (VLAN, segmentacja)
  • Współpraca z działami utrzymania ruchu, automatykami i zespołami IT
  • Wdrażanie i utrzymanie standardów bezpieczeństwa w środowisku OT
  • Tworzenie dokumentacji technicznej oraz procedur operacyjnych

Profile / Requirements

  • 3–5 lat doświadczenia w środowiskach IT/OT lub infrastrukturalnych
  • Praktyczna znajomość systemów przemysłowych (SCADA, MES, PLC – mile widziane)
  • Doświadczenie w pracy w środowisku produkcyjnym (fabryka / zakład przemysłowy)
  • Znajomość sieci przemysłowych i IT (TCP/IP, VLAN, routing, firewall)
  • Systemy operacyjne: Windows Server i/lub Linux
  • Umiejętność pracy w środowisku krytycznym (wysoka dostępność, uptime)
  • Język angielski min. B2

Mile Widziane

  • Znajomość standardów bezpieczeństwa OT (ISA/IEC 62443)
  • Doświadczenie z systemami MES / integracją danych produkcyjnych
  • Chmura (Azure / AWS) w kontekście zbierania i analizy danych z produkcji
  • Podstawy automatyki (PLC – Siemens, Rockwell)
  • Skrypty / automatyzacja (PowerShell, Python)

AO4344

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Administratorka/Administrator aplikacji

PKO Bank Polski

Warsaw

Откликнуться на сайте работодателя

Описание вакансии

Na Co Dzień w Naszym Zespole

  • administrujemy aplikacjami z obszarów zarządzania tożsamością, HR oraz gospodarki własnej Banku,
  • utrzymujemy dostępność aplikacji administrowanych przez nasz zespół poprzez ciągłe monitorowanie zasobów jak i procesów, wgrywanie zmian oraz poprawek aplikacyjnych,
  • realizujemy usługi informatyczne dla jednostek biznesowych banku w zakresie eksploatowanych aplikacji,
  • rozwiązujemy incydenty/problemy zgłaszane przez jednostki banku z zakresu działania administrowanych aplikacji,
  • prowadzimy dokumentację oraz współpracujesz z dostawcami oprogramowania w zakresie obsługi błędów i problemów eksploatacyjnych,
  • bierzemy udział w pracach zespołów projektowych wdrażających usługi IT,
  • współpracujemy z administratorami zasobów informatycznych w zakresie eksploatowanych aplikacji,
  • testujemy i wdrażamy poprawki serwisowe i nowe wersje aplikacji,
  • bierzemy udział w szkoleniach i rozwoju w obszarze zarządzania tożsamością.

To Stanowisko Może Być Twoje, Jeśli

  • masz wiedzę z zakresu dostępnych narzędzi IDM na rynku, najlepiej znajomość OneIdentity Management,
  • masz wiedzę z zakresu zarządzania tożsamością użytkownika, proces J/M/L- doświadczenie w pracy z serwerami działającymi pod kontrolą systemów operacyjnych Microsoft Serwer, Linux (RHEL),
  • posiadasz doświadczenie z narzędziami z obszaru komponentów aplikacyjnych i sieciowych, takimi jak IIS, Windows Services, Certyfikaty SHA, Apache Tomcat, WebServices,
  • masz doświadczenie z bazami danych MS SQL\PostgreSQL w tym umiejętność pisania skryptów SQL, PLSQL,
  • masz doświadczenie w pracy z powłoką shell systemów Linux, pisanie skryptów, edytory,
  • masz doświadczenie w pracy z PowerShell w tym pisanie skryptów,
  • posiadasz umiejętność analizy logów systemowo/aplikacyjnych i na tej podstawie działań proaktywnych w celu eliminacji błędów,
  • znasz język angielski w stopniu umożliwiającym swobodne posługiwanie się dokumentacją techniczną.

Mile Widziane

  • znajomość narzędzi CI/CD: Jenkins, Docker Compose, GITLab etc.
  • podstawowa wiedza na temat technologii chmurowych (Microsoft Azure, Google Cloud).

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Ещё 121 вакансия по этой категории в этой стране

Senior Software Engineer, Fleet Monitoring AnalysisCoreWeave · Poland

Откликнуться на сайте работодателя