Skip to content
Site Reliability Engineer

Site Reliability Engineer

Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind

Krakowfulltimesenior

Apply on the employer's site

Role description

Company Description
Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities – these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.

Job Description
Project – the aim you'll have
We are the AI Experience Framework team that builds the platform powering ServiceNow's AI-first user interfaces - an SSR runtime (karuna) built on Lit and server-rendered web components, running behind a multi-tier proxy/HTTP2 routing chain with sharded V8 isolate pools, paired with a ServiceNow Glide/Java platform layer (karuna-glide) that supplies metadata, ACLs, and service artifacts. This role owns production reliability for that stack end to end: Kubernetes deployment and operations, observability, and hands-on troubleshooting of both the Node.js and JVM sides of the system - not generalist infrastructure work.

Position – How You’ll Contribute

  • Support the deployment, operation, and reliability of production services running on Kubernetes.
  • Monitor service health and investigate production incidents across distributed applications.
  • Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements.
  • Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams.
  • Support CI/CD, GitOps-based deployments, observability, and production monitoring.
  • Work within a client-directed backlog and established priorities.

Qualifications
Expectations – the experience you need

  • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering, or a closely related role, including strong recent hands-on experience supporting Kubernetes-based production services.
  • 3+ years of hands-on production Kubernetes experience strongly preferred. Kubernetes production operations, including deployment, scaling, rollout / rollback, resource tuning, and service-to-service troubleshooting
  • Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene
  • Splunk experience for log aggregation, search, and production troubleshooting
  • Prometheus and Grafana experience, specifically building alert rules and dashboards, not only using existing dashboards
  • CI/CD and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux
  • Strong Linux and networking fundamentals, including DNS, load balancing, TCP / HTTP, HTTP/2, and Kubernetes networking
  • Production troubleshooting experience across Node.js and JVM/Java services, with strong depth in at least one runtime environment. Experience may include Node.js heap snapshots, CPU profiling, event-loop and memory analysis, as well as JVM GC log analysis, thread dumps, JVM tuning, and Java service latency investigation.
  • Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication
  • Very good spoken and written English.

Additional Skills – The Edge You Have

  • Web Components / Lit experience, to perform first-level debugging of UI-related issues
  • Server-side rendering or isomorphic runtime experience
  • Canary rollout / multi-version production operations
  • Distributed tracing and request-context correlation
  • KEDA or event-driven autoscaling
  • Experience with enterprise platform integration layers

Additional Information

Our Offer – Professional Development, Personal Growth

  • Flexible employment and remote work
  • International projects with leading global clients
  • International business trips
  • Non-corporate atmosphere
  • Language classes
  • Internal & external training
  • Private healthcare and insurance
  • Multisport card
  • Well-being initiatives

Position at: Software Mind
#WORLD

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Sr. DevOps Engineer

Concentrix

Krakowfulltimesenior

Apply on the employer's site

Role description

Job Title
Sr. DevOps Engineer

Job Description
We're Concentrix. The intelligent transformation partner. Solution-focused. Tech-powered. Intelligence-fueled.

The global technology and services leader that powers the world’s best brands, today and into the future. We’re solution-focused, tech-powered, intelligence-fueled. With unique data and insights, deep industry expertise, and advanced technology solutions, we’re the intelligent transformation partner that powers a world that works, helping companies become refreshingly simple to work, interact, and transact with. We shape new game-changing careers in over 70 countries, attracting the best talent.

The Concentrix Technical Products and Services team is the driving force behind Concentrix’s transformation, data, and technology services. We integrate world-class digital engineering, creativity, and a deep understanding of human behavior to find and unlock value through tech-powered and intelligence-fueled experiences. We combine human-centered design, powerful data, and strong tech to accelerate transformation at scale. You will be surrounded by the best in the world providing market leading technology and insights to modernize and simplify the customer experience. Within our professional services team, you will deliver strategic consulting, design, advisory services, market research, and contact center analytics that deliver insights to improve outcomes and value for our clients. Hence achieving our vision.

Our game-changers around the world have devoted their careers to ensuring every relationship is exceptional. And we’re proud to be recognized with awards such as "World's Best Workplaces," “Best Companies for Career Growth,” and “Best Company Culture,” year after year.

Join us and be part of this journey towards greater opportunities and brighter futures.

Join us in scaling the infrastructure behind our flagship multilingual platform,
iX Wisdom Translate
, as we expand our
Speech-to-Speech Translation
capabilities.

We are seeking a
Senior DevOps Engineer
with deep expertise in
real-time infrastructure
,
media transport (UDP/RTP)
, and
low-latency streaming
systems.

In this role, you will architect the high-performance backbone for
AI-driven voice orchestration
, ensuring seamless integration with global telephony and
CCaaS platforms
.

You will design, automate, and maintain mission-critical, multi-region environments in
Azure
, utilizing
Terraform, Kubernetes,
and
GitHub Actions
. Your focus will be on optimizing network layers for minimal jitter/packet loss and implementing advanced observability for high-concurrency, real-time workloads.

Key Responsibilities

  • Real-Time Infrastructure Orchestration: Design and scale high-performance Azure environments using IaaC and Kubernetes to support low-latency voice and AI workloads.
  • Network & Media Optimization: Fine-tune networking layers and transport protocols (UDP/RTP) to ensure high-fidelity audio streaming with minimal jitter and packet loss.
  • Observability & Reliability Engineering: Implement advanced telemetry for real-time traffic and manage multi-region failover to maintain high availability for global CCaaS integrations.
  • Automated Delivery Systems: Architect robust CI/CD pipelines for the continuous deployment and testing of complex streaming components across globally distributed environments.

Required
Qualifications & Skills:

  • Experience: 7–10 years in DevOps or Cloud Engineering, specifically supporting real-time, low-latency, or telco-grade production environments.
  • Networking Mastery: Deep expertise in UDP, RTP, and WebRTC protocols, including network performance tuning.
  • Azure & Kubernetes: Proven track record in managing Azure native services (AKS, App Gateway) and advanced Kubernetes orchestration.
  • Automation & IaaC: Advanced proficiency in Terraform for multi-region infrastructure and building complex CI/CD pipelines via GitHub Actions or Azure DevOps.
  • Observability: Hands-on experience implementing full-stack monitoring for real-time traffic (e.g. Application Insights).

Nice To Have

  • Experience with media servers and AI inference pipelines (OpenAI Realtime).
  • Familiarity with telephony APIs or CCaaS integrations (e.g. Genesys Cloud CX, Amazon Connect, or Twilio).
  • Scripting proficiency in Python for custom automation and performance testing tools.
  • Knowledge of secure transport mechanisms such as SRTP, mTLS.

Why Join Us?

  • Pioneer Real-Time AI: Architect the infrastructure for a next-generation Speech-to-Speech translation platform, integrating cutting-edge AI orchestration with global telephony.
  • High-Impact Technical Challenges: Solve unique engineering problems at the intersection of low-latency media streaming (UDP/RTP) and large-scale cloud-native systems.
  • Greenfield Infrastructure Ownership: Shape the foundation of our global platform using a modern stack (Azure, Terraform, AKS) with a focus on automation and high availability.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Distributed Systems Engineer 4 - Data Platform Poland

Netflix

Warsawfulltime

Apply on the employer's site

Role description

Netflix is one of the world's leading entertainment services, with over 300 million paid memberships in over 190 countries enjoying TV series, films and games across a wide variety of genres and languages. Members can play, pause and resume watching as much as they want, anytime, anywhere, and can change their plans at any time.

Our Data Platform team provides centralized data platforms and tools for various business functions at Netflix, enabling them to utilize data to make critical, data-driven decisions. We focus on making it easy for our business partners to work with data efficiently, securely, and responsibly. We aspire to set the industry standard in building a world-class data infrastructure, supporting Netflix’s goal to be the most popular and widely accessible destination for global internet entertainment.

As the Data Platform charter grows, we are expanding our presence to Poland to further increase our impact. We are seeking Distributed Systems Engineers to help evolve and innovate our Product and Infrastructure. We are committed to building a diverse and inclusive team that brings new perspectives as we address the next set of challenges.

Team Spotlights
Data Platform Online Data Stores Team
This team will focus on taking end-to-end ownership of the internal data platform products, including ZooKeeper, one of the critical infrastructure products in our suite, ensuring its continued reliability, scalability and evolution. Over time, this team will expand its charter to more datastores, and to drive broader infrastructure initiatives. This expansion reinforces our mission to boost Netflix engineers’ productivity and innovation by providing a robust, efficient, and secure data platform, while fostering close collaboration between data platform teams to deliver greater organizational throughput and agility.

Experimentation Platform
This team owns the data infrastructure that makes A/B testing work at Netflix. Specifically, they build and maintain the systems that store and govern the record of who was in which experiment, across hundreds of concurrent tests running across streaming, ads, growth, and device platforms. They also own the semantic layer that ties every experiment metric at Netflix back to that allocation data, and the pipelines that keep the whole thing fresh and reliable. As the team matures, the scope grows to include self-service tooling for the teams running experiments and support for AI-driven monitoring and analysis workflows.

Data Movement
The Data Movement Connector EcoSystem at Netflix is a growing domain with an expanding charter to own all the new high-leverage business-critical connectors across Netflix. The team currently offers Frameworks, Libraries, and Control Planes and supports a large set of streaming/incremental and batch/snapshot-based connectors. We believe having the Poland team build and own a set of connectors to move and integrate data across different business domains would be the fastest way to unlock business values. The team will also build the automated systems integration test framework (for batch and streaming connectors). This offers an excellent opportunity to learn the technology stack and deliver direct, relevant impact across the organization.

This could be a great opportunity for you if you enjoy:

  • Solving business needs at scale by applying your software engineering and analytical problem-solving skills
  • Designing and building robust, scalable, and highly available distributed infrastructure
  • Leading cross-functional initiatives and collaborating with engineers, product managers, and technical program managers across teams
  • Sharing experiences with open source communities and contributing to Netflix OSS

About You

  • You have 2 or more years of experience building features or applications for large-scale distributed systems
  • You are proficient in designing and developing RESTful web services
  • You have experience building and operating scalable, fault-tolerant, distributed systems
  • You have experience with Java or other object-oriented programming languages and scripting languages such as Python
  • You are comfortable with multi-threading challenges
  • You have developed your skills through a variety of educational or professional experiences in computer science, engineering, or a related field
  • Nice to have: front-end engineering skills

A Few More Things About Us
As a team, we come from many different backgrounds and countries. Our fields of education range from the humanities to engineering to computer science, and we strive to give people the opportunity to take on different roles, should they choose to. We believe that this diversity and adaptability have helped us build an inclusive and empathetic environment, and we look forward to adding your perspective. Our culture is unique, and we tend to live by our values, so it’s worth learning more about Netflix here.

Inclusion is a Netflix value and we strive to host a meaningful interview experience for all candidates. If you want an accommodation/adjustment for a disability or any other reason during the hiring process, please send a request to your recruiting partner.

We are an equal-opportunity employer and celebrate diversity, recognizing that diversity builds stronger teams. We approach diversity and inclusion seriously and thoughtfully. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior IT Infrastructure Specialist

Nordea

Warsawfulltimesenior

Apply on the employer's site

Role description

Job ID: 2524
Welcome to Group Technology, where we pride ourselves on engineering solutions and direct Nordea’s transformation by providing a holistic technological view and structured understanding of the bank, and its surrounding environment to enable the Customer Vision and the Business Strategy.

We are looking for a full-time highly driven engineer to join our quickly developing Logging and Event Management team. If you’re interested in data ingestion and processing, building value with creative solutions and creating state-of-the-art AI solutions – don’t hesitate to apply. We are seeking an experienced Senior Full-Stack Developer with a strong foundation in frontend development, modern frontend frameworks, and DevOps practices. This is a critical role that blends deep software engineering expertise with infrastructure automation and delivery efficiency.

Nordea is a place where traditions meet tomorrow. We're not just a bank, we're a tech employer on a mission to evolve finance securely and responsibly. Together, we impact millions of people’s daily lives by ensuring they can access our solutions anytime, anywhere, while safeguarding their personal data and wealth. Join us in making an impact on the banking industry.

About Our Team
Meet the Logging and Event Management Team. Our role is to provide Nordea with state-of-an-art Logging and Search solution and by doing so bring further value to Nordea. This role is based in Finland and Poland.

What You'll Be Doing
Main responsibilities in this role:

  • Building and developing Logging and Search services centered around Grafana Enterprise Logs, Elasticsearch and in-house built solutions
  • Design, develop, and maintain robust, scalable frontend User interfaces using React Js, Node JS, spring MVC, Hibernate.
  • Build responsive and user-centric front-end automations/applications using Angular, React, or Vue.js
  • Implement CI/CD pipelines using tools like Jenkins, GitLab CI, GitHub Actions
  • Manage infrastructure using IaC tools (e.g., Terraform, Ansible) and work with container orchestration platforms like Kubernetes
  • Drive performance improvements, automations, and system architecture discussions
  • Maintain a high standard of code quality, test coverage, and documentation

Who You Are
This is the right role for you if you are a driven and curious person that wants to do work that makes a difference. It’s an engineering role, thus don’t expect much repetitive tasks – instead, we look for a team member that wants to focus on building the service with us and is not afraid to take the ownership and initiative. If you have previous technical background and experience with working with logging and monitoring tools – that will make you an even better fit!

Your background and skills include:

  • 5+ years of hands-on software development experience with strong command various Javascript frameworks.
  • Solid experience with modern JavaScript frameworks (React, node js, Angular, or Vue)
  • Strong understanding of RESTful APIs, microservices architecture, and design patterns.
  • 5 + years of experience in DevOps practices including CI/CD, containerization, Kubernetees.
  • 5 + years of experience in incident handling and application support.
  • Proficiency in Linux-based environments, scripting (Bash, Python), and Git
  • Experience with monitoring tools like Prometheus, Grafana, ELK stack is a plus
  • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field
  • Kafka, knowledge of Java Programming for Kafka SMT Development will be an additional asset.
  • Linux knowledge with good understanding of RHEL
  • Docker/Podman and Kubernetes.
  • Ansible
  • Shell Scripting.
  • Go Lang and Python skills will be an additional asset.

It would be ideal if you also:

  • Experience with logging and observability tools, i.e. Elasticsearch.
  • Previous experience with Grafana Enterprise Logs and Object Storage will be an additional asset

What We Offer
Collaboration. Ownership. Passion. Courage. These are the values that guide us in how we work and how we make decisions – and that we imagine you share with us.

People are driven by many different factors. For some, it’s to take their career to the next level. For others, it’s to break new ground within their area of expertise – in other words, with us, you will always move forward.

A culture that fosters performance and growth in one of the largest Nordic banks, offering various opportunities to evolve, develop and learn from brilliant colleagues with diverse backgrounds in a vibrant working environment.

Hybrid working model – we believe in the value of bringing people together and at the same time we embrace the freedom of flexibility.

Diversity and inclusion are a natural part of our daily work. We know that an inclusive workplace is a sustainable one. We genuinely believe that our diverse backgrounds, experiences, characteristics and traits make us stronger together. Every day we strive to find new ways to improve diversity and inclusion within our community e.g. we have signed the European Diversity Charters in the countries where we operate to show our commitment and engage with others to continue learning and improving.

If this sounds like you, get in touch!

Next steps
Submit your application no later than 15/08/2026.

The recruitment process consists of the following steps:

  • Preliminary CV selection
  • Phone conversation with the recruiter
  • Online interview with the hiring leader
  • Background check

We enable dreams and aspirations for a greater good.
We build relationships.
We add a personal touch to everything we do – when advising our customers, collaborating with colleagues, and meeting our potential candidates.

We learn and develop.
We take pride in being experts and thinking ahead. We use our expertise to meet our customers’ needs, from the simplest to the most complex. We bring a growth mindset to our work that enables us to focus on a broader perspective in our daily challenges.

We lead change.
We are responsible and aware of the impact of our decisions, both for our customers and for our local and global communities. Mindful of our responsibility towards current and future generations, we have made sustainability an integrated part of our business strategy.

We are Nordea.
We have a 200-year history of supporting and growing the Nordic economies and our values are deeply rooted in these open, progressive and collaborative societies. As one of the biggest employers in the Nordics, Poland and Estonia, you have excellent opportunities to evolve, develop and move forward with us.

Studies show that members of underrepresented communities don’t apply for jobs unless they tick all the qualification boxes. If this is part of why you hesitate to apply, we would like you to reconsider and give it a chance. Maybe your profile fits our needs much better than you think.

Only for candidates in Finland:
A security clearance will be performed for the person selected for this position.

Only for candidates in Poalnd:
Please include permit for processing personal data in CV as following:

In accordance with art. 6 (1) a and b. Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 April 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46/EC (General Data Protection Regulation) hereinafter ‘GDPR’. I agree to have: my personal data, education and employment history proceeded for the purposes of current and future recruitment processes in Nordea Bank Abp.

The administrator of your personal data is: Nordea Bank Abp operating in Poland through its Branch, address: Aleja Edwarda Rydza Śmiglego 20, 93-281 Łodź. Your personal data will be processed for the recruitment processes in Nordea Bank Abp. You have a right to access your personal data, right to rectify and right to delete. Disclosing the personal data in the scope specified by the provisions of Polish Labour Code from 26 June 1974 and executive acts are mandatory. Providing personal data is necessary to conduct the recruitment processes. The request for the deletion of your personal data means resignation from further participation in recruitment processes and causes the immediate removal of your application. Detailed information concerning processing of your personal data can be found at: https://www.nordea.com/en/doc/nordea\-privacy\-policy\-for\-applicants.pdf

We reserve the right to reply only to selected applications.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Site Reliability Engineer

Akamai Technologies

Krakowfulltimesenior

Apply on the employer's site

Role description

Job Description
Are you passionate about cutting-edge AI infrastructure?
Do you want to build your SRE career on one of the most exciting platforms in cloud computing?
Join the Akamai Inference Cloud Team
The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group. We design, implement, deploy and operate AI platforms that enable customers to run inference models and developers to create AI applications

Partner with the best
As an SRE II, responsibilities include automation, monitoring, incident response, and working collaboratively with skilled team members. Candidates should possess expertise in Linux systems, automation, and SRE practices. Daily activities involve coding, improving dashboards, enhancing alerts, and minimizing repetitive tasks. Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform.

As a Site Reliability Engineer II, you will be responsible for:

  • Building and maintaining dashboards, alerts, and monitoring for inference workloads using Akamai's existing observability platform
  • Writing automation and tooling in Python or Go to reduce operational toil and improve system reliability
  • Participating in on-call rotations, responding to production incidents, and contributing to post-mortem analysis
  • Building and improving runbooks for inference-specific operational procedures, integrating into Akamai's existing incident management processes
  • Contributing to SLO tracking and reporting, identifying trends and areas for improvement
  • Supporting CI/CD pipeline maintenance, deployment safety checks, and rollback procedures
  • Collaborating with product engineering teams to troubleshoot complex problems across the stack

Do What You Love
To be successful in this role you will:

  • Have commercial experience in Site Reliability Engineering
  • Show proficiency in a programming language such as Python or Go, with experience creating automation solutions.
  • Have experience with Linux systems administration and the ability to troubleshoot complex infrastructure issues
  • Show familiarity with Kubernetes and containerization concepts
  • Have experience with monitoring and observability tools such as Prometheus, Grafana, or similar
  • Have exposure to CI/CD pipelines and infrastructure-as-code tools (Terraform, SaltStack, or equivalent)
  • Show a willingness to learn and grow, with genuine curiosity about AI infrastructure and distributed systems

Work in a way that works for you
FlexBase, Akamai's Global Flexible Working Program, is based on the principles that are helping us create the best workplace in the world. When our colleagues said that flexible working was important to them, we listened. We also know flexible working is important to many of the incredible people considering joining Akamai. FlexBase, gives 95% of employees the choice to work from their home, their office, or both (in the country advertised). This permanent workplace flexibility program is consistent and fair globally, to help us find incredible talent, virtually anywhere. We are happy to discuss working options for this role and encourage you to speak with your recruiter in more detail when you apply.

Learn what makes Akamai a great place to work

Connect with us on social and see what life at Akamai is like!

We power and protect life online, by solving the toughest challenges, together.
At Akamai, we're curious, innovative, collaborative and tenacious. We celebrate diversity of thought and we hold an unwavering belief that we can make a meaningful difference. Our teams use their global perspectives to put customers at the forefront of everything they do, so if you are people-centric, you'll thrive here.

Working for you
Benefits
At Akamai, we will provide you with opportunities to grow, flourish, and achieve great things. Our benefit options are designed to meet your individual needs for today and in the future. We provide benefits surrounding all aspects of your life:

  • Your health
  • Your finances
  • Your family
  • Your time at work
  • Your time pursuing other endeavors

Our benefit plan options are designed to meet your individual needs and budget, both today and in the future.

About Us
Akamai powers and protects life online. Leading companies worldwide choose Akamai to build, deliver, and secure their digital experiences helping billions of people live, work, and play every day. With the world's most distributed compute platform from cloud to edge we make it easy for customers to develop and run applications, while we keep experiences closer to users and threats farther away.

Join us
Are you seeking an opportunity to make a real difference in a company with a global reach and exciting services and clients? Come join us and grow with a team of people who will energize and inspire you!

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

109 more openings in this category and country

Senior Site Reliability Engineer (SRE) – KubernetesSoftware Mind · Poland

Apply on the employer's site