Skip to content
Site Reliability Engineer

The week's list

Every role like this one, in one letter

You are reading one posting. There are hundreds like it on the board, and new ones every week. Pick what you want, leave an email, and the list comes to you — no searching, no coming back here.

Counting what came out this past week…

The first letter arrives right away, then one a week. Unsubscribe in one click from any letter — the address goes nowhere else.

Site Reliability Engineer

AI Platform Engineer, Agent Systems IRC296274

GlobalLogic

Krakow

Apply on the employer's site

Role description

Description
The company is building the agent platform for professional music production: the orchestration layer, tool interfaces, skills runtime, and context architecture that allow any AI agent to reason about and act on a music-production workflow.

You will lead the design of the orchestration loop, define how the engine’s capabilities are exposed to models, build the skills runtime that transforms a general-purpose model into a domain specialist, and architect the context and memory systems that keep agents coherent across long creative sessions.

The object model is a song. The users are producers, musicians, and creatives. The domain has real-time constraints, deep semantics, and no existing playbook.

Requirements

  • Five or more years shipping production platform or infrastructure software that other engineers have built on top of.
  • Eighteen or more months of production experience building LLM agent systems, covering orchestration loops, tool use, and context management. We have no preference for a specific framework. We are equally interested in engineers who shipped on a provider-agnostic framework such as LangGraph and engineers who rejected frameworks entirely and built their own harness, provided you can articulate what you learned from the path you took.
  • Demonstrated experience designing tool interfaces for LLM consumption. You can explain what makes a tool schema discoverable and usable by a model versus merely technically correct.
  • Demonstrated experience building context, memory, or state-management systems beyond framework defaults, including compaction, durable memory, or session persistence. You have diagnosed agent failures from raw execution traces and made targeted harness changes in response.
  • Strong proficiency in TypeScript and Python.
  • Experience with the Model Context Protocol (MCP) or similar tool-connectivity standards.

Nice To Have (not Required)

  • Background in music production, audio engineering, or another creative-tool domain, including as a serious hobbyist.
  • Experience with real-time audio systems, professional audio software, or other latency-sensitive environments.
  • Experience making a complex desktop or professional application agent-accessible, in any domain with a rich object model (DAW, IDE, design tool, CAD).
  • Experience building middleware or hook architectures that allow others to customize agent behavior without modifying core code.

Job responsibilities

What You Will Own

  • Tool interfaces. Define how the engine’s capabilities are exposed to LLMs as structured, discoverable tools. This includes schemas, semantic descriptions, scoped tool sets, input validation, and output parsing that a model can reliably produce and the harness can reliably consume. Designing a tool surface that models use well is a distinct discipline from designing an API for human developers, and you will own that discipline.
  • Orchestration and control flow. Design and build the harness: the core loop and the machinery around it. This covers step sequencing, retries, timeouts, error recovery, fallback paths, and multi-agent coordination where a workflow is split across sub-agents with their own tools and context. You will evaluate whether to build this in-house, adopt a framework, or extend an existing one. We have no commitment to any specific framework, and we will not build the platform on top of a single provider or model. A well-reasoned argument for building our own harness is a welcome outcome of that evaluation.
  • Skills runtime. Design the format, packaging, loading, and execution layer for the structured domain knowledge that turns a generic model into a music-production specialist. This is our most distinctive platform primitive and it is largely greenfield.
  • Context, memory, and state. Build the systems that keep agents performant and coherent across long, multi-step creative workflows. This includes context compaction, short-term working memory, durable cross-session memory, session state persistence, continuity across disconnects, and sub-agent delegation in which parent and child contexts remain consistent.
  • Extension points. Design the harness so that new tools, skills, and middleware can be added without modifying the core runtime. Extensibility is an architectural property of the system, not a retrofit.
  • Evaluations, observability, and failure analysis. Evaluations tell us the harness is working; raw execution traces and structured failure logs tell us why it is not. You will build and own the platform-level evaluation surface, the observability that every engineer on the platform depends on, and the feedback loop that converts failed agent runs into targeted harness changes.
  • Ongoing simplification. As frontier models improve, some of the scaffolding we build today will stop earning its keep. You will audit the harness on a regular basis and remove the components that models no longer require.

This Role Is Not

  • LLM integration engineering. This role is not responsible for wiring models to the DAW or building end-user AI features. This role builds the platform those features run on.
  • ML or model engineering. This team does not train models. It builds the systems that agents run on.
  • Research. This team applies current research in production. Original research happens elsewhere in the company.

What we offer

Empowering Projects:
With 500+ clients spanning diverse industries and domains, we provide an exciting opportunity to contribute to groundbreaking projects that leverage cutting-edge technologies. As a team, we engineer digital products that positively impact people’s lives.

Empowering Growth:
We foster a culture of continuous learning and professional development. Our dedication is to provide timely and comprehensive assistance for every consultant through our dedicated Learning & Development team, ensuring their continuous growth and success.

DE&I Matters:
At GlobalLogic, we deeply value and embrace
diversity
. We are dedicated to providing
equal
opportunities for all individuals, fostering an
inclusive
and empowering work environment.

Career Development:
Our corporate culture places a strong emphasis on career development, offering abundant opportunities for growth. Regular interactions with our teams ensure their engagement, motivation, and recognition. We empower our team members to pursue their career goals with confidence and enthusiasm.

Comprehensive Benefits:
In addition to equitable compensation, we provide a comprehensive benefits package that prioritizes the overall well-being of our consultants. We genuinely care about their health and strive to create a positive work environment.

Flexible Opportunities:
At GlobalLogic, we prioritize work-life balance by offering flexible opportunities tailored to your lifestyle. Explore relocation and rotation options for diverse cultural and professional experiences in different countries with our company.

About GlobalLogic
GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Coming to this page

A resume for this role — and a ticket to the draw

We take the posting apart down to the real requirements and rewrite your resume against it — by asking, not inventing: no line appears without your confirmation. Sign in to get it first, and to enter the draw.

  • A resume for this exact role, not a universal one
  • Answers are kept: edit any one, not the whole conversation
  • All in your account — open it from any device

On the wheel

A discount on mentoring

Winners are drawn at random among entries with a confirmed email. The date and the full rules are on the draw page.

Draw rules

Site Reliability Engineer

Software Engineer - Platform Services

Snowflake

Warsawfulltime

Apply on the employer's site

Role description

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done.

Snowflake’s unique architecture allows us to maintain an agile delivery cadence while still providing a very high degree of stability and reliability for customers.

Our intent is to provide a highly available, reliable and scalable tools/services infrastructure that enables our engineering teams to develop, test, debug and release enterprise software quickly. We are champions for code health, testability, maintainability and best practices for development and testing.

At Snowflake our goal is to make each individual feel valued for his or her contributions to the company’s mission. We are looking for smart people who want to do remarkable things. We strive to create an environment of casual intensity where people enjoy coming to work every day.

You will be building infrastructure and automation frameworks for Cloud-based and SaaS Data Warehouse services. These span the full stack including helping set the direction for how we continuously integrate, deploy, verify and monitor our product/services. You will be driving the development of testing infrastructure, automation frameworks, and tools to have robust automated testing pipelines for the Snowflake Data Cloud. This is an awesome opportunity to work with cutting-edge cloud technology in a highly visible role.

AS AN INFRASTRUCTURE AND AUTOMATION ENGINEER AT SNOWFLAKE YOU WILL:

  • Lead/contribute to engineering efforts from planning and organization to execution and delivery to solve complex engineering problems in tools and testing.
  • Define and maintain policies around developer, test, and validation deployments.
  • Design and build advanced CI/CD pipeline frameworks.
  • Design and build tooling and infrastructure to help engineering teams measure and increase their velocity.
  • Drive adoption of best practices in code health, testing, and maintainability.
  • Analyze and decompose complex software systems and collaborate with and influence others to improve the overall design.
  • Collaborate with other teams on effective test coverage, software feature management, and spin-up of resources as required.

ON DAY ONE WE WILL EXPECT YOU TO HAVE:

  • At least 2+ years of experience in software development (SaaS experience preferred).
  • Prefer strong coding skills in Python/Java/C++, NodeJS and other software technologies.
  • Hands-on experience designing and working with modern CI/CD solutions.
  • Comfortable with open systems environments and scripting experience.
  • Experience with Cloud-based infrastructure systems is a plus. (AWS, Azure, GCP).
  • Attention to detail and ability to build reliable and scalable software systems.
  • Effective communication and collaboration skills with a service-oriented mindset.
  • Solid interpersonal skills that are conducive to a team environment.
  • Ability to manage and prioritize multiple requests for competing resources.
  • Able to debug, troubleshoot, and resolve complex technical issues.
  • Strong work ethic and a passion for problem-solving with a self-driven & motivated mindset
  • Experience and knowledge of Git, JIRA, Jenkins, and Snowflake a plus
  • Kubernetes and Docker experience is a plus.
  • Strong database understanding including SQL is a plus

WHY JOIN THE ENGINEERING TEAM AT SNOWFLAKE? AS A MEMBER OF OUR TEAM, YOU WILL :

  • Build an industry-leading data management system that customers love.
  • Measurably impact an innovative product area central to Snowflake’s success.
  • Take charge of your own career- this role has the impact and ability to grow both technically, as well as from a leadership perspective.
  • Ensure the quality, performance, and reliability of a super-robust and secure enterprise SaaS platform that services hundreds of customers and millions of complex queries daily.
  • Learn at scale as you work on a highly scalable and reliable data processing platform that runs on hundreds and thousands of machines and executes Billions of queries.
  • Ensure that we are shipping the highest quality service possible at each release.

Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

How do you want to make your impact?

For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com http://careers.snowflake.com

Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake.

How do you want to make your impact?

For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Sr. DevOps Engineer

Concentrix

Krakowfulltimesenior

Apply on the employer's site

Role description

Job Title
Sr. DevOps Engineer

Job Description
We're Concentrix. The intelligent transformation partner. Solution-focused. Tech-powered. Intelligence-fueled.

The global technology and services leader that powers the world’s best brands, today and into the future. We’re solution-focused, tech-powered, intelligence-fueled. With unique data and insights, deep industry expertise, and advanced technology solutions, we’re the intelligent transformation partner that powers a world that works, helping companies become refreshingly simple to work, interact, and transact with. We shape new game-changing careers in over 70 countries, attracting the best talent.

The Concentrix Technical Products and Services team is the driving force behind Concentrix’s transformation, data, and technology services. We integrate world-class digital engineering, creativity, and a deep understanding of human behavior to find and unlock value through tech-powered and intelligence-fueled experiences. We combine human-centered design, powerful data, and strong tech to accelerate transformation at scale. You will be surrounded by the best in the world providing market leading technology and insights to modernize and simplify the customer experience. Within our professional services team, you will deliver strategic consulting, design, advisory services, market research, and contact center analytics that deliver insights to improve outcomes and value for our clients. Hence achieving our vision.

Our game-changers around the world have devoted their careers to ensuring every relationship is exceptional. And we’re proud to be recognized with awards such as "World's Best Workplaces," “Best Companies for Career Growth,” and “Best Company Culture,” year after year.

Join us and be part of this journey towards greater opportunities and brighter futures.

Join us in scaling the infrastructure behind our flagship multilingual platform,
iX Wisdom Translate
, as we expand our
Speech-to-Speech Translation
capabilities.

We are seeking a
Senior DevOps Engineer
with deep expertise in
real-time infrastructure
,
media transport (UDP/RTP)
, and
low-latency streaming
systems.

In this role, you will architect the high-performance backbone for
AI-driven voice orchestration
, ensuring seamless integration with global telephony and
CCaaS platforms
.

You will design, automate, and maintain mission-critical, multi-region environments in
Azure
, utilizing
Terraform, Kubernetes,
and
GitHub Actions
. Your focus will be on optimizing network layers for minimal jitter/packet loss and implementing advanced observability for high-concurrency, real-time workloads.

Key Responsibilities

  • Real-Time Infrastructure Orchestration: Design and scale high-performance Azure environments using IaaC and Kubernetes to support low-latency voice and AI workloads.
  • Network & Media Optimization: Fine-tune networking layers and transport protocols (UDP/RTP) to ensure high-fidelity audio streaming with minimal jitter and packet loss.
  • Observability & Reliability Engineering: Implement advanced telemetry for real-time traffic and manage multi-region failover to maintain high availability for global CCaaS integrations.
  • Automated Delivery Systems: Architect robust CI/CD pipelines for the continuous deployment and testing of complex streaming components across globally distributed environments.

Required
Qualifications & Skills:

  • Experience: 7–10 years in DevOps or Cloud Engineering, specifically supporting real-time, low-latency, or telco-grade production environments.
  • Networking Mastery: Deep expertise in UDP, RTP, and WebRTC protocols, including network performance tuning.
  • Azure & Kubernetes: Proven track record in managing Azure native services (AKS, App Gateway) and advanced Kubernetes orchestration.
  • Automation & IaaC: Advanced proficiency in Terraform for multi-region infrastructure and building complex CI/CD pipelines via GitHub Actions or Azure DevOps.
  • Observability: Hands-on experience implementing full-stack monitoring for real-time traffic (e.g. Application Insights).

Nice To Have

  • Experience with media servers and AI inference pipelines (OpenAI Realtime).
  • Familiarity with telephony APIs or CCaaS integrations (e.g. Genesys Cloud CX, Amazon Connect, or Twilio).
  • Scripting proficiency in Python for custom automation and performance testing tools.
  • Knowledge of secure transport mechanisms such as SRTP, mTLS.

Why Join Us?

  • Pioneer Real-Time AI: Architect the infrastructure for a next-generation Speech-to-Speech translation platform, integrating cutting-edge AI orchestration with global telephony.
  • High-Impact Technical Challenges: Solve unique engineering problems at the intersection of low-latency media streaming (UDP/RTP) and large-scale cloud-native systems.
  • Greenfield Infrastructure Ownership: Shape the foundation of our global platform using a modern stack (Azure, Terraform, AKS) with a focus on automation and high availability.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Software Test Engineer - Manual

Motorola Solutions

Krakowfulltime

Apply on the employer's site

Role description

Company Overview
At Motorola Solutions, we believe that everything starts with our people. We’re a global close-knit community, united by the relentless pursuit to help keep people safer everywhere. We build and connect technologies to help protect people, property and places. Our solutions foster the collaboration that’s critical for safer communities, safer schools, safer hospitals, safer businesses, and ultimately, safer nations. Connect with a career that matters, and help us build a safer future.

Department Overview
Computer Aided Dispatch center is an essential part of advanced group communication solutions. Dispatching is a central functionality of such a center. Dispatch and Content Logging team is responsible for creating integrated dispatch solutions. The core functionality is delivering audio communication between dispatchers and radio users in the public safety industry - including fire rescue, 911 call takers, police, army and many others. Everywhere reliable, real-time and secure communication is required. On top of the audio services, a variety of features help operators improve their efficiency and effectiveness - like call logging, instant recording, text communication.

See the functionalities of our product : https://www.youtube.com/watch?v\=5PSd7RBb8Ks

Job Description
We are looking for an Manual Tester who will join our software validation team. The team defines test strategy, maintains test environment, performs and executes tests scenarios, defects submission and fixes verification. Manual Tester diagnoses and identifies system integration issues, designs and validates solutions for problems and looks for improvements in daily jobs. An important part of work is knowledge and experience sharing with other team members (test and software engineers).

Basic Requirements

  • software testing methodologies
  • Practical OS knowledge, user perspective (preferable both Windows and Linux or at least Windows)
  • Good troubleshooting skills, an analytic approach to given tasks and problems
  • API testing

Familiarity with some of the following will be a plus:
Experience in front-end testing

Knowledge of TCP/IP, networking protocols and network services at CCNA level

IT security concepts

  • Knowledge of TCP/IP, networking protocols and network services
  • Basic familiarity with programming in Python
  • English language skills at the level allowing efficient communication
  • IT security concepts

In return for your expertise, we’ll support you in this new challenge with coaching & development every step of the way.
Also, To Reward Your Work You’ll Get

  • Private medical & dental coverage, Multisport
  • Life insurance (two annual income),
  • Employee Stock Purchase Plan – 15% discount for buying Motorola’s Stock units,
  • Employee Pension Plan – 3,5 % of the month’s salary gross, which goes to the retirement account
  • IP Tax Relief (up to 80%)
  • Yearly salary increase (depends on individual performance)
  • Yearly bonus (depends on company performance)
  • Flexible working hours (usually day starts between 7-10),
  • 8 hours working day (30 minutes lunch break included).
  • lots of sports activities such as Moto football league, Wakeboarding, Snowboarding, e-gaming league etc.
  • access to wellness facilities and integration events
  • comfortable work conditions (high-class offices, parking space)
  • volleyball field and grill place next to the office
  • training and broad development opportunities
  • Motorola Solutions is supporting CSR activities and encourages employees to participate

Travel Requirements

None

Relocation Provided

International

Position Type

Experienced

Referral Payment Plan

Yes

Company
Motorola Solutions Systems Polska Sp.z.o.o

EEO Statement
Motorola Solutions is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion or belief, sex, sexual orientation, gender identity, national origin, disability, veteran status or any other legally-protected characteristic.

We are proud of our people-first and community-focused culture, empowering every Motorolan to be their most authentic self and to do their best work to deliver on the promise of a safer world. If you’d like to join our team but feel that you don’t quite meet all of the preferred skills, we’d still love to hear why you think you’d be a great addition to our team.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Ridehailing, Site Reliability Engineer

Waymo

Warsawfulltime

Apply on the employer's site

Role description

Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.

Waymo’s software reliability engineers (SRE) are responsible for the stable operation of Waymo’s fully autonomous systems and supporting infrastructure. As an SRE, you combine software and systems engineering techniques to build and run large-scale, fault-tolerant, reliable systems. You focus on optimizing existing systems, building new infrastructure, eliminating manual, error-prone or time-consuming work through automation, and ensuring products that are fast, efficient, and effective.
This role follows a hybrid work schedule and reports to the Tech Lead Manager.
You Will

  • Become a Waymo production expert, and collaborate with other engineers to build reliable systems for autonomous vehicle operations, including depot logistics, automation flow, and critical vehicle state infrastructure
  • Manage end-to-end availability and performance for core fleet services, ensuring we have enough usable vehicle supply available to meet targeted user demand, and developing observability and automation to support this goal
  • Involvement in the whole lifecycle of services - from inception and design, through deployment, operation and refinement
  • Write designs and implement software to improve system architecture, telemetry or deployment for fleet-specific mission-critical services, preventing outages that could hinder vehicle launch or maintenance
  • Write designs and code software/automation for global infrastructure
  • Serve as the first responder for fleet and supply infrastructure by leading incident response efforts. You’ll participate in a sustainable on-call rotation, while championing a culture of blameless retrospectives to drive continuous improvement
  • Be a technical leader -- work with SREs, SWE partners and PMs to develop and set the technical direction and architectural guidelines for Waymo's software development

You Have

  • 6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python
  • Proven ability to conduct deep-dive performance profiling and lead large-scale refactoring efforts to improve system maintainability and latency
  • Demonstrated success managing massive distributed systems and a drive to solve the unique production engineering challenges found at the intersection of software and physical vehicle fleets
  • A proven track record defining SLIs/SLOs/SLA frameworks. Experience designing and deploying observability systems to enhancing system visibility and reliability through sophisticated monitoring aligned to critical service health and the CUJ needs of users, devs and SREs
  • Demonstrated experience leading cross-functional initiatives between Engineering and Dev
  • A history of mentoring junior and mid-level engineers, fostering a culture of operational excellence, and driving technical roadmap decisions for a high-growth department
  • A Bachelors degree in a relevant field or 8+ years similar experience in a high growth environment with leadership experience

We Prefer

  • Demonstrated engineering leadership ability of a mission critical system at scale
  • A demonstrated track record of translating reliability needs into technical roadmaps and rigorous, data-driven SLO frameworks.
  • Proven ability to lead deep-dive architectural investigations and resolve complex, high-impact system failures across the stack
  • A Bachelors of Computer Science (or similar)

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

129 more openings in this category and country

AI Platform Engineer, Agent Systems IRC296274GlobalLogic · Poland

Apply on the employer's site