Skip to content
Site Reliability Engineer

Site Reliability Engineer

Senior Site Reliability Engineer

Valtech

Sofiafulltime

Откликнуться на сайте работодателя

Описание вакансии

Why Valtech? We’re the experience innovation company - a trusted partner to the world’s most recognized brands. To our people we offer growth opportunities, a values-driven culture, international careers and the chance to shape the future of experience.

The opportunity
At Valtech, you’ll find an environment designed for continuous learning, meaningful impact, and professional growth. Whether you're pioneering new digital solutions, challenging conventional thinking or building the next generation of customer experiences, your work will help transform industries.

We Are Proud Of

  • The work we do and the innovation we drive
  • Our values of share, care and dare
  • A workplace culture that fosters creativity, diversity and autonomy
  • Our borderless, global framework, which enables seamless collaboration

The role
As a Site Reliability Engineer (SRE), you are the bridge between software development and operations. You help us deliver reliable speed to our clients, allowing them to leverage the benefits of continuous deployment without losing grip on customer experience. You will work with our multidisciplinary teams in an essential DevOps way of working, where your main responsibility is to keep everyone focused on production while creating the infrastructure to do so.

Role Responsibilities

  • Work with teams to define SLIs and SLOs.
  • Create systems for observability.
  • Work with teams to analyze failure scenarios and possible mitigations.
  • (Assisting to) create runbooks to remediate or prevent failure scenarios.
  • Reduce work that does not add value.
  • Participate and facilitate incident management, including on-call duty.

Must Have Qualifications
You are someone with 5 years of experience in the field of software engineering, DevOps engineering, QA engineering and/or cloud engineering, of which at least the last 2 years as a dedicated Site Reliability Engineer. You feel comfortable taking the lead, making decisions, and know how to mobilize and motivate people to set things in motion. In your current role, people come to you for advice on what to look for to determine the robustness of their production environments, advice for reliable deployment procedures, assistance in analysis of failure scenarios, and ideas on how to mitigate or remediate those.

To be considered for this role, you must meet the following essential qualifications:

  • You are assertive with good communicative skills, capable of taking the lead and coaching a development team to make the right choices.
  • You have experience with incident management in a production environment of a public-facing online service with high business value and preferably high traffic in a 24x7 fashion.
  • You have experience in working in corporate environments.
  • You have experience programming and scripting.
  • You have at least basic knowledge of serverless services in one or more public cloud providers (AWS, Azure, GCP).
  • You have extensive knowledge of and experience with various monitoring systems, amongst which APM systems such as Datadog, New Relic, Dynatrace, Prometheus, and Grafana.
  • You have knowledge of and experience with various pipelining tools, such as GitHub, Azure DevOps, GitLab, Jenkins.
  • You have knowledge of and experience with microservices-related technology: Docker, Kubernetes.
  • You have a good conceptual understanding of software architecture and system thinking.
  • You have worked as an engineer in a DevOps context.
  • You have an excellent command of English (C1 or above).
  • Are familiar with the following technologies:
  • Datadog (or APM equivalent).
  • Argo CI/CD.
  • Java / Springboot.
  • Kafka.
  • Kubernetes / EKS.
  • AWS.
  • Have worked within the context of publicly accessible, highly available eCommerce platforms.
  • Have experience working in an international context with on- and off-shore teams.

Commitment to reaching all kinds of people
We design experiences that work for all kinds of people - and that starts with our own teams. At Valtech, we’re intentional about building an inclusive culture where everyone feels supported to grow, thrive and achieve their goals. No matter your background, you belong here. Explore our Diversity & Inclusion site to see how we’re creating a more equitable Valtech for all.

The Benefits
This is a position based in Sofia, Bulgaria.

Beyond a Competitive Compensation Package, We Offer

  • Flexibility, with hybrid work options and 25 vacation days for a healthy work-life balance.
  • Co-subsidized transportation & Multisport cards.
  • Premium health insurance for fast and easy access to top healthcare services.
  • Training policy for technical and other skills-related events, courses, and certifications.
  • Personal career development roadmap guided by performance evaluations.
  • Self-care program offering psychological consultations & discussions for you and the team.
  • Cozy office space designed for comfort and productivity.
  • Exciting team events and company gatherings.

Your application process
Once you apply, our Talent Acquisition team will review your application. Your CV should cover key information on relevant experiences and expertise. We do not require information such as age, gender, marital status, or a headshot in your application. We review all candidates based on skills, experience, and potential.

⚠️ Beware of recruitment fraud!

We are committed to inclusion and accessibility. If you need reasonable accommodations during the interview process, please either indicate it in your application or let your Talent Partner know.

About Valtech
Valtech is the experience innovation company that exists to unlock a better way to experience the world. By blending crafts, categories, and cultures, we help brands unlock new value in an increasingly digital world.

At the intersection of data, AI, creativity, and technology, we drive transformation for leading organizations, including L’Oréal, Mars, Audi, P&G, Volkswagen Dolby, and more.

At Valtech, we don’t just talk about transformation. We make it happen. Our people are the heart of our success, and we foster a workplace where everyone has the support to thrive, grow and innovate.

Are you ready to create what’s next? Join us.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Site Reliability Engineer

Man Group

Sofiafulltime

Откликнуться на сайте работодателя

Описание вакансии

About Man Group
Man Group is a global alternative investment management firm focused on pursuing outperformance for sophisticated clients via our Systematic, Discretionary and Solutions offerings. Powered by talent and advanced technology, our single and multi-manager investment strategies are underpinned by deep research and span public and private markets, across all major asset classes, with a significant focus on alternatives. Man Group takes a partnership approach to working with clients, establishing deep connections and creating tailored solutions to meet their investment goals and those of the millions of retirees and savers they represent.

Headquartered in London, we manage $227.6 billion* and operate across multiple offices globally. Man Group plc is listed on the London Stock Exchange under the ticker EMG.LN and is a constituent of the FTSE 250 Index. Further information can be found at www.man.com

  • As at 31 December 2025

The Role
Join our high-performing Site Reliability Engineering (SRE) team and play a pivotal role in ensuring the reliability, scalability, and performance of the technology powering Man Group’s hedge funds. You’ll have the autonomy, tools, and support to innovate and shape the future of our platform. This is an opportunity to work on cutting-edge projects, gain mentorship from senior leaders, and develop a deep understanding of both technology and the business.

As an SRE, you’ll take ownership of service reliability and deliver solutions that make a real impact. Your initial focus will include leveraging AI to accelerate incident diagnosis and resolution, improving observability, capacity planning, and automation. Over time, you’ll work across our entire infrastructure stack, operating at scale and driving continuous improvement.

Role Responsibilities

  • Ensure reliability and performance of critical systems across global infrastructure through proactive monitoring and rapid incident response.
  • Design and implement observability solutions using tools like Prometheus, Grafana, ELK, and Loki to provide deep insights into system health.
  • Automate operational tasks and build self-service capabilities to eliminate toil and improve efficiency.
  • Develop and maintain SLIs, SLOs, and error budgets to guide reliability improvements and inform engineering priorities.
  • Participate in incident response efforts, blameless post-mortems, and implement preventive measures to reduce recurrence.
  • Collaborate with development teams to improve system design, deployment practices, and operational excellence.
  • Operate at scale, managing petabyte-level storage, large CPU/GPU deployments, and high-throughput distributed systems.
  • Contribute to capacity planning and performance tuning, ensuring systems meet business demands.
  • Manage multiple ELK clusters hosting hundreds of terabytes of logs, telemetry, and APM data.

Key competencies
Required

  • Strong understanding of SRE principles, including SLIs, SLOs, error budgets, and reliability best practices.
  • Hands-on experience with observability and monitoring tools (Prometheus, Grafana, ELK, Loki, or similar).
  • Proficiency with automation tools (Ansible, Terraform) and scripting/programming languages (Python, Go, PowerShell).
  • Strong troubleshooting and debugging skills across distributed systems, with the ability to diagnose complex production issues under pressure.
  • Experience with incident management, on-call rotations, and post-incident reviews.
  • Familiarity with Kubernetes and container orchestration.
  • A proactive mindset and ability to take ownership of reliability initiatives.

Advantageous

  • Experience with CI/CD pipelines and source control workflows (Git, Jenkins, TeamCity).
  • Administration of Linux and Windows systems and exposure to cloud technologies (AWS/Azure).
  • Understanding of networking concepts, load balancing, and distributed architectures.
  • Knowledge of AI/LLM concepts (context windows, prompt tuning, MCP servers).
  • Interest in FinOps principles, desire to understand the true cost of our decisions.
  • Excellent communication and collaboration skills.

Benefits

  • Modern office located in the OfficeX campus with easy access to transport and amenities.
  • Hybrid working model
  • Competitive compensation package
  • 25 days holiday allowance
  • Premium Health insurance
  • Employee Assistance program
  • Referral Bonus
  • Additional days off for long service and volunteering
  • Multisport card
  • Opportunities for professional development including internal tech talks
  • Conference attendance, and engagement with the open-source community.

Inclusion, Work-Life Balance And Benefits At Man Group
You'll thrive in our working environment that champions equality of opportunity. Your unique perspective will contribute to our success, joining a workplace where inclusion is fundamental and deeply embedded in our culture and values. Through our external and internal initiatives, partnerships and programmes, you'll find opportunities to grow, develop your talents, and help foster an inclusive environment for all across our firm and industry. Learn more at www.man.com/diversity.

You'll have opportunities to make a difference through our charitable and global initiatives, while advancing your career through professional development, and with flexible working arrangements available too. Like all our people, you'll receive two annual 'Mankind' days of paid leave for community volunteering.

Our comprehensive benefits package includes competitive holiday entitlements, pension/401k, life and long-term disability coverage, group sick pay, enhanced parental leave and long-service leave. Depending on your location, you may also enjoy additional benefits such as private medical coverage, discounted gym membership options and pet insurance.

Equal Employment Opportunity Policy
Man Group provides equal employment opportunities to all applicants and all employees without regard to race, color, creed, national origin, ancestry, religion, disability, sex, gender identity and expression, marital status, sexual orientation, military or veteran status, age or any other legally protected category or status in accordance with applicable federal, state and local laws.

Man Group is a Disability Confident Committed employer; if you require help or information on reasonable adjustments as you apply for roles with us, please contact TalentAcquisition@man.com.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Senior Software Engineer (Golang – Pricing Domain)

SumUp

Sofiafulltime

Откликнуться на сайте работодателя

Описание вакансии

At SumUp, we create tools that help small businesses get paid easily, manage money confidently, and grow sustainably. Our mission is to empower merchants around the world, and at the heart of that mission is a robust, scalable Payments Platform developed in Sofia.

We’re now expanding this platform with a new dedicated Pricing team – and we’re looking for Software Engineers (Mid & Senior) who want to build core services from the ground up and make a lasting impact across SumUp.

About The Team
Pricing is an important part of any payments business. The Pricing team will build and own the services that determine how SumUp’s products and transactions are priced worldwide. This includes maintaining existing pricing tools and creating a future-ready pricing platform that enables precision, flexibility, and transparency.

We’ll be a small, high-impact team with strong engineering ownership. You’ll help define how pricing logic is modeled, how it evolves, and how it's applied at scale for millions of transactions across our ecosystem.

About The Role
As a (
Senior) Software Engineer
in our Pricing team, you’ll help design, build, and maintain a scalable and compliant pricing platform for SumUp’s payment ecosystem. You’ll
participate in
architectural discussions, guide technical decisions, and deliver projects independently or in collaboration with your squad.

Your work will directly shape how SumUp evolves its pricing capabilities and supports business growth.

What You’ll Do

  • You’ll design, develop, and maintain scalable backend services for our pricing systems.
  • You’ll develop backoffice tools for pricing management.
  • You’ll write high-quality, maintainable, and scalable code that meets coding standards and best practices.
  • You’ll provide comprehensive documentation and automated good test coverage, improve code quality
  • You’ll collaborate with other software, QA and DevOps engineers to ensure smooth deployment, continuous integration, and support for the software that we deliver in production environments.
  • You’ll actively participate in code reviews with other software engineers to improve code quality and maintainability.
  • You’ll support technical discussions during requirements analysis and scope definition.
  • You’ll participate in team retrospectives and document lessons learned.

You’ll be great for this role if you have

  • You have several years of programming experience, including in Golang
  • You have experience with building and consuming RESTful APIs
  • You’re familiar with Docker, AWS, Kubernetes (user, not admin experience)
  • You have a DevOps mindset
  • You have experience with SQL & NoSQL DBs. We’re working with Postgres and Kafka
  • You understand fundamental system architecture, software design principles, data modeling, and API design
  • You are have an open to feedback, collaborative approach to teamwork and excellent problem-solving skills
  • You take pride in engineering and have a keen sense of ownership of the work that you do

Why you should join SumUp

  • You’ll play a key role in a high-impact team working on one of the most critical services at SumUp, while being part of a global scale-up of 3000+ people from 60+ countries, spread across 4 continents.
  • Shape the future of SumUp’s pricing capabilities and business strategy.
  • You’ll receive 25 days’ paid leave, increasing with tenure, and paid vacation for certain occasions. You’d enjoy a paid 1 month Sabbatical vacation every 3 years.
  • You’ll receive an individual learning budget and can take 10 days paid educational leave to expand your skillset. You'll have access to hackathons.
  • You'd enjoy other great benefits such as additional health and life insurance; funded therapy & coaching sessions with licensed psychotherapists; co-sponsored Multisport card; Child birth/adoption bonus; free shuttle buses from Joliot-Curie metro station; a long list of discounts; attractive referral programme; monthly budget for food vouchers (100 EUR) and flexible benefits (40 EUR) , and more.

About SumUp
Be empowered to do more that matters.
At SumUp, we're on a mission to empower small businesses across the globe by providing simple and affordable tools that allow them to thrive. Today, over 4 million businesses in 37 markets rely on SumUp as their financial partner to manage payments, finance and customer relationships.

Our commitment to small businesses is reflected in our diverse teamOpens in new window of over 3,000 SumUppers from over 90 nationalities, united by global collaboration and an innovative mindset. Our core values lay the foundation for who we are and what we stand for, shaping our work culture and driving our success. We foster inclusivity and a continuous learning culture, providing a safe space for personal and professional growth. Our differences make us unique and strong as we strive to create an environment where everyone belongs and feels supported, no matter how they identify.

SumUp is proud to be an Equal Employment Opportunity employer, actively seeking and embracing diversity in our workforce. We don't make hiring or employment decisions based on race, colour, religion or religious belief, ethnic or national origin, nationality, sex, gender, gender identity, sexual orientation, disability, age or any other basis protected by applicable laws or prohibited by company policy. Our commitment extends beyond recruitment to creating a safe and respectful workplace where harassment of any form is strictly prohibited. Discover more about our culture and opportunities on our careers website, and follow our journey on LinkedInOpens in new window, InstagramOpens in new window, and TikTokOpens in new window.

Job Application Tip
We recognise that candidates feel they need to meet 100% of the job criteria in order to apply for a job. Please note that this is only a guide. If you don’t tick every box, it’s ok too because it means you have room to learn and develop your career at SumUp.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

GOLANG/RUST CLOUD SOFTWARE ENGINEER FOR FAAS (m/f/d)

Schwarz Digits

Sofiafulltime

Откликнуться на сайте работодателя

Описание вакансии

Make an amazing climb in your career in an international team of experts. Our company provides technological services for the whole Schwarz group of more than 30 countries in Europe and the US. Our vision is to be the leading ecosystem for a better life. We built the European sovereign cloud STACKIT. With XM Cyber we set new standards in differing cyber crimes. We run AI better than anyone. With us you will find a variety of opportunities to grow and do your best at your calling – IT. We exist to improve life with our products and services - for today's generation and future generations. We act future proof!

The impact you will create:

  • Design, develop, and maintain high-performance, scalable, and reliable Function-as-a-Service platform components using Golang and/or Rust
  • Collaborate on the development of a secure, container-based runtime environments for serverless functions
  • Participate in the entire software development lifecycle, from requirements analysis to deployment and maintenance
  • Conduct code reviews to ensure code quality, promote knowledge sharing, and drive technical excellence within the team
  • Engage in design and architecture discussions to advance the technical vision and roadmap of our FaaS platform

Experience and skills you will need:

  • Proficiency in Golang and/or Rust programming languages
  • Experience with Kubernetes
  • Experience with MicroVM technology and its ecosystem, including unikernels, Kata Containers, cloud-hypervisor, Firecracker, QEMU, or KVM is a plus
  • Strong understanding of serverless computing concepts, including FaaS, event-driven architectures, and cloud-native design patterns
  • Familiarity with containerization and virtualizationExperience with building and operating large-scale distributed systems, including monitoring, alerting, and error handling
  • Agile mindset, with a passion for collaboration, innovation, and continuous learning

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

End User Automation Engineer

Man Group

Sofiafulltime

Откликнуться на сайте работодателя

Описание вакансии

About Man Group
Man Group is a global alternative investment management firm focused on pursuing outperformance for sophisticated clients via our Systematic, Discretionary and Solutions offerings. Powered by talent and advanced technology, our single and multi-manager investment strategies are underpinned by deep research and span public and private markets, across all major asset classes, with a significant focus on alternatives. Man Group takes a partnership approach to working with clients, establishing deep connections and creating tailored solutions to meet their investment goals and those of the millions of retirees and savers they represent.

Headquartered in London, we manage $227.6 billion* and operate across multiple offices globally. Man Group plc is listed on the London Stock Exchange under the ticker EMG.LN and is a constituent of the FTSE 250 Index. Further information can be found at www.man.com

  • As at 31 December 2025

The Team
We are seeking an experienced End User Systems Engineer to join our Platform Engineering team. You’ll play a critical role in building and maintaining the automation frameworks that underpin our end user operations. Many of these may be newly implemented in the End User space, so you’ll be pivotal in transforming the way systems are designed. This position offers the opportunity to work with cutting-edge technologies at scale, supporting our end users across the entirety of the firm, including trading and fund managers.

As an End User Automation Engineer, you’ll design, implement, and maintain infrastructure-as-code solutions, configuration management and monitoring systems that enable our engineering teams to deliver reliable, scalable services to our end users. You’ll work closely with the entirety of our Platform Engineering team to drive automation initiatives to improve delivery of our services. You should have the enthusiasm for leveraging AI development tools to accelerate delivery and driving their adoption across the team.

Role Responsibilities

  • Building, maintaining and automating end-user platforms and services.
  • Designing infrastructure-as-code solutions using tools such as Terraform and Ansible
  • Providing 3rd line support for end-user services as part of a team rotation.
  • Contributing to the strategic direction of the team and solutions it provides.
  • Collaborating with business stakeholders in a fast-paced, dynamic environment.
  • Management and security of end-user devices.

Key Competencies
Essential

  • Hands-on experience with infrastructure-as-code tools such as Ansible and Terraform.
  • Experience with scripting and automation languages, ideally Python.
  • Good understanding of observability and monitoring practices, including experience with tools such as Prometheus, Grafana, or similar platforms.
  • Understanding of CI/CD principles and experience with tools such as Jenkins or GitHub Actions.
  • Embracing agentic engineering - willingness and ability to work effectively with AI-assisted development tools as part of daily workflow.
  • Self-organised with the ability to prioritise and effectively manage time across multiple projects.
  • An interest in end user platforms and desire to make the platforms and systems used by employees as efficient and performant as possible.
  • Good communication and problem-solving skills with the ability to explain complex technical concepts to diverse audiences.

Advantageous

  • Experience supporting End User platforms or communication systems, such as Microsoft 365 and Slack.
  • Exposure to identity management and associated technologies (Active Directory, Entra ID).
  • Experience with security and secrets management practices.
  • Good knowledge of networking fundamentals, storage, and modern infrastructure architecture.

Benefits

  • Modern office located in the OfficeX campus with easy access to transport and amenities.
  • Hybrid working model
  • Competitive compensation package
  • 25 days holiday allowance
  • Premium Health insurance
  • Employee Assistance program
  • Referral Bonus
  • Additional days off for long service and volunteering
  • Multisport card
  • Opportunities for professional development including internal tech talks
  • Conference attendance, and engagement with the open-source community.

Inclusion, Work-Life Balance And Benefits At Man Group
You'll thrive in our working environment that champions equality of opportunity. Your unique perspective will contribute to our success, joining a workplace where inclusion is fundamental and deeply embedded in our culture and values. Through our external and internal initiatives, partnerships and programmes, you'll find opportunities to grow, develop your talents, and help foster an inclusive environment for all across our firm and industry. Learn more at www.man.com/diversity.

You'll have opportunities to make a difference through our charitable and global initiatives, while advancing your career through professional development, and with flexible working arrangements available too. Like all our people, you'll receive two annual 'Mankind' days of paid leave for community volunteering.

Our comprehensive benefits package includes competitive holiday entitlements, pension/401k, life and long-term disability coverage, group sick pay, enhanced parental leave and long-service leave. Depending on your location, you may also enjoy additional benefits such as private medical coverage, discounted gym membership options and pet insurance.

Equal Employment Opportunity Policy
Man Group provides equal employment opportunities to all applicants and all employees without regard to race, color, creed, national origin, ancestry, religion, disability, sex, gender identity and expression, marital status, sexual orientation, military or veteran status, age or any other legally protected category or status in accordance with applicable federal, state and local laws.

Man Group is a Disability Confident Committed employer; if you require help or information on reasonable adjustments as you apply for roles with us, please contact TalentAcquisition@man.com.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Ещё 30 вакансий по этой категории в этой стране

Senior Site Reliability EngineerValtech · Bulgaria

Откликнуться на сайте работодателя