Skip to content
Site Reliability Engineer

Site Reliability Engineer

Senior Site Reliability Engineer

Quickbase

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

Senior Site Reliability Engineer (SRE) - Azure Platform

Location:
Remote

Department:
Site Reliability Engineering (Bulgaria)

Employment Type:
Full-Time

Level:
Senior Individual Contributor / Technical Lead

Position Summary

We are seeking a senior, hands-on Site Reliability Engineer (SRE) to strengthen our Azure platform capabilities and help mature cloud operations across the organization. This role will serve as a primary technical lead for Microsoft Azure, supporting two existing Azure environments while helping establish standards, operating procedures, and readiness for future development use.

The ideal candidate combines deep Azure expertise, strong SRE discipline, infrastructure automation experience, and practical technical leadership. This person will be expected to lead by example, document and standardize platform practices, mentor other team members, and help train the team so others can support routine Azure administration and operations within the first six months. The role will also provide occasional AWS administration, operational support, and technical guidance as needed.

Key Responsibilities

Cloud Architecture & Platform Engineering

  • Serve as a primary Azure subject matter expert for current Azure environments and related corporate Azure usage.
  • Design, implement, and enhance Azure architecture with emphasis on reliability, security, scalability, observability, cost management, and operational simplicity.
  • Manage and optimize Azure services such as virtual machines, storage, virtual networks, Azure Monitor, Log Analytics, Application Insights, Key Vault, application services, and related platform components.
  • Establish and improve Azure standards for subscriptions, resource groups, networking, naming, tagging, identity, monitoring, backup/recovery, and operational procedures.
  • Provide occasional AWS support as needed, including administration, operational troubleshooting, architecture review, and workload placement guidance.

Infrastructure as Code & Automation

  • Develop, maintain, and govern Infrastructure as Code (IaC) using Terraform, Azure Bicep, ARM templates, or equivalent tooling.
  • Build reusable IaC modules/templates and environment-specific configurations to support secure, repeatable Azure deployments.
  • Build and improve CI/CD automation using Azure DevOps, GitHub Actions, Jenkins, or similar tools, ensuring secure and reliable deployment workflows.
  • Automate provisioning, configuration, monitoring, access management, and operational workflows to reduce manual effort and improve reliability.
  • Use Git-based workflows, pull requests, peer reviews, and change controls to manage infrastructure changes.

Reliability, Observability & Operations

  • Improve logging, monitoring, alerting, and operational dashboards using Azure-native tools and complementary systems.
  • Define and promote reliability practices such as incident response, root-cause analysis, remediation tracking, service health reviews, and operational readiness checks.
  • Act as an escalation resource for significant cloud infrastructure issues and participate as backup in on-call or incident response activities as needed.
  • Identify recurring operational issues and drive long-term fixes through automation, architecture improvements, and better observability.

Security, Compliance & Access Management

  • Administer Azure Key Vault, including secrets management, certificate lifecycle, rotation practices, and governance controls.
  • Implement and manage Azure RBAC, Microsoft Entra ID integrations, managed identities, least-privilege access, and related identity controls.
  • Partner with Security and Compliance teams to ensure infrastructure aligns with SOC 2 controls, internal audit standards, and security best practices.
  • Support policy-driven governance, including tagging, access reviews, configuration standards, and evidence collection for compliance activities.

Technical Leadership, Mentoring & Documentation

  • Lead technical design discussions, platform reviews, and operational improvement efforts across SRE, engineering, security, DevOps, and product teams.
  • Create and maintain clear documentation, runbooks, architecture diagrams, operating procedures, and best-practice guides for Azure operations.
  • Mentor and train other team members through pairing, knowledge-transfer sessions, standards documentation, and operational walkthroughs.
  • Help establish a sustainable support model so routine Azure administration and operations can be shared by additional team members within six months.
  • Communicate technical concepts clearly to both technical and non-technical stakeholders.

Required Qualifications

  • 7+ years of experience in Site Reliability Engineering, Cloud Engineering, Platform Engineering, Systems Engineering, or DevOps roles.
  • 4+ years of hands-on experience designing, operating, and improving Microsoft Azure environments.
  • Strong proficiency with Azure architecture and operational services, including networking, compute, storage, identity, monitoring, Key Vault, and application platform services.
  • Hands-on experience with Infrastructure as Code using Terraform, Azure Bicep, ARM templates, or equivalent tooling.
  • Experience building and maintaining CI/CD workflows using Azure DevOps, GitHub Actions, Jenkins, or comparable platforms.
  • Scripting and automation experience using PowerShell, Python, Bash, or similar languages.
  • Demonstrated expertise in logging, alerting, monitoring, incident response, root-cause analysis, and reliability engineering practices.
  • Knowledge of cloud security practices, RBAC, secrets management, managed identities, and least-privilege access models.
  • Familiarity with SOC 2 or similar compliance-driven environments.
  • Proven ability to lead technical initiatives, influence standards, document best practices, and mentor other engineers.

Preferred Qualifications

  • Experience providing AWS administration or operational support, including IAM, VPC/networking, EC2, S3, CloudWatch, and account/resource governance.
  • Experience with Azure landing-zone concepts, Azure Policy, cost management, backup/recovery patterns, and cloud governance models.
  • Relevant certifications such as Microsoft Azure Solutions Architect Expert (AZ-305), Azure DevOps Engineer Expert (AZ-400), Azure Administrator Associate (AZ-104), or equivalent cloud/DevOps certifications.
  • Experience with containerization and orchestration technologies such as Docker, Kubernetes, or Azure Kubernetes Service (AKS).
  • Familiarity with Zero Trust, identity governance, FinOps/cost optimization, and compliance-driven cloud operations.

Success Criteria

Within the first 6 months, success in this role will include:

  • Documented current-state assessment of the two Azure environments.
  • Defined and communicate Azure operating standards, including naming, tagging, access, monitoring, backup/recovery, and change-management practices.
  • Improved runbooks, support procedures, and operational documentation that enable other team members to assist with routine Azure administration and operations.
  • Meaningful knowledge transfer through mentoring, pairing, and practical training sessions with the SRE and/or engineering teams.
  • Clear ownership of Azure reliability, observability, security, and automation improvement priorities.

Within 6-12 Months, Success In This Role Will Include

  • Measurable improvements in reliability, performance, observability, security, and operational maturity of Azure workloads.
  • Increased automation coverage across provisioning, CI/CD, monitoring, access controls, and recurring operational tasks.
  • Modernized Azure resources and platform patterns aligned with architectural, security, and compliance standards.
  • A more distributed support model in which trained team members can handle routine Azure tasks with reduced dependence on a single subject matter expert.
  • Reliable contribution to occasional AWS administration, operational support, and cross-cloud guidance when needed.

How We Think About AI
At Quickbase, we view AI as a tool to accelerate how work gets done — not replace it. We encourage thoughtful use of AI to improve speed, quality, and decision-making, while maintaining strong judgment, accountability, and data integrity.

Benefits

  • Unlimited remote work policy
  • 25 days of annual leave, 2 additional days off for volunteering
  • Competitive remuneration package incl. an annual bonus
  • Top-notch IT setup.
  • Mental health support, life insurance, food vouchers
  • Additional health insurance - for you and your loved ones
  • Annual wellness support allowance
  • External Professional Learning Opportunities

Equal Opportunity Statement
Quickbase is committed to building a diverse and inclusive workplace. We encourage candidates from all backgrounds to apply – even if you don’t meet every qualification listed. We are proud to be an equal opportunity employer.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Senior Software Engineer

Good Job Games

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

About us

We were founded in 2017 with the motivation to reach people globally by giving them unique and unforgettable experiences with disruptive products. Our games have reached over 3.5 billion people. This could only be done by gathering exceptional talent and creating a culture to enhance team spirit and creativity. We are looking for passionate teammates to join our team!

What you'll be doing

We are looking for a passionate Backend Developer who is excited to architect and implement technology, tools and infrastructure that empower Good Job Games!

  • Building distributed, high-performance shared game services
  • Developing internal tools for operations, data analysis and automation
  • Supporting and empowering teams across Good Job Games
  • Mentoring and guiding junior team members

Desired skills and experiences

  • B.S. or higher preferably in Computer Science, Math or Physics (or equivalent work experience)
  • 4+ years of experience
  • Strong experience in designing, implementing and maintaining distributed, highly scalable, low latency, fault tolerant backend architectures
  • Fluent in using Go programming language and strong understanding of advanced Go syntax and concepts
  • Strong engineering, design and architecture skills
  • Strong experience with AWS, NoSQL/in-memory databases, DevOps practices and CI/CD tools
  • Passion for Match Villains and mobile puzzle games in general

What makes our team so unique

  • Feedback and transparency are at the heart of everything we do
  • Exceptional and passionate people/team members
  • Every idea counts
  • Never-ending learning
  • We never stop asking the questions “why” and “how”

Our Perks

  • Team events and trips
  • Great food
  • On-site gym
  • Full health benefits
  • Compensation for paid military service
  • Relocation support
  • Good Job Games Coin Program that lets you have unforgettable experiences (e.g. Going on a cruise trip to Norway or seeing the Northern Lights)

This is an on-site role in Istanbul, Sarıyer. Unfortunately, we do not offer a fully-remote working option.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Software Engineer

Good Job Games

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

About us

We were founded in 2017 with the motivation to reach people globally by giving them unique and unforgettable experiences with disruptive products. Our games have reached over 3.5 billion people. This could only be done by gathering exceptional talent and creating a culture to enhance team spirit and creativity. We are looking for passionate teammates to join our team!

What you'll be doing

We are looking for a passionate Backend Developer who is excited to architect and implement technology, tools and infrastructure that empower Good Job Games!

  • Building distributed, high-performance shared game services
  • Developing internal tools for operations, data analysis and automation
  • Supporting and empowering teams across Good Job Games
  • Learning, teaching and growing within a strong engineering culture
  • Mentoring and guiding junior team members

Desired skills and experiences

  • B.S. or higher preferably in Computer Science, Math or Physics (or equivalent work experience)
  • 2+ years of experience
  • Understanding distributed, highly scalable, low latency, fault tolerant backend architecture fundamentals
  • Fluent in using Go programming language and strong understanding of advanced Go syntax and concepts
  • Strong engineering, design and architecture skills
  • Experience with AWS, NoSQL/in-memory databases, DevOps practices and CI/CD tools
  • Passion for Match Villains and mobile puzzle games in general

What makes our team so unique

  • Feedback and transparency are at the heart of everything we do
  • Exceptional and passionate people/team members
  • Every idea counts
  • Never-ending learning
  • We never stop asking the questions “why” and “how”

Our Perks

  • Team events and trips
  • Great food
  • On-site gym
  • Full health benefits
  • Compensation for paid military service
  • Relocation support
  • Good Job Games Coin Program that lets you have unforgettable experiences (e.g. Going on a cruise trip to Norway or seeing the Northern Lights)

This is an on-site role in Istanbul, Sarıyer. Unfortunately, we do not offer a fully-remote working option.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

DevOps / Senior Cloud Infrastructure Engineer - Freelancer

Monterail

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

Job Description
We are looking for a DevOps / Senior Cloud Infrastructure Engineer to join us on a freelance basis to build the observability and incident management backbone for autonomous kitchen operations.

4-5 month project | full time | remote
What We're Looking For

  • Senior-level experience as an AWS Cloud Engineer or Site Reliability Engineer.
  • Strong, demonstrable focus on metric-driven observability, monitoring, and alerting at scale.
  • Fluent, hands-on experience in Python for tooling and automation.
  • Hands-on experience with Terraform for Infrastructure as Code.
  • A proven track record of designing, architecting, and owning production systems end-to-end.
  • Ability to work completely independently without a detailed spec sheet or heavy direction.
  • Experience with Jira Service Management or similar ITSM/incident platforms is a plus.
  • Experience with Grafana dashboarding pipelines at scale is a plus.
  • Exposure to AI-assisted ops tooling (AI-Ops, runbook automation) is a plus.
  • Prior experience in a fast-moving hardware, robotics, or IoT fleet environment is a plus.

What you'll do

  • Standardize and own the Tech Ops Incident Management Platform using Jira Service Management across our production fleets.
  • Automate incident resolution workflows with AI tooling, including runbook generation and assignment.
  • Design and implement proper documentation for all Tech Ops incident processes to ensure a clean handover.
  • Build out fleet management and task automation to support global 24/7 remote operations.
  • Own and customize the data-driven observability layer via Grafana across all internal tech teams.
  • Work closely with key leadership stakeholders (Head of Infrastructure & Cloud, VP Engineering) to independently drive architecture decisions.

Requirements
What do we mean by freelance?
Read more at
Monterail Tech Network

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Site Reliability Engineer

Senior Software Engineer (Ruby / GOlang)

Workato

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

About Workato
Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise MCP and trusted by 50% of the Fortune 500, Workato’s cloud-native architecture connects every application, data source, and process to power real-time orchestration at scale. With enterprise-grade security and continuous innovation at its core, Workato provides the trusted foundation for organizations to automate with confidence and operationalize AI across the business. To learn more, visit www.workato.com

Why join us?
Ultimately, Workato believes in fostering a
flexible, trust-oriented culture that empowers everyone to take full ownership of their roles
. We are driven by
innovation
and looking for
team players
who want to actively build our company.

But, we also believe in
balancing productivity with self-care
. That’s why we offer all of our employees a vibrant and dynamic work environment along with a multitude of benefits they can enjoy inside and outside of their work lives.

If this sounds right up your alley, please submit an application. We look forward to getting to know you!

Also, Feel Free To Check Out Why

  • Business Insider named us an “enterprise startup to bet your career on”
  • Forbes’ Cloud 100 recognized us as one of the top 100 private cloud companies in the world
  • Deloitte Tech Fast 500 ranked us as the 17th fastest growing tech company in the Bay Area, and 96th in North America
  • Quartz ranked us the #1 best company for remote workers

Responsibilities
We are looking for an exceptional Senior Backend Developer (Ruby/Go) to join our growing Engine team. The Engine team develops and maintains most things related to Workato Recipe runtime. Everything related to recipe execution: DSL, pulling events, processing webhooks, executing jobs. There are various aspects to it: performance, scaling, storage, durability, atomicity, concurrency guarantees, data protection, and encryption.

In This Role, You Will Also Be Responsible To

  • Build/extend/troubleshot/fix complex heterogeneous GOlang and Ruby applications, Ruby monolith application, as well as small self-contained GOlang microservices.
  • We also consider strong Ruby only candidates, as well as Golang-only candidates who is open to learn some ruby.
  • Improve execution engine of custom third-party code (Ruby DSL, isolation, performance, new features).
  • Write well designed, testable, efficient code in Ruby and GOlang.
  • Integration of data storage solutions Postgres/S3/DynamoDB/Kafka/ClickHouse etc.
  • Contribute in all phases of the development lifecycle.
  • Provide code reviews to your teammates.
  • Evaluate and propose improvements to existing system.
  • Identify bottlenecks and bugs, and devise solutions to these problems.
  • Help maintain code quality, organization and automatization.
  • We always explore new technologies and work with Rust and Wasm can be foreseen.

Requirements
Qualifications / Experience / Technical Skills

  • Strong experience in building scalable distributed backend applications (5+ years).
  • Great understanding of all building blocks of large web applications: databases, load balancers, application servers, message brokers, caching, monitoring, etc.
  • Good understanding of network protocols and stacks.
  • Good understanding of DB technologies: classic databases and modern no-SQL.
  • Knowledge of basic data structures and algorithms and how they are used is a must.
  • Multilingual programming experience: our code base is primarily in Ruby, with trend to migrate to GOlang and Rust.
  • Excellent debugging, analytical, problem solving, and social skills.
  • BS/MS degree in Computer Science, Engineering or a related subject, 7+ years of industry experience.

Optional

  • Background in GO-lang and/or Rust.
  • Background in network programming.
  • Background in application, data security.
  • Deep knowledge of physical DB design.
  • Experience of working with Docker and other isolation technologies.
  • Experience of working with public cloud infrastructure providers(AWS/Azure/Google Cloud).
  • Experience in related fields (DevOps, ML, DBA, Enterprise applications, etc).
  • Experience in building/deploying data processing pipelines is a plus.
  • Experience of working with third-party REST APIs at scale (request throttling, batch processing etc).
  • Familiarity with WASM platform and toolchain.
  • Kotlint-multiplatform

Soft Skills / Personal Characteristics

  • Readiness to work remotely with teams distributed across the world and timezones.

(REQ ID: 2587)

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Ещё 116 вакансий по этой категории в этой стране

Senior Site Reliability EngineerQuickbase · Bulgaria

Откликнуться на сайте работодателя