Skip to content
Cloud Engineer

Список недели

Все такие вакансии — одним письмом

Сейчас вы читаете одну вакансию. Таких же на витрине сотни, и каждую неделю выходят новые. Выберите, что присылать, оставьте почту — список придёт сам, без поисков и без возвращения сюда.

Считаем, сколько вышло за прошлую неделю…

Первое письмо приходит сразу, дальше — раз в неделю. Отписка в один клик из любого письма, адрес больше никуда не уходит.

Cloud Engineer

Lead Site Reliability Engineer - Imunify Reliability Platform

Jobgether

Удалённо

Откликнуться на сайте работодателя

Описание вакансии

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead Site Reliability Engineer - Imunify Reliability Platform based in Slovenia.
This is a greenfield SRE leadership opportunity within a large-scale, security-focused product environment. You will define what “healthy” means across roughly 70 components spanning cloud services and customer-hosted agents. You will establish SLIs, SLOs, error budgets, monitoring standards, alerting, and escalation practices from the ground up. Your work will directly improve the ability to detect silent security-control degradation before it becomes a widespread customer-impacting issue. You will collaborate closely with engineering leads and senior engineers while building the telemetry and reliability platform yourself. The environment is remote-first, async, technically demanding, and focused on measurable outcomes rather than dashboards for their own sake. This is an opportunity to shape the reliability culture and foundations of a major security product.

Accountabilities

  • Define and establish meaningful SLIs for approximately 70 product components, working with squad leads and senior engineers to agree on ownership, measurement, tiering, SLOs, and error budgets.
  • Develop a reliability taxonomy covering service availability and latency, fleet reachability and configuration convergence, security-control efficacy, artifact delivery, and telemetry pipeline health.
  • Ensure reliability indicators are independently measurable and cannot be disabled by the same failure they are intended to detect.
  • Design and build the telemetry collection pipeline for customer-hosted agents and cloud services, balancing push-based collection, sampling, privacy constraints, data quality, and cardinality.
  • Extend instrumentation across Python, Go, and Rust components in collaboration with product engineering teams.
  • Consolidate existing dashboards, queries, and reporting mechanisms into a smaller, more reliable observability platform, retiring tooling that does not provide meaningful operational value.
  • Implement symptom-based, SLO-driven alerting with multi-window burn-rate principles and clear page, ticket, and dashboard classifications.
  • Ensure every production alert has a defined owner, documented failure mode, and actionable runbook.
  • Establish ongoing alert-quality practices, including periodic reviews, measurable actionable-alert rates, and deliberate removal of unnecessary alerts.
  • Build a machine-readable ownership and escalation model that routes incidents to the appropriate engineering squads.
  • Establish severity definitions, acknowledgement expectations, follow-the-sun escalation practices, and clean handoff procedures across multiple time zones.
  • Strengthen incident command and blameless postmortem practices, including reliable timelines, ownership, and follow-through on corrective actions.
  • Coach engineering squads to own their own operational responsibilities and paging rather than becoming a centralized buffer for other teams' alerts.
  • Deliver measurable reliability outcomes over the first year, including complete SLI ownership, production telemetry, tiered alerting, squad on-call adoption, and a significant reduction in the time required to detect silent security-control degradation.

Requirements

  • Substantial production engineering or SRE experience, including experience defining and implementing an SLO framework rather than simply operating within an existing one.
  • Strong Python skills and the ability to read and modify Go or Rust code when implementing instrumentation and reliability improvements.
  • Strong hands-on experience with time-series and event telemetry at scale, including Prometheus/OpenMetrics, Grafana, Alertmanager-class routing systems, and columnar or high-cardinality data stores such as ClickHouse or equivalent technologies.
  • Experience debugging distributed systems running on bare metal and long-lived hosts; this role requires more than Kubernetes-centric operational experience.
  • Practical experience with production-scale configuration management and CI/CD tooling such as Ansible, GitLab CI, Jenkins, or comparable technologies.
  • Strong understanding of telemetry for systems that cannot be directly scraped or fully controlled, including push-based collection, sampling, clock skew, partial reporting, and privacy considerations on customer-managed infrastructure.
  • Excellent written and asynchronous communication skills, with the ability to align multiple engineering teams around measurable definitions of system health.
  • Strong engineering judgment and a pragmatic approach to observability, reliability, alerting, and operational ownership.
  • Experience with security products such as WAF, EDR, antivirus, or vulnerability-management platforms is a strong advantage, particularly an understanding that security-control reliability must measure effective enforcement rather than simple uptime.
  • Familiarity with monitoring and continuous-monitoring requirements related to SOC 2, ISO 27001, NIST SP 800-137, or similar frameworks is valuable.
  • Exposure to OpenTelemetry, eBPF, Sentry, cost-aware telemetry, or cardinality-management techniques is beneficial.
  • Experience working effectively with AI-assisted development tools and modern agentic engineering workflows is a plus.
  • Kubernetes experience is useful for supporting the smaller portion of the platform that runs in Kubernetes.
  • This role is focused on SRE and reliability engineering rather than DevOps ticket management, build-system ownership, cloud cost management, or acting as the on-call team for other engineering squads.

Benefits

  • Fully remote work with flexible working hours, enabling you to work from anywhere worldwide.
  • 24 paid vacation days per year.
  • 10 paid national holidays.
  • Unlimited sick leave.
  • Compensation toward private medical insurance.
  • Co-working space reimbursement.
  • Gym and sports reimbursement.
  • Professional development opportunities through challenging technical projects, learning opportunities, mentoring, and knowledge-sharing programs.
  • Opportunity to receive a reward for an innovative idea that can be patented.
  • Remote-first, asynchronous working environment spanning multiple time zones.
  • Opportunity to define an SRE function, reliability standards, and operational culture from the ground up.

How Jobgether Works
We use an
AI-powered matching process
to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice:
By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Скоро на этой странице

Резюме под эту вакансию — и билет в розыгрыш

Мы разбираем объявление до настоящих требований и переписываем ваше резюме под него — вопросами, а не выдумкой: ни одна строка не появится без вашего подтверждения. Войдите, чтобы получить это первым, — и попасть в розыгрыш.

  • Резюме под конкретную вакансию, а не «универсальное»
  • Ответы хранятся: правится любой, а не весь разговор заново
  • Всё в аккаунте — открывается с любого устройства

Разыгрываем

Скидка на сопровождение

Победителей выбираем случайно среди заявок с подтверждённой почтой. Дата розыгрыша и полные правила — на странице розыгрыша.

Правила розыгрыша

Cloud Engineer

Senior Infrastructure Engineer, AI/ML Systems

Portfolium (Acquired by Instructure)

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

At Instructure
, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning and personal development, facilitate meaningful relationships, and inspire people to go further in their education and careers.

We do this by giving smart, creative, passionate people opportunities to create awesome. And that's where you come in:

Our AI team is where a lot of that gets built: applying advanced AI to real problems in learning, and turning research into product capabilities that educators and students use every day.

This is a hands-on delivery role. You'll build the deployment path that takes models and AI services from prototype to production, and you'll keep them running once they're there. You'll be one of two engineers who own infrastructure for this team, which means wide scope, real ownership, and direct influence over how we build.

You'll work closely with data scientists, applied AI engineers, and product partners to turn advanced AI ideas into reliable product capabilities used at scale.

What You'll Do

  • Build and operate deployment pipelines for AI services, covering CI/CD, infrastructure-as-code, environment promotion, and rollback
  • Deploy and operate model serving, batch scoring, and orchestration pipelines across development, staging, and production
  • Partner with data scientists and applied AI engineers to take prototypes into production, including system design for net-new services
  • Own production reliability for AI services: monitoring, alerting, debugging, performance, and cost
  • Spot repeated patterns and turn them into reusable templates, so the team can ship its second and third variant of something without rebuilding it

What You'll Need

  • Six or more years in infrastructure, DevOps, platform, or ML engineering, with ownership of systems running in production
  • Deep hands-on experience across a wide range of AWS services, including compute, networking, storage, deployment, and monitoring
  • Infrastructure-as-code experience (Terraform, CDK, or CloudFormation)
  • Experience with containers and modern deployment patterns (Docker required, Kubernetes or ECS/EKS a plus), applied to CI/CD pipelines you've designed and operated for production services
  • Experience with orchestration and workflow tooling (Airflow, Dagster, Argo, Step Functions, or similar)
  • Comfort working through ambiguity and collaborating directly with data scientists and researchers

It Would Be a Bonus If You Had

  • Experience with ML platform components and data pipeline orchestration at scale
  • Experience running LLM-based or retrieval-based systems in production
  • Experience operating specialized data stores, including graph databases
  • Experience building internal tooling, templates, or reference implementations that other engineers adopted

Onsite Collaboration Requirement:
This role requires working onsite on Tuesday and Wednesday, with Thursday strongly encouraged as part of our company’s in-person collaboration model.

Why Join Us

Join us and help shape the future of education by turning cutting-edge AI into reliable product capabilities.

At Instructure, we're on a mission to help educators and students learn together, anytime, anywhere, and however works best. You'll join our research-driven team tackling education's biggest challenges with cutting-edge technology.

We value diversity, creativity, and passion, and invest in our teams through mentorship, hack weeks, internal conferences, and a culture where innovation thrives. Here, you'll have the chance to build the next generation of LMS features that make a real impact on students and teachers, and do it in a collaborative, supportive environment that encourages experimentation and growth.

Get in on all the awesome at Instructure!
We offer competitive, meaningful benefits in every country where we operate. While they vary by location, here's a general idea of what you can expect:

  • Competitive compensation, plus all full-time employees participate in our ownership program - because everyone should have a stake in our success.
  • Flexible work culture. Our remote, hybrid and in-office collaboration spaces vary by role, team and location.
  • Generous time off, including local holidays and our annual “Dim the Lights” period in late December, when teams are encouraged to step back and recharge based on departmental needs.
  • Comprehensive wellness programs and mental health support
  • Learning and development resources, including professional development tools and tuition reimbursement, to support your growth
  • The technology and tools you need to do your best work
  • Motivosity employee recognition program
  • A culture rooted in inclusivity, support, and meaningful connection

We believe in hiring great people and treating them right. The more diverse we are, the better our ideas and outcomes.

Instructure is an Equal Opportunity Employer. We comply with applicable employment and anti-discrimination laws in every country where we operate.

All employees must pass a background check as part of the hiring process. To help protect our teams and systems, we’ve implemented identity verification measures. Candidates may be asked to verify their legal name, current physical location, and provide a valid contact number and residential address, in accordance with local data privacy laws.

Any attempt to misrepresent personal or professional information will result in disqualification.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Cloud Engineer

Senior Infrastructure Engineer, AI/ML Systems

Instructure

Удалённоfulltime

Откликнуться на сайте работодателя

Описание вакансии

At Instructure
, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning and personal development, facilitate meaningful relationships, and inspire people to go further in their education and careers.

We do this by giving smart, creative, passionate people opportunities to create awesome. And that's where you come in:

Our AI team is where a lot of that gets built: applying advanced AI to real problems in learning, and turning research into product capabilities that educators and students use every day.

This is a hands-on delivery role. You'll build the deployment path that takes models and AI services from prototype to production, and you'll keep them running once they're there. You'll be one of two engineers who own infrastructure for this team, which means wide scope, real ownership, and direct influence over how we build.

You'll work closely with data scientists, applied AI engineers, and product partners to turn advanced AI ideas into reliable product capabilities used at scale.

What You'll Do

  • Build and operate deployment pipelines for AI services, covering CI/CD, infrastructure-as-code, environment promotion, and rollback
  • Deploy and operate model serving, batch scoring, and orchestration pipelines across development, staging, and production
  • Partner with data scientists and applied AI engineers to take prototypes into production, including system design for net-new services
  • Own production reliability for AI services: monitoring, alerting, debugging, performance, and cost
  • Spot repeated patterns and turn them into reusable templates, so the team can ship its second and third variant of something without rebuilding it

What You'll Need

  • Six or more years in infrastructure, DevOps, platform, or ML engineering, with ownership of systems running in production
  • Deep hands-on experience across a wide range of AWS services, including compute, networking, storage, deployment, and monitoring
  • Infrastructure-as-code experience (Terraform, CDK, or CloudFormation)
  • Experience with containers and modern deployment patterns (Docker required, Kubernetes or ECS/EKS a plus), applied to CI/CD pipelines you've designed and operated for production services
  • Experience with orchestration and workflow tooling (Airflow, Dagster, Argo, Step Functions, or similar)
  • Comfort working through ambiguity and collaborating directly with data scientists and researchers

It Would Be a Bonus If You Had

  • Experience with ML platform components and data pipeline orchestration at scale
  • Experience running LLM-based or retrieval-based systems in production
  • Experience operating specialized data stores, including graph databases
  • Experience building internal tooling, templates, or reference implementations that other engineers adopted

Onsite Collaboration Requirement:
This role requires working onsite on Tuesday and Wednesday, with Thursday strongly encouraged as part of our company’s in-person collaboration model.

Why Join Us

Join us and help shape the future of education by turning cutting-edge AI into reliable product capabilities.

At Instructure, we're on a mission to help educators and students learn together, anytime, anywhere, and however works best. You'll join our research-driven team tackling education's biggest challenges with cutting-edge technology.

We value diversity, creativity, and passion, and invest in our teams through mentorship, hack weeks, internal conferences, and a culture where innovation thrives. Here, you'll have the chance to build the next generation of LMS features that make a real impact on students and teachers, and do it in a collaborative, supportive environment that encourages experimentation and growth.

Get in on all the awesome at Instructure!
We offer competitive, meaningful benefits in every country where we operate. While they vary by location, here's a general idea of what you can expect:

  • Competitive compensation, plus all full-time employees participate in our ownership program - because everyone should have a stake in our success.
  • Flexible work culture. Our remote, hybrid and in-office collaboration spaces vary by role, team and location.
  • Generous time off, including local holidays and our annual “Dim the Lights” period in late December, when teams are encouraged to step back and recharge based on departmental needs.
  • Comprehensive wellness programs and mental health support
  • Learning and development resources, including professional development tools and tuition reimbursement, to support your growth
  • The technology and tools you need to do your best work
  • Motivosity employee recognition program
  • A culture rooted in inclusivity, support, and meaningful connection

We believe in hiring great people and treating them right. The more diverse we are, the better our ideas and outcomes.

Instructure is an Equal Opportunity Employer. We comply with applicable employment and anti-discrimination laws in every country where we operate.

All employees must pass a background check as part of the hiring process. To help protect our teams and systems, we’ve implemented identity verification measures. Candidates may be asked to verify their legal name, current physical location, and provide a valid contact number and residential address, in accordance with local data privacy laws.

Any attempt to misrepresent personal or professional information will result in disqualification.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Cloud Engineer

Senior DevOps Engineer with Golang (f/m/x)

Sii Poland

Удалённо

Откликнуться на сайте работодателя

Описание вакансии

We are looking for a skilled Senior DevOps Engineer to join a project focused on enhancing and modernizing client's cloud platform. This role offers the flexibility of remote work, allowing you to contribute to significant improvements in cloud infrastructure and automation processes. If you are passionate about cloud technologies and automation, this opportunity may be the right fit for you.

Your tasks

  • Building and maintaining platform integrations and automation workflows
  • Developing solutions using Go/Golang to enhance cloud operations
  • Implementing Infrastructure as Code using Terraform
  • Integrating cloud APIs for VMware, Azure, and AWS platforms
  • Creating and managing CI/CD pipelines to streamline deployments
  • Utilizing PowerShell and Bash scripting for automation
  • Managing Kubernetes and container platforms for efficient application deployment
  • Applying configuration management tools like Ansible to maintain system integrity

Requirements

  • Minimum 5 years of experience in DevOps or cloud engineering
  • Hands-on mastery of Go/Golang development
  • Experience with VMware and cloud API integrations
  • Knowledge of AWS or Azure services
  • Advanced level of English

Nice to have

  • Practical experience in Azure and/or AWS platform engineering
  • Familiarity with Ansible or similar configuration management tools
  • Strong command of Terraform and Infrastructure as Code principles
  • Familiarity with CI/CD platforms and automation pipelines
  • Experience with Kubernetes and container orchestration

Job no. JOB-R8PYA

Sii ensures that all hiring decisions are made solely on the basis of qualifications and competence. We are committed to equal and fair treatment of all, regardless of legally protected characteristics. At Sii, we promote a diverse and inclusive work environment, in full compliance with applicable anti-discrimination laws.

Benefits For You

  • Great Place to Work
  • Solid financial situation
  • Contracts with the biggest brands
  • Centre of internal trainings
  • Many experts you can learn from
  • Open and accessible management team
  • Profit sharing
  • Passion Sponsorship program
  • Regular integration events and trips
  • Comfortable and well-equipped offices
  • MySii app
  • Medical care

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Это все вакансии по этой категории в этой стране

Lead Site Reliability Engineer - Imunify Reliability PlatformJobgether

Откликнуться на сайте работодателя