Skip to content
Site Reliability Engineer

Site Reliability Engineer

DevOps/SRE Engineer (Proxy)

SOAX

Tbilisi

Apply on the employer's site

Role description

Overview

SOAX is a proxy infrastructure company with a product customers genuinely love and depend on. Our system is a geo-distributed, high-load platform written in Go, handling HTTP/HTTPS/SOCKS5/TCP/UDP traffic across millions of nodes worldwide.

We've passed the MVP stage, grown fast, and now our infrastructure needs to catch up with the product. We're looking for someone who gets excited by that challenge — not someone who waits for a ticket that spells out exactly what to do (we'll hand you the problem, not the instructions), but someone who will SSH into a box at 2 AM, find the root cause, fix it — with AI in the toolkit, of course — and then build the system so it never happens again.

Fair warning: we have our share of tangled infrastructure — the same problem solved three different ways in three different places over time, all of it needing maintenance until we untangle it. If mismatched patchwork solutions make you twitch and want to fix them properly, you'll fit right in.

This role is for someone who genuinely likes a challenge and is driven by results, who values personal freedom and flexibility (fully remote, four-day work week) — but who also wants real intensity: hard problems, real ownership, and a startup pace.

What you're walking into

  • A geo-distributed system routing proxy traffic across bare-metal nodes globally — millions of nodes, real traffic, real money riding on it.
  • The tangle we mentioned: overlapping fixes, inconsistent patterns, and infrastructure that works but wasn't always built the same way twice.
  • Low-level networking challenges: routing, load balancing, protocol-level debugging.
  • A lean team where your impact is immediately visible.
  • An on-call rotation shared with one other engineer.

What you'll actually do

  • Troubleshoot and fix things.
    Incidents happen. You dig in, find root causes, and resolve them — not "coordinate the response," but actually do the work.
  • Rebuild infrastructure properly.
    Take what we have and make it production-grade: monitoring, alerting, deployment pipelines, disaster recovery.
  • Work at the network level.
    Debug packet flows, optimize routing, understand why latency spikes between nodes in different regions. This is not a "deploy containers and forget" role.
  • Automate everything that should be automated.
    If you're doing something manually twice, build a system for it.
  • Drive your own strategic projects.
    Beyond firefighting, you'll identify what needs fixing at a structural level, propose it, get alignment, and own it end-to-end. This isn't a role where someone else does all the thinking and hands you a backlog.

What we need from you (non-negotiable
)

  • Networking fundamentals you can defend live
    — subnetting, NAT, routing, DNS, TLS. Not textbook definitions, but the ability to reason through a scenario out loud.
  • Real bare-metal experience
    — actual hands-on with physical infrastructure at some point in your career.
  • DevOps as genuine curiosity, not just a job
    — something you built or broke on your own time, not because a ticket told you to.
  • A proven root-cause habit
    — a specific story where you found the actual cause of an issue, not just where you patched the symptom.
  • 5+ years hands-on in DevOps / SRE / Infrastructure roles.
    (3–5 years are welcome to apply, but should be ready to demonstrate senior-level systems thinking live.)
  • Comfort with on-call and incident response.

Strong plus

  • Proxy protocols, traffic routing at scale, or CDN/edge experience.
  • Self-hosted ClickHouse, Redis, Kafka, MongoDB.
  • English good enough for a live technical conversation; Russian a plus but not required.

What we offer

  • Fully remote, four-day work week.
  • Compensation paid in GBP.
  • Reporting directly to the CTO.
  • Paid time off and sick leave.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Workplace System Engineer

TBC

Tbilisifulltime

Apply on the employer's site

Role description

We are TBC - a leading financial institution in the region and the winner of The Banker’s Technology Awards 2026 in Central and Eastern Europe, powered by a courageous and goal-oriented team that fulfills our mission and makes people’s lives easier.

We are setting a new standard of service using data analytics,
artificial intelligence
, and digital products.

We’re here to support your growth and success - just believe in yourself.

Job Description
About the Role
We are seeking a highly qualified and experienced professional to assume the role of Workplace Systems Engineer in the Workplace Systems Engineers Group.

This position requires demonstrated capabilities, deep technical expertise, and the ability to guide complex operations within the WSE domain.

Key Responsibilities

  • Ensure high availability, stability, performance, and reliability of related systems
  • Ensure high-quality IT services that enable productivity and collaboration
  • Design and maintain end‑user computing environments
  • Manage operating system lifecycle, including deployment, updates, and configuration
  • Automate tasks and processes to improve efficiency and reduce manual effort
  • Deploy, update, and maintain applications across organizational devices.
  • Ensure compliance with organizational & security policies and technical standards
  • Collaborate with cross‑functional teams to support broader infrastructure and security initiatives.
  • Contribute to continuous improvement of workplace technologies, processes, and user experience.

Qualifications
Required Qualifications & Experience
The ideal candidate should have strong technical knowledge and be capable of:

  • Administering and supporting Microsoft 365 services, Exchange Online, Microsoft Teams, SharePoint Online, and OneDrive.
  • Managing Microsoft Entra ID, identity security, Conditional Access policies, Single Sign-On (SSO).
  • Deploying and maintaining Microsoft Intune solutions for endpoint management, device compliance, and application deployment.
  • Implementing and supporting Microsoft 365 security and compliance technologies, including Microsoft Defender and related security controls.
  • Troubleshooting and resolving complex technical issues across the Modern Workplace environment while ensuring service stability and user satisfaction.
  • Evaluate new Microsoft technologies and features, recommending enhancements to improve productivity and security
  • Experience working with cross-functional teams, Infrastructure, Information Security, and Business stakeholders.
  • 2+ years in IT operations or digital workplace
  • Fluency in English, with strong written and verbal communication skills
  • Willingness to learn and work with AI tools and modern technologies

Skills & Competencies

  • Strong communication skills with both technical and non-technical stakeholders, including clear incident and status communication.

Education & Certifications

  • Master’s degree in information technology, Computer Science, or related disciplines
  • AWS Certified Solutions Architect: Associate
  • Relevant professional certifications such as Microsoft -MCSA/MCSE
  • Cisco CCNA/CCNP, VMware, or cloud certifications such as Azure/AWS preferred
  • Microsoft certifications (Azure Administrator/Architect, M365 Expert, Security Engineer, etc.) preferred

Additional Information
TBC processes the personal data of the candidate in order to determine the suitability of the candidate for the vacancy, in accordance with the requirements of the Law of Georgia on Personal Data Protection. Information about the candidate may also be processed to determine the suitability of the candidate for future vacancies. Information about the candidate is stored for a maximum of 3 years.
In case you do not want further data processing, want to change or delete data
, please follow the link and contact us through the communication channels located at the same link https://tbcbank.ge/en/privacy\-policy

TBC shares its information with companies included in the TBC Bank Group PLC. Subsidiary companies also ensure personal data processing in accordance with the law.
If you do not wish to share
your data with TBC Group companies, please contact us at the same link https://tbcbank.ge/en/privacy\-policy

TBC conducts the candidate selection process in compliance with the requirements of the Law on the Elimination of All Forms of Discrimination and the principles of equal treatment, eliminating discrimination on any grounds. If you observe any signs of discriminatory treatment, you can report the incident and be assured that your notification will be followed by appropriate action. At TBC, you can raise your voice through various communication channels and report an incident either anonymously or openly:

We will contact only those candidates who pass the first stage of the selection process and are granted candidate status.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

DevOps Engineer (Deployment)

Intermedia Intelligent Communications

Tbilisifulltime

Apply on the employer's site

Role description

Department:
Tech Operations

Location:
Tbilisi, Georgia

Description

  • ALL CANDIDATES MUST BE LOCATED IN TBILISI GEORGIA**

About Intermedia
Are you looking for a company where
YOUR VOICE
is heard? Where you can
MAKE A DIFFERENCE
? Do you
THRIVE
in a
FAST-PACED
work environment? Do you wake every morning
EXCITED
to work with
GREAT PEOPLE
and create
SUCCESS
TOGETHER
? Then Intermedia is the place for you.

Intermedia has established itself as a leading provider of cloud communications and collaboration tech that allows companies to connect better. We have a strong track record of growth, profitability, and creating an environment where everyone matters. Everyone. While we are fast-paced and admittedly a bit intense, we promise that you won’t be bored. You will find Intermedia is a place where you can indulge your passion for creating and supporting great cloud technology. What’s more, we always look to promote from within and have many employees who have been with us 10, 15, and 20+ years!

Culture at Intermedia is built on teamwork and transparency. We hold each other accountable and always have each other’s back!
Are you ready to make your mark?
About The Role
We are looking for a Deployment DevOps Engineer to own and improve application deployment processes across environments. You will build and maintain deployment pipelines primarily using
Octopus Deploy
,
GitHub
, and
Ansible
, ensuring releases are automated, repeatable, secure, and reliable for production.

Key Responsibilities

  • Own end-to-end release and deployment lifecycle: build → package → deploy → verify → rollback
  • Develop and support Octopus Deploy projects, lifecycles, channels, variables, and deployment processes
  • Implement deployment automation with Ansible (playbooks/roles, inventories, idempotent changes)
  • Maintain Git-based release workflows in GitHub (branching, tagging, versioning, release notes)
  • Build/maintain CI pipelines in GitHub Actions (or existing tooling) to produce artifacts and trigger Octopus releases
  • Standardize deployment patterns across applications (templates, shared steps, reusable Ansible roles)
  • Manage environment configuration and secrets in a controlled way (variable sets, permissions, auditing)
  • Improve deployment safety: approvals, health checks, smoke tests, automated validation, and rollback strategies
  • Support production releases, troubleshoot deployment failures, and drive root-cause analysis
  • Maintain release documentation, runbooks, and change management practices
  • Collaborate with developers, QA, and operations to plan releases and reduce downtime

Skills, Knowledge And Expertise

  • Bachelors degree in Computer Science or related field
  • Experience as DevOps / Release / Deployment Engineer supporting production deployments
  • Strong hands-on experience with Octopus Deploy (process design, variables, multi-environment releases)
  • Strong Ansible skills for deployment and configuration (roles, vault/secrets, troubleshooting)
  • Strong working knowledge of GitHub (PR workflows, tagging/releases, permissions)
  • Understanding of CI/CD concepts (artifact management, versioning, promotion between environments)
  • Administration and troubleshooting skills on Windows and/or Linux hosts used for deployments
  • Ability to diagnose deployment issues across app, OS, network, and configuration layers

Nice to have

  • Experience with Terraform and infrastructure provisioning
  • Scripting: PowerShell and/or Bash
  • Monitoring/observability basics (Prometheus/Grafana, logs) to validate releases
  • Experience with IIS/.NET deployments, Nginx, or reverse proxies (depending on stack)
  • Familiarity with container-based deployments (Docker/Kubernetes)
  • Experience in managing VOIP components and protocols (SIP , FreeSwitch, OpenSIP, session border controllers)
  • Experience with load balancing components ( F5 LTM, F5 GTM)
  • Experience with administering AWS or Azure tenants
  • Experience with Virtualization platforms such as VMWare or HyperV

Diversity, Inclusion, and Equal Opportunity
We hire, promote, and compensate employees based on their ability to perform their job responsibilities, without regard to race, color, creed, religion, sex, gender, marital status, national origin, ancestry, age, citizenship, physical or mental disability, sexual orientation, or any other basis protected by applicable law (collectively referred to in our Code of Conduct as “Protected Classes”). We do not tolerate employment discrimination in the workplace, and we are committed to making reasonable accommodations for identified disabilities or other limitations as required by all applicable laws. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

ქსელის მონიტორინგის ინჟინერი (NOC)

Cellfie Mobile

Tbilisi

Apply on the employer's site

Role description

ჩვენ ვაშენებთ TechCo-ს, უფრო მეტს ვიდრე ტრადიციულ ტელეკომ კოპანიას, რომელიც ქმნის ციფრულ პროდუქტებს, მუშაობს მონაცემებზე და სწრაფად გადააქცევს იდეებს რეალურ შედეგებად.

ჩვენთვის კავშირი არ ნიშნავს მხოლოდ ქსელს. ეს ნიშნავს ადამიანებისა და შესაძლებლობების დაახლოებას ციფრული პროდუქტების განვითარებით, მონაცემებზე დაფუძნებული გადაწყვეტილებებით და რაც ყველაზე მნიშვნელოვანია - მომხმარებლისთვის უნიკალური პერსონალიზებული გამოცდილების შექმნით.

ჩვენი მისია
მარტივი და ამბიციურია:

„ვქმნით შესაძლებლობებს ადამიანებისთვის და ვაახლოებთ მათ ერთმანეთთან"
სელფიში მუშაობა ნიშნავს იყო ტრანსფორმაციის მონაწილე, რომელიც განსაზღვრავს როგორი იქნება საქართველოს ციფრული სამყარო მომდევნო წლებში.

სამუშაოს აღწერილობა
სელფიში ვეძებთ გუნდის ახალ წევრს
ქსელის მონიტორინგის ინჟინრის (NOC)
როლზე.

სამუშაო ადგილი: თბილისი

ამ როლში შენ:

  • უზრუნველჰყოფ სატელეკომუნიკაციო ინფრასტრუქტურის 24/7 მონიტორინგს (მაგ.: საბაზო სადგურები და ქსელური მოწყობილობები), რათა ქსელი მუშაობდეს სტაბილურად
  • დროულად დააიდენტიფიცირებ ინციდენტებს და მოახდენ მყისიერ რეაგირებას მათი სწრაფად აღმოსაფხვრელად
  • კოორდინირებას გაუწევ და გააკონტროლებ დაგეგმილ და გადაუდებელ ტექნიკურ სამუშაოებს, რათა მინიმუმამდე დაიყვანო შეფერხება
  • გააანალიზებ ქსელის სტატისტიკას, გამოავლენ ტრენდებს და შესთავაზებ გაუმჯობესების შესაძლებლობებს
  • იმუშავებ შიდა სტანდარტებისა და პროცედურების დაცვით, რათა უზრუნველყო პროცესების ეფექტურობა და ხარისხი

თუ ეს როლი გაინტერესებს, აუცილებელია გქონდეს:

  • უმაღლესი განათლება ტელეკომუნიკაციების, IT-ის ან შესაბამის ტექნიკურ მიმართულებაში
  • ქსელის მონიტორინგის მიმართულებით მუშაობის მინიმუმ 1+ წლიანი გამოცდილება
  • ქსელის მონიტორინგის სისტემებთან მუშაობის გამოცდილება (Grafana, PRTG, Zabbix ან მსგავსი)
  • ამოცანების მართვის სისტემებთან მუშაობის გამოცდილება (ServiceDesk, Jira)
  • ინგლისური ენის ცოდნა სამუშაო დონეზე
  • მაღალი პასუხისმგებლობის გრძნობა
  • სწრაფი სწავლისა და ცვლილებებთან ადაპტაციის უნარი
  • დამოუკიდებლად მუშაობისა და პრიორიტეტების მართვის უნარი
  • ანალიტიკური აზროვნება და პრობლემების სწრაფად იდენტიფიცირების უნარი
  • ოპერატიულად გადაწყვეტილებების მიღებისა და ინციდენტების ეფექტურად მართვის უნარი

გამოგვიგზავნე განაცხადი და გახდი ჩვენი გუნდის წევრი
განაცხადის მიღების ბოლო ვადაა
31 აგვისტო, 2026.
დამატებითი ინფორმაცია
რატომ სელფი
ჩვენ ვმუშაობთ მკაფიო მიზნებით და გაზომვადი შედეგებით. ჩვენთან ყველა ინიციატივას აქვს რეალური ბიზნეს ეფექტი და ქმნის ღირებულებას მომხმარებლებისათვის.

ჩვენ ვაფასებთ:

  • პასუხისმგებლობას (Ownership)
  • საქმის ხახრისხიან შესრულებას
  • პროფესიული ზრდისა და ერთმანეთის განვითარების სურვილს.

სელფი მობაილში შეგიძლია:

  • დასახო ამბიციური მიზნები
  • თამამად გამოხატო შენი აზრი
  • ჩაძიებით მოძებნო უკეთესი გამოსავალი
  • ისწავლო და ასწავლო სხვებს.

ჩვენი კულტურა ემყარება ხუთ პრინციპს:

  • ფოკუსი ადამიანზე - თანამშრომლებზე და მომხმარებლებზე ზრუნვა
  • პასუხისმგებლობა - "შედეგებზე სრული პასუხისმგებლობის აღება, ფორმალური როლის ფარგლებს მიღმაც"
  • სიახლეების ძიება - კითხვები, ჩაძიება და იდეების ქმედებად გარდაქმნა
  • კეთილსინდისიერება - სიტყვა და მოქმედება თანმიმდევრულია
  • გუნდურობა - „ჩვენ“ ყოველთვის უფრო მაღალია „მე“-ზე.

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

Site Reliability Engineer

Senior Software Engineer (MLAI services)

Workato

Tbilisi

Apply on the employer's site

Role description

About Workato
Workato delivers enterprise infrastructure for the agentic era, redefining iPaaS and helping enterprises unify data, applications, processes, and AI into a single, governed platform. A leader in Enterprise MCP and trusted by 50% of the Fortune 500, Workato’s cloud-native architecture connects every application, data source, and process to power real-time orchestration at scale. With enterprise-grade security and continuous innovation at its core, Workato provides the trusted foundation for organizations to automate with confidence and operationalize AI across the business. To learn more, visit www.workato.com

Why join us?
Ultimately, Workato believes in fostering a
flexible, trust-oriented culture that empowers everyone to take full ownership of their roles
. We are driven by
innovation
and looking for
team players
who want to actively build our company.

But, we also believe in
balancing productivity with self-care
. That’s why we offer all of our employees a vibrant and dynamic work environment along with a multitude of benefits they can enjoy inside and outside of their work lives.

If this sounds right up your alley, please submit an application. We look forward to getting to know you!

Also, Feel Free To Check Out Why

  • Business Insider named us an “enterprise startup to bet your career on”
  • Forbes’ Cloud 100 recognized us as one of the top 100 private cloud companies in the world
  • Deloitte Tech Fast 500 ranked us as the 17th fastest growing tech company in the Bay Area, and 96th in North America
  • Quartz ranked us the #1 best company for remote workers

Responsibilities
We are looking for a
Senior Python Engineer
to play a key role in building the core of our AI platform. In this position, you will design and develop production-grade systems that power intelligent automation, agentic workflows, and large-scale retrieval services. This is a highly technical, hands-on role that involves close collaboration with product and platform teams to transform advanced AI concepts into reliable, scalable, and secure solutions used across our enterprise ecosystem. You will also be responsible to:

  • Design, build, and maintain AI-powered services and APIs, leveraging LLMs (OpenAI, Anthropic, Qwen, OSS models) and custom ML models.
  • Develop an enterprise-grade agentic framework that enables orchestration, retrieval, and collaboration between multiple AI agents.
  • Implement and optimize knowledge retrieval systems and agentic search capabilities using vector databases such as Qdrant and ElasticSearch.
  • Write well-structured, efficient, and testable Python code for production services, experimentation, and internal developer tools.
  • Build and maintain shared Python libraries and SDKs used across multiple applications and microservices.
  • Collaborate with cross-functional teams on architecture, internal protocols, and API standards to ensure consistency and reliability across the platform.
  • Develop and enhance monitoring, validation, and observability for production-grade AI solutions.
  • Drive the full software development lifecycle - from design and implementation to deployment, monitoring, and continuous improvement.
  • Identify and resolve performance bottlenecks, reliability issues, and scaling challenges in complex, data-intensive environments.
  • Participate in code reviews and technical discussions, mentoring other engineers and contributing to a culture of excellence.

Example Projects

  • Building an evaluation and observability framework for AI model performance and reliability.
  • Developing an agentic orchestration platform that enables collaboration among multiple AI agents and tools.
  • Implementing semantic retrieval and agentic search capabilities over large enterprise knowledge bases.
  • Designing AI services that process and reason over high-volume real-world data at scale.

Requirements
Qualifications / Experience / Technical Skills

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • 5+ years of experience as a Software Engineer, with strong proficiency in Python.
  • Proven track record of building and maintaining production-grade systems using Python.
  • Strong understanding of distributed systems, API design, and data-driven architectures.
  • Experience with relational and non-relational databases (PostgreSQL, Elastic, Qdrant, or similar).
  • Familiarity with AI/ML system design, including LLM integration and evaluation pipelines.
  • Knowledge of DevOps and observability practices (CI/CD, monitoring, metrics, and model validation).
  • Python
  • FastAPI
  • LLM APIs (OpenAI, Anthropic, Qwen, OSS)
  • LiteLLM
  • Qdrant
  • PostgreSQL
  • ElasticSearch
  • Langfuse
  • Kubernetes
  • GitHub Actions
  • ArgoCD

Preferred Skills

  • Experience working with multiple LLM providers (OpenAI, Anthropic, Qwen, open-source models).
  • Background in developer platforms or AI infrastructure services.
  • Familiarity with vector databases, semantic retrieval, and knowledge graph architectures.
  • Exposure to Langfuse, LiteLLM, LangChain, or similar frameworks.
  • Experience developing enterprise-scale SaaS or distributed backend systems.
  • Contributions to open-source projects in Python, AI, or infrastructure engineering.

Soft Skills / Personal Characteristics

  • Excellent communication skills, with the ability to convey complex technical ideas clearly to both technical and non-technical audiences.
  • Collaborative and proactive approach, comfortable working across teams in a dynamic environment.
  • Strong analytical and problem-solving abilities, with a focus on continuous improvement and innovation.
  • Curiosity and a genuine interest in emerging AI technologies and modern backend architectures.

(REQ ID: 2459)

This is a saved copy of a posting published elsewhere. Postings get taken down without notice — check the employer's site before applying. mentors.coach is not the hiring party.

That is every opening in this category and country

DevOps/SRE Engineer (Proxy)SOAX · Georgia

Apply on the employer's site
DevOps/SRE Engineer (Proxy) — SOAX | mentors.coach