Skip to content
Backend Engineer

Backend Engineer

Research Internship: Efficient Inference for Text-to-Speech Models and LLMs

AGIGO

Zurich

Откликнуться на сайте работодателя

Описание вакансии

QUANTIZATION, PRUNING, SPARSITY

‍

Full-time | Voice & Conversational AI | Enterprise AI | Speech AI Team

Duration:
6 Months (flexible)

Location:
Switzerland (Europe), on-site at AGIGO’s Zurich Office

‍

About AGIGO

AGIGO provides the enterprise-grade conversational AI infrastructure and end-to-end toolchain to design and operate high-agency, human-like AI agents that engage directly with customers over phone, email, and text, handling complete customer interactions across support, bookings, and sales. AGIGO stands out by offering true AI sovereignty through on-premises deployment and zero exposure to third-party services. Powered by AGIGO’s proprietary technology stack, the platform delivers reliable agent operations, execution assurance, ultra-low latency, seamless enterprise integration, and predictable, token-free economics.

Founded in Switzerland in February 2025 by a team of experienced AI pioneers, AGIGO is building the infrastructure for a new generation of enterprise customer interactions, combining human-like communication with the control, reliability, and economics enterprises require at scale.

‍

Your Research Mission

Real-time voice agents require very low-latency decoding and increased serving costs compared to offline models. The latency becomes an extremely differentiating aspect, since a reply from an Voice Agent arriving after one second starts feeling broken. In this internship, you will
build a complete pipeline
for quantizaiton/quantization-aware training, and pruning for our internal LLMs and TTS models. Ideally, one initial checkpoint is compiled into a family of variants, each valid for a particular GPU generation (ADA/Hopper/Blackwell), precision, kernel stack, and batching regime, and every variant has to clear the same speech-aware quality gate before it is allowed out. Therefore, an important question arises: given the hardware and the latency requirements, which build are we allowed to serve? This matters because our stack is not uniform, but rather fluid, with Hopper, Blackwell or ADA machines requested on demand, which also might reward different recipes, so the same model has a different best answer depending on the initial conditions.

‍

What You Will Build

  • The build matrix.
    An initial checkpoint in, let’s say FP16/BF16 precision, then build a matrix from: weight-only 4-bit, activation quantisation with outlier handling, FP8 and NVFP4, KV-cache quantisation, and 2:4 sparsity where the hardware can use it. Each build is tagged with the hardware and workload it is valid for.
  • The quality gate.
    You will implement strong evaluation pipelines to signal issues in a quantized checkpoint, e.g., UTMOS, word error rate or speaker similarity for TTS, or other metrics such as intent accuracy or NER performance for LLMs.
  • Elastic serving and speculative decoding
    . One checkpoint offering several operating points chosen per call rather than per deployment; and speculative decoding for specific LLMs tasks or TTS.

Phase 1: Harness and gate

The benchmark harness and the CI quality gate, plus the schema for a build record: what it is valid for, and what it measured.

Phase 2: Post-training quantisation

With and without in-domain calibration data.

Phase 3: Recovery

Quantisation-aware training and distillation, aiming to beat post-training quantisation at the same bit-width.

Phase 4: Sparsity

2:4 pruning, and investigate whether structured sparsity becomes real throughput.

‍

Key Research Challenges

Does the best build actually differ by hardware, and by how much?
If both generations rank the variants the same way.

Is speculative decoding lossless for audio?
You will investigate in which conditions speculative decoding for audio is lossless.

‍

Your Impact

The resulting recipe of this internship will translate on optimizations in the models deployed in our stack, where each latency point and increase in tok/s really matters.

We value original thinking and encourage you to help shape and redefine the project’s direction as your research uncovers new insights. AGIGO fosters an open, collaborative environment where ideas can evolve freely. Exceptional innovation often emerges where disciplines and perspectives intersect, and we actively support creative exploration that pushes the boundaries of what Voice-AI can achieve.

‍

What You Bring

Required

  • Current Master’s or PhD student (preferred), or recent graduate in Computer Science, Machine Learning, or a related degree field
  • Strong Python programming skills and Git
  • Solid understanding of ML fundamentals and MLOps
  • Hands-on experience with PyTorch
  • Expertise using Claude Code/Codex
  • Fluent in English, highly motivated, willingness to learn

Bonus

  • CUDA, mixed precision, profiling, inference servers, or quantization tools such as: llm-compressor (vLLM) or Model-Optimizer (NVIDIA).

‍

What You Will Gain

  • Direct impact on our product: your code ships in our platform, built alongside our researchers and engineers
  • Mentorship: work closely with our expert team of researchers and engineers
  • Top-tier AI infrastructure: access to GPU clusters with NVIDIA Hopper (H200) and Blackwell RTX 6000 PRO NVIDIA GPUs
  • Research visibility: we will actively support you in publishing your work at a top-tier conference or in a journal paper
  • Disciplined and inspiring research environment: a team of sharp minds grounded in expertise, autonomy, and a shared pursuit of impactful breakthroughs
  • Paid internship:
    market-level salary, flexible hours, unlimited coffee, drinks, fruit and snacks
  • Career path: this internship may lead to a full-time permanent role in AGIGO's world-class AI R&D team

‍

How to Apply

To apply, please send your resume and a brief introduction to internships@agigo.ai with the subject line:

Research Internship – Efficient Inference for Text-to-Speech Models and LLMs – [Your Full Name]

For more information:

👉 https://www.agigo.ai/job\-openings/research\-internship\-efficient\-inference\-for\-text\-to\-speech\-models\-and\-llms

‍

By submitting your application, you agree to allow AGIGO to store and process your data for recruitment purposes. Unless otherwise requested, we may retain your data for up to one year to consider you for this or other future opportunities.

‍

Research in the Field

[1] AWQ: Activation-aware Weight Quantization for LLM Compression, MLSys 2024. https://arxiv.org/abs/2306\.00978

[2] SmoothQuant: Post-Training Quantization for Large Language Models, 2022. https://arxiv.org/abs/2211\.10438

[3] Principled Coarse-Grained Acceptance for Speculative Decoding in Speech, ICASSP 2026. https://arxiv.org/abs/2511\.13732

AGIGO™ is a registered trademark of AGIGO AG, Switzerland.

‍

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Скоро на этой странице

Резюме под эту вакансию — и билет в розыгрыш

Мы разбираем объявление до настоящих требований и переписываем ваше резюме под него — вопросами, а не выдумкой: ни одна строка не появится без вашего подтверждения. Войдите, чтобы получить это первым, — и попасть в розыгрыш.

  • Резюме под конкретную вакансию, а не «универсальное»
  • Ответы хранятся: правится любой, а не весь разговор заново
  • Всё в аккаунте — открывается с любого устройства

Разыгрываем

Скидка на сопровождение

Победителей выбираем случайно среди заявок с подтверждённой почтой. Дата розыгрыша и полные правила — на странице розыгрыша.

Правила розыгрыша

Backend Engineer

Forward-Deployed Software Engineer

Hoshii

Zurichfulltime

Откликнуться на сайте работодателя

Описание вакансии

The inbox was never built to run a business. Orders get missed, RFQs go cold, and every line gets retyped into the ERP by hand — one email, one PDF, one WhatsApp message at a time. We built AI Coworkers that do most of the repetitive work nobody wants to do instead: they read what comes in, sort it, draft linked actions, and post clean entries straight into the customer's systems of record. The team approves; Hoshii does the rest.

We are the ones driving, train-riding and flying to customers, running the workshops, and staying up past the demo to make the PoC actually work — and we can't scale that forever. We are looking for a Forward-Deployed Software Engineer to become the technical face of Hoshii in front of the companies who actually run on our product: someone who can walk into a customer's office, understand exactly where the manual pain lives, build a working proof of concept in days rather than weeks, and take it all the way to production — then stay accountable for it once it's live. You'll follow Amazon's "you build it, you run it" principle: whatever you ship for a customer, you own operationally too. You will join the founding team building the AI Coworker that operations teams actually run their order desk on.

Tasks

  • Run on-site and remote discovery directly with customers: map their order-to-cash workflows, identify the real pain points behind the stated requirements, and translate them into technical scope.

  • Build and demo working proofs of concept fast, iterating directly with customer stakeholders rather than through a requirements document.

  • Own end-to-end deployment: integrate Hoshii's agents with the customer's ERP/CRM systems, configure the workflow, and take it live in production.

  • Follow the
    "you build it, you run it"
    principle: stay accountable for what you deployed, monitor it in production, and be the first responder when something breaks for that customer.

  • Feed field learnings back into the product roadmap — you're often the first to see where the product breaks down against a real workflow, and that signal needs to shape what we build next.

  • Work closely with our channel partners to keep the underlying integrations reliable across every customer running on their systems.

  • Write clean, tested, maintainable code — you're not just configuring, you're building.

Requirements

We are looking for team players with a growth mindset, an entrepreneurial spirit, and genuine comfort being the person in the room with the customer:

  • Strong fullstack coding and debugging skills — comfortable moving across frontend, backend, and infra as the deployment demands — with proven experience building and shipping production software (
    Python
    ,
    Javascript/Typescript
    , or similar).

  • Proven experience building and running
    GenAI PoCs
    end-to-end: prototyping fast with LLMs/agents, then hardening what works into something a customer can actually run in production.

  • Solid understanding of systems and how they integrate — APIs, data pipelines, microservices, event-driven architectures — and the judgment to know where deep integration is worth the effort and where it isn't.

  • A pragmatic
    80/20 instinct
    : you ship the 20% that unblocks the customer's real problem instead of gold-plating the other 80%.

  • Experience working directly with customers — PoCs, technical scoping, on-site deployments, or a similar consulting/pre-sales-adjacent engineering background.

  • Familiarity with
    ERP/CRM systems
    is a strong plus.

  • Genuinely
    product-minded
    : you instinctively separate what a customer says they want from what they actually need, and you enjoy shaping the roadmap, not just executing it.

  • Excellent communication skills in
    English
    and
    German (C1 or above)
    ;
    Swiss German
    is a strong plus.

  • Comfortable with regular travel to customer sites across DACH.

  • Strong sense of ownership
    : you treat what you deploy as yours to run, not just to hand off.

Benefits

  • Autonomy and Ownership
    : You're in charge of your time at Hoshii; decide what you want to do, align with the team, and go for it!

  • Stock Options
    : Ever owned your impact? Well, at Hoshii you get stock options, allowing you to get something back from our collective success story.

  • A World-Class Team
    : We're a team of servant leaders who work hard and put customers first; we hail from top companies and universities, and we hold ourselves to the highest standards.

  • Amazing Office Location
    : We have a lakeside location in Zürich with cozy corners to relax, fostering a comfortable environment where you can thrive.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Backend Engineer

Senior SAP Treasury & Risk Management Consultant (m/w/d) - Freelance

Sidekick Network

Basel

Откликнуться на сайте работодателя

Описание вакансии

Für die Erweiterung eines bestehenden S/4HANA-Transformationsprojekts im Transportation-Umfeld suchen wir einen erfahrenen
SAP Treasury & Risk Management Consultant (m/w/d)
. Im Mittelpunkt stehen die Konzeption und Umsetzung einer neuen Treasury-Landschaft auf Basis von SAP S/4HANA sowie der nachhaltige Know-how-Transfer in das interne Team.

Projektrahmen

  • Start:
    Oktober 2026

  • Laufzeit:
    Ende 2027

  • Auslastung:
    80–100 %

  • Einsatzort:
    überwiegend remote

  • Vor-Ort-Anteil:
    ca. 10–15 % in Bern, Schweiz

  • Vertragsart:
    Freelance

  • Sprache:
    Deutsch (muss) und Englisch

Aufgaben

  • Konzeption und Implementierung von S/4HANA-Lösungen im Bereich Treasury & Risk Mgmt. einschließlich Analyzer und FX Exposure Mgmt.

  • Customizing und Umsetzung der entwickelten Lösungen mit konsequenter Nutzung des SAP-Standards

  • Konzeption und Anbindung von Handelsplattformen sowie angrenzenden SAP- und Non-SAP-Systemen über SAP TPI

  • Analyse und kritische Bewertung fachlicher und technischer Anforderungen sowie Entwicklung geeigneter Lösungsvarianten

  • Aktive Mitarbeit bei Design, Implementierung, Datenmigration, Testing und Fehleranalyse

  • Unterstützung des stabilen Betriebs der bestehenden SAP-Landschaft einschließlich Support und Incident Management

  • Erstellung von Analysen, Entscheidungsgrundlagen sowie strukturierter und nachvollziehbarer Dokumentation

Qualifikation

  • Mehrjährige praktische Erfahrung mit SAP S/4HANA Finance und Schwerpunkt auf Treasury & Risk Management

  • Fundierte Projekterfahrung mit SAP TRM einschließlich Analyzer

  • Nachweisbare praktische Erfahrung mit FX Exposure Mgmt.

  • Sehr gute Kenntnisse im S/4HANA-Customizing sowie in der Implementierung entsprechender Lösungen

  • Erfahrung mit Datenmigration, vorbereitenden Migrationsaktivitäten und Systemintegrationen

  • Sehr gutes Verständnis durchgängiger Treasury-Prozesse sowie die Fähigkeit, komplexe Anforderungen in umsetzbare Lösungsdesigns zu überführen

  • Erfahrung aus komplexen S/4HANA-Transformations- oder Implementierungsprojekten

  • Deutsch mindestens auf C1-Niveau und Englisch mindestens auf B2-Niveau

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Backend Engineer

IT Application Manager 100% (gn)

SPS

Opfikon

Откликнуться на сайте работодателя

Описание вакансии

WIR VERBINDEN DIE PHYSISCHE UND DIE DIGITALE WELT

SPS ist ein führender Outsourcing-Anbieter für Business Process Services, innovativer Dienstleistungen im intelligenten Datenmanagement sowie für die hybride Arbeitswelt. Schweizer Qualität, kombiniert mit globaler Präsenz, tiefer Branchenkenntnis und neusten Technologien, machen SPS zu einem verlässlichen Partner für die Transformation von Unternehmen.

Mit ganzheitlichen End-to-End-Lösungen unterstützt SPS Unternehmen dabei, ihre Geschäftsprozesse effizienter, agiler und zukunftsorientiert zu gestalten. Über 8.500 Mitarbeitende und spezialisierte Partner betreuen Kunden in mehr als 20 Ländern – mit einem Schwerpunkt auf den Branchen Banken, Versicherungen und Gesundheit.

DEINE AUFGABEN

In dieser vielseitigen Funktion bist du für den Betrieb, die Weiterentwicklung und die Optimierung unserer Dokumentenmanagement- und Automatisierungslösungen verantwortlich. Dabei arbeitest du eng mit internen Teams zusammen und leistest einen wichtigen Beitrag zur Digitalisierung der Geschäftsprozesse unserer Kunden.

  • Installation, Konfiguration, Wartung und Weiterentwicklung unserer Dokumentenmanagement- und Automatisierungslösungen
  • Sicherstellung eines stabilen und sicheren Applikationsbetriebs für Kunden aus den Bereichen Banken, Versicherungen und öffentliche Verwaltung
  • Verantwortung für ISAE-relevante Systeme sowie Sicherstellung der Einhaltung von Compliance-, Sicherheits- und Qualitätsstandards
  • Mitarbeit bei internen und externen Audits sowie Umsetzung entsprechender Anforderungen
  • Analyse, Fehlerbehebung und Performance-Optimierung von Applikationen in enger Zusammenarbeit mit den Entwicklungsteams
  • Identifikation von Optimierungspotenzialen in bestehenden Systemen, Services und Betriebsprozessen
  • Mitarbeit in IT-Projekten sowie Unterstützung bei der Einführung neuer Lösungen und Technologien
  • Weiterentwicklung und Automatisierung von Services und Betriebsabläufen
  • Erstellung und Pflege technischer Dokumentationen und Betriebsdokumentationen in Deutsch und Englisch
  • Übernahme von Aufgaben im Application Management sowie kontinuierliche Optimierung der Betriebsprozesse

DEIN PROFIL

  • Abgeschlossene Ausbildung oder Weiterbildung in Informatik oder Wirtschaftsinformatik (HF/FH oder vergleichbar)
  • Mehrjährige Berufserfahrung im Application Engineering, System Engineering oder einer vergleichbaren Funktion
  • Erfahrung in der Planung, Implementierung und dem Betrieb von IT-Systemen und Applikationen
  • Gute Kenntnisse in Windows-Server-Umgebungen sowie Datenbanken; Linux-Kenntnisse sind von Vorteil
  • Analytische und strukturierte Arbeitsweise sowie Freude an der Lösung technischer Problemstellungen
  • Hohe Service- und Kundenorientierung sowie ausgeprägte Kommunikationsfähigkeiten
  • Interesse an modernen IT-Technologien und Bereitschaft zur kontinuierlichen fachlichen Weiterentwicklung
  • Freude an der Zusammenarbeit in interdisziplinären Teams
  • Sehr gute Deutsch- sowie gute Englischkenntnisse in Wort und Schrift
  • Ein einwandfreier Strafregister- sowie Betreibungsregisterauszug wird vorausgesetzt.

Deine Benefits

  • Ein zukunftsorientierter Arbeitgeber, der die Digitalisierung der Schweizer Wirtschaft vorantreibt
  • Ein dynamisches Umfeld in einer Schweizer Unternehmung mit globalen Strukturen
  • Zentraler Arbeitsplatz in Glattbrugg mit Möglichkeit zum Homeoffice
  • Beteiligung an Reka-Checks sowie Gesundheitsvorsorge/ Fitnessabo
  • Zuschüsse für ÖV-Abos und gratis Business Halbtax-Abo
  • Attraktive Mitarbeiterrabatte für viele Marken und Produkte

Wir setzen uns aktiv für Gleichstellung und Diversität ein. Alle qualifizierten Bewerbenden werden unabhängig von Geschlecht, Geschlechtsidentität, Alter, Herkunft, Religion, sexueller Orientierung oder Beeinträchtigung berücksichtigt.

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Backend Engineer

Enterprise Architect - ETRM

Alpiq

Olten

Откликнуться на сайте работодателя

Описание вакансии

Olten - 100% | Permanent

D&C Technology is a functional unit that supports Alpiq’s Business Area Digital & Commerce with management systems for the trading, origination and energy sales businesses, including the delivery of digital solutions for customers from the industrial, commerce and energy sectors.

We are hiring an Enterprise Architect - ETRM who could be based either in Lausanne or Olten, CH.

Your primary responsibilities will be to establish and maintain an architectural vision, provide architecture support and guidance to technology and IT activities in line with the vision, drive the shift towards a simplified, microservices architecture and the reuse of assets and services and to pursue the changes requiredto achieve business and technology.

Your accountabilities

  • Establish and maintain a technology strategy of D&C, including architectural and other design/technology/security standards in line with enterprise standards, policies and guidelines, and represent D&C interests in various enterprise architecture workgroups
  • Establish, maintain and ensure adherence to architectural principles, methodologies, frameworks and approaches, including best practices of cloud architectures, either using the relevant Design Authority, project or line management
  • Ensure architecture activities, decisions and recommendations are motivated by clear business benefits with transparency on risks, ensure follow-up communication activities related to achieved business benefits and/or risk management activities
  • To maintain an architecture repository for D&C in accordance with the agreed architecture framework, capturing architecture documentation and agreed architectural views
  • To contribute to project ideas, projects, demands, and PoCs by-Supporting the scoping of new activities, services and changes to existing services-Analyzing and translating business, information, technical requirements into an architectural blueprint that outlines solutions to achieve the business objectives, ensuring that the blueprint capture both functional and non-functional requirements-Communicating, presenting and iterating architectural views and supporting material with development teams, management, and business stakeholders.-Providing landscape-wide knowledge across multiple technical areas and business segments to ensure dependencies and risks are accurately captured-Ensuring technical solutions and changes on the D&C technology landscape are in line with the architectural vision, principles and standards set out by the relevant design authority-Elaborating and/or reviewing low-level architecture design together with Software Engineers-Perform architectural reviews and prepare conformance reports for development teams, Business Partners, business stakeholders and management
  • Contribute to the continuous improvement of IT governance, policies and maturity within D&C and at an enterprise level
  • To work closely with Business Partners, Business Analysts, Project Managers, technical experts and project resources to achieve common objectives; especially to assist and review the work of peer Architects and Business Analysts
  • Adhere to the D&C Technology Information Security Management System in your area of responsibility

Your profile

  • 8+ years of experience and a background of working with enterprise/solutions architecture,systems integration, and software engineering
  • Minimum of a bachelor’s degree or an equivalent certification, ideally in an engineering, computer science, or similar technical discipline
  • Experience in architecture standards, frameworks and tools, including security best practices; TOGAF (or similar) certification, experience in UML 2.x, BPMN 2.0, and/or ArchiMate is preferred
  • Experience with architecting cloud solutions for the AWS cloud platform; AWS Architect certification and experience with designing secure, microservices and APIs is an advantage
  • Experience of software development methodologies (e.g. Agile, DevOps) and requirements engineering is expected, preferably with certifications, ideally including awareness of project management methodologies
  • Deep knowledge of the trading and/or energy sectors, including related business processes
  • Strong knowledge of how applications, technology and IT services can support business processes
  • Good communications, presentation and inter-personal skills; team and goal oriented, analytical and problem-solving skills are highly valued
  • Fluent in English; competency in German and French is advantageous

Your benefits

Competitive salary package

Market-oriented salary

Training and development

Diverse opportunities for career growth

Flexible work models

Various flexible work models

Это сохранённая копия объявления, опубликованного в другом месте. Вакансии снимают без предупреждения — перед откликом проверьте сайт работодателя. mentors.coach не является нанимающей стороной.

Ещё 259 вакансий по этой категории в этой стране

Research Internship: Efficient Inference for Text-to-Speech Models and LLMsAGIGO · Switzerland

Откликнуться на сайте работодателя