AI Developer with Python for Customer Care AI Platform team (f/m/d)
Über IONOS
Bei IONOS verwalten wir nicht nur Server, wir prägen mit unseren Lösungen die digitale Zukunft für über 6,2 Millionen Kundinnen und Kunden weltweit. Als Europas führender Hosting-Anbieter und Pionier für unabhängige Cloud-Lösungen bauen wir eine Infrastruktur der nächsten Generation: von souveränen Cloud-Architekturen und hochperformanten GPU-Clustern bis hin zu integrierten KI-Automatisierungstools.
Was uns antreibt: Digitale Entscheidungsfreiheit für Europa und echter Impact für kleine und mittlere Unternehmen (KMU) sowie Großunternehmen. Wir arbeiten in agilen Teams, setzen auf transparente Strukturen und glauben daran, dass Exzellenz und Innovation nur mit echtem Teamgeist entstehen.
Bereit für Deinen nächsten Schritt? Werde ein Teil von IONOS und wachse mit uns.
About the team:
Our mission is to build a modern ecosystem used for all IONOS customer support needs. The tools developed by us are used in over 20 locations, by more than 2.000 users, supporting 8 million customer contracts in 10 markets.
The development team has full responsibility for the development lifecycle. This means we plan, develop, test and deploy our software without any other internal or external dependencies.
Our portfolio revolves around an internally built CRM which is now being enhanced with AI capabilities.
About the product you will be building:
We are building a next-generation AI platform designed to redefine how our company interacts with customers. This isn't just a chatbot; it's a high-performance, multimodal AI ecosystem powered by state-of-the-art Speech-to-Speech (S2S) models, advanced Large Language Models (LLMs), and intelligent orchestration frameworks. Our platform will understand, reason, and respond across text and voice — while seamlessly executing real-time actions to resolve customer needs.
We are aiming for a hybrid architecture of Open Source LLMs, industry-leading proprietary models, and Model Context Protocol (MCP) to enable contextual reasoning, tool invocation, and seamless orchestration across systems. The goal is not just to talk to the customer, but to act on their needs.
What makes this project unique:
The Voice Frontier: We are building low-latency, emotive speech-to-speech pipelines for a truly natural voice channel experience.
Deep System Integration: Our platform connects directly to the company's core systems via MCPs, allowing the AI to access real-time customer context and execute complex workflows.
Self-Evolving Logic: We are developing an automated QA and evaluation module that continuously analyzes interactions across channels.By programmatically measuring quality, accuracy, latency, and resolution outcomes, we can close the feedback loop, and adapt system behavior in hours, not weeks.
Hybrid Innovation: You’ll work at the intersection of "build vs. buy," integrating the best of the open-source community with custom-built internal infrastructure.
What's in it for you:
You won't just be shipping code; you’ll be part of making this concept evolve and shift.
You’ll join a friendly, experienced team where your voice matters and your contribution shapes real-world outcomes. You’ll work in a modern environment with technologies and practices that help us ship reliable software efficiently.
Role description:
As an AI Engineer on this team, you will build the core intelligence systems behind our multimodal AI platform.You will be responsible for moving beyond simple chat interfaces to build high-performance, real-time systems that handle complex reasoning, deep context retrieval, LLM orchestration, retrieval-augmented generation (RAG) and seamless voice interactions.
Main responsibilities:
- Design Agentic Workflows: Design and implement LLM-based systems that go behind response generation - enabling structured tool usage, workflow orchestration, and secure interaction with internal services via MCP (Model Context Protocol).
- Build and Optimize RAG & CAG: Develop high-performance Retrieval-Augmented Generation and Context-Augmented Generation pipelines to ensure accurate, relevant, and low-latency responses. Continuously improve context management, ranking strategies, and grounding mechanisms to support complex, multi-step interactions.
- Voice Channel Mastery: Develop and optimize real-time Speech-to-Speech (S2S) pipelines, focusing on streaming architectures, latency reduction (including Time to First Word - TTFW) and maintaining a natural conversational flow.
- Evaluation, Quality & Alignment: Build and maintain an automated QA module, including LLM-as-a-judge patterns, to measure accuracy, safety, latency, and resolution quality at scale. Translate evaluation insights into systematic models and prompt improvements..
- Model Strategy & Hybrid Integration: Integrate and operate both commercial foundation models (e.g., OpenAI, Anthropic, Google) and open-source alternatives (e.g., Qwen, Kimi, DeepSeek, Moonshot, GLM), selecting and optimizing models based on performance, latency, cost, and use-case requirements.
We are looking for some of:
- Strong Python and/or Java Engineering Skills: Advanced-level Python development experience, including asynchronous programming (e.g., FastAPI, asyncio) and building high-performance, production-grade services. Experience with streaming architectures is a strong advantage.
- LLM Application & Multi-Agent Orchestration Experience: Hands-on experience building LLM-powered systems, including multi-step workflows, stateful agents, and tool invocation. Familiarity with orchestration frameworks such as LangChain, LlamaIndex, or LangGraph, particularly in building stateful, multi-turn agents.
- Advanced Retrieval & Context Management: Deep understanding of vector databases (e.g., Weaviate, Qdrant, pgvector, Elasticsearch), semantic search, embedding strategies, and re-ranking techniques. Experience designing and optimizing RAG pipelines.
- Real-Time & Low-Latency Systems: Experience in designing systems that operate under latency constraints, including streaming APIs, event-driven architectures, and performance optimization. Understanding of trade-offs between quality, cost, and response time.
- Evaluation-Driven Development: Experience in implementing evaluation frameworks for LLM-based systems, including automated QA pipelines and LLM-as-a-judge patterns.
- Familiar with API Design: knowledge of RESTful API design, OAuth2
We're looking forward to meeting you.
Hinweis zur Bewerbung
Wir wertschätzen Vielfalt und begrüßen alle Bewerbungen – unabhängig von z. B. Geschlecht, Nationalität, ethnischer und sozialer Herkunft, Religion, Behinderung, Alter sowie sexueller Orientierung und Identität, körperlichen Merkmalen, Familienstand oder einem anderen sachfremden Kriterium nach geltendem Recht.
Empfohlene Jobs
Betriebsleiter / Augenoptikmeister /Bachelor of Optometrie (B.Sc.) (m/w/d)
Betriebsleiter / Augenoptikmeister / Bachelor of Optometrie (B.Sc.) (m/w/d) bei brillen.de – Die Zukunft der Augenoptik Über uns: Die Supervista AG hat ihr Erfolgskonzept in den letzt…
Hüttenleiter (m/w/d) Weihnachtsmarkt Berlin-Spandau
Deine Aufgaben: Du verkaufst unsere Produkte in deiner Hütte mit großer Leidenschaft und bist ein Vorbild für dein Team Nach einer Schulung lernst du die weiteren Verkaufskräfte in deinem Team …
Pflegefachkraft Geriatrie (w/m/d)
Über uns Kompetent. Persönlich. Engagiert. Mit Seele und Sachverstand. Im grünen Lichterfelde verbinden wir moderne Medizin mit persönlicher Zuwendung. Als Krankenhaus mit christlichen Werten be…
TikTok Content Creator (m/w/d)
Allgemeine Infos Standort: Berlin Arbeitszeit: Vollzeit Start: ab sofort StoryMachine ist eine 2017 gegründete Agentur mit dem Schwerpunkt auf strategischem Storytelling für Brands, Persö…
Rechtsanwalts- oder Notarfachangestellte (w/m/d)
Rechtsanwalts- oder Notarfachangestellte (w/m/d) RSG Group GmbH - Head Office Berlin Voll/Teilzeit WER WIR SIND Mit mehr als 4,5 Millionen Mitgliedern in ihren Studios ist die RSG Group eine…
Medizinische Fachangestellte (m/w/d) Orthopädie Unfallchirurgie
PERMACON - Arbeitgeber mit Herz! "Wir bringen Bewegung in Ihre Karriere" Die PERMACON GmbH ist ein Personaldienstleistungsunternehmen, welches seit über 30 Jahren an sechs Standorten bundesweit ver…
Sozialassistent (m/w/d) Förderung der Kreativität - Kita
Sie sind auf der Suche nach einer neuen, beruflichen Herausforderung? Wir haben das perfekte Angebot für Sie. Derzeit suchen wir einen Sozialassistenten m/w/d für eine Berliner Kindertagesstätte in…
Dualer Bachelor of Arts „Online-Marketing und Marketingmanagement“ in Berlin
Shopwise ist eine spezialisierte Agentur mit KI-gestützten Prozessen für den Launch und die Skalierung der nächsten Generation an e-Commerce Unternehmen. Das Studium Der duale Bachelorstudien…
Staplerfahrer (m/w/d)
Über uns JENATEC Industriemontagen – PERSONAL.DIENST:LEISTUNG – Zeitarbeit mit Herz im gewerblichen und technischen Bereich. Richtig gute Arbeit gibt es bei unseren Kunden, zu denen viele namhafte …
Bäckereiverkäufer (m/w/d) - Bio - Backwaren *
Im [Auftrag] einer großen Berliner Bäckerei, welche überwiegend mit Bio-Backwaren arbeitet, suchen wir ab sofort eine Bäckereiverkäuferin oder einen Bäckereiverkäufer auf Vollzeit- oder Teilzeitbasis…