Job
- Level
- Experienced
- Job Field
- Software, Data
- Employment Type
- Full Time
- Contract Type
- Permanent employment
- Location
- Karlsruhe
- Working Model
- Hybrid, Onsite
Job Summary
In this role, you will develop complex AI systems, optimize real-time interactions, implement LLM-based workflows, and ensure accurate outputs through modern evaluation techniques.
Job Technologies
Your role in the team
- As an AI Engineer on this team, you will build the core intelligence systems behind our multimodal AI platform.
- You will be responsible for moving beyond simple chat interfaces to build high-performance, real-time systems that handle complex reasoning, deep context retrieval, LLM orchestration, retrieval-augmented generation (RAG) and seamless voice interactions.
- Design agentic workflows: Design and implement LLM-based systems that go beyond response generation—enabling structured tool usage, workflow orchestration, and secure interaction with internal services via MCP (Model Context Protocol).
- Build and Optimize RAG & CAG: Develop high-performance Retrieval-Augmented Generation and Context-Augmented Generation pipelines to ensure accurate, relevant, and low-latency responses.
- Continuously improve context management, ranking strategies, and grounding mechanisms to support complex, multi-step interactions.
- Voice Channel Mastery: Develop and optimize real-time Speech-to-Speech (S2S) pipelines, focusing on streaming architectures, latency reduction (including Time to First Word - TTFW) and maintaining a natural conversational flow.
- Evaluation, Quality & Alignment: Build and maintain an automated QA module, including LLM-as-a-judge patterns, to measure accuracy, safety, latency, and resolution quality at scale.
- Translate evaluation insights into systematic models and prompt improvements.
- Model Strategy & Hybrid Integration: Integrate and operate both commercial foundation models (e.g., OpenAI, Anthropic, Google) and open-source alternatives (e.g., Qwen, Kimi, DeepSeek, Moonshot, GLM), selecting and optimizing models based on performance, latency, cost, and use-case requirements.
This text has been machine translated. Show original
Our expectations of you
Qualifications
- Vertrautheit mit Orchestrierungs-Frameworks wie LangChain, LlamaIndex oder LangGraph, insbesondere beim Aufbau von zustandsbehafteten, Multi-Turn-Agenten.
- Advanced Retrieval & Context Management: Deep understanding of vector databases (e.g., Weaviate, Qdrant, pgvector, Elasticsearch), semantic search, embedding strategies, and re-ranking techniques.
- Understanding of trade-offs between quality, cost, and response time.
- Familiar with API Design: knowledge of RESTful API design, OAuth2.
Experience
- Strong Python and/or Java Engineering Skills: Advanced-level Python development experience, including asynchronous programming (e.g., FastAPI, asyncio) and building high-performance, production-grade services.
- Experience with streaming architectures is a strong advantage.
- LLM Application & Multi-Agent Orchestration Experience: Practical experience in developing LLM-powered systems, including multi-step workflows, stateful agents, and tool invocation.
- Experience designing and optimizing RAG pipelines.
- Real-Time & Low-Latency Systems: Experience in designing systems that operate under latency constraints, including streaming APIs, event-driven architectures, and performance optimization.
- Evaluation-Driven Development: Experience in implementing evaluation frameworks for LLM-based systems, including automated QA pipelines and LLM-as-a-judge patterns.
This text has been machine translated. Show original
What we offer
- Access to local/international trainings, development and growth opportunities, including access to e-learning platforms, covering both technical and soft skills areas.
- Modern technologies, product responsibility.
- Flexible work schedule.
- Hybrid work option.
- Medical services package from one of two private providers.
- 25 vacation days per year.
- Substitute days off for public holidays that occur on the weekend.
- Meal tickets.
- Internal referral program.
- Team events, networking events organized to promote a passionate, creative and diverse culture.
- Summerfest and Winterfest parties.
This text has been machine translated. Show original
Benefits
Work-Life-Integration
Health, Fitness & Fun
Topics You Will Work On
Job Locations
About Your Employer
IONOS
The 1&1 IONOS product portfolio offers everything businesses need to be successful in the cloud: from domains to classic websites and do-it-yourself solutions, online marketing tools to fully-fledged servers and an IaaS solution.
Description
- Founding Year
- 1988
- Company Type
- Established Company
- Working Model
- Full Remote, Hybrid, Onsite
- Industry
- Internet, IT, Telecommunication