Job
- Level
- Experienced
- Job Field
- IT, DevOps
- Employment Type
- Full Time
- Contract Type
- Permanent employment
- Salary
- 75.000 to 95.000€ gross/year
- Location
- Karlsruhe
- Working Model
- Hybrid, Onsite
Job Summary
In this role, you enhance the reliability of AI-driven healthcare systems by implementing robust infrastructure, workload isolation, and effective automation.
Job Technologies
Your role in the team
- Making systems robust: You ensure workload isolation, redundancy, meaningful autoscaling, and appropriate cluster capacity – enough reserves for peak loads without running unnecessary resources. Disruptions should not even lead to outages.
- Identify known vulnerabilities: You investigate known reliability risks, gather solid evidence, and propose concrete courses of action.
- Implement solutions: You develop improvements yourself - as production code, Infrastructure as Code, automation, testing, runbooks, or technical safeguards.
- Maintain and develop observability: You will take over the maintenance of our alerting infrastructure and improve signals, alert ownership, and routing – real issues reach the right people immediately, while noise reaches no one.
- Keep the platform up to date: You maintain Kubernetes, infrastructure charts, and components on a rhythm that controls lifecycle risks without making every upgrade a top priority.
- Methodically handle incidents: You work calmly and systematically during disruptions, distinguish facts from assumptions, and drive the derived improvements through to completion. The cross-system root cause analysis grows with your system knowledge – it is explicitly not an expectation from day one.
- Strengthen shared responsibility: You work closely with the product teams. They retain responsibility for their systems; you enhance joint reliability practices and drive agreed-upon measures through to completion.
This text has been machine translated. Show original
Our expectations of you
Qualifications
- Troubleshooting and Engineering: You approach unfamiliar distributed systems methodically, think in terms of Failure Modes and Blast Radius, and implement improvements hands-on in code, IaC, and automation.
- Proactivity: You don't just wait for tickets, but notice problems, follow up on them, and bring well-thought-out options.
- Collaboration: You convince through evidence rather than hierarchy, can professionally support a different team decision, and clearly communicate again if new insights change the situation.
- Security awareness: For patient-facing systems, you treat identity, correct routing, tenant separation, and fail-closed behavior as part of reliability — and you ask before crossing an insecure boundary.
- Reflected AI usage: You use AI tools pragmatically but keep verification, context, and responsibility with humans.
- Languages: Very good English. German is helpful but not required for this internal engineering role.
Experience
- Production experience: Typically three to six years of working with production systems, such as in DevOps, SRE, Backend, or Infrastructure Engineering. The quality of your experience is more important than the exact number of years.
This text has been machine translated. Show original
What we offer
- Salary: €75,000-€95,000 gross per year, depending on experience and scope of responsibilities.
- Hybrid: During the onboarding period, two to three days per week in Karlsruhe, then flexible depending on the location, preferably one to two days per week.
This text has been machine translated. Show original
Topics You Will Work On
Job Locations
About Your Employer
Mediform Gmbh
Mediform GmbH is a health-tech startup based in Karlsruhe, founded in 2022. The company has developed MediVoice, an AI-powered telephone assistant capable of autonomously handling routine inquiries in medical practices. Mediform emerged as a spin-off from Innoopract Informationssysteme GmbH.
Description
- Company Type
- Startup
- Industry
- Healthcare, Social Sector