Job
- Level
- Senior
- Job Field
- IT, DevOps, Security
- Employment Type
- Full Time
- Contract Type
- Permanent employment
- Location
- Heilbronn
- Working Model
- Onsite
Job Summary
In this role, you work directly with customers to analyze complex platforms, debug issues collaboratively, and automate recurring tasks with scripts to enhance system stability.
Job Technologies
Your role in the team
- You are both a technical partner and an escalation level - directly alongside the customer on our highly complex platforms (Kubernetes, KubeVirt, Cilium, Ceph, Talos). Tickets organize the work; the actual activity takes place in joint debug sessions, calls, and war rooms - with the customer, not for them.
- You live Shared Fate: In PRRs, you analyze the design, operation, and monitoring of customer workloads, identify reliability gaps, and help define measurable SLOs. With guides, tutorials, and talks, you enable customers to build their workloads correctly from the start—rather than debugging reactive incidents.
- You analyze business-critical customer workloads in depth, debug complex issues together with the customer team, and resolve them permanently. Recurring problems are handled as software tasks and automated with tools and scripts. Insights are fed back to the SRE core team as preventive optimizations.
- You work in highly sensitive environments, with all the resulting security and compliance requirements.
This text has been machine translated. Show original
Our expectations of you
Education
- You have a degree in computer science or a comparable qualification and possess a strong passion for system-related technologies.
Qualifications
- You love diving deep into log files, traces, and metrics to isolate complex error patterns within the system network and to resolve them sustainably.
- You have knowledge in software development and use it to automate recurring tasks permanently and reliably.
- Since we operate in a highly regulated environment, you are willing to comply with our strict compliance requirements and participate in a flexible on-call duty.
- You value close, organization-wide communication and work highly collaboratively with the SRE platform team as well as our software and product developers to create a seamless bridge between platform infrastructure and application development.
Experience
- You have at least 2 years of hands-on experience managing cloud infrastructures with Kubernetes and possess a strong 'Customer-First' mentality — because you know that a stable platform is worthless if the customer cannot utilize it optimally.
This text has been machine translated. Show original
What we offer
- Schwarz Digits creates the technological foundation for digital decision-making freedom in Europe. As the IT and Digital division of the Schwarz Group, we develop and oversee IT infrastructures for the retail divisions Lidl and Kaufland, as well as Schwarz Produktion and PreZero. At the same time, we operate as an independent provider in the external market to support companies across Europe in their digital transformation. We consolidate our core services in the areas of Cloud, Cyber Security, Data & AI, Communication, and Workspace. Contribute to digital decision-making freedom in Europe.
- With us, you work at the intersection of agility and security: you benefit from quick decision-making processes, enjoy real scope for shaping your projects, and build on the solid foundation of the Schwarz Group.
This text has been machine translated. Show original
Benefits
Work-Life-Integration
Topics that you deal with on the job
Job Locations
This is your employer
Schwarz Unternehmenskommunikation GmbH & Co. KG
The Schwarz Group, based in Neckarsulm, is a significant German conglomerate and one of the largest retail groups in Europe. It operates over 13,900 stores under the brands Lidl and Kaufland and employs around 575,000 people.
Description
- Company Type
- Established Company
- Working Model
- Full Remote, Hybrid, Onsite
- Industry
- Trade