Job
- Level
- Experienced
- Location
- Berlin
- Working Model
- Hybrid, Onsite
- Job Field
- IT, DevOps
- Employment Type
- Full Time
- Contract Type
- Permanent employment
Job Summary
In this role, you will develop automation solutions to eliminate manual operations, optimize incident management processes, and enhance system monitoring for smooth production workflows.
Job Technologies
Your role in the team
- Eliminate Toil: Identify manual or repetitive operational friction across R&D and build clean scripts, automation, and internal tools to solve it permanently.
- Pioneer AI-Driven Operations: Design, build, and integrate AI agents to streamline operational workflows, ensuring proper guardrails, monitoring, and human oversight for safe execution.
- Level Up Incident Management: Own and optimize our Incident.io workflows, automation, and integrations. Stay closely engaged with incident response and participate in post-incident reviews to identify friction and turn learnings into improvements to tooling, coordination, and processes, without taking on incident responder responsibilities.
- Enhance Observability & System Health: Maintain and refine monitoring, alerting, and dashboards across our observability stack. You will dive into logs, metrics, and production data to investigate operational edge cases.
- Optimize workflows & runbooks: Collaborate directly with SRE and R&D teams to identify operational pain points, transforming complex procedures into clear, automated runbooks.
- Support Core Production Systems: Collaborate with SREs on database maintenance tasks, health checks, and release/deployment workflows where production reliability is impacted.
This text has been machine translated. Show original
Our expectations of you
Qualifications
- Comfortable working with Linux, command-line tools, logs, and monitoring.
- A structured approach to troubleshooting and solving operational problems.
- Proactive attitude towards improving systems, processes, and tooling.
- Ability to work collaboratively with engineers across different teams.
- Willingness to learn and build deeper expertise in production systems and reliability.
Experience
- 2-4 years of experience in TechOps, DevOps, SRE, Production Engineering, or a similar technical role.
- Experience working with production systems in a SaaS or cloud environment.
- Experience with scripting or automation and a mindset of 'if we do it twice, can we automate it?'
This text has been machine translated. Show original
What we offer
- €1,000 annual learning budget and free German language courses to boost your skills.
- 30 days of annual leave, plus extra paid days for your birthday and moving day.
- Home office setup budget, a monthly home office allowance.
- Freedom to work from abroad for up to 90 days worldwide!
- Mental health support with nilo.health and a discounted Urban Sports Club membership.
- 20% company subsidy on your pension contributions.
- Subsidised BVG public transport ticket and a dog-friendly Berlin office where your furry friend is welcome.
- Leasen Sie Ihr ideales Fahrrad über BusinessBike.
This text has been machine translated. Show original
Benefits
Work-Life-Integration
Topics You Will Work On
Job Locations
About Your Employer
Talon.One GmbH
Talon.One GmbH develops a flexible cloud-based solution for promotion and loyalty programs, enabling businesses to optimize their marketing strategies. The company serves numerous well-known brands worldwide.
Description
- Company Type
- Startup
- Working Model
- Hybrid, Onsite
- Industry
- Internet, IT, Telecommunication