Summary
✨ AI‑Generated
A growing software organization is seeking a hands-on TechOps Engineer to improve the reliability, scalability, and operability of production systems. You will solve production issues, automate repetitive tasks, enhance monitoring and releases, and build practical tooling that helps engineering teams operate more safely.
Highlights
Hands-on production engineering role focused on reliability, scalability, automation, monitoring, and safer operations. The position offers close collaboration with engineering teams and substantial opportunity to improve tooling and production practices.
Description
ABOUT TALON.ONE:
Talon.One is the most powerful incentives engine that unifies loyalty, promotions and gamification into one holistic platform.
Backed by enterprise-grade security and scalability, Talon.One empowers companies to build personalized, profitable promotions and loyalty programs using any data.
Today, over 250 of the world’s most-loved brands including Adidas, Sephora and Carlsberg work with Talon.One to drive deeper engagement and lasting loyalty with their customers.
ABOUT THE TEAM & ROLE
Our SRE / Production Engineering team is responsible for keeping Talon.One reliable, scalable, and easy to operate.
We work closely with engineering teams across R&D to improve how we monitor, release, troubleshoot, and run our production systems.
This is a hands-on role for someone who loves solving production-level problems, automating repetitive work, and building pragmatic tooling to make life safer and easier for the engineers around them.
ONCE YOU ARE HERE, YOU WILL:
Eliminate Toil: Identify manual or repetitive operational friction across R&D and build clean scripts, automation, and internal tools to solve it permanently.Pioneer AI-Driven Operations: Design, build, and integrate AI agents to streamline operational workflows, ensuring proper guardrails, monitoring, and human oversight for safe execution.Level Up Incident Management: Own and optimize our Incident.io workflows, automation, and integrations.
Stay closely engaged with incident response and participate in post-incident reviews to identify friction and turn learnings into improvements to tooling, coordination, and processes, without taking on incident responder responsibilities.Enhance Observability & System Health: Maintain and refine monitoring, alerting, and dashboards across our observability stack.
You will dive into logs, metrics, and production data to investigate operational edge cases.Optimize Workflows & Runbooks: Partner directly with SRE and R&D teams to identify operational pain points, turning messy procedures into clear, automated runbooks.Support Core Production Systems: Collaborate with SREs on database maintenance tasks, health checks, and release/deployment workflows where production reliability is impacted.
WHAT WE NEED YOU TO BRING TO THE TABLE:
2–4 years of experience in TechOps, DevOps, SRE, Production Engineering, or a similar technical roleExperience working with production systems in a SaaS or cloud environmentComfortable working with Linux, command-line tools, logs, and monitoringExperience with scripting or automation and a mindset of "if we do it twice, can we automate it?"A structured approach to troubleshooting and solving operational problemsProactive attitude towards improving systems, processes, and toolingAbility to work collaboratively with engineers across different teamsWillingness to learn and build deeper expertise in production systems and reliabilityNICE TO HAVE
Familiarity with core SRE concepts, such as Service Level Indicators/Objectives (SLIs/SLOs) and error budgets.Experience with observability platforms such as Grafana, Datadog, Prometheus, or SentryExperience with Kubernetes and GCPExperience working with APIs, integrations, CI/CD, or infrastructure automatio
OUR TECH STACK
Cloud & Infrastructure: Google Cloud Platform (GCP), Kubernetes, containerized workloadsInfrastructure as Code: Terraform / HelmObservability & Incident Ops: Grafana, Datadog, Sentry, Incident.ioDatabases: PostgreSQLAutomation & CI/CD: Go, Python, Bash, GitHub Actions / CI/CD tooling
WHAT'S IN IT FOR YOU:120+ team of engineers, product managers and product designers in BerlinLeaders with 8+ years of experience building our promotions engine€1,000 annual learning budget and free German language courses to boost your skills30 days of annual leave, plus extra paid days for your birthday and moving dayHome office setup budget, a monthly home office allowanceFreedom to work from abroad for up to 90 days worldwide!Mental health support with nilo.health and a discounted Urban Sports Club membership20% company subsidy on your pension contributionsSubsidised BVG public transport ticket and a dog-friendly Berlin office where your furry friend is welcomeLease your ideal bike through BusinessBike