← All jobs
TechOps Engineer
Berlin · On-site · Engineering
Apply well, not just fast
Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.
About the role
ObservabilitySREGCPCI/CDKubernetesDatadogGrafanaPythonGoBashPostgreSQLTerraformGitHub ActionsLinuxPrometheusIncident Response
ABOUT TALON.ONE:
Talon.One is the most powerful incentives engine that unifies loyalty, promotions and gamification into one holistic platform. Backed by enterprise-grade security and scalability, Talon.One empowers companies to build personalized, profitable promotions and loyalty programs using any data.
Today, over 250 of the world’s most-loved brands including Adidas, Sephora and Carlsberg work with Talon.One to drive deeper engagement and lasting loyalty with their customers.
ABOUT THE TEAM & ROLE
Our SRE / Production Engineering team is responsible for keeping Talon.One reliable, scalable, and easy to operate. We work closely with engineering teams across R&D to improve how we monitor, release, troubleshoot, and run our production systems.
This is a hands-on role for someone who loves solving production-level problems, automating repetitive work, and building pragmatic tooling to make life safer and easier for the engineers around them.
ONCE YOU ARE HERE, YOU WILL:
- Eliminate Toil: Identify manual or repetitive operational friction across R&D and build clean scripts, automation, and internal tools to solve it permanently.
- Pioneer AI-Driven Operations: Design, build, and integrate AI agents to streamline operational workflows, ensuring proper guardrails, monitoring, and human oversight for safe execution.
- Level Up Incident Management: Own and optimize our Incident.io workflows, automation, and integrations. Stay closely engaged with incident response and participate in post-incident reviews to identify friction and turn learnings into improvements to tooling, coordination, and processes, without taking on incident responder responsibilities.
- Enhance Observability & System Health: Maintain and refine monitoring, alerting, and dashboards across our observability stack. You will dive into logs, metrics, and production data to investigate operational edge cases.
- Optimize Workflows & Runbooks: Partner directly with SRE and R&D teams to identify operational pain points, turning messy procedures into clear, automated runbooks.
- Support Core Production Systems: Collaborate with SREs on database maintenance tasks, health checks, and release/deployment workflows where production reliability is impacted.
WHAT WE NEED YOU TO BRING TO THE TABLE:
- 2–4 years of experience in TechOps, DevOps, SRE, Production Engineering, or a similar technical role
- Experience working with production systems in a SaaS or cloud environment
- Comfortable working with Linux, command-line tools, logs, and monitoring
- Experience with scripting or automation and a mindset of "if we do it twice, can we automate it?"
- A structured approach to troubleshooting and solving operational problems
- Proactive attitude towards improving systems, processes, and tooling
- Ability to work collaboratively with engineers across different teams
- Willingness to learn and build deeper expertise in production systems and reliability
NICE TO HAVE
- Familiarity with core SRE concepts, such as Service Level Indicators/Objectives (SLIs/SLOs) and error budgets.
- Experience with observability platforms such as Grafana, Datadog, Prometheus, or Sentry
- Experience with Kubernetes and GCP
- Experience working with APIs, integrations, CI/CD, or infrastructure automatio
OUR TECH STACK
- Cloud & Infrastructure: Google Cloud Platform (GCP), Kubernetes, containerized workloads
- Infrastructure as Code: Terraform / Helm
- Observability & Incident Ops: Grafana, Datadog, Sentry, Incident.io
- Databases: PostgreSQL
- Automation & CI/CD: Go, Python, Bash, GitHub Actions / CI/CD tooling
WHAT'S IN IT FOR YOU:
- 120+ team of engineers, product managers and product designers in Berlin
- Leaders with 8+ years of experience building our promotions engine
- €1,000 annual learning budget and free German language courses to boost your skills
- 30 days of annual leave, plus extra paid days for your birthday and moving day
- Home office setup budget, a monthly home office allowance
- Freedom to work from abroad for up to 90 days worldwide!
- Mental health support with nilo.health and a discounted Urban Sports Club membership
- 20% company subsidy on your pension contributions
- Subsidised BVG public transport ticket and a dog-friendly Berlin office where your furry friend is welcome
- Lease your ideal bike through BusinessBike