Das ist der Job
Streamline incident response: Partner with the Event C
Darum lohnt es sich
This role sits at the intersection of software engineering and operations — focused on building the tools, automation, and visibility that enable our teams to deliver resilient software at scale.
You’ll work closely with application engineers, DevOps, and QA teams to evolve our infrastructure, CI/CD pipelines, observability frameworks, and reliability practices.
The Outcomes You’ll Deliver Assess and improve visibility: Work with engineering teams to review our current dashboards, metrics, and logs, identify the biggest gaps, and make targeted improvements that help us better understand system health.
Clarify what “reliable” means: Help define initial SLIs and SLOs for a few core user flows, aligning the team on what good performance and availability look like. Playon is looking for an experienced Senior Site Reliability Engineer to help us strengthen the reliability, performance, and scalability of our systems.
This is a hands‑on engineering role with a strong emphasis on automation, performance analysis, and continuous improvement. Tighten monitoring and alerting: Refine alerts and dashboards for the most critical services so we can catch issues earlier and respond faster.
Build observability into delivery: Add instrumentation and telemetry into existing build and deploy processes to make reliability checks part of our normal release workflow.