Platform Engineer (Reliability) - Unannounced Project at Scopely

ES - Spain

on-site · full-time · Visa sponsorship

Apply for this role at Scopely

Responsibilities:
- Define and drive reliability, operational excellence, and SRE practices to maximise uptime and reduce operational toil.
- Automate manual operational work: build tooling, runbooks, and software-driven solutions to improve platform efficiency.
- Lead performance testing, tuning, and capacity planning to meet SLAs and scale multiplayer/MMO systems.
- Implement observability: metrics, traces, and logging to enable proactive detection and fast incident resolution.
- Participate in incident response, lead postmortems, and drive long-term fixes and preventive measures.
- Embed security, compliance, and governance into platform delivery pipelines and operational practices.
- Collaborate cross-functionally to shape the engineering platform roadmap balancing delivery speed, cost, and reliability.

Requirements:
- Strong software engineering background with hands-on SRE or platform engineering experience running production services.
- Proficiency with container platforms (Kubernetes), IaC (Terraform), and cloud environments (AWS/EKS or equivalents).
- Experience in Go or Python and debugging complex distributed systems in production.
- Deep observability and incident management experience (alerts, runbooks, postmortems).
- Understanding of capacity planning, cost optimisation, and cloud governance best practices.
- Excellent communication and collaboration skills across engineering and product teams; prior experience with large-scale games or MMO infrastructure is a plus.