Site Reliability Engineer - Platform Engineering at N26
Berlin
on-site · full-time · Visa sponsorship
Responsibilities:
- Build, operate and improve the Database Platform to enhance reliability, security, and scalability.
- Implement cloud infrastructure, networking and CI/CD best practices for data storage services.
- Automate database provisioning and lifecycle operations in collaboration with product teams.
- Maintain and evolve observability (metrics, logging, tracing) to meet SLOs and availability targets.
- Participate in on-call rotations, incident response and post-incident reviews to reduce toil and outages.
- Contribute to documentation, runbooks and the team roadmap by delivering high-quality code and infrastructure changes.
Requirements:
- Hands-on experience operating cloud infrastructure, preferably AWS, and services like RDS/S3.
- Practical knowledge of operational data storage (PostgreSQL or similar) and database HA patterns.
- Experience with containers and orchestration (Docker, Kubernetes/EKS).
- Proficiency with Infrastructure-as-Code (Terraform, CloudFormation or equivalent).
- Competence in a scripting/programming language (Python preferred) and CI/CD tools (GitHub Actions, ArgoCD, Jenkins, etc.).
- Familiarity with networking, cloud security best practices and observability tools (DataDog, Prometheus, Grafana, OpenTelemetry).
- Nice to have: database internals, performance tuning, replication/failover, sharding and backend engineering experience.