Platform Engineer – Backend & Reliability at Trade Republic
on-site · full-time · Visa sponsorship
Apply for this role at Trade Republic
Responsibilities:
- Design and implement high-throughput, low-latency, fault-tolerant backend services and platform components in Kotlin.
- Treat reliability as a first-class concern: define SLOs, model failure modes, and apply resilience patterns (circuit breakers, bulkheads, graceful degradation, retries) in code.
- Lead load, stress, and chaos testing campaigns and convert findings into architectural and operational improvements.
- Set and enforce backend engineering standards for service design, error handling, resiliency, and safe deployments.
- Drive capacity planning and performance engineering work to keep systems performant ahead of demand.
- Own incident response, run post-mortems, and ensure corrective actions close systemic gaps.
- Make observability first-class with structured logging, metrics, and distributed tracing to support debugging and reliability.
Requirements:
- 5+ years of backend engineering experience building and operating distributed systems at scale.
- Strong proficiency in Kotlin (primary) and experience with Go is a plus.
- Demonstrated expertise in reliability engineering: SLOs, chaos engineering, load testing, and capacity planning.
- Experience with event-driven architectures, container orchestration, and production observability tools and practices.
- Comfortable owning production services, leading incidents, and improving testing, deployment, and resilience practices.
- Excellent collaboration and communication skills to raise engineering standards across teams.