Scalable GmbH
(Senior) Platform Engineer - Real-Time Infrastructure - Banking (m/f/x)
- Location
- Berlin, BE, Germany
- Work model
- Hybrid
- Seniority
- Senior
- Employment
- FullTime
- Posted
- Added to Codestelle
Language requirements
- German
- Not specified
- English alone
- Not specified
Based on explicit wording in the listing. “Not specified” does not mean a language is optional.
About this role
Join us to build and operate the infrastructure foundation behind mission-critical, real-time trading systems with demanding reliability and tail-latency requirements.
This is primarily a platform and infrastructure role. You will work on the runtime and operational environment behind performance-sensitive systems across cloud and physical infrastructure. The scope includes AWS, Linux hosts, networking, containerized runtimes, observability, deployment safety, and, where needed, dedicated hardware.
In this role you will:
- Own infrastructure end to end across cloud and physical environments, including AWS, Linux hosts, networking, containers, and runtime operations
- Design and improve infrastructure for performance-sensitive services, supported by repeatable production-like benchmarks covering throughput, burst behaviour, and p50, p95, p99, and p99.9 latency
- Design, operate, and troubleshoot critical network paths across VPCs, routing, security groups, load balancers, DNS, private connectivity, and service-to-service communication
- Build and improve infrastructure as code, host lifecycle management, and deployment workflows across development, staging, and production
- Improve reliability, failover, recovery, and operational safety across Linux-based runtimes, containers, serverless workloads, event-driven flows, and scheduled processing
- Harden platform capabilities around IAM, KMS, secrets, service connectivity, and secure operational access
- Raise the standard of observability through dashboards, alerting, tracing, health checks, readiness checks, and incident tooling
- Partner with software engineers on reusable infrastructure patterns, production readiness, CI/CD, and incident response
- A university degree in Computer Science, Engineering, or a comparable practical background
- Experience with dedicated host strategies, low-latency networking, CPU pinning, host-level isolation techniques
- Strong experience owning production infrastructure for distributed or real-time systems
- Strong hands-on experience with AWS, especially EC2, VPC networking, IAM, KMS, S3, SQS, SNS, and related operational tooling
- Strong hands-on networking experience, including routing, DNS, load balancers, network security, service connectivity, and debugging complex production traffic paths
- Strong experience with Terraform and infrastructure lifecycle management
- Deep knowledge of Linux systems, containers, runtime debugging, and host-level performance tuning
- Experience operating performance-sensitive workloads where predictability and tail-latency matter, not only average throughput
- Experience with infrastructure patterns for stateful services, active/standby failover, recovery, deployment safety, and controlled cutovers
- Strong analytical thinking, pragmatic decision-making, and clear communication across engineering and operational stakeholders
- Nice to have:
- Experience supporting Rust-based service environments
- Experience in regulated, security-sensitive, or high-reliability environments