Role Description
We are looking for an experienced Senior Site Reliability Engineer to join our highly collaborative and distributed Engineering team. This is a senior individual contributor role with broad technical ownership and significant influence across Engineering. Your mission will be to help design and evolve a platform that is reliable, scalable, secure and efficient, while raising the bar for SRE and cloud-native practices across the organisation.
We are looking for someone who wants to remain deeply hands-on: someone who can design architecture, challenge technical decisions, automate at scale, investigate complex production issues and help other engineers make better technical decisions. You don't need to become a manager to have significant impact here.
-
Architect, build and evolve highly available, scalable and secure infrastructure.
-
Take ownership of major Platform and Reliability initiatives, from technical assessment and architecture through to production.
-
Provide technical leadership and mentorship across the Engineering organisation.
-
Manage and continuously improve our core infrastructure around AWS, Kubernetes, Docker and Terraform.
-
Work with Helm, FluxCD and Kustomize to improve deployment and infrastructure management.
-
Partner closely with development teams to build and improve highly automated CI/CD pipelines, particularly with GitHub Actions.
-
Strengthen observability and proactively identify reliability, performance and capacity issues.
-
Troubleshoot complex production problems and participate in the on-call rotation, incident response and post-mortem analysis.
-
Design and improve our infrastructure security practices and tooling.
-
Work across a rich data and database ecosystem including PostgreSQL, StarRocks, OpenSearch, Redis and MongoDB.
-
Interact with cloud and distributed systems including Confluent, Snowflake and Temporal.
-
Drive automation, architectural reviews and continuous improvements across the platform.
-
Build useful technical documentation, including runbooks, playbooks and architecture documentation.
Qualifications
-
Roughly 8 to 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps or Infrastructure Engineering.
-
Deep experience running Kubernetes and Docker in production at scale.
-
Strong expertise with AWS and complex cloud architectures.
-
Advanced Infrastructure as Code experience, particularly with Terraform.
-
Strong understanding of reliability, scalability, observability and production engineering.
-
Experience designing and operating modern CI/CD environments, ideally with GitHub Actions.
-
Solid experience with production databases, particularly PostgreSQL and/or MongoDB.
-
A strong understanding of modern infrastructure security practices.
-
The ability to investigate complex problems across application, infrastructure and data layers.
-
Experience owning significant infrastructure projects or migrations from design through production, with measurable outcomes.
-
Strong communication skills and the ability to influence technical decisions across Engineering.
-
Experience mentoring engineers and leading technical initiatives without relying on formal management authority.
-
Professional working English and experience working in distributed or international environments.
Requirements
-
Comfortable reading, debugging and writing code when needed.
-
Experience with languages such as Go, Python, Clojure or JavaScript is particularly relevant.
-
Ability to understand unfamiliar systems, go deep when necessary and learn quickly.
Nice to have
-
Experience with Helm, FluxCD, Kustomize, StarRocks, OpenSearch, Redis, Confluent, Snowflake, Temporal, CrowdStrike, Tenable One.
Benefits
-
Competitive compensation and stock options.
-
Health insurance with Alan.
-
Swile meal vouchers (10 EUR per working day, 50% covered by Crossbeam).
-
Remote work support: 50 EUR monthly allowance, 200 EUR home office budget, and up to 50 EUR per week for a coworking space if you live outside Ile-de-France.
-
Annual wellness stipend of 1,000 EUR.
-
Collaborative teammates and a culture built on trust and accountability.