Role Description
You will join the DevOps team responsible for the platforms our engineering teams rely on daily: CI/CD, observability, logging, and developer tooling. Our philosophy is simple: the team owns its systems end-to-end, and every engineer should be able to diagnose and fix issues in their area of responsibility. This is a senior-only role.
We are looking for an engineer we can hand an entire area β observability or CI/CD β and trust to run it: from requirements and technical design through rollout, operations, and mentoring others. You will be the go-to technical reference for your area and the senior escalation point for complex incidents in it.
What Youβll Do
-
Own one of our core platform areas end-to-end: observability (VictoriaMetrics, Grafana, Graylog / VictoriaLogs, fluent bit, exporters, alerting) or CI/CD (Jenkins scripted pipelines, Harbor, Nexus, build agents) β you drive its architecture, reliability, and roadmap.
-
Drive technical initiatives end-to-end: gather requirements, write the design doc, decompose into tasks, implement, deliver to production, and own the operational health afterwards.
-
Drive clarity in ambiguous situations by defining requirements, assumptions, and next steps.
-
Design for reliability and scale: evolve the architecture of our platforms β topology, integration points, scaling approach, and reliability model.
-
Support developers: deploy and monitor applications on both on-premise servers and Kubernetes (Helm), troubleshoot builds and deploys, help teams with metrics, alerts, and logs; participate in chat duty in developer support channels.
-
Automate away toil: repetitive operations, provisioning, and maintenance should be codified, not performed by hand.
-
Investigate production incidents as the senior escalation point for your area: drive resolution, lead post-mortems, implement systemic fixes. Participate in on-call rotations and raise the bar for how on-call works.
-
Mentor less experienced engineers through design discussions, reviews, and pairing; catch debt-inducing shortcuts at the review stage.
-
Use AI in all aspects of day-to-day work: researching, troubleshooting, developing.
Qualifications
-
6+ years as a DevOps Engineer / SRE (or very close responsibilities).
-
Track record of owning technical initiatives end-to-end β from requirements and technical design through production delivery.
-
Confident Linux skills (we use Ubuntu).
-
Working knowledge of the Prometheus stack: metric types, exporters, and how alerting works.
-
Hands-on experience with CI/CD: pipeline design, build orchestration, artifact delivery.
-
Experience with Containers: Docker, image building, registries.
-
Experience with Ansible.
-
Experience with Git.
-
Experience with Bash or Python scripting for automation and observability.
-
Production/on-call experience: diagnosing incidents, restoring service, leading post-mortems.
-
Experience mentoring less experienced engineers.
-
Ownership and attention to detail.
Requirements
-
Solid hands-on experience in two or more of the following areas:
-
VictoriaMetrics / Prometheus stack at scale: architecture, cardinality control, exporters, alerting infrastructure.
-
Log pipelines at scale: Graylog / VictoriaLogs / ELK β collection (fluent bit or similar), retention, sharding, performance.
-
Jenkins scripted pipelines: shared libraries, pipeline infrastructure, build agent fleets.
-
Container registries and artifact management: Harbor, Nexus, base images, image policies.
-
Operating applications on Kubernetes: Helm, workload monitoring and log delivery, deploy troubleshooting.
-
Grafana: dashboards as code, alerting, performance at scale.
Benefits
-
Private medical insurance for the employee and their family.
-
23 paid vacation days per year.
-
11 paid public holidays per year.
-
5 company-paid sick leave days.
-
English learning courses.
-
Relevant professional education.
-
Gym or swimming pool.
-
Home Office Setup Assistance: the company offers assistance with purchasing furniture (office chair, office desk, monitor) and other items to create a comfortable workspace.
-
Co-working.
-
Remote working.