Role Description
We're looking for a
Senior DevOps Engineer
to design, automate, operate, and continuously improve the cloud infrastructure and deployment platform supporting a high-scale cybersecurity product.
-
Work primarily with
AWS, Kubernetes/EKS, Terraform, Docker, CI/CD,
and observability technologies.
-
Contribute to platform engineering, helping engineers ship software faster and more reliably.
-
Participate directly in architectural decisions, incident resolution, platform improvements, and DevOps best practices.
What You'll Be Doing
-
Design, build, and operate highly available and scalable cloud infrastructure primarily on
AWS
.
-
Build and maintain Infrastructure as Code using
Terraform
.
-
Deploy, manage, scale, and troubleshoot workloads running on
Kubernetes
, particularly
Amazon EKS
.
-
Develop automation for infrastructure provisioning, environment configuration, and application deployment.
-
Build and maintain reliable CI/CD pipelines using tools such as
GitHub Actions, Jenkins,
or
GitLab CI
.
-
Improve deployment processes, release reliability, and developer experience across engineering teams.
-
Manage containerized applications using
Docker
and
Kubernetes
.
-
Design and improve observability across infrastructure and applications.
-
Implement monitoring, dashboards, alerts, and centralized logging using technologies such as:
-
Prometheus
-
Grafana
-
AWS CloudWatch
-
OpenSearch / Elasticsearch
-
ELK
-
Participate in production incident response, troubleshooting, mitigation, and root cause analysis.
-
Identify recurring operational issues and automate their resolution.
-
Implement AWS and Kubernetes security best practices.
-
Collaborate with Security teams to strengthen infrastructure security and support compliance initiatives.
-
Improve platform resilience, high availability, scalability, and disaster recovery capabilities.
-
Optimize AWS resource utilization and contribute to cloud cost optimization and FinOps initiatives.
-
Automate operating system and infrastructure configuration using
Ansible
or similar tools.
-
Develop operational tooling and automation using
Python
and
Bash
.
-
Collaborate with Engineering teams on platform architecture and deployment strategies.
-
Continuously evaluate new DevOps, cloud, Kubernetes, observability, and platform engineering technologies.
Qualifications
-
5β8+ years of experience working in DevOps, Platform Engineering, Site Reliability Engineering, Cloud Engineering, or similar roles.
-
Strong hands-on production experience with
AWS
.
-
Strong experience managing
Kubernetes clusters
, ideally
Amazon EKS
.
-
Advanced experience with
Terraform
and Infrastructure as Code.
-
Strong experience with
Docker
and containerized environments.
-
Hands-on experience building and maintaining CI/CD pipelines with:
-
GitHub Actions
-
Jenkins
-
GitLab CI
-
or similar platforms
-
Strong
Linux administration
skills.
-
Experience with configuration management tools such as
Ansible
.
-
Good scripting and automation capabilities with:
-
Experience implementing observability using technologies such as:
-
Prometheus
-
Grafana
-
CloudWatch
-
OpenSearch / Elasticsearch
-
ELK
-
Strong understanding of networking concepts including:
-
VPCs
-
Subnets
-
Load Balancers
-
DNS
-
VPNs
-
Firewalls
-
Routing
-
Security Groups
-
Experience troubleshooting complex distributed systems and production environments.
-
Solid understanding of cloud security and infrastructure security best practices.
-
Experience supporting highly available, production-grade cloud environments.
-
Strong communication skills and ability to collaborate with engineering, security, QA, and business stakeholders.
-
High degree of autonomy, ownership, and analytical problem-solving ability.
-
Bachelor's degree in Computer Science, Software Engineering, Systems Engineering, or a related technical discipline, or equivalent professional experience.
Core Technology Stack
-
AWS
-
Kubernetes / EKS
-
Docker
-
Terraform
-
Ansible
-
GitHub Actions
-
Jenkins
-
Linux
-
Python
-
Bash
-
Prometheus
-
Grafana
-
CloudWatch
-
OpenSearch / Elasticsearch
Nice to Have
-
Experience with
HashiCorp Nomad
.
-
Experience with
HashiCorp Consul
.
-
Broader experience with the HashiCorp ecosystem.
-
Experience operating
Kafka
in production.
-
Experience supporting data infrastructure based on:
-
ClickHouse
-
Elasticsearch
-
MongoDB
-
Redis
-
Experience with distributed systems and high-volume data platforms.
-
Previous experience in cybersecurity, identity security, or security SaaS platforms.
-
Experience working with cloud security controls and compliance frameworks.
-
Experience with
FinOps
and cloud cost optimization.
-
Experience designing internal developer platforms or platform engineering capabilities.
-
Exposure to infrastructure supporting AI/ML or LLM-based applications.
What Success Looks Like
-
Increasing infrastructure reliability and availability.
-
Reducing deployment friction and improving engineering velocity.
-
Increasing infrastructure automation and reducing manual operational work.
-
Improving observability and reducing mean time to detection and resolution.
-
Operating Kubernetes workloads reliably at scale.
-
Strengthening security throughout cloud infrastructure and deployment processes.
-
Improving AWS cost efficiency without compromising reliability.
-
Contributing reusable platform capabilities that allow engineering teams to move faster.
-
Continuously improving production readiness as the platform scales.
Benefits
-
Contractor agreement with payment in USD.
-
100% remote work.
-
Argentina's public holidays.
-
English classes.
-
Referral program.
-
Access to learning platforms.