Role Description
We are seeking a
Principal DevOps Engineer
to serve as a technical leader within our DevOps organization. This role partners closely with other senior technical leaders to translate strategic platform direction into reliable, scalable, and developer-friendly infrastructure capabilities. The Principal will lead the technical execution of platform initiatives, elevate engineering standards, and drive measurable improvements in reliability, cost efficiency, and delivery velocity. This is a hands-on leadership role. You will design and implement platform capabilities across our serverless and Kubernetes environments, guide senior engineers, and influence engineering practices across teams.
Key Responsibilities and Outputs:
-
Technical Leadership & Execution
-
Partner with Architects to implement and operationalize platform strategy.
-
Lead the technical design and delivery of platform capabilities across:
-
Serverless-first AWS workloads
-
Kubernetes (EKS) environments
-
Break down large architectural initiatives into executable milestones.
-
Guide platform implementation patterns and reusable modules.
-
Reliability & Operational Improvement
-
Drive systemic improvements that reduce incident frequency and severity.
-
Strengthen observability standards (logging, metrics, tracing).
-
Support adoption of Product focused SRE frameworks across teams.
-
Improve deployment safety, change management, and incident response practices.
-
Platform Evolution
-
Improve CI/CD automation and deployment workflows.
-
Contribute to Kubernetes cluster design evolution and networking patterns.
-
Improve logging pipelines, cost visibility, tagging standards, and platform guardrails.
-
Support platform modernization efforts with a focus on simplification and predictability.
-
Engineering Enablement
-
Improve developer experience across platform offerings.
-
Establish clear golden paths for deployment and monitoring.
-
Contribute to DORA metric visibility and engineering performance improvements.
-
Support responsible AI tooling adoption in our workflows.
-
Influence & Mentorship
-
Mentor senior engineers and elevate platform engineering standards.
-
Lead technical discussions and design reviews.
-
Make tradeoffs visible and articulate technical risk clearly.
-
Reinforce disciplined intake and prioritization practices to protect strategic capacity.
Qualifications
-
8β12+ years in infrastructure, SRE, or platform engineering.
-
Experience operating serverless and Kubernetes environments at scale.
-
Proven track record improving reliability and reducing operational toil.
-
Experience collaborating closely with multiple engineering teams.
Requirements
-
Strong AWS expertise (Lambda, API Gateway, Step Functions, DynamoDB, S3, IAM, VPC).
-
Strong Kubernetes/EKS experience including networking and observability.
-
Advanced Infrastructure as Code (Terraform preferred).
-
Strong CI/CD pipeline design experience.
-
Experience with observability tooling (OpenTelemetry, OpenSearch, Prometheus/Grafana, CloudWatch).
-
Strong understanding of distributed systems and reliability engineering.
Benefits
-
Remote First Culture
-
Health Care Coverage
-
Education Reimbursement
-
Competitive Paid Time Off
-
Self-Care Days
-
National Holidays
-
2 Founder Days + Juneteenth Observed
-
Paid Volunteer Time Off
-
Charitable Contribution Match
-
Monthly Wellness or Home Office Reimbursement
-
Access to Employee Assistance Program (mental health platform)
-
Parental Leave
-
Retirement Plan with match/contribution