Role Description
Reporting to the Director, Infrastructure Operations, the SRE I (Site Reliability Engineer I) with a strong DevOps focus is a key member of a team responsible for supporting the tooling, pipelines, frameworks, and other technologies that underpin the many platforms deployed within the companyβs infrastructure. This role blends software engineering principles with deep DevOps expertise to automate and streamline the entire software delivery lifecycle.
As SRE I, they will partner closely with Engineering and Information Security peers on developing infrastructure solutions that follow established best practices and design patterns. Responsibilities include:
-
Lead the design, development, and maintenance of secure, scalable, resilient, and cost-effective cloud infrastructure solutions on AWS.
-
Design, implement, manage, and optimize robust CI/CD pipelines using tools like GitHub Actions and AWS CodePipeline.
-
Provide expert DevOps-focused full-stack guidance and support to software engineering teams.
-
Champion and implement DevOps best practices across teams.
-
Participate in grooming and prioritizing development efforts.
-
Research, propose, and implement solutions to improve cloud-based resources.
-
Monitor ticket queues and provide timely updates.
-
Participate in a 24/7 on-call rotation to respond to production incidents.
-
Define and manage Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
-
Evaluate, implement, and support Open Source frameworks and projects.
-
Ensure systems and users follow security standards.
-
Perform other duties as required.
Qualifications
-
6+ years hands-on experience managing and automating UNIX/Linux system environments within a DevOps context.
-
5+ years of experience in a DevOps Engineer or SRE role with a strong DevOps focus.
-
4+ years experience designing and implementing infrastructure as code within the AWS ecosphere using Terraform.
-
Expert-level proficiency with Terraform and Terragrunt for managing AWS infrastructure as code.
-
Strong experience with AWS Cloud services and infrastructure design best practices.
-
Mastery of observability, monitoring, metrics, and alerting at scale.
-
Experience providing DevOps-centric support for applications developed in Java, Python, and Angular.
-
Expert level proficiency in at least one scripting language and one programming language.
-
Mastery of Containerization (Docker) and familiarity with the container ecosystem.
-
Hands-on experience with APM tools.
-
Proficient with Jira, Confluence, and git toolset.
-
Hands-on experience with Agile/Scrum & Waterfall process environments.
Requirements
-
Consistently exhibits personal accountability to outcomes.
-
Able to prioritize and manage multiple projects simultaneously.
-
Self-starter able to work independently with minimal supervision.
-
Driven to learn and stay abreast of the latest technologies and DevOps best practices.
-
Strong analytical and problem-solving skills.
-
Excellent communication and collaboration skills.
Benefits
-
Salary range: $205,000 - $212,500 USD annually.
-
Total compensation package may include annual performance bonus, ESPP, enhanced time off packages, and benefits.
-
Fully remote within the U.S. (Los Angeles or Las Vegas preferred).
-
Up to 10% travel may be required.