Site Reliability / Platform Engineer @Genesis10
Devops
Salary $90.00 - $105.0..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type contract
Posted 6d ago

[Hiring] Site Reliability / Platform Engineer @Genesis10

6d ago - Genesis10 is hiring a remote Site Reliability / Platform Engineer. πŸ’Έ Salary: $90.00 - $105.00 per hour πŸ“Location: USA

Role Description

Genesis10 is currently seeking a Site Reliability / Platform Engineer - Remote position with a Leading Asset Management Firm. This role is open to remote U.S.-based resources, however, candidates who are able to work hybrid in either New York City or Austin, TX are preferred. This is a 12+ month contract opportunity.

We are seeking a highly hands-on SRE / Platform Engineer for a hybrid software development, Site Reliability Engineering, and systems engineering role. This position is ideal for a strong developer who has expanded into infrastructure, cloud, automation, and production engineering and wants to take a more holistic view of how applications and systems operate together.

The team follows an engineering-focused SRE model centered on using software and automation to solve infrastructure and operational problems. Engineers on the team write code every day and work across application and infrastructure layers to improve reliability, performance, scalability, observability, and system integration.

A major initiative for the team is establishing a centralized observability capability across an environment where monitoring and operational data have historically been siloed. The organization is bringing telemetry together using Datadog and enterprise data lake capabilities, creating a common observability foundation that can ultimately support AIOps, agentic AI, automated remediation, and self-healing systems.

This is not a traditional operations or Solutions Architecture position. The successful candidate will be expected to build, deploy, automate, troubleshoot, and improve the environment using technologies such as Kubernetes, Terraform, Ansible, and Datadog.

Qualifications

  • Strong hands-on experience in software development, with some additional experience in SRE, DevOps, Platform Engineering, or Infrastructure Engineering
  • Strong Terraform experience, including the ability to write, deploy, maintain, and troubleshoot Infrastructure-as-Code
  • Strong Ansible experience, including developing playbooks/roles and automating software and infrastructure deployments
  • Hands-on Kubernetes experience, including deployment, configuration, troubleshooting, and production support
  • Strong understanding of observability, including metrics, logs, traces, monitoring, alerting, dashboards, and production troubleshooting
  • Experience with Datadog or a comparable enterprise observability platform; direct Datadog experience is strongly preferred
  • Experience automating the deployment and configuration of infrastructure and/or observability tooling
  • Strong coding or scripting capabilities; Python and/or Java experience is highly desirable
  • Experience working with cloud infrastructure and an understanding of how applications, networks, infrastructure, and data platforms interact
  • Strong troubleshooting and root-cause analysis skills across both application and infrastructure layers
  • Experience supporting highly available, production-scale systems.
  • Demonstrated ability to solve infrastructure and operational problems through code and automation rather than manual processes
  • Ability to work independently, learn unfamiliar technologies quickly, and take ownership of problems from investigation through implementation

Requirements

  • Develop software and automation to improve the reliability, scalability, performance, and operational efficiency of production systems
  • Write and maintain infrastructure and configuration code using Terraform and Ansible as part of day-to-day engineering activities.
  • Deploy, manage, automate, and troubleshoot applications and services running on Kubernetes
  • Automate the deployment and configuration of Datadog and other observability capabilities, including Ansible-based deployments
  • Help establish centralized observability across previously siloed applications, infrastructure, cloud, and data environments
  • Bring together metrics, logs, traces, events, and other telemetry to provide a comprehensive view of system health, dependencies, and potential business impact
  • Build monitoring, dashboards, alerting, and observability capabilities that enable teams to identify and resolve production issues more effectively
  • Help create the observability and operational-data foundation required to support future AIOps and agentic AI capabilities
  • Support longer-term initiatives around intelligent incident detection, root-cause analysis, impact analysis, automated remediation, and self-healing
  • Partner with application developers to improve application performance, resiliency, deployment patterns, and integration with other systems
  • Evaluate how applications interact with infrastructure, networks, APIs, databases, and data platforms and identify opportunities for improvement
  • Troubleshoot complex production issues spanning applications, Kubernetes, cloud infrastructure, networking, and data environments
  • Provide advanced production engineering support across traditional L2/L3 boundaries, focusing on root cause and permanent solutions rather than temporary fixes
  • Support large-scale data environments that may include Databricks, Snowflake, and enterprise data lake technologies
  • Identify repetitive or manual operational processes and replace them with scalable, code-based solutions

Benefits

  • Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years
  • The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years
  • Access to an experienced, caring recruiting team (more than 7 years of experience, on average)
  • Behavioral Health Platform
  • Medical, Dental, Vision
  • Health Savings Account
  • Voluntary Hospital Indemnity (Critical Illness & Accident)
  • Voluntary Term Life Insurance
  • 401K
  • Sick Pay (for applicable states/municipalities)
  • Commuter Benefits (Dallas, NYC, SF)
Before You Apply
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Site Reliability / Platform Engineer @Genesis10
Devops
Salary $90.00 - $105.0..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type contract
Posted 6d ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 127,100+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later