Role Description
At Replicant, we believe AI should work for people, starting with customer service. Our SRE team builds - not just supports - an AI-native platform, and they own exciting domains like:
-
Platform and harness engineering
-
Site reliability
-
Cloud infrastructure
-
CI/CD
-
DevEx
-
Observability
-
Incident management
-
COGS (e.g. cloud spend visibility)
We're looking for a Site Reliability Engineer who has opinions about how these domains should work and wants agency in shaping where they go. If you are energized by enabling teams to succeed through systems- and patterns-level work, come build Replicant’s platform with us!
What You’ll Do
-
Contribute to patterns, design, and implementation of our domains; help shape the future of platform engineering at Replicant.
-
Build and improve systems that help reduce toil and enable Replicant's production infrastructure to remain available and operable under large-scale, real-time conversational AI traffic.
-
Extend and iterate our agent harness: Unsupervised AI agents are currently used by about 10% of the dev team - help us grow that number.
-
Own and improve our CI/CD pipelines and surrounding developer tooling: build and test performance, deployment ergonomics, and paved paths for new services.
-
Participate in on-call rotation and incident management to ensure platform uptime and quality.
Qualifications
-
6+ years’ experience in software development enablement roles.
-
Solid experience owning CI/CD platforms end to end - including domains like caching, architecture, and developer self-service.
-
Effective use of AI tools such as Claude and Cursor for coding, troubleshooting, and reasoning.
-
Familiarity with Node/TypeScript, Python, and Terraform for automation, and developing in a Kubernetes/Helm ecosystem.
-
Practical experience with observability: logs/metrics/tracing, monitoring/alerting, incident management process, and tooling.
-
Experience working in fully remote teams.
Requirements
-
Bonus: Harness engineering experience - building platforms for autonomous agents.
-
Production-at-scale experience with GCP.
-
Telephony and SIP architectures, FreeSWITCH in particular.
Benefits
-
In-person connection that counts: company-wide offsites and smaller team gatherings designed to make remote work feel personal.
-
Tech & learning stipend: Conferences, books, courses — interested? We’ll fund them.
-
Remote by design: We’re distributed — no guilt about life events, we trust you to manage your calendar.
-
Health & wellness: Flexible vacations, paid sabbatical after 5 years, comprehensive benefits, plus a stipend to support your physical and mental well-being.
-
Compensation that matches your impact: competitive salaries in the company you’re helping to build.
-
Equity with upside: We believe in shared ownership—You’ll own a real piece of a fast-growing AI company.