Senior Site Reliability Engineer @Kintsugi AI
Devops
Salary usd 180,000 - 1..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted YDay

[Hiring] Senior Site Reliability Engineer @Kintsugi AI

YDay - Kintsugi AI is hiring a remote Senior Site Reliability Engineer. πŸ’Έ Salary: usd 180,000 - 190,000 per year πŸ“Location: USA

Role Description

Kintsugi is revolutionizing sales tax automation with our AI-powered platform designed specifically for e-commerce and SaaS businesses. Our solution reduces tax preparation time by 75 percent and cuts compliance costs by 50 percent, allowing finance teams to focus on strategic initiatives rather than routine calculations. As we continue to grow and disrupt the tax automation space, we're building a world-class team to help us scale with purpose.

We're looking for a Senior Site Reliability Engineer (DevOps) to help scale and harden the infrastructure that powers Kintsugi. This role sits at the intersection of software engineering and operations:

  • Keep production reliable under real load.
  • Build tooling and automation that reduce manual work.

You'll work closely with Platform Engineering, Product, and QA to:

  • Design resilient architectures.
  • Improve deployment pipelines.
  • Build internal tools and guardrails for quick movement without sacrificing stability.

You'll operate across our managed Kubernetes and AWS-hosted data layer (Postgres, Redis, networking) and shape the reliability and developer-experience foundation of our engineering org.

We're an agentic-coding-first team β€” coding agents already do real engineering and operations work here, not just autocomplete on the side. We expect this role to build and extend that practice, not just adopt it.

What You'll Do

  • Own the reliability of production infrastructure running on managed Kubernetes and AWS, keeping a high-traffic system up and catching issues before customers do.
  • Work agent-first day to day: build, debug, and automate using agentic coding workflows as your default mode.
  • Build internal tools and automation that eliminate recurring manual work (toil) for the team.
  • Develop and operate monitoring, alerting, and observability systems (metrics, tracing, logging) across the stack.
  • Partner with engineering teams to design for reliability and performance from the start.
  • Automate infrastructure management through infrastructure-as-code, and improve CI/CD pipelines and local developer workflows.
  • Lead and evolve incident response practices, including postmortems and blameless learning.
  • Optimize infrastructure for cost efficiency while maintaining high availability and security standards.
  • Contribute to security, compliance, and disaster recovery efforts as the platform scales.
  • Support developer enablement: build and improve the in-house tooling, local dev workflows, and internal platforms other engineers rely on.

Qualifications

  • 5-8 years in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles.
  • Real ownership of a production system at meaningful scale.
  • Fluent working agent-first day to day.
  • Strong foundation in AWS-hosted data and networking services (RDS/Postgres, ElastiCache/Redis, VPC/networking).
  • Experience running workloads on managed Kubernetes.
  • A track record of building tools that removed manual work for a team.
  • Hands-on experience with CI/CD pipelines and infrastructure-as-code (e.g., Terraform, CloudFormation).
  • Expertise in observability stacks (metrics, tracing, logging) and modern monitoring practices.
  • Familiarity with security and compliance in cloud environments (SOC 2, GDPR, etc. a plus).
  • A collaborative mindset with a passion for empowering developers to move fast safely.
  • Experience with (or strong interest in) developer enablement.
  • Nice to have: experience operating across multiple cloud providers.
Before You Apply
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Senior Site Reliability Engineer @Kintsugi AI
Devops
Salary usd 180,000 - 1..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted YDay
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 128,600+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later