Staff Site Reliability Engineer @Filevine
Software Development
Salary $235,000 - $275..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 2mths ago

[Hiring] Staff Site Reliability Engineer @Filevine

2mths ago - Filevine is hiring a remote Staff Site Reliability Engineer. πŸ’Έ Salary: $235,000 - $275,000 πŸ“Location: USA

Role Description

As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership. You don’t just maintain systems β€” you shape engineering culture, define the technical standard for how Filevine runs in production, and bridge the gap between high-level business goals and robust, internet-scale technical execution. You bring a forward-looking perspective β€” actively shaping how AI and machine learning drive the future of reliability practice.

  • You own the roadmap across two critical SRE domains β€” Observability & Alerting and Platform Infrastructure.
  • You are accountable for ensuring the team solves reliability problems permanently rather than absorbing them as toil.
  • You operate as the senior IC counterpart to the Engineering Manager: technical correctness lives with you.
  • You partner with the Reliability Architect and engineering leadership on significant technical decisions.
  • You mentor engineers across experience levels and influence reliability strategy across the broader organization.
  • You are the senior technical voice responsible for ensuring that uptime, incident response, and every production change meet the operational standard the business demands.
  • This role does not participate in on-call rotation, but you are deeply invested in the engineers who do β€” shaping the on-call strategy, tooling, and culture that make production support sustainable and effective.

Qualifications

  • 12+ years of experience in software engineering, infrastructure, platform engineering, or SRE, including 6+ years in SRE and 3+ years leading complex, cross-functional technical initiatives for distributed production systems.
  • Expert-level depth in observability and platform infrastructure, with broad expertise in incident response, capacity planning, automation, and reliability engineering.
  • Advanced experience with a major container-orchestration platform, preferably Kubernetes, and an observability platform such as New Relic, Datadog, or equivalent.
  • Strong software-engineering ability in Python, Go, Bash, or another general-purpose language, with experience building production tooling, automation, or platform capabilities.
  • Proven ability to mentor engineers and communicate technical risk clearly to engineering, product, and executive audiences.
  • Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is strongly preferred.

Requirements

  • Define and execute the technical strategy for Observability & Alerting, Platform Infrastructure, and operational excellence.
  • Lead the evolution of reliable, scalable, secure, and efficient cloud platforms and distributed systems.
  • Champion SLIs, SLOs, error budgets, capacity planning, operational readiness, and automation across the service lifecycle.
  • Lead the organization through complex production incidents and turn post-incident learning into permanent engineering improvements.
  • Build self-service platform capabilities that reduce toil, improve engineering safety and velocity, and make every team more capable of owning their own reliability.
  • Mentor engineers and serve as a trusted technical authority for long-term reliability and platform direction.

Benefits

  • A dynamic, rapidly growing company, focused on helping organizations thrive.
  • Medical, Dental, & Vision Insurance (for full-time employees).
  • Competitive & Fair Pay.
  • Maternity & paternity leave (for full-time employees).
  • Short & long-term disability.
  • Opportunity to learn from a dedicated leadership team.
  • Top-of-the-line company swag.
Before You Apply
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Staff Site Reliability Engineer @Filevine
Software Development
Salary $235,000 - $275..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 2mths ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 126,845+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later