Senior Database Reliability Engineer @Trumid
All Others
Salary usd 225,000 - 2..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 4wks ago

[Hiring] Senior Database Reliability Engineer @Trumid

4wks ago - Trumid is hiring a remote Senior Database Reliability Engineer. 💸 Salary: usd 225,000 - 265,000 per year 📍Location: USA

Role Description

This role owns data resilience and continuity, and the scope test is simple: if losing it loses data, or makes data unavailable, it’s yours. The job exists so that data-layer failure modes — storage contention, replica lag, region loss — are found and retired in drills, not discovered in production. Our reliability doctrine is to assume failure and concentrate statefulness into a small core of systems proven against specific failure modes. Postgres is the heart of that core: everything around it gets to be disruptible because the data layer is not.

You’d join the databases side of our SRE team, working alongside deep Postgres expertise. Some of what you’d walk into:

  • A production RDS fleet backing a live trading venue, with performance work that goes deep: we’ve characterized WAL-write contention under concurrent commits down to the fsync level, and are weighing group-commit tuning, dedicated log volumes, and storage-class changes against actual measurements.
  • Disaster recovery as an engineering discipline: automated cross-region failover with promotion measured in minutes, and restore paths (snapshot, point-in-time, logical) validated by timed, documented drills on a fixed cadence.
  • Database observability and AI-assisted tooling: engine performance telemetry exported into Prometheus and Grafana, modular database health-check skills, and an automated reviewer for schema-migration PRs.

What you’ll do?

  • Own the durability, recoverability, and performance of the Postgres/RDS fleet across every environment: replication, failover, backup and restore, and storage behavior under load.
  • Make recovery provable: restore-tested coverage of every production database, timed failover drills, and measured RPO/RTO per tier — evidence, not assertion.
  • Own the data lifecycle end to end: retention and cleanup policies that preserve recoverability, access control at the data layer, encryption posture, and knowing where sensitive data lives.
  • Hunt performance pathologies at the engine level: lock contention, WAL throughput, replica lag, bloat, index hygiene, write amplification.
  • Build database observability with the team, and extend our AI-assisted operations tooling (health-check skills, migration PR review).

Qualifications

  • Five or more years running production PostgreSQL at meaningful scale — ideally on RDS or Aurora — with depth in the internals: replication, WAL mechanics, MVCC and vacuum behavior, query planning and performance.
  • You’ve owned the full lifecycle of data somewhere, not just the query path: retention, backup and restore, access control, and security posture.
  • Disaster-recovery experience you can talk through concretely — failovers you designed, drills you ran, and what they changed.
  • Infrastructure fluency: infrastructure-as-code (Terraform or similar), scripting (Python, bash, SQL), Linux, and cloud storage/IOPS characteristics.
  • An SRE sensibility: SLOs, blameless postmortems, and a preference for rehearsed over improvised.
  • Clear writing — you leave runbooks and decision records behind you.

Nice to have

  • Warehouse and pipeline experience (BigQuery, AlloyDB, Kafka-based pipelines) — over time we intend to extend the same reliability guarantees beyond Postgres to every store that holds business data.
  • Experience in regulated or fintech environments; exposure to data classification and compliance review.
  • Kubernetes; Prometheus/Grafana/ELK-style observability stacks.
  • Interest in AI-assisted operations tooling.

Benefits

  • Highly competitive compensation
  • Fully paid medical, dental and vision coverage
  • Team-oriented and collaborative company culture
  • Lively and dynamic office space with fully stocked kitchen
  • In compliance with New York City Pay Transparency Law, the base salary range for this role in New York City is between $225,000 – $265,000. This range does not include discretionary bonus or other forms of compensation or benefits offered in connection with this job. Several factors are considered when determining a candidate’s compensation.
Before You Apply
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Senior Database Reliability Engineer @Trumid
All Others
Salary usd 225,000 - 2..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 4wks ago
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
×
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 ★★★★★ from 500+ reviews

⚡ 127,028+ remote jobs, refreshed hourly

🔔 Real-time alerts: Apply first, direct to employer

🛡️ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later