Senior Database Reliability Engineer (DBRE) Remote (U.S.)
Job Title: Senior Database Reliability Engineer (DBRE)
Employment Type: Full-Time
Work Location: 100% Remote (United States)
Work Authorization: Only U.S. Citizens, Green Card Holders (GC), and GC-EAD candidates will be considered. No sponsorship is available for this position.
About the RoleWe are seeking a highly skilled Senior Database Reliability Engineer (DBRE) to join our engineering team. In this role, you will be responsible for designing, operating, and optimizing highly available, secure, and scalable database platforms. You will work closely with Site Reliability Engineering (SRE), Infrastructure, Security, and Application Engineering teams to improve the reliability, performance, and automation of mission-critical data systems.
Required Experience6 8 years of overall IT experience.
5+ years of hands-on experience administering and troubleshooting PostgreSQL in production environments.
Experience with cloud infrastructure and managed database services, including:
PostgreSQL on Kubernetes
Amazon RDS
AWS
Terraform
Service discovery
Secrets management
5+ years managing production databases or distributed data systems across application, database, operating system, storage, and networking layers.
3+ years of Linux systems engineering, including:
Performance tuning
Memory management
I/O optimization
Security and networking
Experience with infrastructure automation tools such as Terraform, Ansible, Chef, or Puppet.
2+ years of scripting or programming experience using Python, Bash, Go, Ruby, or Perl.
Experience supporting at least one additional distributed data platform such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.
Design, administer, maintain, and secure PostgreSQL infrastructure to ensure high availability and performance.
Build and maintain ETL pipelines and develop database migration procedures and automation scripts.
Perform database installation, upgrades, patching, backup and recovery, monitoring, capacity planning, and architectural improvements.
Collaborate with SRE and engineering teams to improve platform reliability and operational excellence.
Participate in on-call rotations, incident response, root cause analysis, and post-incident reviews.
Automate database operations and reduce operational overhead through infrastructure-as-code and scripting.
Monitor database performance and proactively identify and resolve production issues.
Support scalable, secure, and resilient database platforms in AWS cloud environments.
Experience building self-service database platforms and automation tooling.
Hands-on experience operating Kafka/MSK, ClickHouse, or Redis clusters at scale.
Strong understanding of disaster recovery, backup validation, SLOs, failover testing, and reliability engineering practices.
Deep knowledge of PostgreSQL internals, including:
Replication
Concurrency
Transaction consistency
Query optimization
Indexing
Backup and recovery
Excellent communication skills with experience creating technical documentation and collaborating across engineering teams.
PostgreSQL
AWS
Amazon RDS
Kubernetes
Terraform
Linux
Python or Bash
Database Performance Tuning
Backup & Recovery
Database Replication
ETL
Monitoring & Alerting
Infrastructure Automation
By continuing you agree to our Terms & Privacy Policy.