Senior Database Reliability Engineer (DBRE) Remote (U.S.)

Job Title: Senior Database Reliability Engineer (DBRE)

Employment Type: Full-Time

Work Location: 100% Remote (United States)

Work Authorization: Only U.S. Citizens, Green Card Holders (GC), and GC-EAD candidates will be considered. No sponsorship is available for this position.

About the Role

We are seeking a highly skilled Senior Database Reliability Engineer (DBRE) to join our engineering team. In this role, you will be responsible for designing, operating, and optimizing highly available, secure, and scalable database platforms. You will work closely with Site Reliability Engineering (SRE), Infrastructure, Security, and Application Engineering teams to improve the reliability, performance, and automation of mission-critical data systems.

Required Experience
  • 6 8 years of overall IT experience.

  • 5+ years of hands-on experience administering and troubleshooting PostgreSQL in production environments.

  • Experience with cloud infrastructure and managed database services, including:

    • PostgreSQL on Kubernetes

    • Amazon RDS

    • AWS

    • Terraform

    • Service discovery

    • Secrets management

  • 5+ years managing production databases or distributed data systems across application, database, operating system, storage, and networking layers.

  • 3+ years of Linux systems engineering, including:

    • Performance tuning

    • Memory management

    • I/O optimization

    • Security and networking

  • Experience with infrastructure automation tools such as Terraform, Ansible, Chef, or Puppet.

  • 2+ years of scripting or programming experience using Python, Bash, Go, Ruby, or Perl.

  • Experience supporting at least one additional distributed data platform such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.

Key Responsibilities
  • Design, administer, maintain, and secure PostgreSQL infrastructure to ensure high availability and performance.

  • Build and maintain ETL pipelines and develop database migration procedures and automation scripts.

  • Perform database installation, upgrades, patching, backup and recovery, monitoring, capacity planning, and architectural improvements.

  • Collaborate with SRE and engineering teams to improve platform reliability and operational excellence.

  • Participate in on-call rotations, incident response, root cause analysis, and post-incident reviews.

  • Automate database operations and reduce operational overhead through infrastructure-as-code and scripting.

  • Monitor database performance and proactively identify and resolve production issues.

  • Support scalable, secure, and resilient database platforms in AWS cloud environments.

Preferred Qualifications
  • Experience building self-service database platforms and automation tooling.

  • Hands-on experience operating Kafka/MSK, ClickHouse, or Redis clusters at scale.

  • Strong understanding of disaster recovery, backup validation, SLOs, failover testing, and reliability engineering practices.

  • Deep knowledge of PostgreSQL internals, including:

    • Replication

    • Concurrency

    • Transaction consistency

    • Query optimization

    • Indexing

    • Backup and recovery

  • Excellent communication skills with experience creating technical documentation and collaborating across engineering teams.

Required Technical Skills
  • PostgreSQL

  • AWS

  • Amazon RDS

  • Kubernetes

  • Terraform

  • Linux

  • Python or Bash

  • Database Performance Tuning

  • Backup & Recovery

  • Database Replication

  • ETL

  • Monitoring & Alerting

  • Infrastructure Automation

Email:
[email protected]

Similar jobs

Senior Database Reliability Engineer (DBRE)

Apply Now
Back to search page