Fyerx logo

Site Reliability Engineer (SRE) - Database Focus

Fyerx
Posted 4 hours ago
Probably WorldwideRemoteEngineering & Development
Is this job info correct?

Job Description

This is a remote position.

Site Reliability Engineer (SRE) - Database Focus

Job Details
  • Employment Type: Contract
  • Work Mode: Remote
  • Location: Offshore
  • Total Experience Required: 5 to 9 years
  • Relevant Experience Required: 4+ years of dedicated database reliability engineering or database administration (DBA) experience in high-throughput cloud environments
  • Mandatory Certification: AWS Certified Database - Specialty, Certified Postgres Professional, or Oracle Database Administration Certified Professional

Job Summary
We are seeking an experienced Database Site Reliability Engineer (DB SRE) to oversee the stability, performance, and scaling limits of our high-volume transaction databases. The ideal candidate will build automated cluster management tools, orchestrate sub-millisecond query optimizations, manage multi-region replication layers, and design fault-tolerant configurations to ensure maximum database uptime and zero data loss.

Key Responsibilities
  • Design and maintain highly available database cluster topologies across public clouds, leveraging automated scaling, sharding mechanisms, and read-replica strategies.
  • Optimize sub-millisecond database query performance, conducting rigorous index evaluations, identifying locking contentions, and rewrites of inefficient SQL scripts.
  • Orchestrate automated backup and disaster recovery validation sweeps, configuring point-in-time recovery (PITR) parameters and multi-region failover tests.
  • Build automated telemetry dashboards and proactive alerts utilizing monitoring frameworks (e.g., Prometheus, Grafana, Datadog) to track database health metrics (IOPS, CPU utilization, connections).
  • Write robust automation scripts using Python, Go, or Bash to manage recurring database maintenance routines, schema migration deployments, and resource adjustments.
  • Diagnose and remediate production database performance bottlenecks, resolving replication lags, connection pool limits, memory usage leaks, and deadlocks.
  • Enforce data safety and governance protocols, configuring encryption-at-rest policies, fine-grained access parameters, and compliance masking routines to secure sensitive records.


Requirements

  • 5 to 9 years of core systems engineering or database administration experience, with 4+ dedicated years actively managing high-scale, cloud-native relational databases (e.g., PostgreSQL, MySQL, Aurora) or NoSQL databases.
  • Strong technical mastery of database engine configurations, connection poolers (e.g., PgBouncer), infrastructure automation (Terraform), and system telemetry structures.
  • Deep structural understanding of write-ahead logging (WAL), isolation levels, transaction mechanics, distributed storage bounds, and network latency impacts.
  • Mandatory certification: AWS Database Specialty, Certified Postgres Professional, or Oracle Database Admin Certified Professional.

Preferred Qualifications
  • Prior experience implementing large-scale live database data migrations with minimal production operational windows.
  • Familiarity with distributed database engines or streaming message queues (e.g., CockroachDB, Kafka, Debezium) for real-time change data capture (CDC).


Similar jobs