Remote job
Manager of Database Reliability Engineering
Job details
About this role
Role overview A manager-of-managers role leading a database reliability engineering function for a high-traffic daily fantasy sports platform that sees seasonal and event-driven traffic peaks. The role owns cloud database infrastructure, data lifecycle strategy, and operational readiness for major sporting moments throughout the year. The team is open to remote candidates across the United States, with a preference for Atlanta-based applicants.
Responsibilities - Own cloud database infrastructure, including AlloyDB, Redis, and Kafka, along with capacity planning, vendor management, and spend oversight including committed-use discount strategy. - Support identity and access modernization, migrating to a modern identity platform with workflow-aware controls and break-glass mechanisms. - Build data lifecycle services that enable aging, archiving, and retention across a multi-year horizon. - Lead database observability work, producing clear, actionable signals rather than noise, and own the data side of business continuity and disaster recovery. - Run high-availability operations for peak live moments, including on-call design, war-room readiness, staffing, and vendor escalation coordination. - Partner with infrastructure engineering on shared surfaces such as caching, messaging, and queues, and represent the team in cross-functional planning for launches, load testing, and migrations. - Develop technical leads, delegate real ownership, and build depth so critical work never depends on a single person.
Requirements - 5+ years managing database, DevOps, or SRE teams for consumer products with real-time, high-concurrency, or seasonal traffic patterns. - Deep hands-on background in cloud infrastructure (GCP strongly preferred; AWS or similar also relevant), AlloyDB, PostgreSQL, Redis, MongoDB, and infrastructure-as-code tools such as Crossplane, Cuelang, Terraform, or Ansible. - Experience running identity and access programs, modern observability stacks, and incident management practices that hold up under real pressure. - Track record managing cloud cost, committed-use or reserved capacity strategy, vendor contracts, and budget ownership. - Solid grounding in security best practices and compliance requirements relevant to a fast-moving consumer platform. - Demonstrated ability to mentor technical leads and delegate effectively, with strong communication and willingness to flag risk early.
Benefits and work setup - Typical salary range of $200,000 to $220,000 depending on role, level, and work location. - Company-subsidized medical, dental, and vision plans, 401(k) with company match, and an annual bonus. - Flexible PTO with two weeks strongly encouraged, 16-week paid parental leave, and disability benefits. - Company equipment provided (Windows or Mac options), company-wide events, and an annual performance review cadence focused on growth.