Remote job
Senior DevOps Engineer – Infra Services
Job details
About this role
Role overview A senior-level infrastructure engineering role focused on building and operating large-scale DNS and Kubernetes platforms inside a private cloud. The position blends deep networking knowledge with Infrastructure-as-Code and GitOps automation, supporting high-throughput, multi-tenant, mission-critical services that span public and private regions.
Responsibilities - Design, deploy, and operate large-scale DNS architectures spanning private and public cloud across multiple regions. - Implement and support enterprise DNS tooling including Infoblox DDI (DNS, IPAM, DHCP), Cloudflare DNS, and managed DNS services on GCP and AWS. - Run multi-tenant Kubernetes clusters on bare metal servers backed by software-defined storage protocols. - Build Infrastructure-as-Code using Ansible, Terraform or Pulumi, plus Python and Bash, to automate provisioning of DNS services and supporting components. - Stand up CI/CD pipelines for infrastructure changes, including patching, service upgrades, testing, and rollback procedures. - Develop automated monitoring, alerting, and self-healing using GitOps workflows and observability stacks such as Prometheus, Loki, and Grafana. - Harden services for high availability, disaster recovery, and scale-out, and perform deep troubleshooting across hypervisors, storage, and networking layers. - Take part in on-call rotation, incident response, and global change management.
Requirements - Five or more years managing global DNS and Kubernetes environments on Linux (RHEL, CentOS, or Ubuntu). - Hands-on expertise with Infoblox DDI or comparable enterprise DNS systems, BIND9, Cloudflare, and GCP or AWS DNS services. - Expert-level Kubernetes operations experience, including cluster lifecycle on Ubuntu Linux. - Proven track record automating infrastructure with Ansible, Terraform, Python, and CI/CD pipelines. - Strong Linux systems engineering background with advanced scripting in Python or Bash. - Familiarity with hypervisor technologies (KVM, vSphere), software-defined storage protocols (iSCSI, NFS, Ceph), and L2/L3 networking. - Experience operating 24x7 mission-critical production environments. - Ability to produce technical documentation and contribute to internal knowledge bases.
Nice to have - Background in telco, edge cloud, or large enterprise infrastructure environments. - Familiarity with security compliance frameworks such as CIS or NIST. - Experience building automated test environments for system upgrades and DNS migrations. - Bachelor's degree in computer science, IT, engineering, or related field, or equivalent experience with relevant certifications.
Benefits and work setup - Fully remote for candidates outside a 30-mile office radius; hybrid (3 days in office) for those within range. - Compensation includes performance bonus eligibility, equity, health, dental, and vision coverage from day one, mental health support resources, 401k matching, ESPP, paid holidays, volunteer time, and 12 weeks of paid parental leave.