Remote job
Senior SRE or DevOps engineer in networking-heavy infrastructure (Remote)
Job details
About this role
Role overview This is a senior, hands-on position focused on running and evolving large-scale production infrastructure for a proxy and networking-heavy platform. The work spans reliability engineering and infrastructure automation, with a notable mandate to migrate a sizable estate from Docker Swarm to an on-premises Kubernetes setup. It is a remote role suited to an engineer who enjoys owning systems end-to-end and partnering closely with backend developers.
Responsibilities - Own and continuously improve a production environment spanning hundreds of servers and roughly fifty services. - Lead the end-to-end migration from Docker Swarm to Kubernetes running on bare-metal infrastructure. - Maintain high availability targets and reduce operational toil across the platform. - Drive observability practices in collaboration with the development team, including dashboards, alerts, and instrumentation. - Establish and enforce Infrastructure-as-Code standards and keep CI/CD pipelines reliable. - Participate in an on-call rotation, lead incident response, run blameless retrospectives, and drive systematic follow-ups. - Build platform tooling that improves developer experience and reduces infrastructure friction.
Requirements - Demonstrated experience operating highly available infrastructure at scale: hundreds of servers, dozens of services, and real production load. - Hands-on Kubernetes expertise in self-hosted or bare-metal environments. - Strong working knowledge of Infrastructure-as-Code tooling and patterns. - Proactive communication style, surfacing issues early and keeping the team informed without prompting. - Solid scripting and development skills for automation and tooling work. - Comfort participating in on-call duties and owning production reliability.
Nice to have - Prior experience planning, executing, and stabilising at least one major infrastructure migration. - Working familiarity with Python and/or Go (the backend is Python, edge services are Go). - Background with proxy or other networking-heavy infrastructure. - Experience in small teams where developers share infrastructure responsibility. - Familiarity with edge clusters or split compute/edge architectures.
Benefits and work setup - Gross salary starting from 6000 EUR, with a quarterly KPI-based performance bonus; salary is open for discussion based on skills and experience. - Remote work setup. - Learning budget with over forty internal learning options, plus support for external conferences, mentorship, and ongoing knowledge-sharing. - Private health insurance, gym allowance, and access to a wellness app. - Team events, an overseas workation, and regular milestone celebrations.