Remote job
Mid/Senior Software Engineer - Cloud/SRE
Job details
About this role
Role overview Join a cloud infrastructure team building a next-generation data platform designed for the serverless and edge computing era. The role focuses on Site Reliability Engineering, keeping production databases and services running smoothly on Kubernetes while continuously improving deployment and observability. It is a full-time, remote position open to candidates worldwide with at least three years of relevant experience.
Responsibilities - Operate and maintain the production environment running on Kubernetes, ensuring stability and performance for database services. - Improve the deployment lifecycle for databases and supporting services, making releases faster and more reliable. - Implement automation and infrastructure-as-code scripts to eliminate manual work and reduce human error. - Strengthen monitoring and alerting systems so that issues are detected and resolved before they reach end users. - Participate in an on-call rotation, handling incident response and post-incident improvements across the team.
Requirements - Hands-on experience running Docker and Kubernetes in production systems. - Familiarity with at least one widely used automation tool such as Terraform or Ansible. - Working knowledge of a major cloud provider, ideally AWS or GCP. - Intermediate-level programming skills in Golang and/or Python. - Solid understanding of Linux and its networking layer, with a genuine interest in distributed systems. - A commitment to software quality and curiosity to explore unfamiliar technologies.
Nice to have - Contributions to open-source projects. - AWS or Kubernetes certifications. - Experience with monitoring and logging stacks inside Kubernetes environments. - Deep expertise in Linux internals beyond the basics.
Benefits and work setup - Competitive salary plus stock options. - Fully remote, with the ability to work from anywhere in the world. - Dedicated budget for attending conferences, events, and tech talks.