Remote job
DevOps Engineer
Job details
About this role
Role overview
Own the operational foundation of a cloud-based infrastructure platform used by enterprise customers with consumption-based business models. This remote, full-time role combines reliability engineering, cloud operations, release management, security, automation, and incident troubleshooting. You will help ensure dependable performance while improving scalability, resilience, and infrastructure efficiency.
Responsibilities
- Maintain service reliability, availability, and application performance standards. - Plan infrastructure capacity and identify opportunities to optimize operating costs. - Design and maintain disaster-recovery and high-availability capabilities. - Manage CI/CD pipelines, deployment workflows, and release processes. - Build automation and internal tooling that reduce operational effort and improve consistency. - Monitor infrastructure, strengthen security and compliance controls, and provide second-level troubleshooting support.
Requirements
- At least five years of experience in DevOps, site reliability engineering, or a closely related discipline. - Bachelor’s degree in engineering or a related field. - Experience operating cloud infrastructure at scale, ideally in a SaaS environment. - Practical proficiency with observability and monitoring platforms such as CloudWatch, Datadog, or Splunk. - Familiarity with AWS scaling patterns and Kubernetes. - Strong debugging, troubleshooting, database performance-tuning, CI/CD automation, and Python scripting skills.
Nice to have
- Experience working with event-streaming systems.
Benefits and work setup
- Remote, full-time position.