Remote job
Platform Engineer
Job details
About this role
Role overview A platform engineering position responsible for the delivery and reliability backbone that supports product engineering teams. The work spans pipeline automation, runtime observability, and the discipline of service-level objectives, with infrastructure expressed as code so environments stay reproducible. The role participates in a follow-the-sun on-call rotation, meaning incidents and operational ownership are shared across time zones rather than left to a single region.
Responsibilities - Design, build, and maintain continuous integration and delivery pipelines, including enclave-specific build and deploy workflows. - Develop observability tooling such as metrics, logs, and traces that surface the health of services and infrastructure. - Operate and evolve the SLO and error-budget framework, translating reliability targets into actionable signals for engineering teams. - Define infrastructure as code using AWS CDK and GitHub Actions so provisioning, configuration, and rollout stay repeatable. - Take part in a 24/7 follow-the-sun on-call rotation, responding to incidents, writing post-incident notes, and updating runbooks. - Partner with product and security engineers to keep build environments and deploy targets compliant and dependable.
Requirements - Hands-on experience building and operating CI/CD pipelines at scale. - Familiarity with infrastructure-as-code patterns, ideally with AWS CDK and GitHub Actions. - Working knowledge of observability stacks and the practice of defining SLOs with error budgets. - Comfort participating in on-call rotations and handling production incidents calmly and methodically. - Strong scripting or programming background for automation glue. - Ability to collaborate across teams and write documentation that other engineers actually use.