Remote job
Software Engineer
Job details
About this role
Role overview
Join a cloud and production engineering team responsible for the internal platforms that help software teams release changes safely and quickly. The work centers on progressive delivery, feature flags, experimentation, configuration management, and change analysis, with a strong emphasis on reliability, observability, and operational quality. This is an individual-contributor role focused on delivering well-defined backend and platform capabilities with moderate guidance.
Responsibilities
- Design, document, review, and audit technical solutions for platform features and components. - Build and operate scalable backend services, reusable components, infrastructure modules, and deployment pipelines across cloud environments. - Create comprehensive unit, component, and end-to-end testing plans, while debugging issues across services, environments, and external dependencies. - Implement metrics, logs, traces, dashboards, and service-level objectives so systems remain observable and operationally safe. - Participate in on-call rotations, investigate incidents, perform root-cause analysis, improve runbooks, and use operational data to strengthen reliability. - Work within Agile processes, collaborate across teams, identify delivery risks, and contribute to continuous improvement.
Requirements
- Bachelor’s or master’s degree in computer science, software engineering, or equivalent practical experience. - At least three years of professional experience building backend, platform, or reliability-focused services. - Experience with a modern programming language such as C#, Java, or Go. - Working knowledge of distributed systems, data structures, production reliability, web services, REST or gRPC APIs, and structured data. - Experience with Git, complex CI/CD workflows, automated testing, and Agile engineering practices. - Experience using coding assistants for problem solving, debugging, profiling, or technical documentation.
Nice to have
- Experience operating services in Azure or GCP, including Terraform and infrastructure-as-code in production. - Background in progressive delivery, configuration management, feature-flag systems, or experimentation platforms. - Familiarity with OpenTelemetry, Prometheus, Grafana, incident response, root-cause analysis, and operational readiness reviews. - Experience owning features through design, rollout planning, launch support, and post-release improvement.
Benefits and work setup
- Hybrid arrangement requiring access to an office and typically at least two in-office days per week, with the exact cadence determined by the team.