Remote job
AWS Engineer
Job details
About this role
Role overview This position leads the AWS side of a consulting practice that serves enterprise clients committed to staying on Amazon Web Services. The work spans serverless and containerized architectures, event-driven systems, and infrastructure as code, with a strong emphasis on environments that must pass compliance audits. It suits an engineer who treats infrastructure as a product: defensible, automated, and observable, not just functional.
Responsibilities - Architect serverless and containerized workloads on Lambda, Fargate, and ECS, selecting the right runtime for each use case and recognizing when each is the wrong fit - Design event-driven systems on SNS, SQS, and EventBridge, including dead-letter queues and replay paths that hold up under failure - Model data on DynamoDB and Aurora, with explicit access pattern design and awareness of how modeling choices affect cost at scale - Codify infrastructure in Terraform or AWS CDK using reusable modules that can be shared across client accounts - Build CI/CD pipelines with staged rollouts, automated verification, and a rollback path that has actually been tested in production - Operate multi-account AWS Organizations, defining IAM boundaries, VPC topology, and Transit Gateway networking - Own observability through CloudWatch, X-Ray, and existing client tooling, building dashboards that answer operational questions rather than just collect metrics - Drive cost optimization through right-sizing, savings plans, and reclaiming unused resources - Lead incident response on production systems and ensure follow-through with durable fixes
Requirements - Four or more years architecting and operating production AWS environments - Strong infrastructure-as-code practice with Terraform or AWS CDK, with no reliance on console-driven administration - Deep working knowledge of IAM, VPC networking, and the AWS security model, including common pitfalls - Scripting fluency in Python, TypeScript, or Go, with the ability to read application code when debugging an infrastructure symptom
Nice to have - DynamoDB optimization at scale, including patterns for large-volume deletion and migration - Hands-on experience with GCP alongside AWS - Exposure to compliance-constrained environments such as SOC 2, PCI, or HIPAA