Job details
About this role
Role overview A senior or staff-level platform engineering role building the runtime infrastructure that lets AI agents execute safely and at scale. The position is deeply technical and high-ownership, focused on distributed systems, isolation, and performance rather than typical product or feature engineering.
Responsibilities - Design and build agent runtime infrastructure using microVMs, Rust, and Go - Define and enforce security boundaries for running untrusted code in production - Architect scheduling and orchestration systems inspired by but not limited to existing orchestrators - Build a globally distributed systems architecture spanning multiple regions and bare-metal providers - Own performance, isolation, and reliability characteristics at scale - Set technical direction across the engineering organization and make high-stakes architectural decisions with incomplete information
Requirements - Experience building and operating distributed systems or infrastructure at scale, ideally from 0-to-1 or early-stage scale - Strong systems background in concurrency, isolation, networking, and performance - Proficiency in Go and/or Rust, or the ability to ramp quickly - Deep intuition for security and reliability, with attention to failure modes rather than only happy paths - Comfort making high-stakes architectural decisions under uncertainty
Nice to have - Experience with Firecracker, microVMs, containers, or sandboxing - Background with Kubernetes or building custom schedulers - Deep expertise in reducing I/O latency and maximizing throughput for data-intensive workloads
Benefits and work setup Small, globally distributed team based in San Francisco. Position can be remote or on-site, with direct collaboration with the founder and significant technical authority. California pay range listed at $165,000-$230,000 USD.