Remote job
Engineering Manager (Distributed P2P Systems)
Job details
About this role
Role overview
An engineering manager is needed to lead a team of senior engineers building the distributed systems that power a decentralized compute and storage platform. The role blends people leadership with hands-on technical ownership of P2P networking, distributed storage, compute orchestration, and core platform services. The work happens in a remote, async, writing-first engineering culture with European time-zone overlap.
Responsibilities
- Lead, mentor, and grow a team of distributed systems and backend engineers - Drive technical direction and execution for P2P, compute, and core platform systems - Stay hands-on with architecture, design, implementation, and critical production code - Guide development of high-throughput distributed protocols, including peer discovery, replication, and data distribution - Oversee compute orchestration, provisioning, scaling, and workload distribution - Champion reliability, observability, and production readiness across the platform - Establish engineering practices around testing, CI/CD, releases, and incident response - Run 1:1s, retrospectives, performance reviews, and career development conversations - Translate technical and business priorities into clear execution plans and measurable outcomes
Requirements
- 7+ years of software engineering experience with 3+ years in an engineering manager or technical leadership role - Proven track record shipping distributed systems or infrastructure software to production - Strong understanding of P2P concepts such as DHTs, gossip protocols, NAT traversal, replication, consensus, or content-addressed systems - Backend and systems programming experience with Go, or alternatively C/C++ or Rust - Solid grasp of concurrency, networking, distributed communication, and fault-tolerant design - Familiarity with technologies such as libp2p, gRPC, Protobuf, IPFS/CIDs, Kubernetes, or containerized infrastructure - Experience with observability tooling such as Prometheus, Grafana, or OpenTelemetry - Comfort working in a remote, async environment with strong written communication
Nice to have
- Experience with distributed storage, object storage, or distributed filesystems (e.g., S3 APIs, NFS, FUSE, POSIX semantics, block storage, or storage engines) - Knowledge of erasure coding, sharding, garbage collection, or large-scale data migration - Experience managing multi-cloud, bare-metal, or geographically distributed infrastructure - Familiarity with GPU infrastructure, distributed AI compute, model serving, or inference routing - Experience with infrastructure-as-code such as Terraform or Pulumi - Background in decentralized, blockchain, Web3, or other P2P architectures