Remote job
Staff Scanning Engineer
Job details
About this role
Role overview
A Staff Software Developer is needed to design, build, and operate the large-scale systems that power internet-wide scanning, DNS resolution, attribution, and data pipelines for an internet-intelligence platform. The role sits at the intersection of backend engineering, distributed systems, and applied internet research, working on production systems that discover, collect, validate, and enrich data across IPs, domains, certificates, services, and cloud infrastructure. The goal is improving the accuracy, coverage, freshness, and reliability of the foundational internet map that powers downstream products and research.
Responsibilities
- Design, build, and operate internet-scale scanning, DNS, crawling, attribution, and data collection systems. - Own production backend services and pipelines that collect, process, validate, and enrich massive internet datasets. - Improve the quality, freshness, reliability, and coverage of the platform's internet map. - Build systems that support IP scanning, DNS lookups, protocol fingerprinting, asset attribution, and large-scale data ingestion. - Lead technical design for distributed services, streaming pipelines, storage systems, caching layers, scheduling systems, and scan orchestration infrastructure. - Partner with research, product, security, and engineering teams to turn research ideas into safe, scalable production systems. - Make thoughtful tradeoffs around reliability, cost, scale, latency, data quality, and operational complexity. - Participate in operational ownership including observability, incident response, and on-call rotation. - Mentor engineers, review designs, and raise the technical bar for backend and scanning infrastructure.
Requirements
- 10+ years of software engineering experience building production systems. - Strong backend engineering experience with data ingestion pipelines, scanning systems, crawlers, schedulers, APIs, storage systems, or high-throughput services. - Experience designing and operating systems that handle large-scale data, high request volume, or internet-scale workloads. - Strong programming experience, with Go as the primary language. - Experience with cloud infrastructure such as AWS, GCP, or Azure. - Experience with message queues or streaming systems such as Google Pub/Sub, Kafka, Kinesis, or similar. - Experience with distributed databases or large-scale storage such as Bigtable, Spanner, HBase, Cassandra, Postgres, or similar. - Strong understanding of reliability, observability, testing, and operability for production systems.