Remote job
Senior Software Engineer, Data Platform & Backend
Job details
About this role
Role overview
Senior software engineering role focused on building dependable backend and distributed data systems for evidence-based, time-aware economic relationship data. The work spans production services, data processing, historical versioning, cloud infrastructure, and reliable delivery of datasets and APIs. You will own important technical problems from investigation through deployment, verification, and ongoing operation.
Responsibilities
- Design and operate services and processing pipelines that move extracted information through validation, storage, historical versioning, analytics, and delivery. - Build safeguards for data quality, lineage, provenance, reconciliation, and publication control. - Develop reliable asynchronous workflows using queues, workers, orchestration, retries, checkpoints, timeouts, idempotency, and safe reprocessing. - Improve query performance, processing throughput, system observability, deployment workflows, and cloud cost efficiency without weakening correctness. - Ensure current workflows can coexist with point-in-time reconstruction and that later corrections do not alter the historical meaning of earlier data. - Collaborate with AI-focused engineers so probabilistic outputs enter deterministic production systems safely.
Requirements
- Senior-level software engineering ownership, typically supported by six or more years of relevant experience or equivalent technical scope. - Strong production experience with Python, Java, Scala, or a comparable backend language, with the ability to work effectively in Python. - Strong SQL and relational data modeling skills, plus experience building backend services, APIs, distributed processing systems, or data-intensive applications. - Practical experience with asynchronous processing, queues, workers, workflow orchestration, or large-scale batch systems. - Demonstrated ability to design for idempotency, retries, failure recovery, safe replay, automated testing, and production debugging. - Experience with Git, Linux, Docker, CI/CD, and operating production systems on AWS or another major cloud platform; strong written technical communication.
Nice to have
- Experience with Java or Scala in distributed or high-throughput systems; temporal or bitemporal data, event sourcing, change data models, or point-in-time datasets. - Familiarity with data lakes, large-scale object storage, partitioned analytical datasets, workflow frameworks, lineage systems, financial or market data, and audit-heavy enterprise products. - Experience with services such as S3, EC2, ECS, Lambda, SQS, Step Functions, RDS/PostgreSQL, Athena, Glue, and CloudWatch, or equivalent technologies.
Benefits and work setup
- Full-time position based in Delhi NCR, India. The team is currently remote while a collaborative office is being established, with regular in-person collaboration expected afterward. - Requires meaningful and consistent working overlap with the United States Eastern time zone.