About the role
What You’ll Do
- Design, build, and maintain large-scale data pipelines (batch and streaming) for robotics foundation model training and evaluation at petabyte scale
- Own core data infrastructure: data model, storage systems, ingestion pipelines, transformation frameworks, and orchestration layers
- Standardize data models and unify processing pipelines across real-world teleoperation and synthetic simulation datasets
- Collaborate with a team of driven individuals committed to building general-purpose Physical AI
What You’ll Bring
- Excellent software engineering skills (Python, Go, or similar)
- Extensive experience designing, building, and maintaining large-scale data pipelines (8+ years)
- Deep understanding of distributed systems (Spark, Kafka, or similar)
- Extensive experience with data storage technologies (data lakes, warehouses, object stores like S3)
- Experience running and maintaining production-grade infrastructure (Kubernetes, Terraform)
- Bonus: Experience supporting AI systems, in particular embodied AI like self-driving
About this listing
Screened by Joboru
This role passed our automated spam and quality filters and was active in our feed when last checked. Joboru is an aggregator — here is how we screen listings. If anything looks off, tell us.
Similar jobs you may like
Plant Equipment Trainer
1 day agoTalent Finder
Software Engineering Manager
1 day agoHalian Technology Limited
OpenShift Engineer
1 day agoTeksystems
Senior Systems Engineer
1 day agoEclectic Recruitment Ltd
Senior Safety Engineer
1 day agoMeridian Business Support
SRE Engineer
1 day agoTeksystems
CMM Programmer / Inspector
1 day agoProdrive
Activities Coordinator
1 day agoCare UK
CMM Programmer (Composites)
1 day agoThe Collective Network