← All verified jobs

Data Infrastructure

genesis Bay AreaFullTime

See all open roles at genesis

Ghost-risk verdict

Some ghost-posting signals

  • open for 94 days (90+ without a fill is a strong ghost signal)
  • 40 open roles at this company in 30 days (mass-hiring blitz)
  • no salary disclosed (correlates with ghost postings)

How we score ghost risk →

See your fit for this role and apply with a truthfully tailored résumé.

See my fit, free

About the role

What You’ll Do

Design, build, and maintain large-scale data pipelines (batch and streaming) for robotics foundation model training and evaluation at petabyte scale

Own core data infrastructure: data model, storage systems, ingestion pipelines, transformation frameworks, and orchestration layers

Standardize data models and unify processing pipelines across real-world teleoperation and synthetic simulation datasets

Collaborate with a team of driven individuals committed to building general-purpose Physical AI

What You’ll Bring

Excellent software engineering skills (Python, Go, or similar)

Extensive experience designing, building, and maintaining large-scale data pipelines (8+ years)

Deep understanding of distributed systems (Spark, Kafka, or similar)

Extensive experience with data storage technologies (data lakes, warehouses, object stores like S3)

Experience running and maintaining production-grade infrastructure (Kubernetes, Terraform)

Bonus: Experience supporting AI systems, in particular embodied AI like self-driving

Stop applying to ghosts.

OyaPilot surfaces only verified, real jobs, scores your fit, and tailors your application truthfully.