Data Infrastructure

Genesis

  • Bay Area, California, United States
  • Posted Apr 30, 2026
Sign up — let your agent apply Sign in

BestApply tailors your resume and applies for you.

GoS3TerraformPythonSparkKubernetesKafka

Job description

About the role

Design and operate large‑scale data pipelines and infrastructure for robotics foundation model training and evaluation, supporting both batch and streaming workloads at petabyte scale while collaborating with a Physical AI team.

Requirements

  • Excellent software engineering skills in Python, Go, or similar languages
  • 8+ years experience designing, building, and maintaining large-scale data pipelines
  • Deep understanding of distributed systems such as Spark, Kafka, or similar
  • Extensive experience with data storage technologies including data lakes, warehouses, and object stores like S3
  • Experience running and maintaining production‑grade infrastructure using Kubernetes and Terraform
  • Bonus: Experience supporting AI systems, particularly embodied AI like self‑driving