Rivian

Staff Software Engineer, ML Data Infrastructure, Autonomy

Rivian
Apply
4 days ago
Palo Alto, CA, USAStaff+
H1B sponsor

Base Salary

$207k - $258k/yr

Responsibilities

  • Own the unified autonomy data layer spanning document metadata, columnar analytics, real-time search, and object-stored sensor payloads.
  • Design the data lake, metadata access layer, APIs, schemas, formats, and migration strategy.
  • Build batch and streaming pipelines and storage systems for evaluation data from on-road and simulation runs.
  • Develop indexing, semantic search, scenario discovery, and embedding-based retrieval capabilities across fleet data.
  • Create documented tools and APIs that allow engineers to find, slice, and materialize data without writing pipelines.
  • Establish schema validation, producer-consumer contracts, freshness and completeness monitoring, alerting, and continuous platform testing.
  • Set data engineering standards, provide technical leadership, and mentor engineers across Autonomy.
  • Partner with security and privacy teams on retention, access control, consent, and regional requirements for vehicle data.

Requirements

  • Bachelor’s degree in Computer Science, Electrical Engineering, or a related field, or equivalent experience.
  • 8+ years of software engineering experience focused on data infrastructure or data platforms.
  • 5+ years owning large-scale ML or analytics data platforms, including storage and access models.
  • 5+ years with columnar and analytical stores such as ClickHouse, Pinot, Druid, BigQuery, Databricks, or Snowflake.
  • 3+ years designing storage and query layers for heterogeneous, high-volume metrics.
  • 3+ years with document or NoSQL stores, including schema design, indexing, sharding, and operations at scale.
  • 3+ years with Parquet, Arrow, Iceberg, Delta, and modern data or table formats.
  • 3+ years building batch and streaming pipelines with production orchestration and data quality gates.
  • 3+ years of hands-on Python and production experience with Go, C++, or Rust.
  • 3+ years with queueing and event-driven systems such as SQS, Kafka, or Kinesis.
  • 2+ years building production monitoring and alerting with tools such as Prometheus, Grafana, Datadog, or CloudWatch.
  • 2+ years migrating production data platforms between storage or format architectures while maintaining service availability.
  • Demonstrated technical leadership, system design ability, implementation depth, schema and interface design, and mentoring experience.
  • Preferred experience includes applied ML for data problems, robotics or autonomous-vehicle data, vector databases, ML training-data pipelines, feature stores, data catalogs, or lineage systems.

Benefits

  • Comprehensive benefits are available to eligible full-time and part-time employees and their families, including paid vacation, paid sick leave, life, medical, dental, vision, short-term disability, and long-term disability insurance.
  • Eligible employees may participate in Rivian’s 401(k) Plan and Employee Stock Purchase Program.
  • Full-time employee coverage begins on the first day of employment; part-time coverage begins the first of the month after 90 days.
  • The posted salary range applies to San Francisco Bay Area applicants.

Tech Stack

Apache KafkaC++ClickHouseDatabricksDatadogGoGoogle BigQueryGrafanaPrometheusPythonRustSnowflake

Categories

Data Engineering
Rivian

About Rivian

10,000+ employees

Rivian designs and manufactures electric vehicles, including the R1T pickup, R1S SUV, and commercial delivery vans, selling directly to consumers and fleet operators. Founded in 2009 and headquartered in Irvine, California, it operates a manufacturing plant in Normal, Illinois, and is publicly traded on NASDAQ. The company also develops in-vehicle software and electrical architecture and offers charging and service support for its owners.

Contact me