19 hours ago
Base Salary
$174k - $252k/yr
Responsibilities
- Own BigQuery’s technical strategy and roadmap for proprietary and open-source storage formats.
- Design storage-layer innovations supporting zero-copy column updates, efficient random access, and deep-learning training pipelines.
- Partner with Dremel, BigLake, and GCS teams to integrate next-generation formats without regressing SQL concurrency.
- Guide Google’s influence in the open data lakehouse ecosystem and the evolution of emerging data formats.
- Lead technical and ROI evaluations of open-source formats and assess Arrow-native execution integration with BigQuery pipelines.
Requirements
- Bachelor’s degree in Computer Science or a related technical field, or equivalent practical experience.
- At least 5 years of software development experience with C++ or Go.
- At least 3 years of experience designing and optimizing distributed systems, storage architectures, or database internals.
- Experience with specialized file formats such as Vortex, Lance, Nimble, or FastLanes is preferred.
- Experience with approximate nearest-neighbor vector search and storage layers optimized for RAG and multimodal data is preferred.
- Understanding of machine-learning data pipelines, PyTorch/TensorFlow data consumption, and AI-training I/O bottlenecks is preferred.
- Successful contributions to major open-source data projects such as Apache Arrow, Parquet, or Iceberg are preferred.
Benefits
- Google benefits are available; the posting also lists equity and a 15% bonus target, though compensation figures are excluded here.
About Google
Google builds consumer and enterprise software and services including Search, Android, YouTube, Chrome, Maps, Gmail, and Google Cloud. Its business model centers on digital advertising and paid cloud, software, and hardware offerings (e.g., Pixel and Nest) for consumers, developers, and organizations. Founded in 1998 and headquartered in Mountain View, California, Google operates globally as a subsidiary of Alphabet Inc.
