
Staff Software Engineer, Data Infrastructure
Peregrine Technologies3 days ago
Base Salary
$200k - $275k/yr
Responsibilities
- Design and operate a high-throughput, real-time data integration platform across diverse customer environments.
- Architect a scalable open table format layer for reliable petabyte-scale data storage.
- Build and optimize distributed batch and streaming data-processing pipelines.
- Drive performance, reliability, and cost efficiency across the data infrastructure stack.
- Collaborate with platform and product engineering teams on data contracts, schemas, and integration patterns.
- Establish best practices, tooling, and patterns that improve data infrastructure quality across the organization.
Requirements
- 8+ years of experience architecting and operating large-scale data infrastructure systems in production.
- Deep expertise with open table formats, particularly Apache Iceberg, including schema evolution, partitioning, compaction, and time travel.
- Extensive hands-on experience with Apache Spark for batch and streaming data processing at scale.
- Strong real-time data integration and stream-processing experience with Apache Kafka, Apache Flink, or equivalent technologies.
- Experience with data pipeline orchestration using Airflow or similar tools.
- Strong software engineering fundamentals in Python and/or Scala with production-quality coding experience.
- Extensive experience with AWS or comparable cloud platforms, including S3-based data lake architectures.
- Experience with Kubernetes and containerized deployment of data workloads.
- Degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
- Located in San Francisco, New York, or Washington DC and open to working in an office.
Benefits
- Benefits, equity if applicable, and bonus if applicable are offered.
- Salary is listed as $200,000–$275,000 annually.
- The role requires working in an office in San Francisco, New York, or Washington DC.
Tech Stack
Categories
Data Engineering