12 hours ago
Base Salary
$217k - $304k/yr
Responsibilities
- Design, write, and deliver reliable software for distributed data movement across streaming and batch workloads.
- Own the architecture and evolution of the ingestion platform’s control plane, data plane, APIs, controllers, connectors, schemas, and sink integrations.
- Expand production-ready ingestion paths to S3, GCS, and Apache Iceberg, including transformations, deduplication, and dead-letter queues.
- Build modular connectors and abstractions for Kafka, BigQuery, S3, GCS, Iceberg, Flink, and other data stores.
- Improve self-service platform usage through APIs, safe defaults, automated provisioning, documentation, onboarding, observability, and alerting.
- Lead safe, incremental migrations from bespoke and legacy ingestion systems in partnership with Ads Data Platform, ML Indexing, and other teams.
- Establish reliability, security, and operational practices for pipelines running across Kubernetes clusters.
- Identify architectural and product-experience gaps and lead redesigns that improve developer velocity and platform scalability.
- Collaborate cross-functionally on roadmaps and durable solutions across Infrastructure, Data Platform, Product, Ads, ML, Storage, and partner teams.
- Mentor and guide backend and data infrastructure engineers.
Requirements
- 10+ years of hands-on experience building internet-scale distributed systems, data infrastructure, or platforms used by other developers.
- BS, MS, or PhD in Computer Science or a related field, or equivalent practical experience.
- Strong software development experience in one or more general-purpose languages such as Go, Python, Java, or Scala.
- Deep experience designing and operating high-throughput, fault-tolerant data pipelines or platform services across streaming and batch processing.
- Experience with data movement systems such as Kafka, BigQuery, S3, GCS, Apache Iceberg, Flink, or comparable technologies.
- Experience designing APIs and platform abstractions, including schema evolution, Protocol Buffers, and gRPC.
- Familiarity with Kubernetes and cloud-native platform concepts such as controllers, custom resources, workload identity, infrastructure automation, and multi-cluster deployments.
- Strong understanding of observability and operational excellence, including metrics, logging, alerting, SLOs, incident response, data quality, and safe migration practices.
- Ability to advocate for platform users and translate ambiguous customer and partner needs into scalable solutions.
- Demonstrated technical leadership guiding teams through complex, cross-functional initiatives.
- Exceptional written and verbal communication skills.
- Experience building low-code or self-service developer platforms, or curiosity to learn, is a strong plus.
Benefits
- U.S.-based employees may receive medical, dental, and vision insurance, a 401(k) program with employer match, generous vacation time, and parental leave.
- The position may include equity in the form of restricted stock units and, depending on the position offered, commission.
- The posting states that interviews may be recorded, transcribed, and summarized by AI, with an opportunity to opt out.
Tech Stack
Categories
BackendData Engineering
About Reddit
Reddit builds a social news and discussion platform organized into user-created forums (“subreddits”) for consumers, creators, and communities. It monetizes primarily through advertising tools for brands and self-serve advertisers, plus a Premium subscription; developers can access data via its API. Founded in 2005 and headquartered in San Francisco, Reddit became a public company on the NYSE in 2024 and hosts over 100,000 active communities.
