Responsibilities
- Design, build, and operate a high-quality data lake and supporting software systems.
- Write production-grade software for services, APIs, tooling, and automation.
- Innovate solutions using Trino and Starburst for complex data management challenges.
- Collaborate with technical leads and product managers to develop data products.
- Leverage AI to enhance dataset access for users across Starburst.
- Enable dataset preparation and model evaluation for AI projects.
- Define and evolve engineering processes and best practices.
- Iterate on data architecture and software systems with a focus on quality.
- Identify emerging patterns in data management and software engineering.
Requirements
- At least 7 years of experience in software and/or data engineering.
- Strong fundamentals in software engineering with proficiency in Java, Python, or Scala.
- Experience building and optimizing data pipelines using Trino, Spark, and dbt.
- Experience designing and building backend services and APIs.
- Familiarity with managing data infrastructure in public clouds, especially AWS.
- Experience with orchestration frameworks like Apache Airflow or Dagster.
- Knowledge of AI application design patterns.
- Fluency in SQL and ability to switch between SQL and programming languages.
- Experience with API integrations for data extraction.
- Knowledge of modern data lake modeling techniques.
- Proficiency with Infrastructure-as-Code tools like Terraform or Ansible.
- Strong communication skills and ability to coordinate across teams.
- Willingness to travel 25% for various company events.
Benefits
- Competitive pay and attractive stock grants.
- Flexible paid time off.
- Supportive and inclusive work environment.
Tech Stack
About Starburst
Starburst was founded by the inventors of OS Trino so that data-driven companies could have a full-featured open data lakehouse platform powered by the #1 SQL analytics engine. Our end-to-end analytics platform includes the capabilities needed to discover, organize, consume, and share data with industry-leading price-performance for cloud and on-premises workloads. We believe the lakehouse should be the center of gravity, but support accessing data outside the lake when needed. With Starburst, teams can access more complete data, run scalable analytics, lower the cost of infrastructure, use tools best suited to their needs, and avoid vendor lock-in. Trusted by companies like Apache Corporation, Comcast, Doordash, DBS Bank, and VMware, Starburst helps you make better decisions faster on all data. Join us! -- Our company was founded in an unusual way; with customers and revenue from the beginning! Our growth is already ahead of some of the most successful software start-ups, and we don’t plan on slowing down. We believe our opportunity is huge. Every large company in the world suffers from a data silo problem. Traditional data warehouse products approach the problem with old solutions that breed inefficiency and ultimately can’t help business analysts run fast analytics on all their data. This prevents the business from making better decisions to improve their company’s performance. Starburst provides a modern solution that addresses these data silo & speed of access problems. Starburst helps enterprises harness the value of Trino, the fastest distributed query engine available today, by adding the tools and 24x7 support that meet the needs for big data access at scale. Ultimately, Starburst helps organizations run analytics anywhere to make better business decisions. We’re building a team of all-stars: engineers, customer success, sales, and marketing pros who are at the top of their game. If this sounds like you, check out our open roles: starburst.io/careers
