9 hours ago
Remote, EMEAEntry Level
Responsibilities
- Implement, test, debug, and maintain query execution-engine components such as query operators, expression evaluation, execution scheduling, and resource management.
- Write clear, maintainable Java code and add unit and integration tests following established review, testing, and documentation practices.
- Investigate query correctness, reliability, and performance issues using query profiles, logs, telemetry, benchmarks, and production data.
- Develop familiarity with vectorized processing, concurrency, memory usage, parallelism, data exchange, and disk spilling.
- Collaborate with senior engineers and partner teams to understand requirements, break work into manageable tasks, and deliver incremental improvements.
- Use AI-assisted development tools for code comprehension, debugging, test generation, and documentation while validating results through testing and review.
- Participate in design discussions and code reviews, seek feedback, and take on broader technical responsibility over time.
Requirements
- Up to two years of professional software development experience, or equivalent experience through internships, coursework, or substantial personal projects.
- Proficiency in Java or another object-oriented programming language and willingness to learn a large Java codebase.
- Understanding of data structures, algorithms, and basic concurrency concepts.
- Ability to solve well-defined technical problems, ask questions, apply feedback, communicate progress, and surface blockers.
- Commitment to writing tested, maintainable software and learning established development, testing, and review practices.
- Bachelor’s degree in Computer Science or a related technical field, or equivalent practical experience.
- Preferred experience includes database internals, query processing, distributed systems, large-scale data processing, profiling, memory management, concurrency, vectorized processing, C++, Apache Arrow, Parquet, Apache Iceberg, Spark, Hadoop, cloud object stores, or open-source contributions.
Tech Stack
Categories
About Dremio
Dremio builds a unified data lakehouse platform for self-service analytics and AI, enabling SQL-based access and federated queries across cloud, hybrid, and on‑prem data, powered by Apache Iceberg and Apache Arrow. It sells enterprise software and cloud services to data engineering, BI, and data science teams; customers include Maersk and S&P Global. Founded in 2015 and headquartered in Austin, Dremio is now part of SAP.