3 months ago
Base Salary
$165k - $276k/yr
Responsibilities
- Develop query planning, query execution, columnar storage, encoding and compression, caching, distributed compute, and cluster-level resource management capabilities in Nova.
- Extend Nova to support additional warehouse-imported data types, including metrics, profiles, and dimensions.
- Build for sustained, concurrent, and programmatic query workloads generated by AI agents.
- Lead projects that reduce compute, storage, network, and memory costs while improving or maintaining latency and throughput.
- Profile and optimize JVM performance through garbage collection tuning, memory management, concurrency improvements, and data layout decisions.
- Build guardrails and observability to identify expensive or pathological queries before they affect the system.
- Improve reliability by identifying failure modes, implementing durable fixes, participating in on-call incident response, and strengthening production detection and response.
- Contribute to capacity planning, safe rollout practices, and operational tooling.
- Lead the design and execution of multi-week to multi-month projects and contribute through design documents, architecture discussions, and code reviews.
- Collaborate with Product, Middleware, Data Pipeline, and engineering teams to connect Nova capabilities to customer value.
Requirements
- At least 3 years of industry experience in backend or infrastructure engineering with exposure to distributed data systems.
- Hands-on experience building or extending distributed data systems such as query engines, columnar storage, large-scale data processing frameworks, streaming systems, or storage engines.
- Experience improving cost or performance on cloud infrastructure, including compute, storage, or network resources.
- Strong fundamentals in distributed systems, partitioning, replication, consistency, failover, data structures and algorithms, concurrency, multithreading, and performance optimization.
- Production experience with AWS, S3, DynamoDB, EC2, Kafka, Redis or ElastiCache, Kubernetes, and Terraform, or strong equivalents.
- Proficiency in Java, C++, or Python.
- Ability to own and ship significant portions of complex systems and collaborate effectively with engineers and partner teams.
- Preferred experience with Druid, ClickHouse, Presto/Trino, BigQuery, Snowflake, or similar OLAP or query engine systems.
- Preferred JVM expertise, including garbage collection tuning, profiling, and memory optimization.
- Preferred experience with Arrow, Parquet, ORC, or custom columnar data formats and encodings.
- Preferred familiarity with product analytics, experimentation platforms, or event-driven data systems.
- Open-source data infrastructure contributions or published work in data systems is preferred.
Benefits
- Medical, dental, and vision insurance, including 100% employer-paid premiums for employees on select plans.
- 401(k) retirement plan with employer matching of up to 1% of eligible pay, capped at $2,000 annually.
- Flexible time off and paid holidays.
- Monthly wellness and commuter transit/parking stipends, quarterly learning and development stipends, and new-hire home office equipment.
- Twelve weeks of paid parental leave, fertility, adoption, surrogacy, and backup childcare support.
- Mental health and wellness benefits, including no-cost access to Modern Health coaching and therapy sessions.
- Employee Stock Purchase Program, charitable giving grant, and paid volunteer time off.
- Hybrid work arrangement; Washington employees receive unlimited PTO, 10 to 13 annual holidays, medical, dental, and vision PPO/CDHP plans, and a company-sponsored 401(k).
Tech Stack
Amazon DynamoDBApache KafkaAWSC++ClickHouseGoogle BigQueryJavaKubernetesPrestoPythonRedisSnowflakeTerraform
Categories
BackendData Engineering
About Amplitude
We help companies unlock the power of their products. ⚡ • Join the team: http://amplitude.com/careers • Join the community: http://community.amplitude.com