1 day ago
Base Salary
$220k - $256k/yr
Responsibilities
- Define the 3–5 year technical roadmap and architecture for the enterprise data lakehouse.
- Build internal SDKs, CLI tools, control planes, custom operators, and automated orchestration frameworks.
- Design high-performance systems for schema evolution, multi-engine compute optimization, data discovery, and data movement.
- Prototype and benchmark emerging data technologies and specialized Spark extensions.
- Automate data lineage, PII masking, fine-grained access control, and compliance-by-design across petabyte-scale environments.
- Implement streaming support for OpenTelemetry data from AI agents and develop AI evaluation metrics and reporting.
- Lead architecture reviews, establish software and system-design standards, and serve as the escalation point for outages and performance bottlenecks.
- Mentor engineers, influence hiring and development standards, and drive technology adoption across autonomous teams.
Requirements
- 15+ years of experience in software engineering and distributed systems, including at least four years in a principal- or staff-level capacity leading platform-scale initiatives.
- Strong foundations in software engineering, algorithms, data structures, system architecture, and scalable production applications.
- Professional development experience with Java/J2EE, Spring Boot, microservices, and Python.
- Experience building data platforms, framework engines, and APIs that power data movement rather than only building ETL/ELT pipelines.
- Deep knowledge of Apache Iceberg, Hudi, or Delta Lake, including metadata management, manifest files, and compaction strategies.
- Experience contributing to or deeply customizing open-source data projects such as Spark or dbt.
- Extensive experience designing high-throughput, fault-tolerant data pipelines and orchestration frameworks.
- Extensive cloud architecture experience with AWS services and Azure, plus infrastructure automation using Terraform and Git Actions.
- Experience with CI/CD automation, modern development tools, and production-quality, testable, high-performance code.
- Experience using AI code-generation platforms in the software development lifecycle for code generation, reviews, unit tests, and root-cause analysis.
- Ability to architect infrastructure supporting AI-agent development, including vector database integration.
- Demonstrated ability to lead by influence, communicate architectural trade-offs, and drive adoption across multiple autonomous teams.
- Bachelor’s or master’s degree in Computer Science with a distributed-systems focus is preferred, or equivalent deep industry experience.
Benefits
- Health, dental, and vision insurance.
- Retirement savings plan, paid time off, health savings account, flexible spending accounts, life insurance, and disability insurance.
- Tuition reimbursement and additional benefits supporting personal and professional well-being.
- Non-sales roles are typically eligible for a quarterly or annual bonus under the applicable plan.
Tech Stack
Categories
BackendData Engineering
About WEX
WEX is a public financial technology company that provides corporate payment solutions to fleets, healthcare/benefits administrators, and travel businesses. Its products include fuel and mobility cards, virtual card payments, and platforms for HSAs/FSAs/HRAs, with revenue from payment processing and software/service fees. Founded in 1983 and headquartered in Portland, Maine, WEX trades on the NYSE under WEX and was formerly known as Wright Express.
