
Fireworks AI
Fireworks is the fastest way to build, tune, and scale AI on open models. Ship production-ready AI in seconds on our globally distributed cloud infrastructure, optimized for your use case. Fireworks powers production workloads at companies like Uber, Doordash, Notion, and Cursor—delivering 15× faster speed, 4× lower latency, and 4× more concurrency than closed models.
Open Positions at Fireworks AI
21 open positions
Own reliability across Fireworks AI’s cloud and AI infrastructure, from SLOs and observability to incident response, failure testing, and automation. You’ll solve cross-system production problems and make high-throughput AI systems dependable at scale.
Own foundational enterprise capabilities that let large customers securely run Fireworks in production, spanning backend systems, identity, authorization, billing, encryption, and cloud deployment. You’ll turn customer and security requirements into end-to-end features that operate safely across live AI infrastructure.
Build and own end-to-end AI developer products for Fireworks Nexus, including intelligent routing, integrations, observability, APIs, and enterprise controls. This high-autonomy role combines React and TypeScript frontend work with Python and Go backend development for technical users.
Build the distributed infrastructure that powers large-scale LLM and multimodal model training. You’ll work with AI researchers and engineers to optimize GPU-based training systems, pipelines, storage, and orchestration.
Build and operationalize tailored machine learning solutions that connect cutting-edge AI research with real customer outcomes. This hands-on, customer-focused role spans PoCs, AI applications, model enablement, and ML platform improvements.
Build and deploy production-scale GenAI systems directly with enterprise customers as a hands-on technical partner. This role combines AI infrastructure engineering, model serving and fine-tuning with customer discovery, executive engagement, and product feedback.
Build, fine-tune, and operationalize machine learning models and AI applications for enterprise customers. This hands-on role combines applied ML engineering, customer-facing solution delivery, platform development, and performance optimization.
Embed with ambitious enterprise customers to build, optimize, and deploy production-grade AI systems. This hands-on role combines software delivery, model serving and fine-tuning, technical discovery, and executive-level customer partnership.
Join Fireworks as a Research Engineer working at the intersection of machine learning research and distributed training infrastructure. You’ll design and extend advanced deep learning methods while building the systems needed to run them at massive GPU scale.
Build and optimize the backend and machine-learning infrastructure powering Fireworks AI’s scalable generative AI platform. You’ll work on high-performance systems spanning compute, storage, networking, and model-serving workloads.
Unlock 11 More Matching Jobs
Choose a plan to access all jobs matching your criteria.
First to know
Discover the latest jobs before everyone else
Zero Spam
No ghost jobs, reposts, or sponsored listings
AI-powered filters
Find the most relevant jobs for you