
AI Field Engineer - Enterprise
Fireworks AI3 months ago
Base Salary
$200k - $260k/yr
Responsibilities
- Build end-to-end POCs, MVPs, and production integrations alongside customer engineering teams.
- Architect inference foundations, size deployments, and tune serving patterns for enterprise-scale GenAI products.
- Run load tests and establish latency, throughput, and cost baselines against realistic traffic profiles.
- Deploy and validate model families using inference frameworks, including optimization of model shapes and quantization configurations.
- Advise customers on model selection, fine-tuning strategies, and evaluation methodology.
- Build fine-tuning pipelines and production-quality evaluation frameworks with customers.
- Own technical customer relationships from discovery through production deployment and spend time on-site with customer teams.
- Translate recurring customer pain points into product proposals, internal tooling, documentation, platform improvements, and roadmap feedback.
Requirements
- At least 5 years of hands-on experience in a customer-facing technical role such as Forward Deployed Engineer, Applied AI Engineer, Solutions Architect, ML Engineer with field exposure, or technical founder.
- Demonstrated ability to build production software with customers and ship code running in another organization’s production environment.
- Strong Python skills and familiarity with Kubernetes and infrastructure engineering.
- Working knowledge of LLM inference trade-offs, model serving, and fine-tuning workflows, with SFT required and DPO/RFT preferred.
- Experience with AWS, Azure, or GCP cloud infrastructure and deploying models on GPU infrastructure.
- Ability to conduct discovery conversations, present to executives, and troubleshoot technical issues with ML engineers.
- Preferred: 10+ years in technical field or engineering roles.
- Preferred: experience with vLLM, SGLang, TensorRT-LLM, hyperscaler AI platforms, agentic systems, tool-use chains, or AI-native developer toolchains.
Categories
Forward DeployedML Engineering
About Fireworks AI
Fireworks AI builds a generative AI platform for developers and enterprises to train, fine-tune, and serve open models for production use across text, image, audio, embeddings, and multimodal workloads. It offers managed inference and tooling via APIs on globally distributed infrastructure, with a usage-based SaaS model. Founded in 2022 and headquartered in San Mateo, CA, Fireworks AI is a privately held, Series D company.