Fireworks AI

AI Field Engineer - Enterprise

Fireworks AI
Apply
3 months ago
Remote, United States +2 moreSenior
H1B sponsor

Base Salary

$200k - $260k/yr

Responsibilities

  • Build end-to-end POCs, MVPs, and production integrations alongside customer engineering teams.
  • Architect inference foundations, size deployments, and tune serving patterns for enterprise-scale GenAI products.
  • Run load tests and establish latency, throughput, and cost baselines against realistic traffic profiles.
  • Deploy and validate model families using inference frameworks, including optimization of model shapes and quantization configurations.
  • Advise customers on model selection, fine-tuning strategies, and evaluation methodology.
  • Build fine-tuning pipelines and production-quality evaluation frameworks with customers.
  • Own technical customer relationships from discovery through production deployment and spend time on-site with customer teams.
  • Translate recurring customer pain points into product proposals, internal tooling, documentation, platform improvements, and roadmap feedback.

Requirements

  • At least 5 years of hands-on experience in a customer-facing technical role such as Forward Deployed Engineer, Applied AI Engineer, Solutions Architect, ML Engineer with field exposure, or technical founder.
  • Demonstrated ability to build production software with customers and ship code running in another organization’s production environment.
  • Strong Python skills and familiarity with Kubernetes and infrastructure engineering.
  • Working knowledge of LLM inference trade-offs, model serving, and fine-tuning workflows, with SFT required and DPO/RFT preferred.
  • Experience with AWS, Azure, or GCP cloud infrastructure and deploying models on GPU infrastructure.
  • Ability to conduct discovery conversations, present to executives, and troubleshoot technical issues with ML engineers.
  • Preferred: 10+ years in technical field or engineering roles.
  • Preferred: experience with vLLM, SGLang, TensorRT-LLM, hyperscaler AI platforms, agentic systems, tool-use chains, or AI-native developer toolchains.
Fireworks AI

About Fireworks AI

201-500 employees

Fireworks AI builds a generative AI platform for developers and enterprises to train, fine-tune, and serve open models for production use across text, image, audio, embeddings, and multimodal workloads. It offers managed inference and tooling via APIs on globally distributed infrastructure, with a usage-based SaaS model. Founded in 2022 and headquartered in San Mateo, CA, Fireworks AI is a privately held, Series D company.

Contact me