
Senior Forward Deployed Engineer I (AI Inference)
DigitalOcean8 days ago
Responsibilities
- Act as an AI Inference lead on the FDE team.
- Architect and deploy production-grade, multi-tenant LLM inference engines.
- Embed with external tech leads to debug latency spikes and profile GPU memory utilization.
- Implement strategies for distributed-systems problems unique to LLM serving.
- Guide customers on compute efficiency techniques.
- Translate customer edge cases into reusable internal blueprints.
Requirements
- 6+ years in AI/ML systems with a focus on cluster-scale serving.
- Hands-on experience with vLLM, llm-d, SGLang, or similar frameworks.
- Expert-level proficiency in Python or GoLang.
- Strong communication skills and a customer-facing engineering mindset.
- Experience in Forward Deployed Engineering or Technical Consulting roles.
Benefits
- Career development resources including reimbursement for conferences and training.
- Access to LinkedIn Learning's 10,000+ courses.
- Competitive benefits including an Employee Assistance Program and flexible time off.
- Potential for bonuses and equity compensation.
Tech Stack
About DigitalOcean
DigitalOcean is the AI-Native Cloud purpose-built for the inference and agentic era. Its five-layer integrated platform—spanning GPU and CPU infrastructure, core cloud, inference, data, and managed agent orchestration—is open throughout with no vendor lock-in, giving builders everything they need to start fast, scale production AI workloads, and improve unit economics. More than 650,000 customers and millions of developers globally trust DigitalOcean to build, ship, and scale their applications.