
Senior Forward Deployed Engineer I (AI Inference)
DigitalOceanabout 3 hours ago
Responsibilities
- Act as an AI Inference lead on the FDE team.
- Architect and deploy production-grade, multi-tenant LLM inference engines.
- Embed with external tech leads to debug latency spikes and profile GPU memory utilization.
- Implement strategies for distributed-systems problems unique to LLM serving.
- Guide customers on compute efficiency techniques.
- Translate customer edge cases into reusable internal blueprints.
Requirements
- 6+ years in AI/ML systems with a focus on cluster-scale serving.
- Hands-on experience with vLLM, llm-d, SGLang, or similar frameworks.
- Expert-level proficiency in Python or GoLang.
- Strong communication skills and a customer-facing engineering mindset.
- Experience in Forward Deployed Engineering or Technical Consulting roles.
Benefits
- Career development resources including reimbursement for conferences and training.
- Access to LinkedIn Learning's 10,000+ courses.
- Competitive benefits including an Employee Assistance Program and flexible time off.
- Potential for bonuses and equity compensation.