Responsibilities
- Design scalable, highly available cloud infrastructure for AI platform deployments, including compute, storage, networking, security, and enterprise integrations.
- Develop Infrastructure as Code and platform deployment solutions using Terraform and Helm, including multi-region high-availability and disaster-recovery strategies.
- Design Kubernetes clusters, autoscaling, multi-zone deployments, observability, and CI/CD pipelines for infrastructure and applications.
- Build multi-agent systems and agent logic using LangChain, LangGraph, or similar frameworks.
- Design AI application evaluation frameworks, optimize prompts through A/B testing, and guide deployment and operations.
- Conduct technical maturity assessments and infrastructure audits, understand enterprise requirements, and present recommendations to customers.
- Partner cross-functionally with engagement managers, product teams, engineering teams, and enterprise customers.
Requirements
- 7+ years of experience in hands-on technical customer-facing roles such as Solutions Architect, Deployed Architect, or Forward Deployed Engineer.
- 3+ years of experience designing and deploying production infrastructure on GCP, AWS, or Azure.
- Strong Kubernetes experience, including GKE, EKS, or AKS cluster design, autoscaling, and multi-zone deployments.
- Experience with Terraform, Helm, GitOps practices, high-availability and disaster-recovery solutions, database systems, networking, security, and observability.
- Experience with CI/CD pipelines for infrastructure and applications.
- 1+ years of experience building production AI/ML applications or agents.
- Strong experience with LLM frameworks such as LangChain and LangGraph, plus experience with state management, evaluation frameworks, prompt engineering, A/B testing, vector stores, RAG patterns, tool integration, API design, and error handling.
- Strong Python and/or TypeScript development skills.
- Enterprise customer-facing experience and experience conducting technical assessments or infrastructure audits.
- Strong communication, problem-solving, consultative, cross-functional collaboration, and hands-on engineering skills.
Benefits
- Medical, dental, and vision coverage.
- Flexible vacation.
- 401(k) plan.
- Meals on in-office days in the US.
- Competitive locally aligned benefits for APAC team members.
- Location: Singapore.
Tech Stack
Categories
About LangChain
At LangChain, our mission is to make intelligent agents ubiquitous. We build the foundation for agent engineering in the real world, helping developers move from prototypes to production-ready AI agents that teams can rely on. What began as widely adopted open-source tools has grown into a platform for building, evaluating, deploying, and operating agents at scale. LangChain provides the agent engineering platform and open source frameworks developers need to ship reliable agents fast. LangSmith offers observability, evaluation, and deployment for rapid iteration. Our open source frameworks, LangGraph, LangChain, and Deep Agents, help developers build agents with speed and granular control. LangSmith is trusted by leading AI teams at Zip, Vanta, Klarna, Workday, Linkedin, Cloudflare, and more.
