Responsibilities
- Partner with Sales and enterprise customers to identify business challenges and technical requirements.
- Design and deliver scalable AI solutions, technical demonstrations, architecture discussions, RFI/RFP responses, and PoCs.
- Build and integrate production-grade Generative AI applications, including LLM, RAG, agent, multimodal, knowledge-base, vector-database, and API-integrated systems.
- Design end-to-end architectures and optimize prompts, inference workflows, caching, latency, cost, scalability, and system stability.
- Evaluate production solutions against accuracy, latency, cost, scalability, and customer business KPIs.
- Lead technical customer engagements, support negotiations and deal closure, and contribute to upsell and cross-sell initiatives.
- Own technical quality and commercial success for assigned projects while developing long-term enterprise relationships.
- Provide customer and market feedback to Product and Engineering teams and support product localization for Japan.
- Develop repeatable industry solution frameworks and drive scalable adoption across accounts.
Requirements
- Bachelor’s degree or higher in Computer Science, Engineering, or a related technical field.
- At least 5 years of experience in AI solutions, technical consulting, or related technical roles, including at least 2 years of people management experience.
- Experience delivering enterprise-level customer projects.
- Strong software engineering capability with the ability to build and integrate production-grade applications.
- Familiarity with Generative AI technologies, including LLMs, RAG architectures, agent systems, or multimodal models.
- Experience deploying applications in cloud environments using technologies such as Docker, Kubernetes, or major public cloud platforms.
- Understanding of Transformer architecture fundamentals and prompt engineering principles.
- Experience building knowledge-base systems and working with vector databases.
- Familiarity with LangChain or LlamaIndex and with inference optimization, latency control, and cost management.
- Experience designing scalable, production-ready end-to-end AI architectures and applying Generative AI or multimodal models in production.
- Ability to translate technical capabilities into measurable business value and convert PoCs into commercial deployments.
- Experience building scalable solutions from 0 to 1 and replicating them across multiple accounts.
- Ability to bridge technical and commercial discussions with enterprise stakeholders.
- Proficiency in Python is preferred.
Tech Stack
Categories
About ByteDance
ByteDance is a global incubator of platforms at the cutting edge of commerce, content, entertainment and enterprise services - over 2.5bn people interact with ByteDance products including TikTok. Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible. Together, we inspire creativity and enrich life - a mission we aim towards achieving every day. At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve. We are committed to building a safe, healthy and positive online environment for all our users. We have over 110,000 employees based in more than 30 countries globally. Join us.
