28 days ago
Base Salary
$180k - $224k/yr
Responsibilities
- Partner with enterprise customers to understand AI use cases, architectures, success criteria, and deployment requirements.
- Lead technical discovery and solution design for pre-sales and expansion opportunities.
- Scope, build, and deliver proofs of concept demonstrating business and technical value.
- Design and implement production-ready Tavily API integrations, including RAG pipelines, agent workflows, internal tools, and industry-specific GenAI applications.
- Work with customer engineering, data, product, and AI teams to move prototypes into production.
- Monitor API usage and recommend improvements to reliability, latency, coverage, and overall value.
- Translate customer needs, blockers, and technical patterns into product and roadmap input.
- Create reusable reference architectures, integration templates, deployment guides, demo environments, and technical documentation.
- Represent Tavily in architecture reviews, executive technical discussions, implementation check-ins, and post-deployment reviews.
- Collaborate with Sales, Customer Success, Product, and Engineering to drive adoption, retention, and expansion.
Requirements
- 5+ years of software engineering experience, ideally in a customer-facing technical role such as Forward Deployed Engineer, Solutions Architect, Solutions Engineer, Sales Engineer, or Technical Consultant.
- Strong hands-on engineering skills with Python, APIs, backend systems, and production software development.
- Experience building with LLMs, Retrieval-Augmented Generation, agent architectures, context engineering, and modern AI application stacks.
- Experience with enterprise technical discovery, proofs of concept, solution design, stakeholder management, and production rollout.
- Strong understanding of how enterprises evaluate, deploy, secure, and scale AI systems.
- Ability to communicate with technical and executive stakeholders and explain complex technical concepts clearly.
- High autonomy, ownership, and comfort working in a fast-moving startup environment.
- Experience with agent or LLM orchestration frameworks such as LangChain, LlamaIndex, LangGraph, OpenAI Agents SDK, or CrewAI is preferred.
- Experience with vector databases such as Pinecone, Weaviate, pgvector, or Qdrant is preferred.
- Experience with enterprise AI use cases, large-customer production rollouts, internal tools, technical playbooks, demos, or customer-facing assets is preferred.
- Based in New York City or willing to relocate; applicants must be authorized to work in the country of application.
Benefits
- Full-time, on-site work in the New York office.
- Competitive compensation and benefits package.
- Career growth and learning opportunities.
- Flexibility and ownership.
- Collaborative, innovative, and international work environment.
- Opportunity to work on impactful AI projects.
Tech Stack
Categories
Forward Deployed
About Nebius
Nebius builds a full-stack AI cloud offering GPU compute, storage, and tools for training and deploying ML models for startups, enterprises, and research labs. It sells consumption-based cloud infrastructure (IaaS/PaaS) and managed services tailored to generative AI workloads, including large-scale model training and inference. The company is headquartered in Amsterdam and operates as an independent provider.
