3 months ago
Sunnyvale, CA, USAMid Level
H1B Sponsor
Base Salary
$140k - $210k/yr
Responsibilities
- Serve as the named technical partner for existing financial services customers across CoreWeave’s infrastructure, Models, Weave, observability, and inference platform.
- Deepen platform adoption, identify and drive expansion opportunities, strengthen technical stakeholder relationships, and advise customers as they scale AI workloads in production.
- Help customers address model development, research velocity, evaluation, production deployment, governance, performance, reliability, and infrastructure efficiency.
- Partner with Account Managers on commercial activity and Specialist Field Engineers on domain expertise.
- Ensure customers receive maximum value from CoreWeave and proactively address technical blockers and business-critical needs.
- Represent the voice of financial services customers internally and surface product feedback to product and engineering teams.
- Build demos or proofs of concept and translate complex technical concepts for customer engineering and executive audiences.
Requirements
- At least four years of relevant experience in solutions engineering, AI-oriented solutions consulting, or technical field engineering.
- Proficiency in Python.
- Hands-on experience training, fine-tuning, evaluating, and deploying deep learning models, including modern LLM architectures.
- Experience designing and deploying production LLM-powered applications for customer use cases.
- Experience working with financial services customers such as quantitative trading firms, hedge funds, asset managers, banks, or other enterprise finance organizations.
- Familiarity with running AI workloads on at least one major cloud platform: AWS, GCP, or Azure.
- Ability to solve complex and novel technical problems with enterprise customers.
- Excellent written, verbal communication, and presentation skills for engineering and executive audiences.
- Preferred knowledge of cloud infrastructure for AI workloads, GPU compute, high-performance networking, and storage.
- Preferred familiarity with PyTorch and modern LLM tools including vLLM, LangChain, or LlamaIndex.
- Preferred experience with Slurm or Kubernetes for ML job orchestration.
- Preferred experience with hyperparameter optimization and experiment tracking tools.
- Preferred background in ML Engineering, AI Engineering, MLOps, or LLMOps.
- Preferred technical pre-sales or solutions architecture experience focused on net-new logos or greenfield accounts.
- Preferred familiarity with NVIDIA H100, H200, or B200 GPUs, InfiniBand networking, and parallel file systems.
Benefits
- Medical, dental, and vision insurance fully paid by CoreWeave for US-based employees.
- Company-paid life insurance, voluntary supplemental life insurance, and short- and long-term disability insurance.
- Flexible Spending Account and Health Savings Account.
- Tuition reimbursement and eligibility to participate in the Employee Stock Purchase Program.
- Mental wellness benefits through Spring Health and family-forming support through Carrot.
- Paid parental leave and flexible, full-service childcare support through Kinside.
- 401(k) with an employer match.
- Flexible paid time off.
- Catered lunch at office and data center locations.
- Casual work environment and a culture focused on innovative disruption.
- The posting specifies US-based benefits, with benefits varying by location.
- The role requires access to export-controlled information and applicable US person or export authorization eligibility.
Categories
Solutions Engineering
About CoreWeave
CoreWeave is the Essential Cloud for AI. CoreWeave is a cloud purpose-built for scaling, supporting, and accelerating GenAI. We’re a comprehensive platform and strategic partner designed to tackle today—and tomorrow’s—challenges of deploying AI at scale. We manage the complexities of AI growth to make supercomputing accessible and push the limits of what’s possible. Our teams create modern solutions to support modern technology. Get the premier choice for working with GenAI workloads.
