19 days ago
Base Salary
$180k - $300k/yr
Responsibilities
- Build and scale research infrastructure including experiment harnesses, evaluation pipelines, and trajectory tooling.
- Productionize promising research prototypes by making them fast, reliable, and scalable.
- Partner with researchers to frame hypotheses, select measurements, and design reliable experiments.
- Collaborate with AI engineers, infrastructure teams, and product leads to bring research into production.
- Track developments in LLMs, agent architectures, and evaluation methodology.
- Prototype and evaluate prompting strategies, reasoning workflows, and tool-use policies for production agents.
Requirements
- 5+ years of software engineering experience focused on backend development and scaled systems.
- Hands-on experience running experiments, evaluating results, and iterating on them, including through side projects or personal work.
- Ability to lead in fast-paced startup environments with limited resources and minimal structure.
- Strong experience with Python and web frameworks such as FastAPI.
- Experience deploying applications on AWS and working with ECS or Kubernetes, Postgres, and S3.
- Familiarity with observability and monitoring tools.
- Research engineering experience, experience embedded in a research team, agentic AI systems, large-scale data-driven applications, AI or LLM-powered products, or evaluation and benchmarking tooling are preferred.
Benefits
- U.S. base salary range of $180,000–$300,000 plus startup equity, health insurance, fertility benefits, a technology setup stipend, and flexible time off.
- Additional perks include in-office snacks, team happy hours and outings, an annual company offsite, and built-in collaboration time.
- Full-time, fully in-office role five days per week in New York near Madison Square Park.
Tech Stack
Categories
AI ResearchBackend