2 months ago
Bengaluru, IndiaMid Level

Responsibilities

  • Build and maintain LLM harnesses with agent loops, tool/function calling, context construction, memory, retries, and failure handling.
  • Design, iterate, version, and evaluate prompts using structured experiments and evaluation datasets.
  • Build and operate RAG pipelines covering document extraction, chunking, indexing, retrieval, reranking, prompt assembly, and response handling.
  • Generate embeddings and manage vector indexes such as OpenSearch, Pinecone, and pgvector, tuning them for quality and cost.
  • Build data extraction, parsing, and curation pipelines for documents, structured records, and customer datasets.
  • Write accuracy-validation scripts, maintain golden and regression evaluation datasets, and report quality metrics.
  • Apply classification, clustering, lightweight fine-tuning, data analysis, sample inspection, and error analysis to improve AI features.
  • Collaborate with AI Architects, Product Engineers, Senior AI Engineers, and Strike Teams to deliver production AI features on schedule.

Requirements

  • Bachelor's or Master's degree in Computer Science, Data Science, Statistics, or a related technical field.
  • 3+ years of AI or data science experience, including hands-on LLM-based application development.
  • Strong Python skills and experience with pandas, numpy, and scikit-learn.
  • Experience building LLM harnesses with agent loops, tool/function calling, and structured outputs using Anthropic, OpenAI, or Bedrock APIs.
  • Strong prompt engineering practice, including structured iteration, prompt versioning, and prompt evaluation against datasets.
  • Experience with at least one of LangChain, LlamaIndex, or LangGraph and at least one vector database.
  • Knowledge of embeddings, similarity search, retrieval evaluation, and classical machine learning.
  • Ability to write clean, tested production code and familiarity with Git, code review, and CI/CD.
  • Comfort using Claude Code and performing data analysis, error inspection, and iterative experimentation.
  • Preferred qualifications include production experience with LangGraph, Claude Agent SDK, OpenAI Agents, custom agent harnesses, structured evaluation harnesses, AWS, Bedrock, SageMaker, OpenSearch, model fine-tuning or distillation, Guidewire or insurance products, and document-heavy, regulated, insurance, or finance datasets.

Benefits

  • The role is based in Bangalore and joins Guidewire's Professional Services AI Engineering team.
  • The position offers cross-functional Strike Team collaboration with AI Architects, Product Engineers, and Senior AI Engineers.
  • Guidewire is an equal opportunity and affirmative action employer.
  • Employment offers are contingent upon applicable criminal history and other background checks.

Tech Stack

AWSGitNumPyPandasPythonscikit-learn

Categories

Guidewire Software

About Guidewire Software

1,001-5,000 employees

Guidewire is the platform P&C insurers trust to engage, innovate, and grow efficiently. More than 570 insurers in 43 countries, from new ventures to the largest and most complex in the world, rely on Guidewire products. With core systems leveraging data and analytics, digital, and artificial intelligence, Guidewire defines cloud platform excellence for P&C insurers. We are proud of our unparalleled implementation record, with 1,700+ successful projects supported by the industry’s largest R&D team and SI partner ecosystem. Our marketplace represents the largest partner community in P&C, where customers can access hundreds of applications to accelerate integration, localization, and innovation. For more information, please visit https://www.guidewire.com/.

Contact me