2 hours ago
Boston, MA, USA +3 moreMid Level
H1B Sponsor
Base Salary
$134k - $247k/yr
Responsibilities
- Own maintenance, support, and operational readiness for internal AI-enabled applications and tools.
- Productionize prototypes and inherited applications by improving deployment configuration, monitoring, error handling, documentation, testing, access controls, and supportability.
- Own and improve Vercel-hosted application operations and support applications across Azure, AWS, GCP, and other cloud environments.
- Build and maintain GitHub Actions CI/CD workflows and infrastructure-as-code patterns.
- Manage secrets, environment variables, deployment automation, access control, logging, alerting, monitoring, and operational documentation.
- Participate in testing, UAT, rollout, incident response, and long-term maintenance for internal applications.
- Fix bugs, update dependencies, address security patches, improve reliability and observability, and reduce operational toil.
- Create production-readiness checklists, UAT checklists, CI/CD templates, infrastructure patterns, runbooks, support handoffs, and internal tool lifecycle standards.
- Support infrastructure and operations for applications using LLMs, AI agents, RAG workflows, automation frameworks, and enterprise integrations.
- Collaborate with Corporate AI, IT, Enterprise Data, Security, and business stakeholders to communicate risks, tradeoffs, and infrastructure needs.
Requirements
- At least 4 years of experience in platform engineering, DevOps, infrastructure engineering, internal tools engineering, automation engineering, software engineering, or a related technical role.
- Strong cloud infrastructure experience with Azure, AWS, or GCP, with Azure especially helpful.
- Experience with infrastructure-as-code tools such as Terraform, Bicep, Pulumi, or CloudFormation.
- Experience building, maintaining, or supporting CI/CD pipelines, especially with GitHub Actions.
- Understanding of deployment patterns, environments, secrets management, access controls, monitoring, logging, and production support.
- Ability to read, maintain, and improve application code in Python, TypeScript, JavaScript, Node.js, or similar languages.
- Experience supporting production or production-like systems, including bug fixes, incident response, dependency updates, documentation, and reliability improvements.
- Familiarity with LLM APIs, prompt engineering, RAG, agents, model evaluation, AI security risks, and responsible AI practices.
- Preferred experience includes Vercel operations, Azure infrastructure and security, authentication and authorization, observability, modern web stacks, backend services, APIs, integrations, serverless applications, containers, AI-powered internal tools, enterprise integrations, regulated environments, lifecycle standards, UAT, rollout processes, and mentoring or enabling engineers.
- Strong ownership, communication, service orientation, and comfort with brownfield engineering, maintenance, support, and operational excellence are expected.
Benefits
- Benefits include options supporting employees physically, financially, and emotionally.
- Additional benefits details are available through Axon’s careers site.
Tech Stack
AWSAzureGitHub ActionsGoogle Cloud PlatformJavaScriptNode.jsPythonReactSnowflakeTerraformTypeScriptVercel
