8 hours ago
Remote, Indonesia +5 moreStaff+
Responsibilities
- Design and scale production ML systems for LLM-based applications.
- Build training and evaluation pipelines for continuous model improvement.
- Fine-tune foundation models using LoRA, QLoRA, SFT, and DPO.
- Optimize inference performance, latency, and GPU utilization.
- Develop training datasets and evaluation frameworks.
- Own deployment, monitoring, and reliability of ML services in production.
- Define engineering standards and mentor a small ML team.
- Partner with product and backend engineers.
Requirements
- Experience building and shipping production ML systems used by real users.
- Strong experience with Python and PyTorch or JAX.
- Hands-on experience with LLM fine-tuning and inference optimization.
- Deep understanding of ML infrastructure, distributed training, or GPU-based systems.
- Strong engineering mindset focused on scalability, reliability, and code quality.
- Comfort taking ownership in a fast-moving startup environment.
Benefits
- Remote-first work environment.
- Founding engineering role with significant technical ownership and influence over product architecture.
- Competitive cash compensation plus equity.
- Long-term growth opportunities.
- Opportunity to work on challenging AI infrastructure problems within a small, high-talent engineering team.
Categories
About OnHires
OnHires is a global recruitment and staffing agency that helps companies hire tech and creative talent, with focus areas including AI, Web3, blockchain, and DeFi. It provides IT recruiting and HR consulting services for startups and established firms, covering roles from engineers and designers to executives. Founded in 2020 and headquartered in Sheridan, Wyoming, OnHires is privately held.
