Software Engineer – AI Inference Engine
FriendliAI5 months ago
Seoul, Korea, SouthSenior
Responsibilities
- Design and optimize custom GPU kernels for transformer and diffusion AI workloads.
- Develop core inference-engine components including the kernel compiler, memory planner, runtime, and related infrastructure.
- Collaborate with cloud and infrastructure engineers to optimize end-to-end inference performance.
- Analyze software- and hardware-level performance bottlenecks and implement targeted optimizations.
- Add support for new model architectures and tensor-compute patterns.
- Maintain production-grade profiling, benchmarking, and validation tools.
Requirements
- At least 5 years of experience in production or high-impact research environments.
- Production-level expertise in Python and C++.
- Bachelor’s or master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent.
- Experience developing machine-learning frameworks or performance-critical runtime systems.
- Hands-on experience writing, optimizing, and profiling GPU kernels.
- Experience with generative AI models such as transformer and diffusion models.
- Preferred: experience with machine-learning compilers, code-generation systems, dynamic-shape compilation, memory planning, kernel fusion, inference engines, compilers, high-performance numerical libraries, and multi-GPU or distributed inference strategies.
Benefits
- Flexible working hours.
- Daily lunch and dinner, unlimited snacks and beverages.
- Health check-up support and top-tier equipment and hardware support.
- Startup equity, health insurance, and other benefits.
- Supportive and highly collaborative work environment.