FriendliAI

Software Engineer – AI Inference Engine

FriendliAI
Apply
5 months ago
Seoul, Korea, SouthSenior

Responsibilities

  • Design and optimize custom GPU kernels for transformer and diffusion AI workloads.
  • Develop core inference-engine components including the kernel compiler, memory planner, runtime, and related infrastructure.
  • Collaborate with cloud and infrastructure engineers to optimize end-to-end inference performance.
  • Analyze software- and hardware-level performance bottlenecks and implement targeted optimizations.
  • Add support for new model architectures and tensor-compute patterns.
  • Maintain production-grade profiling, benchmarking, and validation tools.

Requirements

  • At least 5 years of experience in production or high-impact research environments.
  • Production-level expertise in Python and C++.
  • Bachelor’s or master’s degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent.
  • Experience developing machine-learning frameworks or performance-critical runtime systems.
  • Hands-on experience writing, optimizing, and profiling GPU kernels.
  • Experience with generative AI models such as transformer and diffusion models.
  • Preferred: experience with machine-learning compilers, code-generation systems, dynamic-shape compilation, memory planning, kernel fusion, inference engines, compilers, high-performance numerical libraries, and multi-GPU or distributed inference strategies.

Benefits

  • Flexible working hours.
  • Daily lunch and dinner, unlimited snacks and beverages.
  • Health check-up support and top-tier equipment and hardware support.
  • Startup equity, health insurance, and other benefits.
  • Supportive and highly collaborative work environment.

Tech Stack

Categories

FriendliAI

About FriendliAI

51-200 employees
Contact me