or apply directly on KRAFTON's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 30 days KRAFTON's roles stay open a median of 45 days
- Reposted
- No
- Salary listed
- No 0% of KRAFTON's roles list one
- Ghost-job risk at KRAFTON
- high 44 stale, 1 reposted of 61 open
- Hiring momentum
- 92 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-17
Measured from postings appearing on and disappearing from KRAFTON's own greenhouse board since 2026-08-03. Full hiring picture for KRAFTON.
About this role
The ML Serving Engineer will design and implement high-performance ML serving architectures for game applications, ensuring stable and efficient operation of AI models in production environments. Responsibilities include building real-time serving infrastructure for large language models (LLMs), optimizing performance, and establishing CI/CD pipelines for continuous deployment and monitoring.
- benefits
- 1/5
- freshness
- 4/5
- career value
- 5/5
- role clarity
- 5/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of KRAFTON as an employer.
What you need
- 5+ years of experience in designing, building, and operating AI/LLM model serving systems in production environments
- Experience applying high-performance LLM inference engines like vLLM or TensorRT-LLM for system optimization
- Knowledge of memory and computation efficiency techniques such as KV Cache Offloading/Loading and quantization
- Strong analytical skills to address latency and performance issues across inference engines and GPU hardware
- Experience building and enhancing monitoring and observability systems for serving stability and CI/CD automation
- Ability to travel internationally without restrictions
Nice to have
- Experience in making architectural decisions and understanding trade-offs in performance, cost, and stability
- Contributions to open-source serving frameworks like vLLM or LMCache
- Deep understanding of high-performance LLM inference stacks and experience in custom kernel development and CUDA acceleration
Worth weighing
- No salary or specific benefits mentioned
- The role requires a significant level of expertise and experience, which may limit applicants
- The position involves a probation period of 5 months with no changes in employment type or salary during this time
Summarised from KRAFTON's posting. Read the full original.
Listed by KRAFTON on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.
More roles at KRAFTON
- [Game Research & Insights Dept.] Jr. 시선 데이터 분석 연구원 (경력무관 / 인턴)Seoul; Seoul, South Korea
- [HR Div.] ER Manager (10년 이상)Seoul
- [Studio Support Div.] Partner Relationship Manager (5~10년 / 계약직)Seoul
- [KRAFTON JUNGLE] 크래프톤 정글 SW Engineer (1년 이상 / 계약직)Bundang
- [Studio Support Div.] Studio Relationship Manager (3~6년 / 계약직)Seoul
- [Studio Support Div.] Sr. Client Engineer (7년 이상)Seoul
- [Publishing Services Management Div.] 글로벌 라이브 서비스 운영 담당자 (5년 이상)Seoul
- [PUBG Franchise] PUBG: BATTLEGROUNDS Global Marketer (1~3년 / 계약직)Seoul