LLM Engineer (Optimization)
About the Role
ABOUT THE TEAM & MISSION LLM Engineer (Optimization)는 대규모 언어 모델(LLM)의 추론 성능(Inference Performance)을 극대화하여 실제 서비스 환경에 최적화된 AI 시스템을 개발합니다. 서버(GPU Cluster)부터 Edge 및 On-device 환경까지 다양한 하드웨어에서 최고의 성능과 효율을 달성할 수 있도록 Inference Engine, Runtime, Compiler 및 Model Optimization 기술을 연구·개발합니다. 최신 LLM Serving 기술과 GPU/Accelerator 최적화를 활용하여 고성능·저지연·저비용 AI 서비스를 구현하는 핵심 역할을 수행합니다. RESPONSIBILITIES - LLM Inference Optimization - 대규모 언어 모델(LLM)의 추론 성능(Latency, Throughput, Memory Efficiency)을 최적화합니다. - 다양한 모델 구조 및 추론 환경에 맞는 최적화…
Unlock the apply link on every AI job
Browsing is free. Membership is $9 a year or $4.90 every three months and unlocks the apply link and the full description on every job. Cancel any time from your billing page.
- The apply link on every job, straight to the official posting
- Full job descriptions instead of the preview
- Every AI job, every day: thousands of listings from 1,000+ company career pages
- Save jobs and get daily alerts for your categories
Secure checkout by Dodo Payments·Cancel anytime from your billing page