Recruitment background

Software Engineer - GPU Kernel

프렌들리에이아이|2026. 1. 7. 게시|
16

공고 원문

경력2년~10년
채용 유형정규직
학력학사
지역서울
마감일마감
출처그룹바이

주요업무

About the job FriendliAI is looking for a GPU Kernel Engineer to design, build, and optimize the low-level compute kernels that power our large-scale, GPU-accelerated AI inference platform. You will be delivering world-class inference speed across NVIDIA and AMD GPUs. With our recent $20M funding, we are scaling our team to meet market demand. This is a deeply technical, high-impact role where you will write GPU code, implement advanced optimizations. As part of our engine team, you will contribute directly to the company's proprietary inference engine which supports over 450,000 models on Hugging Face. You will work with the inventors of continuous batching and collaborate with the platform team to deploy your work into production. Key Responsibilities Design, implement, and optimize high-performance GPU kernels for AI inference (e.g., GEMM, attention, routing) Develop and maintain GPU code in CUDA and C++, including low-level assembly when needed Implement reduced-precision and quantized kernels (FP8/FP4) for low-latency or high-throughput inference Benchmark and ensure cross-vendor performance parity between NVIDIA and AMD hardware Contribute to internal GPU libraries and tune performance of performance-critical components Accelerate multi-modal model pipelines Investigate and integrate next-generation GPU features

자격요건

3+ years of experience in GPU programming, HPC, or performance-critical systems Bachelor's or Master's degrees in Computer Science, Computer Engineering, Electrical Engineering, or a related field Strong proficiency in CUDA for NVIDIA GPUs or ROCm/HIP for AMD GPUs Deep understanding of GPU architecture: warps, threads, memory hierarchy, synchronization, and latency-throughput trade-offs Proficiency in C++ Experience with GPU profiling and performance tuning Strong numerical background with understanding of precision trade-offs and quantization techniques

우대사항

Experience optimizing transformer, multi-modal, or Mixture-of-Experts (MoE) architectures at the kernel level Familiarity with the latest GPU libraries and frameworks (CUTLASS, Triton, …) Inter-GPU communication programming experience Open-source contributions related to GPU performance or ML acceleration Research or conference presentations on GPU optimization, HPC, or numerical computing

기술스택

C++, Python, CUDA

채용절차

서류 전형 → 1 차 인터뷰 → 2 차 인터뷰 → 컬처핏 인터뷰 → 최종 합격 근무 형태: 정규직 근무 환경: 유연한 원격 근무 및 필요시 오피스 출근을 병행하는 하이브리드 근무 Aggressive Compensation: 업계 최고 수준의 기본급과 성과에 따른 상한선 없는(Uncapped) 파격적인 커미션 구조 제공 Stock Option: 회사의 성장과 함께 자산 증식이 가능한 스톡옵션 부여 채용 절차는 상황에 따라 일부 변경될 수 있습니다

이런 공고는 어때요?

메텍홀딩스

메텍홀딩스

10년 이하 · 정규직 · 서울

펌웨어엔지니어#RTOS  #펌웨어설계  #임베디드
상시
그룹바이그룹바이
33
클루닉스

클루닉스

5년~20년 · 정규직 · 서울

[서울] GPU/HPC 시스템 개발 경력직 채용#컨테이너  #시스템개발  #HPC
상시
그룹바이그룹바이
20
써냅스

써냅스

10년 이하 · 정규직 · 서울

C#/Python 소프트웨어 개발자 (건축 분야)#Rhino플러그인  #건축설계자동화  #파라메트릭디자인
상시
그룹바이그룹바이
53
님버스테크

님버스테크

20년 이하 · 정규직 · 대구 외 2개

2026년 제2차 범정부 정보자원 통합구축 사업 인재 영입 (광주,대구,대전)#자원통합구축  #광주근무  #공공인프라
상시
그룹바이그룹바이
61
제너레잇

제너레잇

3년~20년 · 정규직 · 서울

[제너레잇] Python Algorithm Engineer#성능최적화  #프롭테크  #기하연산
상시
그룹바이그룹바이
9
타인에이아이

타인에이아이

10년 이하 · 정규직 · 서울

AI-Native Developer (경력무관)#AI에이전트  #AI엔터테인먼트  #MVP개발
상시
그룹바이그룹바이
146
코리아포트원

코리아포트원

5년~20년 · 정규직 · 서울

AI Agent Engineer(5년이상)#에이전트설계  #커머스
상시
그룹바이그룹바이
23
피엔에스테크놀러지

피엔에스테크놀러지

20년 이하 · 정규직 · 경기

[피엔에스테크놀러지] 기술연구소 머신비전 SW 팀원 모집#머신비전  #영상처리  #제조
상시
그룹바이그룹바이
86
타이드풀

타이드풀

5년~10년 · 정규직 · 경기

Product Dev. Team Lead#팀리딩  #HaaS  #EdgeAI
상시
그룹바이그룹바이
3
디사일로

디사일로

20년 이하 · 체험형인턴 · 서울

[DESILO] 시스템 개발자 인턴#Python경험  #API설계  #동형암호
상시
그룹바이그룹바이
132
팀카이

팀카이

20년 이하 · 체험형인턴 · 서울

FDE(Forward Deployed Engineer) 모집 공고#B2B  #고객상담  #AIAgent
상시
그룹바이그룹바이
133
알엑스

알엑스

5년~10년 · 정규직 · 서울

AI Agent 서비스 개발 경력 채용#플랫폼개발  #AI플랫폼
상시
그룹바이그룹바이
111
위시켓

위시켓

3년~10년 · 정규직 · 서울

AIDP FDE(Forward Deployed Engineer) (3년 ~ 12년)#BI구축  #데이터솔루션  #원격근무
상시
그룹바이그룹바이
43
토모도모

토모도모

3년~20년 · 정규직 · 서울

소프트웨어 엔지니어 정규직 채용 (당산) / 아이티레이#금융IT  #솔루션지원
상시
그룹바이그룹바이
77
AMD

AMD

3년 이하 · 정규직 · 해외

AI Framework Eng.#멀티GPU  #ROCm  #LLM
상시
직행직행
6
현대자동차

현대자동차

5년 이상 · 정규직 · 서울

[SDF] 제조 DX 솔루션/시스템 아키텍트#스마트제조  #아키텍처설계  #IIOT
상시
직행직행
153
AWS

AWS

10년 이상 · 정규직 · 서울

Senior AI Application Architect, Professional Services#생성형AI  #RAG  #클라우드
상시
직행직행
10
CJ대한통운

CJ대한통운

3년 이상 · 정규직 · 서울

[TES물류기술연구소] 최적화 솔루션 엔지니어 경력사원 모집#최적화모델  #AI최적화  #물류
상시
직행직행
17
아마존

아마존

2년 이상 · 정규직 · 서울

Data Center Operations Program Technician - People with Disability (PwD), ICN#랙설치  #안전점검  #클라우드
상시
직행직행
0
AMD

AMD

3년 이하 · 체험형인턴 · 해외

Short Term 2027 Software Engineering Intern/ Co-Op#하이브리드  #AI  #Python
상시
직행직행
3