Recruitment background

Software Engineer - AI Inference System

프렌들리에이아이|2026. 1. 7. 게시|
25

공고 원문

경력2년~10년
채용 유형정규직
학력학사
지역서울
마감일마감
출처그룹바이

주요업무

About the job We are seeking a highly technical Inference Engine Engineer to optimize the performance and efficiency of our core inference engine. You will focus on designing, implementing, and optimizing GPU kernels and supporting infrastructure for next-generation generative and agentic AI workloads. Your work will directly power the most latency-critical and compute-intensive systems deployed by our customers. The ideal candidate is an exceptional engineer with a strong foundation in GPU programming and compiler infrastructure. You enjoy pushing the performance boundaries and have experience supporting production-scale machine learning applications. Key Responsibilities Design and optimize custom GPU kernels for AI (e.g., transformer and diffusion) workloads Contribute to the development of FriendliAI's kernel compiler, memory planner, runtime, and other core components Collaborate with cloud and infrastructure engineers to ensure end-to-end inference performance Analyze performance bottlenecks across the software and hardware stack, and implement targeted optimizations Drive support for new model architectures and tensor compute patterns Maintain production-grade performance infrastructure, including profiling, benchmarking, and validation tools

자격요건

Qualifications 5+ years of experience in production or high-impact research environments Production-level expertise in Python and C++ Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent Experience developing machine learning frameworks or performance-critical runtime systems Hands-on experience writing and optimizing GPU kernels Hands-on experience profiling GPU kernels Experience working with generative AI models such as transformer and diffusion models

우대사항

Preferred Experience Experience developing machine learning compilers or code generation systems Familiarity with dynamic shape compilation, memory planning, and kernel fusion Contributions to inference engines, compilers, or high-performance numerical libraries Understanding of multi-GPU and distributed inference strategies

기술스택

C++, Python

채용절차

서류 전형 → 1 차 인터뷰 → 2 차 인터뷰 → 컬처핏 인터뷰 → 최종 합격 근무 형태: 정규직 근무 환경: 유연한 원격 근무 및 필요시 오피스 출근을 병행하는 하이브리드 근무 Aggressive Compensation: 업계 최고 수준의 기본급과 성과에 따른 상한선 없는(Uncapped) 파격적인 커미션 구조 제공 Stock Option: 회사의 성장과 함께 자산 증식이 가능한 스톡옵션 부여 채용 절차는 상황에 따라 일부 변경될 수 있습니다

이런 공고는 어때요?

트레독스

트레독스

3년~10년 · 정규직 · 서울

AI 엔지니어 구인 · 경력#OCR고도화  #온프레미스  #LLM
상시
그룹바이그룹바이
141
스몰티켓

스몰티켓

7년~15년 · 정규직 · 서울

Video AI & Vision ML Engineer#모빌리티  #AI모델개발  #영상처리
상시
그룹바이그룹바이
54
큐픽스

큐픽스

3년~8년 · 정규직 외 1개 · 경기

머신러닝 엔지니어(LLM, RAG, MLOps, AI Agent) - 공간 지능 플랫폼#AI에이전트  #MLOps  #디지털트윈
상시
그룹바이그룹바이
42
두노소프트

두노소프트

9년~20년 · 계약직 외 1개 · 서울

증권사 퇴직연금 개발 및 운영 SM (여의도/고급/JAVA/장기 근무 가능)계약직,프리랜서#금융IT  #계약관리
상시
그룹바이그룹바이
18
YesPlz AI

YesPlz AI

3년~8년 · 정규직 · 기타

YesPlz AI 에서 같이 성장할 백앤드 엔지니어를 찾습니다! (경력 3년 이상)#재택근무  #이커머스  #API아키텍처
상시
그룹바이그룹바이
436
블랙피그에이아이

블랙피그에이아이

3년~7년 · 정규직 · 서울

[스타트업] 풀스택 개발자 모집#API연동  #프론트엔드개발  #소프트웨어
상시
그룹바이그룹바이
168
르몽

르몽

7년~20년 · 정규직 · 서울

백엔드 엔지니어 겸 개발 PM#백엔드  #개발PM  #RAG
상시
그룹바이그룹바이
84
엘리스그룹

엘리스그룹

5년~10년 · 정규직 · 서울

시니어 클라우드 풀스택 엔지니어#시스템아키텍처  #FastAPI  #클라우드
상시
그룹바이그룹바이
44
이노바이드

이노바이드

5년~20년 · 정규직 · 서울

AI 엔지니어#B2BSaaS  #데이터파이프라인
상시
그룹바이그룹바이
17
몬드리안에이아이

몬드리안에이아이

2년~20년 · 정규직 · 인천

AI/LLM Agent Engineer#AI에이전트  #RAG설계  #제안서작성
상시
그룹바이그룹바이
146
키즐링

키즐링

4년~10년 · 정규직 · 서울

앱 백엔드 개발자(5년 이상)#API설계  #영상플랫폼
상시
그룹바이그룹바이
420
인포뱅크

인포뱅크

4년~8년 · 정규직 · 경기

[코스닥 상장사] CRM Dev Engineer(JAVA) - 4년 이상#대량발송  #CRM
상시
그룹바이그룹바이
82
재미스튜디오

재미스튜디오

10년 이하 · 정규직 · 대전

(주)재미스튜디오 AI Service/MLOps Engineer 모집
상시
그룹바이그룹바이
147
펄크럼테크놀로지스

펄크럼테크놀로지스

3년~20년 · 정규직 · 서울

Product Engineer#이커머스  #AIAgent  #원격근무
상시
그룹바이그룹바이
155
두노소프트

두노소프트

3년~20년 · 계약직 외 1개 · 서울

SaaS 시스템 연동 개발업무 (초중급/JAVA/4개월) 계약직,프리랜서#B2B서비스  #ERP연동
상시
그룹바이그룹바이
47
벙커키즈

벙커키즈

20년 이하 · 정규직 · 서울

[위프(WHIF)] Tech- Backend Engineer#백엔드  #AI캐릭터  #LLM
상시
그룹바이그룹바이
361
지엔에이컴퍼니

지엔에이컴퍼니

2년~10년 · 정규직 · 서울

[플레이오] Backend Engineer (Python)#API개발  #게임플랫폼
상시
그룹바이그룹바이
53
어쎈드(ASCEND)

어쎈드(ASCEND)

10년 이하 · 정규직 · 서울

Quantitative Research Engineer#퀀트트레이딩  #트레이딩개발  #재택가능
상시
그룹바이그룹바이
785
모다플

모다플

3년~10년 · 정규직 · 서울

머신러닝 엔지니어(3년이상)#머신러닝모델  #모빌리티  #MLOps
상시
그룹바이그룹바이
51
래빗홀컴퍼니

래빗홀컴퍼니

10년 이하 · 정규직 외 1개 · 서울

DevOps / 인프라 엔지니어#자율출퇴근제  #글로벌인프라  #서버보안
상시
그룹바이그룹바이
348