
Senior Software Engineer, LLM Inference
Description
מהנדס/ת בכיר/ה בתוך צוות ה-LLM Inference ב-NVIDIA יעבוד/תה על אופטימיזציה של אלגוריתמים להסקה של מודלים שפה גדולים, כולל ארכיטקטורות היברידיות ו-Mixture-of-Experts, על פני חומרת NVIDIA המגוונת מנתונים גדולים עד devices edge. התפקיד כולל כתיבת ותיוניון של GPU kernels ב-CUDA ו-Triton, פתרון בעיות של Distributed Inference, ותרומה לספריות open-source כמו vLLM ו-SGLang. דרוש/ה בעל/ת לפחות 5 שנים של ניסיון בהנדסת תוכנה במערכות קריטיות לביצועים, הבנה עמוקה של ארכיטקטורות deep learning, וניסיון בתכנות GPU ותחומים קשורים לחומרה כולל networking וחישובים מבוזרים. מיקום: תל אביב · משרה מלאה להגשת מועמדות: באתר החברה, דרך הכפתור למטה
Details
Location
false
Location
Tel Aviv
Senior Software Engineer, LLM Inference
This role is published on NVIDIA's own site. You apply there; our part is getting it seen.
Apply on the company siteAsk about this listingSource
This listing was collected from a public external source. To contact the poster, visit the original post.
Original page at nvidia.wd5.myworkdayjobs.com