עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
We are building the first Neuron performance engineering team in Tel Aviv. As a Machine Learning Performance Engineer, you'll help shape the direction of this team from the ground up - profiling and optimizing workloads across the full ML software stack, writing high-performance kernels, and improving the Neuron SDK that external developers depend on. You'll work at the boundary between software and hardware, collaborating directly with compiler, runtime, and chip design engineers to close performance gaps customers care about.
The team is new and small, which means broad scope, direct ownership, and real influence over the technical direction we take. If you enjoy digging into performance bottlenecks and turning analysis into measurable wins, this role is for you.
Key job responsibilities
Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.
Profile ML workloads end-to-end to identify bottlenecks - memory
דרישות:
Basic Qualifications
- 3+ years of non-internship professional software development experience.
- Knowledge of Python and/or C++ programming.
- Knowledge of computer architecture, operating systems, and parallel computing.
- Experience with PyTorch, TensorFlow, and/or JAX.
Preferred Qualifications
- Master's degree in Computer Science, Engineering, Mathematics, or a related field.
- Experience optimizing performance for LLM, Vision, or other deep-learning models.
- Experience with kernel writing or parallel programming (CUDA, Triton, CUTLASS, Pallas, Mojo, SIMD, MPI).
- Experience with compiler optimization or hardware-software co-design. המשרה מיועדת לנשים ולגברים כאחד.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
שאלות ותשובות עבור משרת Machine Learning Performance Engineer
כמהנדס ביצועי למידת מכונה באמזון, תהיה חלק מצוות ה-Neuron הראשון בתל אביב, ותעזור לעצב את כיוונו. תפקידך יכלול אופטימיזציה של עומסי עבודה בכל ערימת התוכנה של למידת מכונה, כתיבת ליבות בעלות ביצועים גבוהים ושיפור ה-SDK של Neuron, המשמש מפתחים חיצוניים. תעבוד בצומת שבין תוכנה לחומרה, בשיתוף פעולה עם מהנדסי מהדרים, זמן ריצה ותכנון שבבים כדי לפתור פערי ביצועים קריטיים ללקוחות.
משרות נוספות מומלצות עבורך
-
Applied AI Engineer
-
תל אביב - יפו
PhaseV
-
-
AI Engineer
-
תל אביב - יפו
אביבית דבוש השמה
-
-
מצטרפים לצוות מומחי הבינה המלאכותית של ארד טק
-
תל אביב - יפו
Arad Tech - Technology, Engineering & AI Recruitment
-
-
AI Engineer
-
תל אביב - יפו
Twine Security
-
-
הפניקס מגייסת ML ENGINEER
-
ראשון לציון
הפניקס חברה לביטוח
-
-
Data Scientist
-
פתח תקווה
HighTech Company
-
ערב