jobify_logo ×
  • מִשׁתַמֵשׁ
  • התחברות/הרשמה
  • עמוד הבית
  • מי אנחנו
  • מעסיקים מובילים
  • פרסום משרה חינם
  • צרו קשר
  • תנאי שימוש
  • מדיניות פרטיות
  • הצהרת נגישות
קרן עזריאלי טקסט בעברית עם סמל אינסוף social_security the_israeli_employment_service work_office המקום
jobify_logo
  • מי אנחנו
  • מעסיקים מובילים
  • פרסום משרה חינם
  • צרו קשר
דילוג לתוכן

עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!

במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.

מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.

Machine Learning Performance Engineer, Annapurna Labs

Amazon Web Services (AWS)

Amazon Web Services (AWS) Amazon Web Services (AWS)

  • תל אביב - יפו
  • LinkedIn
LinkedIn

Machine Learning Performance Engineer, Annapurna Labs

Amazon Web Services (AWS)

Amazon Web Services (AWS) Amazon Web Services (AWS)

  • תל אביב - יפו
  • bag_icon מלאה, עבודה מהבית
  • coins_icon 25,000-35,000 ₪ (הערכה מבוססת AI)
    הערכה מבוססת AI ולא שכר של המעסיק
  • LinkedIn
LinkedIn


Description

The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and generative AI workloads on Amazon's custom machine learning accelerators — Inferentia and Trainium. These chips power workloads for thousands of AWS customers, from large language model training runs to real-time inference serving billions of daily predictions.

We are building the first Neuron performance engineering team in Tel Aviv. As a Machine Learning Performance Engineer, you'll help shape the direction of this team from the ground up — profiling and optimizing workloads across the full ML software stack, writing high-performance kernels, and improving the Neuron SDK that external developers depend on. You'll work at the boundary between software and hardware, collaborating directly with compiler, runtime, and chip design engineers to close performance gaps customers care about.

The team is new and small, which means broad scope, direct ownership, and real influence over the technical direction we take. If you enjoy digging into performance bottlenecks and turning analysis into measurable wins, this role is for you.

Key job responsibilities

Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.

Profile ML workloads end-to-end to identify bottlenecks — memory, compute, or communication — and drive optimizations through to a measured improvement.

Enhance the programming model and tooling that kernel and model developers rely on, improving usability and debugging workflows.

Identify and drive optimization opportunities across the Neuron software stack (compiler, runtime, frameworks).

Document software designs, operational runbooks, and performance findings so the broader team can build on your work.

A day in the life

You might start your morning reviewing profiling data from a customer's large diffusion model training job, tracing a utilization gap back to a specific kernel. After a design discussion with compiler engineers about a new operator fusion strategy, you spend the afternoon writing and benchmarking a kernel prototype. Later, you review a teammate's pull request for a runtime optimization and share your findings in a short write-up for the broader Neuron organization. Your work directly translates into faster model execution and lower cost for AWS customers running ML workloads at scale.

About The Team

The Neuron Performance Engineering team in Tel Aviv is part of Annapurna Labs within AWS. Our mission is to make sure every ML workload running on Inferentia and Trainium chips reaches its full performance potential. We partner closely with compiler, framework, and hardware teams across Annapurna Labs, and we work directly with AWS customers to understand their models and unblock their adoption.

We are a newly formed group which is part of the larger Neuron organization. If you want to shape a team's technical culture from its earliest days while working on problems that matter to the future of AI infrastructure, we'd love to hear from you.

Basic Qualifications

  • 3+ years of non-internship professional software development experience
  • Knowledge of Python and/or C++ programming
  • Knowledge of computer architecture, operating systems, and parallel computing
  • Experience with PyTorch, TensorFlow, and/or JAX

Preferred Qualifications

  • Master's degree in Computer Science, Engineering, Mathematics, or a related field
  • Experience optimizing performance for LLM, Vision, or other deep-learning models
  • Experience with kernel writing or parallel programming (CUDA, Triton, CUTLASS, Pallas, Mojo, SIMD, MPI)
  • Experience with compiler optimization or hardware-software co-design

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Company - Annapurna Labs Ltd.

Job ID: A10555765


במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.

מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.

שאלות ותשובות עבור משרת Machine Learning Performance Engineer, Annapurna Labs

כמהנדס ביצועי למידת מכונה ב-Annapurna Labs ב-AWS, תהיה אחראי על אופטימיזציה של עומסי עבודה של למידת מכונה על מאיצי ה-ML המותאמים אישית של אמזון, Inferentia ו-Trainium. זה כולל פרופיל וזיהוי צווארי בקבוק, כתיבת ליבות בעלות ביצועים גבוהים ושיפור ה-SDK של Neuron, המשמש להאצת למידה עמוקה ועומסי עבודה של AI גנרטיבי.

מהנדס ביצועי למידת מכונה ב-Annapurna Labs תורם ישירות ללקוחות AWS על ידי הבטחת שעומסי העבודה של למידת המכונה שלהם יגיעו לפוטנציאל הביצועים המלא שלהם. עבודתך מתורגמת לביצוע מודלים מהיר יותר ועלויות נמוכות יותר עבור לקוחות המריצים עומסי עבודה של ML בקנה מידה גדול, על ידי אופטימיזציה של ערימת התוכנה של Neuron ושיתוף פעולה עם צוותי חומרה וקומפיילר.

לתפקיד מהנדס ביצועי למידת מכונה בצוות החדש של Annapurna Labs בתל אביב נדרשות לפחות 3 שנות ניסיון בפיתוח תוכנה מקצועי, ידע ב-Python ו/או C++, הבנה בארכיטקטורת מחשבים, מערכות הפעלה ומחשוב מקבילי, וניסיון עם PyTorch, TensorFlow ו/או JAX. ניסיון באופטימיזציית ביצועים עבור מודלי LLM או למידה עמוקה אחרים, וכתיבת ליבות, מהווים יתרון.

משרות נוספות מומלצות עבורך
  • רשימת משאלות

    Machine Learning Performance Engineer, Annapurna Labs

    • map_icon תל אביב - יפו
    Annapurna Labs Ltd.

    Annapurna Labs Ltd.

  • רשימת משאלות

    AI Engineer

    • map_icon תל אביב - יפו
    Autofleet

    Autofleet

  • רשימת משאלות

    AI Developer

    • map_icon תל אביב - יפו
    Acentecom

    Acentecom

  • רשימת משאלות

    Machine Learning Backend Engineer

    • map_icon תל אביב - יפו
    mylo AI

    mylo AI

  • רשימת משאלות

    Software Engineer - Institutional AI

    • map_icon רמת גן
    HighTech Company

    HighTech Company

  • רשימת משאלות

    Data Scientist

    • map_icon תל אביב - יפו
    mylo AI

    mylo AI

לכל המשרות של מהנדס למידת מכונה

הכשרות רלוונטיות

מכללת INT

מכללת INT

קורס דאטה סיינס / Data Science

הטכניון -  מכון טכנולוגי לישראל

הטכניון - מכון טכנולוגי לישראל

Generative AI and LLM Hands On

  • ערב
  • clk_icon 7 חודשים
Google Reichman Tech School

Google Reichman Tech School

פיתוח מודלים של AI ו-Deep Learning

  • ערב
  • clk_icon 4 חודשים
Google Reichman Tech School

Google Reichman Tech School

פיתוח מודלים של AI ו-Deep Learning

  • ערב
  • clk_icon 4 חודשים

ניתן לצפות במשרות שסימנת בכל שלב תחת התפריט הראשי בקטגוריית 'משרות שאהבתי'

המקום קרן עזריאלי טקסט בעברית עם סמל אינסוף
  • מי אנחנו
  • מעסיקים מובילים
  • צרו קשר
  • תנאי שימוש
  • מדיניות פרטיות
  • הצהרת נגישות

2026 Ⓒ ג'וביפיי - כל הזכויות שמורות

קרן עזריאלי טקסט בעברית עם סמל אינסוף social_security the_israeli_employment_service israel_innovation_authority work_office המקום
המערכת בונה את הפרופיל התעסוקתי שלך

עוד רגע...

המערכת זיהתה ששינית את הנתונים באזור האישי ומעדכנת את ההמלצות על תפקידים ומשרות בהתאם.

מצטערים, לא הצלחנו לנתח בהצלחה את הנתונים שהזנת.
אתם מוזמנים לנסות להזין שוב או להעלות קובץ קורות חיים במידה ויש לכם.
בהצלחה

הגעת להגבלה היומית של שלושה עדכונים בפרופיל האישי ביום

loader

הבקשה שלך נשלחה בהצלחה!

יש באפשרותך לשלוח בקשה לקבלת ייעוץ אישי ללא עלות מיועצת קריירה.

באפשרותך לשלוח בקשה לקבלת ייעוץ אישי ללא עלות

  • בעיה טכנית

  • סיוע בכתיבת קורות חיים או בהכנה לראיון עבודה

  • התאמה של משרות

  • אחר:

פנייתך נשלחה בהצלחה. נציג מטעם ארגון נכי צהל ייצור איתך קשר בהקדם