עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
By joining us, you will be part of a strategic effort to establish us as the definitive platform for high-performance LLM inference. You will engage with our skilled problem-solvers and top organizations, crafting AI technology advancements.
What you'll be doing:
Research, invent, and implement groundbreaking algorithms for LLM inference to advance the state of the art in both low-latency and high-throughput scenarios.
Translate research into practical software solutions that directly impact our products and customers.
Collaborate with internal research, engineering, and product teams across the globe to drive the development of advanced inference technologies.
Analyze the performance of new algorithms on our latest hardware, identifying bottlenecks and opportunities for algorithmic optimizations.
Partner with leading scientific organizations and industry pioneers to remain at the forefront of technological advancements and integrate the latest innovations into practical applications.
What we need to see:
MSc/PhD in Computer Science, Electrical Engineering, or a closely related field.
At least 5 years of relevant experience in deep learning research or applied research.
Publications in a top-tier AI/ML conference (e.g., NeurIPS, ICLR, ICML).
Deep understanding of LLM architectures coupled with hands-on experience in training large-scale models.
Excellent programming skills, particularly in Python and deep learning frameworks like PyTorch, and experience with software engineering standards.
A strong problem-solving mentality and a proactive attitude, driven by the ambition to deliver solutions with real-world impact.
Ways to stand out from the crowd:
Hands-on research experience in LLM inference optimization algorithms such as speculative decoding or parallelization strategies.
Proven experience with High-Performance Computing (HPC) environments, including training or running inference on large-scale GPU clusters (tens to hundreds of GPUs).
Deep familiarity and experience with popular LLM inference systems (e.g., vLLM, TensorRT-LLM).
Experience from a world-class industrial research group or a top-tier institution.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
משרות נוספות מומלצות עבורך
-
Senior Deep Learning Researcher
-
תל אביב - יפו
Buildots
-
-
Senior Deep Learning Researcher, LLM Inference
-
תל אביב - יפו
NVIDIA AI
-
-
Senior Deep Learning Researcher
-
תל אביב - יפו
Buildots
-
-
Senior Deep Learning Researcher, Diffusion
-
תל אביב - יפו
NVIDIA
-
-
Senior Deep Learning Researcher, Diffusion
-
מיקום לא צוין
NVIDIA
-