עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
Why is this role so important?
Our Retail Intelligence products help leading brands and retailers understand how their products, brands and categories perform online. Behind them is data collected from retailers and marketplaces around the world: product pages, brands and categories, each described differently by every site.
Your mission is to turn that data into a single, trusted view: classifying products into a unified taxonomy, normalizing brands and attributes, and matching the same entities across sources. And doing it at scale, across a catalog of more than a billion records that keeps growing and changing every day.
This is an applied ML role within data engineering. Youll build with LLMs, agentic frameworks such as LangGraph, embeddings and classical ML, and ship them as production pipelines. Its hands-on work, not research for its own sake, but it takes a real understanding of classification and NLP methods to choose the right tool for each problem and prove that it works.
So, what will you be doing all day?
Your daily responsibilities may include:
Designing and building LLM-powered and ML-based pipelines that classify, normalize, structure and match product, brand and category data
Building agentic workflows (LangGraph or similar) that automate complex data tasks end to end
Choosing the right approach for each problem (LLMs, embeddings, fine-tuned models, classical classifiers or rules), balancing accuracy, cost and latency
Scaling solutions to run efficiently over billions of records, using Spark, Databricks and our cloud infrastructure
Building evaluation frameworks: ground-truth datasets, labeling processes, quality metrics and ongoing monitoring
Taking solutions from POC to production, and owning them after launch
Working closely with Product to define requirements and shape the roadmap
Collaborating with data engineers, data scientists and other R&D teams on infrastructure and best practices.
This is the perfect job for someone who:
Holds a B.Sc. or M.Sc. in Computer Science, Data Science, Mathematics or another relevant field
Has 4+ years of hands-on experience as a data engineer, ML engineer or data scientist, with solutions running in production
Has strong Python skills and writes production-quality code
Has hands-on experience building LLM-based applications in production (prompt engineering, structured outputs, RAG, embeddings, evaluation)
Has worked with the modern LLM stack: LLM provider APIs (OpenAI, Anthropic, etc.), LangGraph or LangChain, Hugging Face and vector stores
Has a solid grasp of text classification and NLP methods, both classical and modern, and knows when to use each
Has experience processing large-scale data with Spark/PySpark, Databricks or similar, on AWS or another cloud
Understands evaluation and data quality well: precision/recall trade-offs, building ground truth, error analysis
Is pragmatic and delivery-focused, comfortable with ambiguity, and communicates clearly with Product and business stakeholders
Has experience with taxonomies, entity resolution or product/e-commerce data (advantage)
Has experience with fine-tuning or deploying open-source models (advantage).
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
ערב