עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
About the Role:
As a SRE Technical Lead at our company, you will play a critical role in ensuring the reliability, scalability, and performance of the platform that powers our customers operations. Youll operate at the intersection of software engineering and production operations, taking full ownership of the systems you build and run.
This is a hands-on individual contributor role - not a managerial position. You'll be deep in the technical work, driving impact through engineering excellence rather than people management.
This role is not just about responding to incidents - its about fundamentally improving how our platform behaves under real-world conditions. You will drive reliability initiatives end-to-end: defining measurable service goals, shaping engineering priorities through error budgets, and implementing solutions that prevent issues before they occur.
Youll work closely with teams across the organization, embedding reliability and observability into every layer of the stack. At the same time, youll leverage automation, modern infrastructure practices, and emerging AI capabilities to continuously evolve how we operate and scale.
What you will do:
Develop deep product knowledge across our platform - understanding its internals, failure modes, and operational behavior well enough to own incident resolution end-to-end.
Define and track SLAs/SLOs/SLIs across critical platform services, and use error budgets to drive engineering decisions.
Own production reliability - including on-call rotations, incident response, and post-mortems - with a focus on minimizing MTTR and preventing recurrence through systemic fixes, not just firefighting.
Work hand-in-hand with engineering teams across the stack - infrastructure, application, and business layers - to embed reliability requirements everywhere.
What skills and experience youll bring to our company:
5+ years of experience as an SRE or Platform Developer (or similar) in a high-scale production environment, with hands-on ownership across the full stack - infrastructure and application layers.
Experience introducing or scaling AI-powered systems in real-world products (ML, LLMs, agents, or decision systems)
Strong coding skills and a software engineering mindset - you build your own tools rather than waiting for someone else to.
A true owner - you take responsibility for systems end-to-end and proactively drive improvements without waiting for direction.
Business-level reliability experience is a strong advantage.
Experience with infrastructure-as-code and modern container orchestration platforms.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
משרות נוספות מומלצות עבורך
-
Sr. SRE AI Engineer
-
תל אביב - יפו
Navan
-
-
Service Reliability Engineer MI
-
אור יהודה
AudioCodes
-
-
Agentic AI Developer
-
רמת גן
Empire Media Network
-
-
Agent Operations Engineer
-
רמת גן
Empire Media Network
-
-
Site Reliability Engineer
-
תל אביב - יפו
BMC Software
-
-
SaaS Site Reliability Engineer
-
תל אביב - יפו
BMC Software
-
ערב
באר שבע