jobify_logo ×
  • מִשׁתַמֵשׁ
  • התחברות/הרשמה
  • עמוד הבית
  • מי אנחנו
  • מעסיקים מובילים
  • פרסום משרה חינם
  • צרו קשר
  • תנאי שימוש
  • מדיניות פרטיות
  • הצהרת נגישות
קרן עזריאלי טקסט בעברית עם סמל אינסוף social_security the_israeli_employment_service work_office המקום
jobify_logo
  • מי אנחנו
  • מעסיקים מובילים
  • פרסום משרה חינם
  • צרו קשר
דילוג לתוכן

עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!

במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.

מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.

Staff Incident Manager

Jobgether

Jobgether Jobgether

  • תל אביב - יפו
  • LinkedIn
LinkedIn

Staff Incident Manager

Jobgether

Jobgether Jobgether

  • תל אביב - יפו
  • bag_icon מלאה, היברידית
  • coins_icon 25,000-35,000 ₪ הערכה מבוססת AI ולא שכר שהתקבל מהמעסיק
    הערכה מבוססת AI ולא שכר של המעסיק
  • LinkedIn
LinkedIn


This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Incident Manager based in Israel.

As a Staff Incident Manager, you will lead the end-to-end response to critical production incidents, ensuring fast resolution while strengthening long-term operational resilience. You will coordinate cross-functional engineering teams during high-impact events, establish best practices for incident management, and continuously improve reliability processes. Working in a remote, globally distributed environment, you will drive operational excellence through structured communication, measurable performance improvements, and blameless postmortems. This role combines technical understanding, leadership, and strategic thinking to help maintain highly available, business-critical systems that support customers worldwide.

Accountabilities

  • Lead the end-to-end management of major and critical production incidents, coordinating stakeholders from detection through resolution while ensuring structured decision-making and timely escalation.
  • Drive improvements in operational reliability by reducing key metrics such as Mean Time to Detect (MTTD), Mean Time to Engage, and Mean Time to Recover (MTTR).
  • Manage and optimize the on-call program, including rotations, escalation policies, alert quality, and paging processes to improve engineer effectiveness and reduce alert fatigue.
  • Coordinate clear and timely communication during incidents, including internal stakeholder updates and customer-facing status communications in partnership with support teams.
  • Facilitate blameless postmortems, ensuring actionable follow-up items are documented, assigned, and completed to prevent recurring incidents.
  • Translate recurring operational issues into long-term reliability initiatives by collaborating with engineering teams on roadmap planning.
  • Define and maintain incident management standards, including severity frameworks, operational procedures, response playbooks, and governance.
  • Monitor and report on incident trends, service reliability, SLA and SLO performance, and operational health to engineering leadership.

Requirements

  • 7+ years of experience in incident management, site reliability, production operations, or a related technical role supporting customer-facing systems.
  • Proven experience leading major incident response in production environments, ideally as an Incident Commander or within a dedicated incident management function.
  • Strong understanding of distributed systems, cloud-native infrastructure, and modern production operations.
  • Experience designing or evolving incident management frameworks, including severity models, runbooks, on-call programs, and operational processes.
  • Hands-on familiarity with observability platforms covering monitoring, logging, tracing, alerting, and incident response; experience with Datadog is strongly preferred.
  • Experience using incident management and paging platforms such as PagerDuty, OpsGenie, or similar tools, along with status page solutions.
  • Demonstrated ability to conduct blameless postmortems that drive meaningful operational improvements.
  • Excellent written and verbal communication skills, with the ability to communicate effectively during high-pressure situations.
  • Strong leadership, organizational, and decision-making skills, with the ability to coordinate multiple technical teams during critical events.
  • Advanced English proficiency is required.
  • Familiarity with PCI-DSS, SRE practices, error budgets, SLO-driven operations, payments, fintech, or other highly regulated industries is considered an advantage.
  • Business-level Spanish proficiency and experience supporting globally distributed engineering organizations are beneficial but not mandatory.
  • Willingness to participate in an on-call rotation and respond to critical incidents outside standard business hours.

Benefits

  • Competitive compensation package.
  • Fully remote work with the flexibility to work from anywhere.
  • One-time home office allowance to set up a productive remote workspace.
  • Company-provided work equipment.
  • Stock options.
  • Comprehensive health insurance coverage.
  • Flexible paid time off.
  • Professional development opportunities, including language, technical, and personal growth courses.
  • Opportunity to work within a globally distributed engineering organization focused on innovation, reliability, and operational excellence.

How Jobgether Works

We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.

We appreciate your interest and wish you the best!

Why Apply Through Jobgether?

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.


במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.

מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.

שאלות ותשובות עבור משרת Staff Incident Manager

כ-Staff Incident Manager ב-Jobgether, תפקידך הוא להוביל את התגובה מקצה לקצה לאירועי ייצור קריטיים, להבטיח פתרון מהיר ולחזק את החוסן התפעולי לטווח ארוך. זה כולל תיאום צוותי הנדסה חוצי-פונקציות במהלך אירועים בעלי השפעה גבוהה, קביעת שיטות עבודה מומלצות לניהול אירועים ושיפור מתמיד של תהליכי אמינות.

לתפקיד Staff Incident Manager ב-Jobgether נדרשים 7+ שנות ניסיון בניהול אירועים או תפקיד טכני דומה, עם ניסיון מוכח בהובלת תגובה לאירועים גדולים בסביבות ייצור. נדרשת הבנה חזקה של מערכות מבוזרות, תשתית ענן-נייטיב ותפעול ייצור מודרני, יחד עם ניסיון בפלטפורמות ניטור ותגובה לאירועים כמו Datadog ו-PagerDuty.

Staff Incident Manager ב-Jobgether תורם לשיפור האמינות התפעולית על ידי הפחתת מדדים מרכזיים כמו Mean Time to Detect (MTTD) ו-Mean Time to Recover (MTTR). בנוסף, הוא מנהל ומייעל את תוכנית הכוננות, כולל רוטציות, מדיניות הסלמה ואיכות התראות, כדי לשפר את יעילות המהנדסים ולהפחית עומס התראות.

לכל המשרות של מנהל אירועים

הכשרות רלוונטיות

ג׳ון ברייס

ג׳ון ברייס

קורס DevOps

  • ערב
ג׳ון ברייס

ג׳ון ברייס

קורס DevOps

  • ערב
NAYA College

NAYA College

מהנדס DevOps בסביבת הענן – Cloud DevOps Engineer – test only

ג׳ון ברייס

ג׳ון ברייס

קורס DevOps

  • בוקר

ניתן לצפות במשרות שסימנת בכל שלב תחת התפריט הראשי בקטגוריית 'משרות שאהבתי'

המקום קרן עזריאלי טקסט בעברית עם סמל אינסוף
  • מי אנחנו
  • מעסיקים מובילים
  • צרו קשר
  • תנאי שימוש
  • מדיניות פרטיות
  • הצהרת נגישות

2026 Ⓒ ג'וביפיי - כל הזכויות שמורות

קרן עזריאלי טקסט בעברית עם סמל אינסוף social_security the_israeli_employment_service israel_innovation_authority work_office המקום
המערכת בונה את הפרופיל התעסוקתי שלך

עוד רגע...

המערכת זיהתה ששינית את הנתונים באזור האישי ומעדכנת את ההמלצות על תפקידים ומשרות בהתאם.

מצטערים, לא הצלחנו לנתח בהצלחה את הנתונים שהזנת.
אתם מוזמנים לנסות להזין שוב או להעלות קובץ קורות חיים במידה ויש לכם.
בהצלחה

הגעת להגבלה היומית של שלושה עדכונים בפרופיל האישי ביום

loader

הבקשה שלך נשלחה בהצלחה!

יש באפשרותך לשלוח בקשה לקבלת ייעוץ אישי ללא עלות מיועצת קריירה.

באפשרותך לשלוח בקשה לקבלת ייעוץ אישי ללא עלות

  • בעיה טכנית

  • סיוע בכתיבת קורות חיים או בהכנה לראיון עבודה

  • התאמה של משרות

  • אחר:

פנייתך נשלחה בהצלחה. נציג מטעם ארגון נכי צהל ייצור איתך קשר בהקדם