עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
At JFrog, we're reinventing DevOps to help the world's greatest companies innovate - and we want you along for the ride. This is a special place with a unique combination of brilliance, spirit, and just all-around great people. If you're willing to do more, your career can take off. And since software plays a central role in everyone's lives, you'll be part of an important mission. Thousands of customers, including the majority of the Fortune 100, trust JFrog to manage, accelerate, and secure their software delivery from code to production - a concept we call "liquid software." Wouldn't it be amazing if you could join us on our journey?
We're hiring an SRE to help improve the availability, performance, scalability, and operational excellence of our SaaS environments. You'll work closely with Engineering and Cloud teams to automate operations, scale JFrog's large-scale, multi-cloud, Kubernetes-based SaaS environments, strengthen observability, and improve incident response using modern SRE practices (SLOs/SLIs, error budgets, postmortems).
This role is hands-on, collaborative, and impact-focused. If you're eager to make a significant impact in a fast-paced, high-growth environment, we encourage you to apply.
As a Site Reliability Engineer at JFrog, you will…
Support the reliability, availability, performance, and scalability of JFrog's large-scale, multi-cloud, Kubernetes-based SaaS environments
Investigate and troubleshoot production issues across distributed systems, infrastructure, Kubernetes, and cloud environments in close collaboration with Engineering teams
Design and develop backend services, internal platforms, and production engineering tools using Python, Go, or similar technologies
Improve reliability, observability, and operational readiness through SRE practices, monitoring and alerting, capacity awareness, postmortems, and safer CI/CD and production change processes
Evaluate and contribute to AI-assisted and agentic automation solutions that improve operational efficiency, troubleshooting, and production workflows
Support resilience initiatives, including disaster recovery validation, service readiness, health checks, and production readiness reviews
Participate in on-call rotations, lead incident response when needed, and drive follow-up actions to prevent recurrence
Continuously learn and evaluate new technologies that can improve reliability, automation, and operational excellence
To be a Site Reliability Engineer at JFrog, you need…
2-4 years of experience in SRE, Production Engineering, DevOps, or a similar role with hands-on production exposure
Strong troubleshooting and analytical skills, with the ability to investigate production issues in a structured and methodical way
Hands-on experience with Kubernetes-based containerized workloads
Experience with at least one public cloud provider: AWS, GCP, or Azure
Experience developing backend services, internal platforms, automation, or production engineering tools using Python, Go, or another programming language
Practical understanding of Linux fundamentals, networking concepts, HTTP, DNS, service connectivity, and production troubleshooting
Familiarity with CI/CD tools such as Jenkins, ArgoCD, GitHub Actions, or similar
Exposure to observability tools covering metrics, logs, and traces, such as Prometheus, Grafana, Coralogix, New Relic, or similar platforms
Understanding of incident management processes, alerting systems, and production support workflows
Ability to learn quickly, take ownership, communicate clearly, and work well in a collaborative production environment
Experience using AI-assisted operational workflows such as log analysis, incident summarization, triage support, or troubleshooting – an advantage
Familiarity with agentic automation frameworks such as LangGraph, LangChain, CrewAI, or similar – an advantage
Experience using AI-assisted development tools such as Cursor, Claude Code, GitHub Copilot, ChatGPT, or similar tools – an advantage
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
שאלות ותשובות עבור משרת Site Reliability Engineer
כמהנדס/ת Site Reliability Engineer ב-JFrog, תהיו אחראים/יות לתמוך באמינות, זמינות, ביצועים ומדרגיות של סביבות ה-SaaS מרובות העננים ומבוססות Kubernetes של JFrog. התפקיד כולל חקירה ופתרון בעיות ייצור במערכות מבוזרות, תכנון ופיתוח שירותי בקאנד וכלים הנדסיים, שיפור אמינות באמצעות שיטות SRE, והשתתפות בתורנויות כוננות.
משרות נוספות מומלצות עבורך
-
Service Reliability Engineer MI
-
אור יהודה
Stratacom
-
-
Site Reliability Engineer (SRE) On-Prem
-
תל אביב - יפו
Dream
-
-
Sr. SRE AI Engineer
-
תל אביב - יפו
Navan
-
-
Service Reliability Engineer MI
-
אור יהודה
AudioCodes
-
-
Agentic AI Developer
-
רמת גן
Empire Media Network
-
-
Agent Operations Engineer
-
רמת גן
Empire Media Network
-
ערב
באר שבע