עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
עוד על התפקיד
the role involves practical tasks:
Implement monitoring, logging, tracing, and metrics for On-Premise and Cloud
Manage SLOs, SLAs, SLIs, and Error Budgets
Monitor, troubleshoot, and resolve performance issues. Handle incidents and
perform root cause analysis
Develop automation for operational processes. Manage infrastructure and write
maintenance scripts
Build intuitive dashboards and optimize alert management
Cross-Team collaboration: Work with Development, Infrastructure, Security, and
DevOps teams
Apply SRE principles to enhance system reliability and performance.
Experience: 7+ years hands-on in Observability, Monitoring, SRE, On-Premise
&Cloud
Tools: Prometheus, Grafana, ELK, Splunk, Datadog, Terraform, Ansible, etc.
Programming: Python, Java
Big Data: Hadoop, Spark, Kafka
Cloud: AWS/GCP/Azure, CloudWatch, Stackdriver
SLO/SLA/SLI/ Error Budgets: Practical experience
Incident Management: RCA, incident handling
automation: Scripting and process automation
כישורים נדרשים
Degree in Computer Science or a related field
Kubernetes, Docker experience
AI& Machine Learning knowledge.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
משרות נוספות מומלצות עבורך
-
Site Reliability Engineer (SRE) On-Prem
-
תל אביב - יפו
Dream
-
-
Sr. SRE AI Engineer
-
תל אביב - יפו
Navan
-
-
Service Reliability Engineer MI
-
אור יהודה
AudioCodes
-
-
Agentic AI Developer
-
רמת גן
Empire Media Network
-
-
Agent Operations Engineer
-
רמת גן
Empire Media Network
-
-
Site Reliability Engineer
-
תל אביב - יפו
BMC Software
-
ערב