עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
Senior SRE Engineer
R&D | Tel Aviv, Israel | Full-time | Senior
About the Role
We’re looking for a Senior SRE Engineer with experience developing tools to join a fast-growing R&D team in the SaaS and AI security space.
This role blends deep infrastructure expertise with strong software engineering skills, focusing on building scalable, reliable, and observable systems.
You’ll play a key role in building SRE practices from the ground up — designing reliability pipelines, observability frameworks, and foundational infrastructure that will scale with rapid company growth.
This is a hands-on position with real ownership, impact, and leadership opportunities.
What You’ll Do
- Design and implement scalable, reliable cloud infrastructure on AWS using Infrastructure as Code (Terraform / Pulumi).
- Build and maintain robust CI/CD pipelines using GitOps methodologies.
- Develop internal tooling and automation in Python / Go to improve reliability and reduce operational toil.
- Architect and implement observability solutions (metrics, logs, tracing, alerting).
- Define and manage SLIs, SLOs, and Error Budgets to ensure system reliability.
- Lead incident response, post-mortems, and long-term reliability improvements.
- Optimize cloud costs through architectural and data-driven decisions.
- Collaborate closely with development teams to improve application performance and resilience.
- Mentor engineers and promote SRE best practices across the organization.
Tech Stack
AWS · Kubernetes (EKS) · Python · Kafka · RabbitMQ · Pulumi · PostgreSQL · Databricks · GitHub Actions
Requirements
- 5+ years of experience in DevOps / SRE roles in production environments.
- Strong programming skills in at least one language (Python, Go, Java, or similar).
- Deep understanding of SRE principles: reliability engineering, capacity planning, incident management.
- Strong Kubernetes expertise (EKS preferred), including networking and advanced workloads.
- Proven experience implementing Infrastructure as Code at scale.
- Hands-on experience with observability tools (Prometheus, Grafana, ELK, Datadog, or similar).
- Experience working with distributed systems and complex architectures.
- Excellent problem-solving and debugging skills.
- Strong communication skills and ability to work cross-functionally.
What Sets You Apart
- You solve operational challenges with code, not just configuration.
- You think system-wide and excel at identifying root causes.
- You’re passionate about automation and reducing manual work.
- You balance engineering excellence with pragmatism.
- You stay up to date with cloud-native technologies and best practices.
- You can clearly explain complex technical concepts to diverse audiences.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
ערב
באר שבע