עדיין מחפשים עבודה במנועי חיפוש? הגיע הזמן להשתדרג!
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.
About the Role
We are looking for a Senior DevOps Engineer / Technical Lead to own our cloud infrastructure long-term across AWS and GCP — production reliability, cloud operations, observability, security, FinOps, and infrastructure strategy.
Reporting to VP R&D.
Key Responsibilities
- Own and evolve cloud infrastructure across AWS and GCP, including Kubernetes-based platforms (EKS/GKE), networking, IAM, storage, and core infrastructure services.
- Lead infrastructure integration efforts during acquisitions, platform consolidations, and cloud migration projects.
- Design, deploy, and maintain Infrastructure-as-Code using Terraform.
- Act as the primary escalation point for infrastructure and production issues.
- Lead incident response, post-mortems, and continuous operational improvements.
- Build and maintain observability platforms using Prometheus, Grafana, Datadog, and related tools, including monitoring standards, alerting strategies, SLOs, and SLAs.
- Support large-scale data pipelines, real-time event processing systems, and high-throughput production environments.
- Troubleshoot and optimize large-scale distributed systems, including capacity planning and performance tuning.
- Lead cloud cost optimization initiatives across AWS and GCP, including FinOps practices, resource governance, and cost visibility.
- Support SOC2, ISO27001, and infrastructure security initiatives, implementing operational controls and security best practices.
- Set technical standards and direction for infrastructure; mentor engineers through design and code review.
Requirements
What You'll Bring
- 5+ years operating large-scale, mission-critical production on AWS, plus architecture-level knowledge of GCP — GKE, VPC, IAM and resource hierarchy, and the managed-service equivalents you'd map to.
- Proven technical ownership of a large-scale infrastructure initiative — target architecture, execution, and delivery to production together with other engineers.
- Deep expertise in Kubernetes (EKS/GKE), cloud networking, infrastructure security, and Infrastructure-as-Code using Terraform.
- Strong experience operating highly available distributed systems, large-scale data pipelines, and real-time event processing environments, including at least two stateful services in production (Kafka, Redis, OpenSearch, ClickHouse, or similar).
- Hands-on experience with observability and production operations — monitoring, alerting, incident response, root cause analysis, and performance optimization under pressure — with deep experience in at least one modern stack (Prometheus/Grafana, Datadog, or similar).
- Experience with capacity planning, cloud cost optimization (FinOps), and infrastructure governance.
Bonus Points For
- Experience leading a large-scale cloud migration — cutover strategy, state migration, running dual-cloud during transition.
- Experience supporting SOC2, ISO27001, or similar security and compliance frameworks.
- Experience with AdTech, MarTech, Gaming, Analytics, or other high-scale data-driven platforms.
- Experience with ClickHouse, BigQuery, Redshift, Snowflake, or similar analytics platforms.
- Experience with VictoriaMetrics, Thanos, Cortex, ArgoCD, Flux, or other modern observability and GitOps tools.
- Proficiency in Python, Go, or Bash for automation and tooling.
במקום לעבור לבד על אלפי מודעות, Jobify מנתחת את קורות החיים שלך ומציגה לך רק משרות שבאמת מתאימות לך.
מעל 80,000 משרות • 4,000 חדשות ביום
חינם. בלי פרסומות. בלי אותיות קטנות.