דרושים » ניהול ביניים » Senior DevOps Engineer REF320H

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 2 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
At our company Software Technologies, we secure the world - and now we're securing the AI revolution. We're building a Workforce AI Security Platform: a cloud-native, multi-tenant SaaS platform that governs, protects, and enables safe AI adoption across global enterprises. This is a greenfield opportunity to help architect the infrastructure backbone of a platform that will define how organizations adopt AI at work securely.
We're looking for a Senior DevOps Engineer who owns infrastructure end-to-end, ships with confidence, and raises the reliability bar without being asked. You will work closely with backend, full-stack, and security teams, as well as DevOps teams across other our company organizations, to build a highly available, reliable, and secure production environment. If you get energized by building systems that scale, pipelines that teams love, and platforms that never sleep - this role is for you.
Key Responsibilities
Own and evolve our cloud infrastructure across multi-region production environments, end-to-end.
Lead our GitOps deployment model - designing and maintaining declarative, automated deployment workflows with zero manual gates.
Build, maintain, and optimize CI/CD pipelines with a strong focus on developer experience, reliability, and speed.
Initiate, implement, and champion an AI-first DevOps & SRE ecosystem - identifying opportunities, building AI agents and intelligent automation, and driving their adoption across engineering operations.
Develop automation frameworks for provisioning, scaling, observability, and incident response, leveraging AI-powered tooling and agentic workflows to reduce toil.
Operate and improve our observability platform: metrics, logs, alerting, dashboards, SLOs/SLIs, and on-call tooling.
Champion zero-trust secrets management and credential-less authentication patterns across the stack.
Partner with architects and engineering leadership on cloud cost optimization, availability, and performance.
Build internal tooling and automation that multiplies engineering velocity across the organization.
דרישות:
5+ years of hands-on DevOps experience in a SaaS product environment - Must.
Demonstrated initiative in applying AI to engineering operations - designing and building AI agents, agentic workflows, LLM-powered automation, or Model Context Protocol (MCP) integrations that reduced operational toil and improved production reliability, quality, or velocity - Must.
Strong scripting and programming skills - Python and Bash for automation, tooling, and AI agent development; Go is a plus.
Strong motivation to continuously learn and adopt emerging technologies, and to share that knowledge across the team.
Deep, hands-on AWS expertise; multi-cloud (AWS, GCP, Azure) experience is a strong plus - Must.
Strong understanding of containers and orchestration - Docker, Kubernetes, including workloads, networking, service mesh (Istio), Helm/Kustomize, and autoscaling (KEDA, HPA, VPA).
Strong experience with:
Infrastructure-as-Code - Terraform, Crossplane, and/or cloud-native declarative tooling.
GitOps principles and tooling (ArgoCD or equivalent).
CI/CD platforms - building reusable, scalable, security-hardened pipeline templates (GitHub Actions or equivalent).
Secrets management - dynamic injection, IRSA/Workload Identity, avoiding long-lived credentials.
Experience embedding security into CI/CD: vulnerability scanning, SBOM generation, and supply chain security (Trivy, Grype, Syft, JFrog Xray).
Solid observability knowledge - OpenTelemetry, Prometheus, Grafana, Datadog, ELK/OpenSearch, distributed tracing.
Hands-on experience with AI/ML workloads or LLMOps infrastructure - a significant advantage.
Cost-awareness (FinOps) - treating cloud spend as a core engineering metric.
Clear communication skills - able to align engineers, security teams, and leadership around infrastructure decisions.
A strong sense of ownership - proactively identifying gaps and driving improvemen המשרה מיועדת לנשים ולגברים כאחד.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8841149
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at our company and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the company Technology team might be right for you.
We're looking for a Senior DevOps Engineer with a strong security orientation to join our company's DevOps team.
our company's platform runs at significant scale, and our DevOps team sits at the core of keeping it fast, reliable, and secure. As we grow, so does the scope of what we own - we need an engineer who can step in as a strong pillar of our company's production ecosystem.
This is a hands-on role for someone who takes security seriously, moves fast, and knows how to get things done in a complex, high-scale environment.
our company's Technology Stack:
AWS, GCP, CLoudFlare ,Kubernetes, Terragrunt, Ansible, Jenkins, ArgoCD, Argo Workflows, Kong & Nginx, HashiCorp Vault, Kafka, RabbitMQ, Mongodb, Aurora Postgresql & Mysql, Prometheus, Grafana, VictoriaMetrics
Programming languages: Python, NodeJS, Go, Kotlin
What am I going to do?
Full Ownership: Drive infrastructure initiatives through their entire lifecycle, taking accountability from initial design to delivery and long-term operations.
Kubernetes Orchestration: Architect, implement, and maintain production-grade Kubernetes clusters, ensuring they remain scalable and resilient under high-scale demand.
Cloud Architecture: Design and manage robust AWS environments, including VPC, IAM, and EKS, to support a highly available platform architecture.
Infrastructure as Code: Utilize Terraform to build and evolve our environment, applying configuration management principles to all IaC workflows.
CI/CD Excellence: Support and improve our deployment pipelines using Jenkins and GitHub Actions to maintain a fast development velocity.
Observability: Implement comprehensive monitoring solutions with Prometheus and Grafana to ensure deep visibility into platform health.
Operational Resilience: Join the DevOps on-call rotation, taking responsibility for mitigating production issues and maintaining site reliability.
Tooling & Innovation: Continuously evaluate and adopt tools - security and otherwise - that raise the bar on engineering efficiency and security posture at our company.
Requirements:
6+ years of hands-on DevOps / Platform Engineering experience in large-scale production environments on a public cloud (AWS preferred).
Proven leadership mindset - able to own projects end to end and be accountable for outcomes.
Strong, production-grade Kubernetes experience across design, deployment, scaling, and troubleshooting.
Solid AWS experience with VPC, IAM, EC2, EKS, Load Balancers, and DNS.
Experience designing and operating highly available, scalable infrastructure systems.
Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis).
Hands-on experience with Infrastructure as Code and configuration management (Terraform required; Terragrunt and Ansible a plus).
Experience with Docker and containerized workloads.
2+ years building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
Experience with monitoring and observability tools (Prometheus, Grafana).
Experience with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8806869
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have:
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796409
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
30/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior DevOps Engineer at our company, youll build and scale the cloud-native platform behind an AI-native analytics product processing millions of inference requests across multiple LLM providers.
This is a role for someone that can own infrastructure projects end-to-end from Kubernetes and CI/CD to observability and security, while improving developer velocity and reliability at scale.
Responsibilities
Build and scale production Kubernetes infrastructure for AI workloads
Own CI/CD standards and delivery pipelines across teams
Improve reliability, observability, and incident response at scale
Implement secure, scalable IaC foundations (Terraform/Crossplane)
Build internal tooling and self-service workflows to improve DevEx
Drive cloud architecture decisions across availability, performance, and cost.
Requirements:
6+ years in Platform Engineering, DevOps, or SRE roles
Strong cloud experience (GCP preferred) with Kubernetes at scale
Production Kubernetes expertise: HPA, Helm, ArgoCD, Keda, secrets management
Strong IaC foundations with Terraform and/or Crossplane
CI/CD ownership experience (GitHub Actions, GitLab CI, or similar)
Strong scripting/programming skills (Python, Go, or Bash), automation-first mindset
Experience improving developer experience via internal platforms or tooling
Clear communicator who documents decisions, owns outcomes, and drives execution
Advantages
GPU / ML infrastructure experience, scaling compute for AI workloads
OpenTelemetry and observability at scale (Datadog preferred), distributed tracing, metrics, and monitoring
FinOps mindset - cost visibility, rightsizing, and commitment management.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8802446
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
חברה חסויה
Location: Tel Aviv-Yafo and Ra'anana
Job Type: Full Time
As a Senior DevOps Engineer, youll help turn agentic AI capabilities for diagnosing and troubleshooting network and GPU infrastructure into secure, scalable, production-ready services. This role stands out through its end-to-end ownership across cloud and customer-managed environments, close partnership with software and AI engineers, and direct influence on the reliability of NVIDIAs AI infrastructure.


What You'll Be Doing:
Own the DevOps, infrastructure, security, release, and reliability lifecycle - from development environments and CI/CD through deployment, production readiness, and sustained operations.
Build and operate Kubernetes environments and Helm-based deployments for a Python, FastAPI, Node.js, and React microservices platform across SaaS and on-premises footprints.
Engineer GitLab CI/CD pipelines with automated testing, container builds, vulnerability scanning, and versioned image and Helm chart publication through JFrog Artifactory.
Automate infrastructure provisioning, configuration, upgrades, and routine operational workflows to accelerate delivery and improve engineering productivity.
Operate PostgreSQL, Temporal workflow services, and S3-compatible object storage with disciplined capacity planning, backups, recovery testing, and safe migrations.
Strengthen release reliability through deployment validation, reduced-downtime strategies, persistent-state protection, and recovery plans for active workflows.
Deliver actionable observability and security using OpenTelemetry, Datadog/Grafana, Langfuse, secrets management, identity integration, TLS, Kubernetes RBAC, network policies, and container hardening.
Partner with software and AI engineers to troubleshoot distributed systems, investigate incidents, define reliability targets, and improve platform performance, resource efficiency, and customer outcomes.
Requirements:
What We Need to See:
Bachelors degree in Computer Science, Software Engineering, or a related field, or equivalent experience.
5+ years of experience in DevOps, site reliability engineering, or platform engineering supporting distributed applications and microservices.
Strong hands-on experience with Kubernetes, Docker, and Helm, including networking, storage, workload scheduling, scaling, and troubleshooting.
Strong Linux administration skills and proficiency in Python and Bash for automation, plus experience with infrastructure as code and configuration tooling such as Terraform and Ansible.
Experience building and maintaining CI/CD pipelines, including runners, container registries, artifact management, automated quality gates, and secure release practices.
Practical experience operating PostgreSQL or comparable relational databases, including SQL, migrations, backup and restore, and performance troubleshooting.
Strong networking and observability fundamentals across TCP/IP, DNS, HTTP, TLS, load balancing, ingress, metrics, logs, traces, dashboards, and actionable alerting.
Sound understanding of secure infrastructure operations and incident response, with demonstrated ownership, cross-functional collaboration, and prioritization in an evolving environment.


Ways To Stand Out From the Crowd:
Experience operating AI applications, agent platforms, or LLM services, including monitoring latency, failures, token usage, and cost.
Familiarity with Temporal, LangGraph, Model Context Protocol (MCP), Langfuse, ClickHouse, Redis/Valkey, or S3-compatible storage.
Deep experience with OpenTelemetry instrumentation and collectors, Datadog APM, or Prometheus/Grafana.
Experience with self-hosted Kubernetes, OpenShift, Kubernetes operators, CloudNativePG, or GPU clusters and AI data centers.
Experience building reproducible AMD64 and ARM64 container images, optimizing BuildKit pipelines, and securing the software supply chain.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8837920
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
5 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a Senior Platform Engineer to join our CTI Group. We're a startup that builds a lot, so you'll work across a wide range of systems and tools, build new platforms and services, and have real influence over where the platform goes next.
You'll work on the platform behind our CTI products: the microservices running on Kubernetes, the cloud and on-prem environments around them, and the infrastructure behind the systems. There's a lot still to build, and you'll have a real say in what comes next.
Responsibilities
Own what you build end to end, from design through implementation to running it in production.
Design and build the internal platforms, tools, and services other engineering teams rely on, and add new ones as we grow.
Make architectural decisions across infrastructure, applications, networking, security, and data.
Design, build, and maintain our cloud and on-prem environments.
Build and operate the infrastructure behind our AI/ML systems at scale.
Manage and evolve our Kubernetes environments.
Build and maintain CI/CD pipelines, deployment automation, and Infrastructure-as-Code with Terraform.
Build security into the platform, from IAM and secrets management to network policies.
Set up the monitoring and observability that show us how the platform is doing.
Design for scalability, availability, performance, and resilience in high-scale production systems.
Apply Infrastructure-as-Code practices (Terraform) end to end.
Requirements:
5+ years in DevOps or Platform Engineering roles.
Strong software engineering skills, including production-grade Python development using OOP, and proficiency in Bash scripting.
Experience designing and building platforms and infrastructure from scratch.
A strong architectural sense: you understand how infrastructure, applications, networking, security, and data fit together.
Hands-on AWS experience.
Solid Kubernetes experience, including deploying, operating, and troubleshooting clusters.
Experience with Terraform and Docker, and comfort with Linux, networking, and security fundamentals.
Experience with high-scale production systems and with the infrastructure behind AI/ML workloads.
A startup mindset: hands-on, curious, and comfortable with broad scope and a lot of room to shape your own work.
Bonus Points:
Experience in the cyber security field.
On-prem, customer private cloud, or air-gapped deployment experience.
Model serving and ML tooling such as MLFlow, and GPU scheduling on Kubernetes.
Observability stacks such as Prometheus, Grafana, or Loki.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8836113
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're hiring a Senior/Principal Site Reliability Engineer to own production reliability for Cortex Agentix Endpoint Security (following an acquisition of KOI Start Up) as it scales. You'll define and operate our SLOs and error budgets, lead high-severity incident response, and ensure our Kubernetes and AWS infrastructure stays stable under growth. You'll also build and supervise the AI agents that handle routine alert triage and monitor tuning, focusing your own time on the reliability engineering that requires human judgment. This role is a strong fit for someone who treats reliability as an engineering discipline and enjoys ownership, incident command, and applying AI to operational work.
Your Impact:
Own reliability as an engineering discipline - define SLIs, set SLOs, and run error-budget-based decision-making so "how reliable are we" becomes a number that governs how fast we ship.
Own production incidents end-to-end - lead response, mitigation, and resolution for high-severity incidents, and drive blameless postmortems that feed real fixes back into the system.
Own the reliability and capacity of production infrastructure as we scale - forecasting headroom, validating scaling behavior under load, and keeping latency and error rates within SLO.
Run and evolve Kubernetes environments so releases and infra changes are safe by default across hundreds of tenant apps.
Own, build, and supervise our SRE AI agents that triage alerts, review monitors, resolves and summarize incidents. Set and expand the trust ladder that governs what the agents do autonomously, what needs approval, and what stays human. This is a core part of the role.
Requirements:
Your Experience:
5+ years operating production cloud infrastructure, with a strong reliability focus (SRE, or DevOps/platform engineering with reliability ownership).
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Experience defining and operating SLIs, SLOs, and error budgets - or a clear grasp of the discipline and the drive to establish it from scratch.
Strong observability and alerting experience in Datadog or comparable platforms, including raising signal-to-noise in production.
Proven incident-response instincts - comfortable owning high-severity incidents and a genuine believer in blameless postmortems.
Proven ability to own platform and reliability projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work - and in building and supervising agents, not just using them.
Ownership-driven - You take responsibility for the reliability of the systems you build and operate, from SLO definition through incident command and continuous improvement.
Reliability as engineering - You treat reliability as a software problem to be solved with code, measurement, and automation - not an ops queue to be worked by hand.
Collaboration - You work effectively across engineering, security, product, and leadership to align on reliability priorities and drive shared outcomes.
Innovation balanced with pragmatism - You actively explore new approaches, particularly AI-assisted operations and agent supervision, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate reliability, risk, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8834235
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
7 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
Principal DevOps Engineer (Cortex Agentix Endpoint Security)
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities:
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8834189
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Responsibilities
- Design, build, and operate the internal engineering platform powering our company's build, test, deployment, and security validation workflows at scale
- Write and maintain production-grade Python and shell tooling that drives platform automation - this is a hands-on coding role, not just pipeline configuration
- Architect and manage hybrid cloud/on-prem execution infrastructure, including large-scale Kubernetes runner pools across multiple AWS regions
- Own and evolve CI/CD pipelines at scale using GitHub Actions, including reusable workflows, ARC-based runner orchestration, and build caching strategies (BuildKit, sccache, Valkey)
- Operate and tune DinD environments (Sysbox, EBS/NVMe, overlay storage, MTU/networking) for build, test, and release workloads
- Connect and manage self-hosted and on-prem runners, routing physical device (wbox) test jobs by site and device type
- Implement DevSecOps controls including least-privilege IAM, OIDC, isolated runner groups, container signing, and automated security scans
- Drive platform observability, cost optimization, and reliability improvements across the engineering infrastructure
- Collaborate cross-functionally with hundreds of engineers to improve engineering velocity and release confidence
- Take end-to-end ownership of complex infrastructure problems and drive them to resolution.
Requirements:
Technical Skills
- 5+ years of hands-on DevOps experience with a strong software development background - prior development experience is a must
- B.Sc. in Computer Science or equivalent practical experience
- Strong programming skills in Python (or a similar high-level language); ability to write and own production tooling
- Proven experience designing and building scalable systems, automation frameworks, and infrastructure as code using Terraform and Helm
- Solid understanding of Linux, containers (Docker), and Git-based workflows
- Hands-on experience with CI/CD at scale using GitHub Actions or similar - including reusable actions, workflow design, and automation frameworks
- Deep experience with hybrid cloud infrastructure (AWS and on-prem), including EKS, ARC, Karpenter, ECR, S3, Direct Connect, VPC endpoints, IAM/OIDC, and Secrets Manager
- Experience operating spot and on-demand runner pools for builds, DinD tests, releases, and security scans across multiple AWS regions
- Experience with DinD environments (Sysbox, EBS/NVMe, memory limits, overlay storage, MTU/networking) and build caching (BuildKit, sccache, Valkey)
- Experience connecting on-prem/self-hosted runners and routing physical device (wbox) test jobs by site and device type
- Experience implementing DevSecOps controls and improving platform observability, cost efficiency, and reliability
- Platform & tooling familiarity: Kubernetes (EKS, on-prem) GitHub Actions ARC Karpenter Terraform Helm Docker/DinD Sysbox containerd BuildKit ECR S3 ElastiCache (Valkey) sccache Direct Connect VPC endpoints IAM/OIDC Secrets Manager self-hosted runners
Soft Skills
- Strong system-level thinking and troubleshooting skills; able to diagnose and resolve complex infrastructure issues independently
- Takes end-to-end ownership and drives problems to resolution without hand-holding
- Excellent communication and cross-team collaboration skills; comfortable working alongside large engineering organizations
Nice to Have / Advantage
- Experience with Jenkins
- Familiarity with GitHub merge queue
- Experience with MinIO or on-prem S3 caching
- Hardware-in-the-loop CI experience
- MTU/VPC networking tuning expertise
- Monorepo CI optimization experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8828046
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
Were looking for a DevSecOps Manager to lead our Israel-based DevSecOps team and own the security posture of our cloud infrastructure. Youll inherit a strong, established team of DevOps and DevSecOps engineers with meaningful ownership and a mature security program.
This is a hands-on leadership role that combines people management, technical depth, and architectural influence. Youll define the cloud security roadmap, remain close to the code and systems, and partner with Product Security, R&D, and Cloud Engineering to make security part of the platform by default.
Youll own:
The security posture of our multi-cloud infrastructure across AWS, Azure, and GCP.
Identity governance, least-privilege access, credential lifecycle management, and cloud vulnerability management.
Security controls embedded directly into the cloud platform, enabling R&D teams to build and operate securely by default.
Developer toolchain security, including CI/CD pipelines, software supply chain risks, and developer-facing guardrails.
The implementation of security as code, replacing manual processes and ticket-based controls with scalable automation.
The leadership, development, and technical growth of a team of DevOps and DevSecOps engineers.
Teams roadmap, priorities, and quarterly planning in alignment with Cloud Engineering and the broader Security organization.
The teams role as the primary cloud security partner for Product Security, R&D, and Cloud Engineering.
Clear security standards, ownership boundaries, and operating practices across the cloud platform.
Youll solve:
Security challenges across a large, evolving environment with multiple production cells, several cloud providers, and more than 500 engineers.
Gaps in cloud identity, permissions, credentials, workload protection, and vulnerability management.
Security risks within Kubernetes, including RBAC, pod security, network policies, and runtime protection.
Risks across GitHub, CI/CD pipelines, third-party dependencies, and the software supply chain.
Friction between strong security controls and the speed, usability, and autonomy R&D teams need.
Repetitive security work by turning manual controls into automated, reusable platform capabilities.
Cross-functional challenges that require alignment between Cloud Engineering, Product Security, R&D, and compliance stakeholders.
Emerging cloud security gaps by proactively identifying them, proposing practical solutions, and driving remediation through completion.
Requirements:
You are:
An experienced engineering manager with at least two years of experience leading a DevSecOps, DevOps, platform, or cloud security team.
A hands-on technical leader with at least 7 years of experience in DevOps and cloud security who can build and troubleshoot solutions - not only review them.
Deeply experienced with AWS security services and controls (including IAM, SCPs, GuardDuty, CloudTrail, and KMS), and across Azure and GCP as well.
Strong in Kubernetes security, including RBAC, pod security, network policies, and runtime tooling.
Experienced in securing GitHub organizations and CI/CD pipelines built with technologies such as GitHub Actions and Jenkins.
Familiar with cloud security posture management platforms such as Wiz, Orca Security, or Prisma Cloud.
Fluent in infrastructure as code using Terraform, Crossplane, or equivalent technologies.
Comfortable balancing security, developer experience, operational needs, and business priorities, Proactive and accountable, with a track record of identifying gaps and driving solutions through completion.
An advantage:
Experience supporting compliance programs such as SOC 2, GDPR, or ISO 27001.
Security certifications such as CKS, CISSP, or CISM.
Experience with GitOps practices and tools such as Argo CD or Argo Workflows.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8838175
סגור
שירות זה פתוח ללקוחות VIP בלבד