דרושים » ניהול ביניים » Senior DevOps Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are transforming our internal infrastructure into a high-velocity, self-service product. This is a senior opportunity to architect the future of our platform, setting the technical standard for how we deliver services while ensuring reliability and developer independence across diverse global markets.

Responsabilities:

Self-service enablement: Build secure, automated pathways that accelerate development while keeping organizational standards intact.
Architectural modularization: Design Kubernetes-native abstractions - controllers and operators - and decouple infrastructure into independent service lifecycles that support rapid iteration.
Everything as code: Treat infrastructure, policy, and configuration with the same SDLC as application software - versioned, reviewed, tested, and CI-gated.
Proactive reliability: Own production through robust observability, clear SLOs, and high operational hygiene; define alerting that fires correctly and lead blameless postmortems that fix the system.
Requirements:
Strong GCP experience at production scale (GKE, IAM, networking, Cloud SQL/BigQuery). No compromise here.
Kubernetes-native mindset: declarative-first, comfortable with controllers and operators, and a platform-engineering instinct for self-service and separation of concerns.
Everything-as-code discipline: any code ships through a defined SDLC (versioned, reviewed, tested, CI-gated). Terraform (KCC a plus), Helm, GitOps (ArgoCD or similar) in your toolkit.
Software engineering background, or proven strong development ability in Go and/or Python - you write production code, not just glue scripts.
Security awareness baked into how you build: least-privilege IAM, secret hygiene, policy and supply-chain gates, not bolted on after.
Production-first mindset: you design for reliability, observability, and recoverability, and you own what you ship.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8818237
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're hiring a Senior/Principal Site Reliability Engineer to own production reliability for Cortex Agentix Endpoint Security (following an acquisition of KOI Start Up) as it scales. You'll define and operate our SLOs and error budgets, lead high-severity incident response, and ensure our Kubernetes and AWS infrastructure stays stable under growth. You'll also build and supervise the AI agents that handle routine alert triage and monitor tuning, focusing your own time on the reliability engineering that requires human judgment. This role is a strong fit for someone who treats reliability as an engineering discipline and enjoys ownership, incident command, and applying AI to operational work.
Your Impact:
Own reliability as an engineering discipline - define SLIs, set SLOs, and run error-budget-based decision-making so "how reliable are we" becomes a number that governs how fast we ship.
Own production incidents end-to-end - lead response, mitigation, and resolution for high-severity incidents, and drive blameless postmortems that feed real fixes back into the system.
Own the reliability and capacity of production infrastructure as we scale - forecasting headroom, validating scaling behavior under load, and keeping latency and error rates within SLO.
Run and evolve Kubernetes environments so releases and infra changes are safe by default across hundreds of tenant apps.
Own, build, and supervise our SRE AI agents that triage alerts, review monitors, resolves and summarize incidents. Set and expand the trust ladder that governs what the agents do autonomously, what needs approval, and what stays human. This is a core part of the role.
Improve observability and incident response - raise signal quality, cut alert noise, and own the monitoring the triage agents depend on.
Eliminate toil - relentlessly identify manual, repetitive operational work and remove it through automation and agents, protecting engineering time for reliability work that only humans can do.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, capacity risks, and automation opportunities.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Requirements:
Your Experience:
5+ years operating production cloud infrastructure, with a strong reliability focus (SRE, or DevOps/platform engineering with reliability ownership).
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Experience defining and operating SLIs, SLOs, and error budgets - or a clear grasp of the discipline and the drive to establish it from scratch.
Strong observability and alerting experience in Datadog or comparable platforms, including raising signal-to-noise in production.
Proven incident-response instincts - comfortable owning high-severity incidents and a genuine believer in blameless postmortems.
Proven ability to own platform and reliability projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work - and in building and supervising agents, not just using them.
Ownership-driven - You take responsibility for the reliability of the systems you build and operate, from SLO definition through incident command and continuous improvement.
Reliability as engineering - You treat reliability as a software problem to be solved with code, measurement, and automation - not an ops queue to be worked by hand.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781551
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
06/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're hiring a Senior/Principal Site Reliability Engineer to own production reliability for Cortex Agentix Endpoint Security (following an acquisition of KOI Start Up) as it scales. You'll define and operate our SLOs and error budgets, lead high-severity incident response, and ensure our Kubernetes and AWS infrastructure stays stable under growth. You'll also build and supervise the AI agents that handle routine alert triage and monitor tuning, focusing your own time on the reliability engineering that requires human judgment. This role is a strong fit for someone who treats reliability as an engineering discipline and enjoys ownership, incident command, and applying AI to operational work.
Your Impact:
Own reliability as an engineering discipline - define SLIs, set SLOs, and run error-budget-based decision-making so "how reliable are we" becomes a number that governs how fast we ship.
Own production incidents end-to-end - lead response, mitigation, and resolution for high-severity incidents, and drive blameless postmortems that feed real fixes back into the system.
Own the reliability and capacity of production infrastructure as we scale - forecasting headroom, validating scaling behavior under load, and keeping latency and error rates within SLO.
Run and evolve Kubernetes environments so releases and infra changes are safe by default across hundreds of tenant apps.
Own, build, and supervise our SRE AI agents that triage alerts, review monitors, resolves and summarize incidents. Set and expand the trust ladder that governs what the agents do autonomously, what needs approval, and what stays human. This is a core part of the role.
Improve observability and incident response - raise signal quality, cut alert noise, and own the monitoring the triage agents depend on.
Eliminate toil - relentlessly identify manual, repetitive operational work and remove it through automation and agents, protecting engineering time for reliability work that only humans can do.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, capacity risks, and automation opportunities.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Requirements:
Your Experience:
5+ years operating production cloud infrastructure, with a strong reliability focus (SRE, or DevOps/platform engineering with reliability ownership).
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Experience defining and operating SLIs, SLOs, and error budgets - or a clear grasp of the discipline and the drive to establish it from scratch.
Strong observability and alerting experience in Datadog or comparable platforms, including raising signal-to-noise in production.
Proven incident-response instincts - comfortable owning high-severity incidents and a genuine believer in blameless postmortems.
Proven ability to own platform and reliability projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work - and in building and supervising agents, not just using them.
Ownership-driven - You take responsibility for the reliability of the systems you build and operate, from SLO definition through incident command and continuous improvement.
Reliability as engineering - You treat reliability as a software problem to be solved with code, measurement, and automation - not an ops queue to be worked by hand.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771789
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at our company and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the company Technology team might be right for you.
We're looking for a Senior DevOps Engineer with a strong security orientation to join our company's DevOps team.
our company's platform runs at significant scale, and our DevOps team sits at the core of keeping it fast, reliable, and secure. As we grow, so does the scope of what we own - we need an engineer who can step in as a strong pillar of our company's production ecosystem.
This is a hands-on role for someone who takes security seriously, moves fast, and knows how to get things done in a complex, high-scale environment.
our company's Technology Stack:
AWS, GCP, CLoudFlare ,Kubernetes, Terragrunt, Ansible, Jenkins, ArgoCD, Argo Workflows, Kong & Nginx, HashiCorp Vault, Kafka, RabbitMQ, Mongodb, Aurora Postgresql & Mysql, Prometheus, Grafana, VictoriaMetrics
Programming languages: Python, NodeJS, Go, Kotlin
What am I going to do?
Full Ownership: Drive infrastructure initiatives through their entire lifecycle, taking accountability from initial design to delivery and long-term operations.
Kubernetes Orchestration: Architect, implement, and maintain production-grade Kubernetes clusters, ensuring they remain scalable and resilient under high-scale demand.
Cloud Architecture: Design and manage robust AWS environments, including VPC, IAM, and EKS, to support a highly available platform architecture.
Infrastructure as Code: Utilize Terraform to build and evolve our environment, applying configuration management principles to all IaC workflows.
CI/CD Excellence: Support and improve our deployment pipelines using Jenkins and GitHub Actions to maintain a fast development velocity.
Observability: Implement comprehensive monitoring solutions with Prometheus and Grafana to ensure deep visibility into platform health.
Operational Resilience: Join the DevOps on-call rotation, taking responsibility for mitigating production issues and maintaining site reliability.
Tooling & Innovation: Continuously evaluate and adopt tools - security and otherwise - that raise the bar on engineering efficiency and security posture at our company.
Requirements:
6+ years of hands-on DevOps / Platform Engineering experience in large-scale production environments on a public cloud (AWS preferred).
Proven leadership mindset - able to own projects end to end and be accountable for outcomes.
Strong, production-grade Kubernetes experience across design, deployment, scaling, and troubleshooting.
Solid AWS experience with VPC, IAM, EC2, EKS, Load Balancers, and DNS.
Experience designing and operating highly available, scalable infrastructure systems.
Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis).
Hands-on experience with Infrastructure as Code and configuration management (Terraform required; Terragrunt and Ansible a plus).
Experience with Docker and containerized workloads.
2+ years building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
Experience with monitoring and observability tools (Prometheus, Grafana).
Experience with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8806869
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Platform and infrastructure (K8s + IaC)
Architect and operate our Kubernetes platform on AWS - scaling, networking, cost, and reliability
Own infrastructure as code (Terraform) so environments are reproducible, auditable, and fast to evolve
Build the foundations that let the team provision and scale services without friction


AI agent platform - managing and scaling agents at volume
Operate and scale the orchestration layer (Temporal) that runs dozens of agents in parallel
Tune the platform for the unusual load profile of agent workloads - bursty, long-running, data-heavy, latency-sensitive
Give engineers the primitives to deploy, version, observe, and roll back agents safely under enterprise SLAs


CI/CD and developer experience
Own build and deploy pipelines end-to-end - fast, safe, boring releases
Invest in DX as a first-class product: the team treats developer experience as leverage, and you set the bar
Reduce the time from merged to in production and from idea to running experiment


Observability and reliability
Build monitoring, alerting, and tracing that make production legible - for services and for agents
Own incident response and the reliability practices that keep enterprise customers SLAs intact
Turn incidents into systemic fixes, not repeated firefighting


Security and compliance
Own secrets management, hardening, and the day-to-day security posture of the platform
Support our compliance commitments (we are Mastercard-certified, operating in fintech - the bar is high)
Build security into the pipeline so it is the default, not a gate
Requirements:
Strong DevOps/platform/SRE experience at a company with a real engineering culture (FAANG, unicorn, or a well-established startup with high standards)
Deep Kubernetes and AWS - you have architected, scaled, and debugged production clusters, not just deployed to them
Infrastructure as code in your bones - Terraform or equivalent, with strong opinions on reproducibility and auditability
CI/CD ownership - you have built pipelines that engineers trust and rarely think about
Production observability and incident response at meaningful scale
Comfort with and curiosity about AI/LLM workloads. We are an AI-native company; you should be using AI in how you work, and excited to operate the infrastructure agents run on
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8814945
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
27/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the Technology team might be right for you. We're looking for a Senior DevOps Engineer with a strong security orientation to join our DevOps team. our platform runs at significant scale, and our DevOps team sits at the core of keeping it fast, reliable, and secure. As we grow, so does the scope of what we own - we need an engineer who can step in as a strong pillar of our production ecosystem. This is a hands-on role for someone who takes security seriously, moves fast, and knows how to get things done in a complex, high-scale environment. Our Technology Stack: AWS, GCP, CLoudFlare,Kubernetes, Terragrunt, Ansible, Jenkins, ArgoCD, Argo Workflows, Kong & Nginx, HashiCorp Vault, Kafka, RabbitMQ, Mongodb, Aurora Postgresql & Mysql, Prometheus, Grafana, VictoriaMetrics Programming languages: Python, NodeJS, Go, Kotlin


What am I going to do?:

* Full Ownership: Drive infrastructure initiatives through their entire lifecycle, taking accountability from initial design to delivery and long-term operations.
* Kubernetes Orchestration: Architect, implement, and maintain production-grade Kubernetes clusters, ensuring they remain scalable and resilient under high-scale demand.
* Cloud Architecture: Design and manage robust AWS environments, including VPC, IAM, and EKS, to support a highly available platform architecture.
* Infrastructure as Code: Utilize Terraform to build and evolve our environment, applying configuration management principles to all IaC workflows.
* CI/CD Excellence: Support and improve our deployment pipelines using Jenkins and GitHub Actions to maintain a fast development velocity.
* Observability: Implement comprehensive monitoring solutions with Prometheus and Grafana to ensure deep visibility into platform health.
* Operational Resilience: Join the DevOps on-call rotation, taking responsibility for mitigating production issues and maintaining site reliability.
* Tooling & Innovation: Continuously evaluate and adopt tools - security and otherwise - that raise the bar on engineering efficiency and security posture.
Equal opportunities:
We're not about checklists. If you don't meet 100% of the requirements for this role but still feel passionate about the position and think you have the right skills and qualifications to excel at it, we want to hear from you. We prioritize diversity. We celebrate difference and embed it into every aspect of our workplace and product, as well as our community. We are proud and committed to providing equal opportunity employment to all individuals regardless of race, color, religion, sex, sexual orientation, citizenship, national origin, disability, Veteran status, or any other characteristic protected by law. In addition, we will provide accommodation to individuals with disabilities or a special need.
Requirements:
* 6+ years of hands-on DevOps / Platform Engineering experience in large-scale production environments on a public cloud (AWS preferred).
* Proven leadership mindset - able to own projects end to end and be accountable for outcomes.
* Strong, production-grade Kubernetes experience across design, deployment, scaling, and troubleshooting.
* Solid AWS experience with VPC, IAM, EC2, EKS, Load Balancers, and DNS.
* Experience designing and operating highly available, scalable infrastructure systems.
* Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis).
* Hands-on
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8799862
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
our company harnesses the power of AI to transform how revenue teams win. The company Revenue AI Operating System unifies data, insights, and workflows into a single, trusted system that observes, guides, and acts alongside the worlds most successful revenue teams. Powered by the company Revenue Graph, AI-powered intelligence, specialized agents, and trusted applications, our company helps more than 5,000 companies around the world deeply understand their teams and customers, automate critical sales workflows, and close more deals with less effort.
At our company, you will join a company built on innovative products, ambitious goals, and passionate people. We are shaping the future of revenue intelligence and we want people who are excited to build what comes next. You will work with a team that dreams big, moves fast, and cares deeply about the craft and about each other. Here, transparency and trust are core to how we operate, and every person has the opportunity to make a visible impact. If you want to grow, stretch, and do work that truly matters, we are the place to do the best work of your career.
At our company, were on a mission to help companies unlock reality and reach their full potential. As a Senior Site Reliability Engineer (SRE), youll play a key role in shaping our new Production Reliability domain. Youll drive reliability initiatives, lead cross-team projects, and ensure our SaaS platform stays robust, scalable, and efficient. This is a high-impact, hands-on role that demands technical expertise and a proactive approach.
You'll Own
Design, build, and maintain scalable, fault-tolerant systems.
Define and enforce reliability processes, SLOs, SLIs, and SLAs.
Lead complex incident responses, including on-call rotations and postmortems.
You'll Solve
Challenges related to observability, testing, production stability, and development productivity.
Reliability improvements through data-driven decisions.
Complex production incidents.
You'll Impact
Build automation, tooling, and self-service capabilities.
Collaborate with engineering, product, and support teams to embed reliability into everything we do.
Mentor engineers and promote operational excellence across the organization.
Requirements:
You have 7+ years of experience in SRE, DevOps, or Production Engineering roles, ideally in SaaS environments.
You have a deep understanding of distributed systems, failure modes, resiliency patterns, observability, and operating large-scale production services running on Kubernetes.
You are hands-on with building and owning monitoring tools.
You are experienced with CI/CD tools.
You are proficient with infrastructure-as-code tools.
You have solid experience with cloud platforms (AWS preferred).
Advantage: Experience with Java.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8788349
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8779592
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Senior Developer (DevOps and Infrastructure)
We're revolutionizing how the world moves money with our unified global payments platform. Our team is at the forefront of ensuring the security and resilience of this critical infrastructure, and we're looking for a strong DevOps engineer to play a pivotal role in designing and implementing robust solutions across our infrastructure and applications.
If you have a passion for building secure-by-design systems, a deep understanding of both infrastructure and application principles, and the ability to translate complex requirements into actionable blueprints, we want to hear from you. You'll be instrumental in shaping our infra and security landscape, ensuring our platform remains a trusted and secure environment for our global users.
What You'll Be Doing:
You will work with Python and APIs, codifying infrastructure, and lead the architectural design and implementation of solutions for our cloud infrastructure, network, and applications.
Constantly push optimizations and best practices. Define and maintain security standards, frameworks, and best practices across the organization.
Collaborate closely with engineering, product, and operations teams to integrate security seamlessly into the development lifecycle and infrastructure deployments.
Evaluate and recommend security technologies and tools to enhance our security posture.
Develop security reference architectures and patterns to guide engineering teams in building secure solutions.
Participate in threat modeling and risk assessment activities to proactively identify and mitigate potential security threats.
Provide expert guidance and mentorship to engineering teams on security-related topics.
Stay current with the latest security trends, threats, and technologies, and translate them into actionable strategies for us.
Requirements:
8+ years of experience in DevOps / Software, with a strong focus on security architecture for both infrastructure and applications.
5+ years of experience in designing, implementing and leading infra lifecycles in cloud environments (e.g., AWS, GCP, Azure).
Solid coding skills: terraform/ python a must: Ability to identify IAC areas that should become a module and taking this module all the way to production
A strong and advanced expert in terraform - A MUST
Expertise in low-level networking - A MUST
Excellent communication and collaboration skills, with the ability to articulate complex technical concepts to technical and non-technical audiences from different cultures, across the globe
A proactive and strategic mindset with a passion for building secure and scalable systems.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8787145
סגור
שירות זה פתוח ללקוחות VIP בלבד