דרושים » ניהול ביניים » Senior SRE Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 13 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
our company harnesses the power of AI to transform how revenue teams win. The company Revenue AI Operating System unifies data, insights, and workflows into a single, trusted system that observes, guides, and acts alongside the worlds most successful revenue teams. Powered by the company Revenue Graph, AI-powered intelligence, specialized agents, and trusted applications, our company helps more than 5,000 companies around the world deeply understand their teams and customers, automate critical sales workflows, and close more deals with less effort.
At our company, you will join a company built on innovative products, ambitious goals, and passionate people. We are shaping the future of revenue intelligence and we want people who are excited to build what comes next. You will work with a team that dreams big, moves fast, and cares deeply about the craft and about each other. Here, transparency and trust are core to how we operate, and every person has the opportunity to make a visible impact. If you want to grow, stretch, and do work that truly matters, we are the place to do the best work of your career.
At our company, were on a mission to help companies unlock reality and reach their full potential. As a Senior Site Reliability Engineer (SRE), youll play a key role in shaping our new Production Reliability domain. Youll drive reliability initiatives, lead cross-team projects, and ensure our SaaS platform stays robust, scalable, and efficient. This is a high-impact, hands-on role that demands technical expertise and a proactive approach.
You'll Own
Design, build, and maintain scalable, fault-tolerant systems.
Define and enforce reliability processes, SLOs, SLIs, and SLAs.
Lead complex incident responses, including on-call rotations and postmortems.
You'll Solve
Challenges related to observability, testing, production stability, and development productivity.
Reliability improvements through data-driven decisions.
Complex production incidents.
You'll Impact
Build automation, tooling, and self-service capabilities.
Collaborate with engineering, product, and support teams to embed reliability into everything we do.
Mentor engineers and promote operational excellence across the organization.
Requirements:
You have 7+ years of experience in SRE, DevOps, or Production Engineering roles, ideally in SaaS environments.
You have a deep understanding of distributed systems, failure modes, resiliency patterns, observability, and operating large-scale production services running on Kubernetes.
You are hands-on with building and owning monitoring tools.
You are experienced with CI/CD tools.
You are proficient with infrastructure-as-code tools.
You have solid experience with cloud platforms (AWS preferred).
Advantage: Experience with Java.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8788349
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a Senior Infrastructure Engineer who views "Infrastructure as Software." In 2026, we dont just manage servers; we build high-performance environments that allow multi-agent systems to operate at scale.
You will be a core member of the R&D team, blending deep DevOps expertise with the coding rigor of a Backend Engineer. You arent just "configuring" AWS; you are architecting the distributed systems and data pipelines that power our autonomous security brain. Your mission is to ensure that while our agents are evolving and taking actions, our underlying platform remains immutable, observable, and infinitely scalable.
What You'll Do
Design, build, and operate our company's cloud infrastructure using AWS, Kubernetes, and Infrastructure as Code.
Build internal tools and platform services using Python and Go to improve developer productivity and system reliability.
Own infrastructure automation with Terraform, Pulumi, and modern cloud-native tooling.
Partner closely with Backend, Data Science, and Security Engineering teams to build scalable, reliable platforms.
Improve observability, monitoring, and incident response across distributed production systems.
Design and optimize infrastructure for performance, scalability, security, and cost efficiency.
Help shape engineering best practices, platform architecture, and developer experience as our company continues to grow.
Requirements:
5+ years of experience in Infrastructure, DevOps, Platform Engineering, or Backend Engineering.
Strong software engineering skills with hands-on experience building production systems in Python or Go.
Deep hands-on experience with AWS, including services such as EKS, RDS, VPC, and IAM.
Strong experience designing, operating, and scaling production Kubernetes environments.
Experience with Infrastructure as Code, CI/CD, GitOps, and modern cloud-native development practices.
A systems mindset with the ability to solve architectural challenges across infrastructure and application layers.
Comfortable using modern AI-powered developer tools and agentic workflows to improve engineering productivity.
The company Mindset: You take ownership, act with accountability, collaborate openly, and focus on delivering meaningful impact. You thrive in fast-moving environments, embrace ambiguity, and enjoy solving hard problems together.
Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
Full professional fluency (written and verbal) in both Hebrew and English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're hiring a Senior/Principal Site Reliability Engineer to own production reliability for Cortex Agentix Endpoint Security (following an acquisition of KOI Start Up) as it scales. You'll define and operate our SLOs and error budgets, lead high-severity incident response, and ensure our Kubernetes and AWS infrastructure stays stable under growth. You'll also build and supervise the AI agents that handle routine alert triage and monitor tuning, focusing your own time on the reliability engineering that requires human judgment. This role is a strong fit for someone who treats reliability as an engineering discipline and enjoys ownership, incident command, and applying AI to operational work.
Your Impact:
Own reliability as an engineering discipline - define SLIs, set SLOs, and run error-budget-based decision-making so "how reliable are we" becomes a number that governs how fast we ship.
Own production incidents end-to-end - lead response, mitigation, and resolution for high-severity incidents, and drive blameless postmortems that feed real fixes back into the system.
Own the reliability and capacity of production infrastructure as we scale - forecasting headroom, validating scaling behavior under load, and keeping latency and error rates within SLO.
Run and evolve Kubernetes environments so releases and infra changes are safe by default across hundreds of tenant apps.
Own, build, and supervise our SRE AI agents that triage alerts, review monitors, resolves and summarize incidents. Set and expand the trust ladder that governs what the agents do autonomously, what needs approval, and what stays human. This is a core part of the role.
Improve observability and incident response - raise signal quality, cut alert noise, and own the monitoring the triage agents depend on.
Eliminate toil - relentlessly identify manual, repetitive operational work and remove it through automation and agents, protecting engineering time for reliability work that only humans can do.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, capacity risks, and automation opportunities.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Requirements:
Your Experience:
5+ years operating production cloud infrastructure, with a strong reliability focus (SRE, or DevOps/platform engineering with reliability ownership).
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Experience defining and operating SLIs, SLOs, and error budgets - or a clear grasp of the discipline and the drive to establish it from scratch.
Strong observability and alerting experience in Datadog or comparable platforms, including raising signal-to-noise in production.
Proven incident-response instincts - comfortable owning high-severity incidents and a genuine believer in blameless postmortems.
Proven ability to own platform and reliability projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work - and in building and supervising agents, not just using them.
Ownership-driven - You take responsibility for the reliability of the systems you build and operate, from SLO definition through incident command and continuous improvement.
Reliability as engineering - You treat reliability as a software problem to be solved with code, measurement, and automation - not an ops queue to be worked by hand.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781551
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
06/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're hiring a Senior/Principal Site Reliability Engineer to own production reliability for Cortex Agentix Endpoint Security (following an acquisition of KOI Start Up) as it scales. You'll define and operate our SLOs and error budgets, lead high-severity incident response, and ensure our Kubernetes and AWS infrastructure stays stable under growth. You'll also build and supervise the AI agents that handle routine alert triage and monitor tuning, focusing your own time on the reliability engineering that requires human judgment. This role is a strong fit for someone who treats reliability as an engineering discipline and enjoys ownership, incident command, and applying AI to operational work.
Your Impact:
Own reliability as an engineering discipline - define SLIs, set SLOs, and run error-budget-based decision-making so "how reliable are we" becomes a number that governs how fast we ship.
Own production incidents end-to-end - lead response, mitigation, and resolution for high-severity incidents, and drive blameless postmortems that feed real fixes back into the system.
Own the reliability and capacity of production infrastructure as we scale - forecasting headroom, validating scaling behavior under load, and keeping latency and error rates within SLO.
Run and evolve Kubernetes environments so releases and infra changes are safe by default across hundreds of tenant apps.
Own, build, and supervise our SRE AI agents that triage alerts, review monitors, resolves and summarize incidents. Set and expand the trust ladder that governs what the agents do autonomously, what needs approval, and what stays human. This is a core part of the role.
Improve observability and incident response - raise signal quality, cut alert noise, and own the monitoring the triage agents depend on.
Eliminate toil - relentlessly identify manual, repetitive operational work and remove it through automation and agents, protecting engineering time for reliability work that only humans can do.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, capacity risks, and automation opportunities.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Requirements:
Your Experience:
5+ years operating production cloud infrastructure, with a strong reliability focus (SRE, or DevOps/platform engineering with reliability ownership).
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Experience defining and operating SLIs, SLOs, and error budgets - or a clear grasp of the discipline and the drive to establish it from scratch.
Strong observability and alerting experience in Datadog or comparable platforms, including raising signal-to-noise in production.
Proven incident-response instincts - comfortable owning high-severity incidents and a genuine believer in blameless postmortems.
Proven ability to own platform and reliability projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work - and in building and supervising agents, not just using them.
Ownership-driven - You take responsibility for the reliability of the systems you build and operate, from SLO definition through incident command and continuous improvement.
Reliability as engineering - You treat reliability as a software problem to be solved with code, measurement, and automation - not an ops queue to be worked by hand.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771789
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Senior AI Engineer to design and build production-grade, LLM-powered systems. You'll work at the intersection of software engineering and applied AI - shipping agents, RAG pipelines, and tool-using systems that solve real problems at scale. This is a hands-on, high-ownership role for someone who thrives at the frontier of what's possible with modern LLMs and isn't afraid to write the glue, the infrastructure, and the prompts that make it all work.
This is a **cross-functional, company-wide role**. You won't be embedded in a single product team - instead, you'll partner with every department to identify high-leverage opportunities and build AI-powered tools and workflows that boost productivity and efficiency across the entire organization.
This is a great opportunity to be part of one of the fastest-growing infrastructure companies in history, an organization that is in the center of the hurricane being created by the revolution in artificial intelligence.
"our company's data management vision is the future of the market."- Forbes
we are the data platform company for the AI era. We are building the enterprise software infrastructure to capture, catalog, refine, enrich, and protect massive datasets and make them available for real-time data analysis and AI training and inference. Designed from the ground up to make AI simple to deploy and manage, our company takes the cost and complexity out of deploying enterprise and AI infrastructure across data center, edge, and cloud.
Our success has been built through intense innovation, a customer-first mentality and a team of fearless workers who leverage their skills & experiences to make real market impact. This is an opportunity to be a key contributor at a pivotal time in our companys growth and at a pivotal point in computing history.
What You'll Do:
- Design, build, and operate LLM-powered applications, agents, and workflows end-to-end - from prototype to production.
- Architect retrieval, context engineering, and tool-use strategies that make models reliable, accurate, and cost-efficient.
- Integrate LLMs with internal services, third-party APIs, and data stores to automate complex business and engineering workflows.
- Build, evaluate, and continuously improve evaluation harnesses for non-deterministic systems.
- Collaborate closely with product, research, and platform teams to translate ambiguous problems into shipped capabilities.
- Stay ahead of the rapidly evolving LLM ecosystem (models, frameworks, agentic patterns) and bring the best ideas into our stack.
Requirements:
Engineering Foundations:
- Strong Python skills- you write clean, idiomatic, well-tested code and understand the language deeply.
- Hands-on experience using coding agents(Cursor, Claude Code, GitHub Copilot, or similar) to build complex software systems. You know how to delegate effectively to AI assistants and review their output critically.
- Experience with multiple database paradigms- both SQL (PostgreSQL, MySQL) and NoSQL (MongoDB, Redis, DynamoDB, or similar). You can choose the right tool for the job.
- Experience designing and integrating with third-party APIs- REST and gRPC. Comfortable building robust clients, handling auth, retries, rate limits, and schema evolution.
- Production experience with Docker and Kubernetes- containerizing services, writing manifests, and debugging deployments.
- Strong Linux fundamentals- confident in bash and the terminal; you can navigate, script, and troubleshoot a server without reaching for a GUI.
- Experience building cloud-native tools on AWS, GCP, or Azure (compute, storage, queues, serverless, IAM).
AI / LLM Expertise:
- Solid understanding of what an LLM is and how it works- tokenization, attention, context windows, sampling, and the practical implications of each for system design.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8744445
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking a skilled and motivated DevOps engineer with deep familiarity in the streaming ecosystem to join our elite infrastructure team. If you're excited by the challenge of operating mission-critical systems at scale and optimizing the developer experience through automation and tooling, wed love to hear from you.
What you will do:
Automate Deployment and Operation:
Oversee deployment of Kafka and RabbitMQ clusters (including Confluent Cloud & CFK). Build automation pipelines to ensure repeatability and resiliency across environments.
Monitor and Support Production Systems:
Own production stability of global Kafka clusters. Handle on-call rotations, incident management, troubleshooting, and scaling challenges.
Improve Infrastructure Observability
Build and maintain observability systems: dashboards, alerting pipelines, metrics collection (Prometheus, Grafana, etc.).
Optimize System Performance:
Collaborate with peers on benchmarking and optimization initiatives. Work on tuning Kafka brokers, cluster configurations, and runtime parameters.
Provide Developer Support and Training (Infra-focused)
Help developers configure topics, quotas, and consumers appropriately. Train service owners to interpret monitoring data and avoid pitfalls.
Develop and Maintain Infrastructure:
Contribute to building infrastructure tools and scripts (IaC, Helm charts, etc.) that make provisioning and managing clusters reliable and efficient.
Secure Infrastructure Access:
Configure and maintain secure access patterns across streaming infrastructure, ensuring proper authentication and role-based access controls are enforced for both developers and services.
Requirements:
8+ years of experience in DevOps, SRE, or Infrastructure Engineering roles.
Deep hands-on Kafka experience, including deploying, maintaining, scaling, and monitoring clusters.
Experience with RabbitMQ.
Extensive experience with Docker, Kubernetes, Helm, and GitOps-style deployments.
Infrastructure as Code experience (Terraform, Pulumi, etc.).
Strong skills in scripting and automation (Python, Bash, etc.).
Familiarity with Confluent Cloud, Confluent for Kubernetes, and similar tools.
Solid understanding of authentication and authorization mechanisms in distributed systems.
Production support mindset - with proven troubleshooting and incident resolution history.
Collaboration and communication skills - especially with dev teams depending on platform support.
Experience with Istio Service Mesh (bonus).
Experience with GovCloud (bonus).
Bonus Qualities:
Mentorship and leadership experience in infrastructure or SRE teams.
Contributions to automation or monitoring open-source tooling.
Active participant in SRE or DevOps communities.
Conference speaker or internal tech trainer.
Technical writing about infrastructure automation or reliability.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8754250
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Are you a results-driven backend or full-stack engineer passionate about building scalable, cloud-native microservices? Do you thrive in an agile, startup-like environment while having the backing of a market leader? At our company, you wont just be a coder; you'll be an architect of innovation, shaping the future of identity security.
Our engineering team is the core of our success. We build high-quality, professional-grade products on a mature, event-driven microservices architecture hosted in AWS. We're now seeking a Senior Backend Software Engineer to be a foundational member of a new team building a cutting-edge, cloud-based SaaS identity analytics product from the ground up.
Your Mission: The First 12 Months

First 3 Months: You will be fully integrated into our agile team, actively contributing to the design and development of core microservices for our new SaaS product. You will have a comprehensive understanding of the product architecture and roadmap, and you'll be writing high-quality, test-covered code in Go.

First 6 Months: You will take ownership of significant features, from design and estimation to implementation and deployment. You'll be expected to contribute to our continuous delivery pipeline and improve code quality by producing unit and end-to-end tests, aiming to increase code coverage by 15-20%.

First 12 Months: You will be a key contributor to the product's evolution, mentoring junior engineers and collaborating with product management to define and implement new features. You will have made a measurable impact on the product's performance and scalability, and you'll be a go-to expert for critical components of the system.

What You'll Do

Design, develop, and deploy backend microservices in Go. Success will be measured by the delivery of well-tested, scalable, and maintainable code that meets product requirements and is deployed to production in a timely manner.

Integrate AI-assisted tooling into day-to-day DevOps and engineering workflows. Success will be measured by improvements in productivity, scalability, and operational efficiency. You will achieve this by leveraging AI tools to generate initial configuration drafts, validate infrastructure code, and recommend workflow improvements, while utilizing AI-driven automation to reduce repetitive manual tasks and accelerate engineering execution without sacrificing high-quality standards.

Collaborate on designs, code reviews, and testing. Success will be measured by your active participation in team meetings, providing constructive feedback on peer code reviews, and contributing to a collaborative and positive team environment.

Produce unit and end-to-end tests. Success will be measured by your consistent contributions to our test suites, with a focus on increasing code coverage and improving the overall quality and reliability of the product.

Create and refine design and engineering best practices. Success will be measured by tangible improvements to our development processes and documentation, leading to increased team efficiency and code quality.
Requirements:
5+ years of professional software development experience.
3+ years of hands-on experience with Go.
A Bachelor of Science in Computer Science, a related field, or equivalent practical knowledge.
Advanced experience with object-oriented analysis and design.
Strong communication skills, with the ability to articulate complex technical concepts to both technical and non-technical audiences.

Preferred Qualifications:
Proven experience with AWS.
A deep understanding of Continuous Delivery principles and practices.
Prior experience working on Big Data or Machine Learning products.
Familiarity with instrumenting code for production performance metrics.
While this is a backend-focused role, a solid understanding of modern JavaScript frameworks (React, Angular, Backbone) and ES6+ is a plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764428
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Senior Developer (DevOps and Infrastructure)
We're revolutionizing how the world moves money with our unified global payments platform. Our team is at the forefront of ensuring the security and resilience of this critical infrastructure, and we're looking for a strong DevOps engineer to play a pivotal role in designing and implementing robust solutions across our infrastructure and applications.
If you have a passion for building secure-by-design systems, a deep understanding of both infrastructure and application principles, and the ability to translate complex requirements into actionable blueprints, we want to hear from you. You'll be instrumental in shaping our infra and security landscape, ensuring our platform remains a trusted and secure environment for our global users.
What You'll Be Doing:
You will work with Python and APIs, codifying infrastructure, and lead the architectural design and implementation of solutions for our cloud infrastructure, network, and applications.
Constantly push optimizations and best practices. Define and maintain security standards, frameworks, and best practices across the organization.
Collaborate closely with engineering, product, and operations teams to integrate security seamlessly into the development lifecycle and infrastructure deployments.
Evaluate and recommend security technologies and tools to enhance our security posture.
Develop security reference architectures and patterns to guide engineering teams in building secure solutions.
Participate in threat modeling and risk assessment activities to proactively identify and mitigate potential security threats.
Provide expert guidance and mentorship to engineering teams on security-related topics.
Stay current with the latest security trends, threats, and technologies, and translate them into actionable strategies for us.
Requirements:
8+ years of experience in DevOps / Software, with a strong focus on security architecture for both infrastructure and applications.
5+ years of experience in designing, implementing and leading infra lifecycles in cloud environments (e.g., AWS, GCP, Azure).
Solid coding skills: terraform/ python a must: Ability to identify IAC areas that should become a module and taking this module all the way to production
A strong and advanced expert in terraform - A MUST
Expertise in low-level networking - A MUST
Excellent communication and collaboration skills, with the ability to articulate complex technical concepts to technical and non-technical audiences from different cultures, across the globe
A proactive and strategic mindset with a passion for building secure and scalable systems.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8787145
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
2 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As the worlds leading vendor of Cyber Security, facing the most sophisticated threats and attacks, weve assembled a global team of the most driven, creative, and innovative people. At our company, our employees are redefining the security landscape by meeting our customers real-time needs and providing our cutting-edge technologies and services to an ever-growing customer base.
our company Software Technologies has been honored by Time Magazine as one of the Worlds Best Companies and recently Gartner rated our company email security as a market leader for product, detection and innovation. We've also earned a spot on the Forbes list of the Worlds Best Places to Work for five consecutive years (2020-2024) and recognized as one of the Worlds Top Female-Friendly Companies. If you're passionate about making the world a safer place and want to be part of an award-winning company culture, we invite you to join us.
our company Harmony Email Security and Collaboration (Previously AVANAN) is a unique email solution that fully secures cloud email and cloud platforms using AI.
we are seeking a promising and talented Senior DevFinOps Engineer to join our DevOps group. If you thrive in a fast-paced, dynamic environment, can handle multiple requests simultaneously, and enjoy working independently as part of a cutting-edge DevOps team, this is your opportunity to help make the world a safer place!
Key Responsibilities
Act as a DevFinOps Engineer within a highly skilled team, bridging engineering and finance to drive cloud cost efficiency across large-scale operations from development to production.
Design, develop, and maintain Avanan's cloud cost visibility, allocation, and optimization solutions - including tagging strategies, cost dashboards, budgets, and anomaly detection across accounts and services.
Implement tools and procedures for cost monitoring, forecasting, and alerting across our SaaS multi-tenant product family.
Embed FinOps practices into the CI/CD lifecycle - surfacing the cost impact of changes early, and enforcing cost guardrails as part of deployment automation.
build AI-based FinOps agents for cloud services at the infrastructure and application levels
Continuously identify and execute cost-optimization opportunities (right-sizing, reserved capacity/savings plans, spot usage, storage tiering, idle-resource cleanup) without compromising performance, reliability, or security.
דרישות:
Hands-on mindset - we all write code daily!
3+ years of relevant DevOps/Cloud experience building and operating CI/CD pipelines for both development and production - must.
2+ years of AWS Cloud experience working with high-traffic systems and multiple services, with a strong grasp of AWS pricing models and cost management tooling (Cost Explorer, CUR, Budgets, Compute Optimizer) - must.
Strong scripting skills, with fluency in Python - must.
Experience working with AI tools to achieve cloud or application cost control, identify and investigate the root cause of cost changes (differentiating between organic growth / infrastructure change / config change / application change).
Demonstrated experience driving measurable cloud cost reductions and building cost-optimization tooling or automation - must.
Experience with containers and orchestration tools (Docker, Kubernetes, or ECS) and understanding of their cost drivers - must.
Familiarity with FinOps principles and practices (FinOps Foundation framework, showback/chargeback, unit economics) - an advantage.
Experience with CI integration tools such as Jenkins.
Familiarity with AWS CloudFormation and infrastructure-as-code - an advantage.
Exposure to a wide range of open-source technologies (Redis, Nagios, Grafana, Prometheus, etc.) and cost-analytics tooling.
Knowledge of best practices in security, performance, monitoring, and cost governance.
Proven ability to research, evaluate, and implement new technologies, including running proofs of concept and cost analysis.#ENGL המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785679
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
03/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a Senior Data Engineer to help build next-generation data platform - the lakehouse foundation that will power data processing across the entire product. This is not a "write pipelines on top of someone else's platform" role, and it's not a pure infrastructure role either. It's both, deliberately.

You'll own the platform end to end: the infrastructure it runs on (Spark on Kubernetes, Apache Iceberg, AWS Glue, Airflow), the frameworks and tooling that let dozens of other engineers build on it without reinventing the wheel, and the design of the data pipelines themselves. Everything you build becomes leverage for the teams around you - your abstractions, base images, CI/CD flows, and operational patterns are what make the platform usable at scale.

You'll also own one of the hardest ongoing trade-offs in a high-scale data platform: balancing cost and performance. Compute sizing, storage layout, partitioning and compaction strategy, job scheduling - every decision has a price tag and a latency profile, and you'll be the one making those calls with data.

This role is ideal for an engineer who is equally comfortable debugging a Spark executor OOM on Kubernetes at 10am, designing a clean Python framework API at noon, and modeling the cost impact of a table layout change in the afternoon.



What You'll Do

Platform & Infrastructure

- Design, deploy, and operate our Spark-on-Kubernetes compute platform, including autoscaling, resource tuning, and multi-tenancy considerations.

- Own the lakehouse storage layer built on Apache Iceberg and AWS Glue catalog - table design, partitioning, compaction, schema evolution, and retention.

- Build and operate orchestration on Airflow: DAG standards, deployment flows, environment promotion, and reliability.

- Own production operations of the platform: monitoring, alerting, incident response, and continuous hardening.

Frameworks & Developer Enablement

- Build the code frameworks, libraries, and templates that other engineers use to write pipelines - so that spinning up a new production-grade Spark job is measured in hours, not weeks.

- Define and enforce standards for pipeline structure, testing, observability, and deployment across teams.

- Own CI/CD for data workloads: image builds, artifact promotion, and GitOps-based delivery.

- Act as a technical partner to product and research teams building on the platform - your customers are other engineers.

Data Pipelines & Architecture

- Design and build scalable batch and streaming pipelines processing complex, high-volume datasets from diverse sources.

- Lead large-scale backfills and migration initiatives, ensuring data consistency and integrity across evolving storage and compute platforms.

- Design event-driven data flows over large-scale queue systems (Kafka) for reliable, efficient data movement.

Cost & Performance

- Continuously balance cost against performance: right-size compute, tune queries and jobs, optimize storage layout and file sizes, and choose the correct engine for each workload.

- Build cost visibility and attribution into the platform so trade-offs are made with data, not guesswork.
דרישות:
- 5+ years of experience in software engineering, with meaningful time spent building and operating large-scale data platforms.

- Strong hands-on experience with distributed processing engines (Spark strongly preferred), including performance tuning and debugging in production.

- Practical experience deploying and operating workloads in Kubernetes-based environments - you're not afraid of infra work; you enjoy it.

- Experience building shared frameworks, libraries, or internal tooling used by other engineers, with the product mindset that comes with it (clean APIs, docs, versioning, backward compatibility).

- Strong proficiency in SQL and data modeling: complex analytical queries, query tuning, partitioning strategies.

- Solid software engineerin המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8766023
סגור
שירות זה פתוח ללקוחות VIP בלבד