דרושים » תוכנה » QA Performance / COST team

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 9 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As QA Engineer of COST team you will focus on cross team integration testing performed by imitating real customer environments and scenarios.

your main responsibilities will be:

Maintaining and extending large environments configured with real-life data loaded by automated and manual activity. Preform stability and integration tests on complicated customer environments.
Requirements:
3+ year experience in E2E complex distributed client-server testing
Experience writing and/or executing complex scripts: Shell script, Batch files, Python or similar language
Experience in writing automation tests
Strong Linux administration experience
Deep knowledge and experienced with different firewalls/routers configuration (Palo Alto/Check Point/Cisco/Juniper/Fortinet)
Must have strong troubleshooting skills and ability to perform root cause analysis with solving skills
Excellent communication skills, both written and verbal
Ability to interact successfully with multiple teams across the global organization
Quick learner with a desire to learn new tools and techniques
Experience with applications using SQL, must be able to query databases.
Knowledge in Docker and Kubernetes
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8839497
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo and Ra'anana
Job Type: Full Time
As a Senior DevOps Engineer, youll help turn agentic AI capabilities for diagnosing and troubleshooting network and GPU infrastructure into secure, scalable, production-ready services. This role stands out through its end-to-end ownership across cloud and customer-managed environments, close partnership with software and AI engineers, and direct influence on the reliability of NVIDIAs AI infrastructure.


What You'll Be Doing:
Own the DevOps, infrastructure, security, release, and reliability lifecycle - from development environments and CI/CD through deployment, production readiness, and sustained operations.
Build and operate Kubernetes environments and Helm-based deployments for a Python, FastAPI, Node.js, and React microservices platform across SaaS and on-premises footprints.
Engineer GitLab CI/CD pipelines with automated testing, container builds, vulnerability scanning, and versioned image and Helm chart publication through JFrog Artifactory.
Automate infrastructure provisioning, configuration, upgrades, and routine operational workflows to accelerate delivery and improve engineering productivity.
Operate PostgreSQL, Temporal workflow services, and S3-compatible object storage with disciplined capacity planning, backups, recovery testing, and safe migrations.
Strengthen release reliability through deployment validation, reduced-downtime strategies, persistent-state protection, and recovery plans for active workflows.
Deliver actionable observability and security using OpenTelemetry, Datadog/Grafana, Langfuse, secrets management, identity integration, TLS, Kubernetes RBAC, network policies, and container hardening.
Partner with software and AI engineers to troubleshoot distributed systems, investigate incidents, define reliability targets, and improve platform performance, resource efficiency, and customer outcomes.
Requirements:
What We Need to See:
Bachelors degree in Computer Science, Software Engineering, or a related field, or equivalent experience.
5+ years of experience in DevOps, site reliability engineering, or platform engineering supporting distributed applications and microservices.
Strong hands-on experience with Kubernetes, Docker, and Helm, including networking, storage, workload scheduling, scaling, and troubleshooting.
Strong Linux administration skills and proficiency in Python and Bash for automation, plus experience with infrastructure as code and configuration tooling such as Terraform and Ansible.
Experience building and maintaining CI/CD pipelines, including runners, container registries, artifact management, automated quality gates, and secure release practices.
Practical experience operating PostgreSQL or comparable relational databases, including SQL, migrations, backup and restore, and performance troubleshooting.
Strong networking and observability fundamentals across TCP/IP, DNS, HTTP, TLS, load balancing, ingress, metrics, logs, traces, dashboards, and actionable alerting.
Sound understanding of secure infrastructure operations and incident response, with demonstrated ownership, cross-functional collaboration, and prioritization in an evolving environment.


Ways To Stand Out From the Crowd:
Experience operating AI applications, agent platforms, or LLM services, including monitoring latency, failures, token usage, and cost.
Familiarity with Temporal, LangGraph, Model Context Protocol (MCP), Langfuse, ClickHouse, Redis/Valkey, or S3-compatible storage.
Deep experience with OpenTelemetry instrumentation and collectors, Datadog APM, or Prometheus/Grafana.
Experience with self-hosted Kubernetes, OpenShift, Kubernetes operators, CloudNativePG, or GPU clusters and AI data centers.
Experience building reproducible AMD64 and ARM64 container images, optimizing BuildKit pipelines, and securing the software supply chain.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8837920
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have:
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796409
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
26/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are constantly striving to make our systems reliable, scalable, and simple to operate so our services are available to travelers when they need them most. With our continued growth, we have exciting challenges ahead and we're looking for a Senior Site Reliability Engineer to join our team in Tel Aviv. This role blends classic SRE ownership with pragmatic AI SRE work: you will build and operate the platforms, automation, observability, and incident response practices that keep Navan reliable, while helping teams use AI solutions, AI providers, and their APIs safely and dependably.



This is a hands-on engineering role, not a research role. You will partner with product, platform, data, security, support, and incident response teams to make production systems and AI-powered experiences more resilient. You will use software engineering, infrastructure as code, SLOs, telemetry, provider observability, and automation as your main tools, and you will apply AI where it creates measurable reliability value rather than novelty.



This position is based out of our new Tel Aviv office.



What You'll Do:

Support AI-based application solutions where reliability matters. Partner with the development teams building AI-powered travel experiences to support the development and production operation of their solution.
Work with AI solutions, providers, and APIs. Partner with teams integrating AI capabilities and providers, with attention to API reliability, authentication, quotas, rate limits, latency and provider-specific operational constraints.
Troubleshoot AI tools and provider issues. Diagnose failures across AI-powered workflows, provider APIs, configuration, permission errors, degraded responses and related areas.
Operate reliable production platforms. implement and run cloud infrastructure,and help product teams move quickly without compromising reliability.
Improve observability. Build dashboards, alerts, traces, logs, and runbooks that make service health clear, actionable, and tied to SLOs and customer impact.
Apply AI to SRE workflows. Prototype and productionize AI-assisted systems that create effective and efficient operations
Automate operational toil. Create tools, workflows, and automation that remove repetitive manual work and make operational knowledge easier to use.
Requirements:
5+ years of experience as a Senior SRE, Infrastructure Software Engineer, Production Engineer, or DevOps Engineer.
3+ years of experience operating production, 24x7 customer-facing systems.
Hands-on experience delivering production infrastructure, platform tooling, and automation used by engineering teams.
Strong software engineering skills in Python, Go, Java, or a similar language, with a bias toward production-quality code, tests, monitoring, and documentation.
Experience with cloud infrastructure, container orchestration, Linux systems, networking, CI/CD, and infrastructure as code such as Terraform or CloudFormation.
Experience building, tuning, and automating observability systems such as Grafana, Prometheus, New Relic, Datadog, Splunk, or similar tools.
Familiarity with SLOs, incident response, on-call practices, root cause analysis, and blameless postmortems.
Practical experience or strong interest in AI solutions, AI providers, agents, AI APIs, provider integrations, or AI-assisted internal tools.
Ability to troubleshoot AI tools and provider/API issues, including rate limits, quota, auth, permission errors, latency, SDK or API contract changes, content quality issues, and service degradations.
Excellent communication skills and the ability to work with stakeholders and domain experts across the company.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797911
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are seeking a passionate QA Engineer to join our FW testing team. This position will be part of our growing team, helping NVIDIA products meet industry-leading benchmarks for efficiency and quality. As a Firmware QA Engineer, you will be immersed in the latest advancements and actively contribute to developing world-class products. Your efforts will directly impact the performance and quality of our offerings, ensuring they consistently meet industry-leading standards.


We are looking for ideal candidates who will actively test industry-leading features in the FW sector, quickly learn new features, conduct both manual and automated testing, and contribute to the development of automation systems.


What you'll be doing:
Review arch design and requirements documents for new features introduced in the Nvlink program.
Join the effort of QA shift left towards simulation for NVL7 generation.
Design, develop, and perform tests for the new features, as part of FW GA or update releases.
Automate newly added tests in the existing automation framework, and add new capabilities and features to the framework.
Report bugs found during execution, assist with reproduction and debugs to understand root cause, verify bug fixes provided by R&D team, and raise if not fixed.
Fully collaborate with different external teams like : :product marketing, R&D, and verification.
Define and build setups topologies for appropriate product coverage.
Requirements:
What we need to see:
Practical / BSc in Computer Science or Electrical/Electronics Engineering.
0-3 years of proven experience.
Clear verbal and written communication, proficient in written and spoken English.
Self-reliant individual, capable of planning and carrying out tasks within their assigned area.
Personal: good interpersonal skills, quick learner, proactive, innovative, highly motivated, and committed.

Ways to stand out from the crowd:
Scripting skills and proven experience in Python or other programming languages.
Proven experience in QA - methodologies, and test design.
Working in a fast-paced environment with multiple parallel running activities.
Knowledge/experience in Networking protocols.
Knowledge of Linux/Windows OS.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8837853
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Senior DevEx/DevOps Engineer.
As a Senior DevEx Engineer , you will join a core team dedicated to empowering our engineering organization by streamlining the entire software development lifecycle. Operating at the intersection of DevOps and Developer Experience, you will collaborate with R&D, Product, and Security to eliminate bottlenecks, designing and delivering robust CI/CD and infrastructure solutions that directly optimize developer productivity and velocity.
Responsibilities:
Project Leadership: Lead complex infrastructure projects from initial design through implementation and delivery.
CI/CD Optimization: Design and build high-performance pipelines using GitHub Actions and TeamCity.
Infrastructure Management: Scale cloud environments using Terraform, Ansible, and IaC best practices.
Requirements:
DevOps Experience: 5+ years of professional experience in DevOps/DevEx roles with a strong understanding of the software development lifecycle (SDLC).
IaC & Automation: Proven expertise in Infrastructure as Code (Terraform, Ansible) and advanced scripting for process automation.
Advanced CI/CD: Extensive hands-on experience designing and maintaining high-level CI/CD workflows and tools (e.g., GitHub Actions, TeamCity).
Cloud & Containerization: Solid experience managing cloud environments (AWS/GCP) and Linux-based containerization (Docker).
Communication: Strong professional communication skills in both English and Hebrew.
Preferred Qualifications (PQ):
Advanced Orchestration: Proficiency in Kubernetes ecosystems and container orchestration tools.
Coding & Observability: Strong programming skills in Python and experience with modern monitoring/logging suites.
FinOps & Architecture: Advanced AWS architectural knowledge and cloud cost optimization experience.
Testing Awareness: Understanding of developer workflows, testing frameworks, and integration tools.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8830451
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo and Herzliya
Job Type: Full Time
In this role, you will join a world-class team of engineers within the Surgical Operating unit responsible for developing medical devices that focus on lung cancer diagnosis.
Our Surgical Operating Unit is one new, powerful operating unit bringing together the people and product portfolio of Surgical Robotics and Surgical Innovations. With our Mission as our North Star, we will build on our legacy of proven surgical solutions and advance the promise of robotics and digital solutions to benefit our customers and patients.
Make your impact by exploring a career with the worlds leading Medical Device company, striving to alleviate pain, restore health, and extend life.

We are seeking a highly qualified and detail-oriented Software Verification Engineer to join our software verification team. This position involves verification and integration of complex, multidisciplinary medical systems. The role requires direct collaboration with cross-functional teams, including software, algorithm, and hardware engineering, to ensure the system meets all functional and performance requirements.

Responsibilities may include the following and other duties may be assigned:
Designing and executing SW verification activities: Proactively identify gaps in performance, accuracy, and quality across the software and algorithmic components of the system. Develop and prioritize test strategies based on risk, impact, and business needs. Create comprehensive test plans to validate functionality, accuracy, and performance.
Automation and continuous improvement: Contribute to the development and expansion of automated testing solutions. Continuously look for ways to improve test efficiency, coverage, and reliability by adopting new tools, scripting approaches, and creative testing methods.
Performing script-based testing: Develop and maintain test scripts to automate and enhance testing processes using Python, MATLAB, or other tools.
Collaborating with cross-functional teams: Work closely with cross-functional teams, including software, algorithm, hardware engineering, quality, and regulatory, to ensure successful testing outcomes.
Requirements:
Required Knowledge and Experience:
Education: Bachelors degree in engineering, with a preference for Biomedical Engineering.
At least 1 year of experience in software verification, algorithm validation, or system integration, preferably in the medical device industry.
Hands-on experience with multidisciplinary systems involving software, hardware, and algorithms.
Basic programming (MATLAB, Python) skills, hardware understanding, and experience with complex, cross-functional systems.
Experience with test automation is a strong advantage.
Strong analytical skills, a methodical approach to problem solving, and good documentation abilities.
Excellent interpersonal and communication skills; ability to collaborate effectively with cross-functional teams.
Comfort working in lab environments, including live/tissue animal, and cadaver labs, as well as with fluoroscopy systems.
Availability to work on-site in Herzliya at least 4 days per week.


Physical Job Requirements:
The above statements are intended to describe the general nature and level of work being performed by employees assigned to this position, but they are not an exhaustive list of all the required responsibilities and skills of this position. 
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8813235
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a Software Engineer to join our R&D team. In this role, you'll design and build scalable, high-performance systems that bring together rich backend capabilities and intuitive frontend experiences. You'll be part of a mission-driven team, building innovative technologies that enable security teams to detect, analyze, and respond to threats faster than ever before.


WHAT YOU WILL DO
Develop and own end-to-end features across the entire stack, from backend services to frontend interfaces.
Design and implement distributed backend systems that handle massive volumes of security data in real-time.
Build modern, responsive, and accessible user interfaces using React, TypeScript, Remix, Tailwind, and Shadcn.
Architect and maintain robust APIs and data contracts between the frontend and backend.
Collaborate with product, design, and security experts to deliver impactful user facing features.
Ensure high code quality by writing comprehensive unit, integration, and E2E tests using tools like Vitest, Playwright, and React Testing Library, as well as backend testing tools
Optimize system performance, reliability, and scalability on both ends of the stack.
Take part in critical architectural discussions and influence the technical direction of the platform.
Requirements:
WHAT YOU WILL BRING
7+ years of professional software engineering experience, with significant hands-on experience in both backend and frontend development.
Deep knowledge of Go or Python for backend services, and React with TypeScript for frontend applications.
Experience designing and working with distributed systems and microservices architecture.
Proven experience with backend testing frameworks such as Pytest or unittest for Python, and Gos testing package along with tools like testify or gomega.
Strong understanding of web fundamentals, API design, system scalability, and data processing.
Experience with performance monitoring and visualization tools like Grafana or Datadog.
A product-driven mindset with an eye for good UX and intuitive design.
Proven experience with testing practices across the stack.
Experience with CI/CD pipelines, Docker, and cloud-native infrastructure.
Strong sense of ownership, collaboration, and accountability.
Excellent communication skills and ability to work cross-functionally.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8838114
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
We are on an expedition to find an On-Premise Site Reliability Engineer (SRE)- someone who is passionate about building rock-solid, high-performance infrastructure and bringing order to complex environments. In this role, you will own the end-to-end reliability, automation, and deployment of our platform across customer sites, working hands-on with cutting-edge AI, bare-metal, and hybrid cloud architectures.
Youll collaborate closely with Product, R&D, and Architecture teams while serving as the ultimate technical authority for our customer deployments. From designing automated Ansible workflows and mastering Kubernetes to troubleshooting complex network topologies, you will eliminate toil, streamline cluster operations, and ensure every deployment is scalable, seamless, and mission-ready.
Responsibilities
Lead End-to-End On-Prem & Hybrid Deployments: Own the technical delivery and reliability of our platform in close collaboration with Product, R&D, and customer technical teams.
Architect, Execute & Improve K8s Deployments: Take a definitive hands-on role in deploying, configuring, operating, and continuously improving our platform using advanced, enterprise-grade Kubernetes architectures.
Helm Chart Management: Design, modify, and manage Helm charts to package, version, and streamline complex application deployments across different environments.
Drive Automation & Simplification: Design, implement, and maintain robust deployment automation using Ansible. You must have a passion for turning complex manual tasks into reliable, repeatable, single-click operations.
Manage Infrastructure as Code: Utilize Git as the single source of truth to manage configurations, manifests, and automation playbooks, enforcing modern engineering best practices.
Bridge On-Prem and Cloud: Leverage AWS resources (specifically EC2 and S3) for hybrid components, staging environments, or cloud-to-on-prem data flows.
Technical Tier-3 Escalation: Serve as the ultimate technical authority for deployment, Linux networking, and Kubernetes orchestration issues.
Continuous Improvement: Constantly refine our delivery pipelines, optimize bootstrap processes, and create rock-solid technical documentation.
Requirements:
SRE / Delivery Mindset: 3-5 years of hands-on experience in enterprise infrastructure deployment, systems engineering, or an on-prem operational reliability role.
Kubernetes & Helm Expert: Deep, production-grade experience with Kubernetes architecture, deployment, advanced troubleshooting, and CNI networking. Proven working experience creating, maintaining, and deploying applications using Helm charts.
Ansible Mastery: Proven experience writing clean, scalable Ansible roles and playbooks for configuration management, automation, and infrastructure provisioning.
Modern Workflows (Git & AWS): Solid working experience using Git for version control and collaborating on code/infrastructure. Practical experience provisioning and managing AWS resources (EC2 and S3).
Core Systems & Linux: Strong Linux background (Ubuntu) with a deep understanding of system internals, containerized runtimes, and troubleshooting distributed applications.
Solid Networking Knowledge: Hands-on experience with routing, firewalls, and switching topology (mainly Cisco)
Storage Foundations: Working knowledge of storage protocols (iSCSI, SAN, local NVMe) and enterprise storage arrays (like DELL) interacting with Kubernetes Persistent Volumes.
GPU & Accelerated Compute: Working knowledge of managing GPU-enabled Kubernetes nodes, including NVIDIA drivers/runtime and basic troubleshooting.
Air-Gapped Deployments: Experience deploying and maintaining software in air-gapped or offline environments, including registry mirroring and artifact staging.
Problem-Solver: Strong debugging and problem-solving skills in complex, distributed environments with an intense ownership and accountability mindset.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8836134
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
26/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a DevOps Engineer to join the R&D team and spread our power. In this role, you will design and implement scalable systems that will keep us running smoothly and support our significant business growth. You will join an innovative, high-performance team and work with cutting-edge technologies in a dynamic and agile environment.
Responsibilities:
Take an active part of all DevOps areas: our infrastructure and cloud environment, tools, services and up-time.
Be part of our products architectural and infrastructure design, examine and implement new cloud technologies, and open-source tools to improve the delivery and availability of the product.
Plan and push forward the growth and scale of data capacity for various products.
Be responsible for the smooth production-grade execution of provided solutions.
Requirements:
Minimum Qualifications
5 years of experience as a DevOps engineer on a high-scale distributed system, working in a Linux environment.
Hands-on experience with containerized environments and microservices, specifically Docker and Kubernetes.
Experience working in a multi-cloud environment.
Knowledge of build/release systems and CI/CD pipelines.
Scripting or programming skills with Python, Bash, or Go.
Full professional fluency (written and verbal) in both Hebrew and English.
Preferred Qualifications:
The mindset and approach for automating away from manual efforts.
An innovative approach, with the ability to quickly learn technologies.
A strong sense of ownership and accountability.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797750
סגור
שירות זה פתוח ללקוחות VIP בלבד