דרושים » מחשבים ורשתות » DevOps Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Responsibilities:
Own operational responsibility for the application and platform layers.
Oversee and drive production operations end-to-end: deployments, maintenance, scaling, high availability and enhancements across AWS and GCP.
Operate Kubernetes in productions at scale across AWS EKS and GCP GKE
Manage all infrastructure as code.
Participate in on-call, lead and participate in incident management: triage, mitigation, and clear stakeholder communication during live incidents.
Own deployment tooling, troubleshooting, and performance tuning across compute, networking, and data layers.
Contribute to cost visibility and FinOps practices across cloud providers.
Requirements:
Requirements
5+ years running production systems, hands-on.
Strong AWS experience (GCP a plus).
Production Kubernetes - networking, autoscaling, resource management, etc
Terraform (Terragrunt/Pulumi) as a daily tool, including modules and state.
Helm , GitOps with ArgoCD
GitHub and GitHub Actions; pipeline-as-code.
Solid Linux and networking: DNS, TLS, load balancing, VPC/routing.
Prometheus/Cortex, Grafana, Loki, or equivalents
Scripting in Python, Go, or Bash; comfortable reading application code.
Clear written English;
Experience in a distributed, global team.

Advantage
Development background.
Kafka, MQTT, or similar messaging at scale.
Data platform experience (Databricks, Spark,).
FinOps: cost allocation and unit economics.
Cloud security and compliance (SOC 2, CIS, CSPM).
Effective use of AI-assisted engineering tooling.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8763613
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a DevOps Engineer.
our DevOps team operates the infrastructure that powers our AI and Computer Vision platform across construction sites in 15+ countries. From data pipelines and ML workloads to backend services - you'll work with a diverse, modern, Kubernetes-based stack and have real influence on how we build, deploy, and operate.
What you'll do:
Own Multi-Cloud Infrastructure: Work alongside the team to design, scale, and operate our high-scale, multi-region production infrastructure across AWS and GCP, powering construction sites globally.
Drive Kubernetes at Scale: Manage and evolve our Kubernetes platform on EKS and leveraging GitOps practices with ArgoCD and Helm to enable safe, fast, and reliable deployments.
Build Robust CI/CD: Design and maintain CI/CD pipelines that empower dozens of engineers to ship confidently - with automation, testing, and progressive delivery built in.
Tackle Diverse Infrastructure Challenges: Work hands-on with a wide variety of workloads - from heavy data processing and Computer Vision pipelines to backend services and ML inference - each with unique scaling, performance, and reliability requirements.
Ensure Reliability & Observability: Build and maintain world-class observability (metrics, logs, tracing, alerting) so that issues are caught early and resolved fast. Performance, reliability, and scalability are at the core of what you do.
Security & Cost: Partner with the team to strengthen our security posture, identity and access management, compliance, and cloud cost optimization across both clouds.
Ownership from 0 to 1: You will have real influence over our architecture and tooling. We want engineers who care about shaping what we build and how we build it, ensuring performance, security, and observability are baked in from day one.
Requirements:
A seasoned DevOps / Infrastructure engineer (5+ years) with strong hands-on experience in production cloud environments.
Proven expertise operating large-scale, distributed systems - with deep understanding of Kubernetes, networking, and cloud-native architecture.
Strong experience with multi-cloud environments (AWS and/or GCP), Infrastructure-as-Code (Terraform), and GitOps workflows (ArgoCD, Flux, or similar).
Hands-on experience with CI/CD systems (Jenkins, GitHub Actions, etc.).
Solid scripting and automation skills (Python, Bash, or Go).
Proven track record of being a collaborative team player who partners closely with developers, ML engineers, and cross-functional stakeholders across the organization.
Experience with observability stacks (Prometheus, Grafana, OpenTelemetry, Logz.io, or similar).
Experience with databases (relational and/or NoSQL) - including operational aspects like backups, migrations, and performance tuning.
AI-Native Engineering: You are an AI-native engineer who leverages LLMs and agentic tools (like Cursor, Copilot, or Claude) not just for command completion, but as a core operational partner - automating diagnostics, runbooks, and infrastructure workflows so you can focus on the critical things
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8729164
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 11 שעות
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are seeking a promising and talented Senior DevFinOps Engineer to join our DevOps group. If you thrive in a fast-paced, dynamic environment, can handle multiple requests simultaneously, and enjoy working independently as part of a cutting-edge DevOps team, this is your opportunity to help make the world a safer place!
Key Responsibilities
Act as a DevFinOps Engineer within a highly skilled team, bridging engineering and finance to drive cloud cost efficiency across large-scale operations from development to production.
Design, develop, and maintain Avanan's cloud cost visibility, allocation, and optimization solutions - including tagging strategies, cost dashboards, budgets, and anomaly detection across accounts and services.
Implement tools and procedures for cost monitoring, forecasting, and alerting across our SaaS multi-tenant product family.
Embed FinOps practices into the CI/CD lifecycle - surfacing the cost impact of changes early, and enforcing cost guardrails as part of deployment automation.
build AI-based FinOps agents for cloud services at the infrastructure and application levels
Continuously identify and execute cost-optimization opportunities (right-sizing, reserved capacity/savings plans, spot usage, storage tiering, idle-resource cleanup) without compromising performance, reliability, or security.
Partner with engineering teams to drive cost accountability - providing unit-economics insight (cost per tenant/feature/service) and making cost a first-class engineering metric.
Plan capacity and model the financial impact of scaling decisions, balancing cost efficiency with fault tolerance and growth.
Design and shape our cost reporting, chargeback/showback, and budgeting solutions.
Execute all tasks with cloud infrastructure security and financial governance as guiding principles.
Requirements:
Hands-on mindset - we all write code daily!
3+ years of relevant DevOps/Cloud experience building and operating CI/CD pipelines for both development and production - must.
2+ years of AWS Cloud experience working with high-traffic systems and multiple services, with a strong grasp of AWS pricing models and cost management tooling (Cost Explorer, CUR, Budgets, Compute Optimizer) - must.
Strong scripting skills, with fluency in Python - must.
Experience working with AI tools to achieve cloud or application cost control, identify and investigate the root cause of cost changes (differentiating between organic growth / infrastructure change / config change / application change).
Demonstrated experience driving measurable cloud cost reductions and building cost-optimization tooling or automation - must.
Experience with containers and orchestration tools (Docker, Kubernetes, or ECS) and understanding of their cost drivers - must.
Familiarity with FinOps principles and practices (FinOps Foundation framework, showback/chargeback, unit economics) - an advantage.
Experience with CI integration tools such as Jenkins.
Familiarity with AWS CloudFormation and infrastructure-as-code - an advantage.
Exposure to a wide range of open-source technologies (Redis, Nagios, Grafana, Prometheus, etc.) and cost-analytics tooling.
Knowledge of best practices in security, performance, monitoring, and cost governance.
Proven ability to research, evaluate, and implement new technologies, including running proofs of concept and cost analysis.
Huge advantage: measuring and controlling AI cost.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8723227
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a Senior Infrastructure Engineer who views "Infrastructure as Software." In 2026, we dont just manage servers; we build high-performance environments that allow multi-agent systems to operate at scale.
You will be a core member of the R&D team, blending deep DevOps expertise with the coding rigor of a Backend Engineer. You arent just "configuring" AWS; you are architecting the distributed systems and data pipelines that power our autonomous security brain. Your mission is to ensure that while our agents are evolving and taking actions, our underlying platform remains immutable, observable, and infinitely scalable.
What You'll Do
Design, build, and operate our company's cloud infrastructure using AWS, Kubernetes, and Infrastructure as Code.
Build internal tools and platform services using Python and Go to improve developer productivity and system reliability.
Own infrastructure automation with Terraform, Pulumi, and modern cloud-native tooling.
Partner closely with Backend, Data Science, and Security Engineering teams to build scalable, reliable platforms.
Improve observability, monitoring, and incident response across distributed production systems.
Design and optimize infrastructure for performance, scalability, security, and cost efficiency.
Help shape engineering best practices, platform architecture, and developer experience as our company continues to grow.
Requirements:
5+ years of experience in Infrastructure, DevOps, Platform Engineering, or Backend Engineering.
Strong software engineering skills with hands-on experience building production systems in Python or Go.
Deep hands-on experience with AWS, including services such as EKS, RDS, VPC, and IAM.
Strong experience designing, operating, and scaling production Kubernetes environments.
Experience with Infrastructure as Code, CI/CD, GitOps, and modern cloud-native development practices.
A systems mindset with the ability to solve architectural challenges across infrastructure and application layers.
Comfortable using modern AI-powered developer tools and agentic workflows to improve engineering productivity.
The company Mindset: You take ownership, act with accountability, collaborate openly, and focus on delivering meaningful impact. You thrive in fast-moving environments, embrace ambiguity, and enjoy solving hard problems together.
Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
Full professional fluency (written and verbal) in both Hebrew and English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
19/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a senior-level engineer who is deeply fluent in modern cloud-native technologies and can leverage them to design, build, and continuously improve the systems that power Identity Protection products. This role is open in Tel Aviv, Israel.

What You'll Do:

Design and implement new services and integrations across a hybrid environment - primarily AWS and an on-premises data center, with additional presence on GCP and OCI - leveraging shared platforms such as Kubernetes, Kafka, and modern monitoring stacks.

Build and maintain CI/CD pipelines (Jenkins) and configuration management workflows using Chef and related tooling.

Write automation, tooling, and glue logic in Bash and Python to reduce toil and accelerate delivery.

Utilize AI and automation platforms (e.g., Claude, n8n) to drive engineering efficiency and intelligent operations.

Contribute architectural insight to platform decisions - understanding system trade-offs and designing for scale and reliability.

Participate in on-call rotations and incident response for production Identity Protection services
Requirements:
Deep knowledge and hands-on experience with Linux-based systems, including administration, performance tuning, and troubleshooting in production environments.

Deep familiarity with Kubernetes - enough to design workloads, debug failures, and make informed architectural decisions.

Strong understanding of distributed messaging systems (Kafka and/or RabbitMQ)

Deep experience with MongoDB (or comparable NoSQL): schema design, query optimization, index tuning, maintenance operations, and understanding of internals - working closely with the data layer.

Proficiency with Docker and containerized application design.

Working knowledge of observability tooling - Prometheus, Grafana, and centralized logging stacks - and the ability to instrument and interpret systems using them.

Solid scripting in Bash and Python; comfortable writing production-grade automation.

Experience with configuration management tools, Chef preferred.

Ability to reason at the architectural level: understand big-picture system design and communicate trade-offs clearly.

On-call availability on a rotational basis.

Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8743577
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8739903
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
6 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior DevOps Engineer , youll play a critical role in scaling and evolving our infrastructure as we grow. Youll work alongside experienced engineers to drive automation, optimize cloud operations, and ensure our systems are secure, resilient, and high-performing. This role is perfect for a hands-on engineer who thrives in fast-paced environments and wants to shape infrastructure practices at a product-focused security startup.



What Youll Do

Own and Evolve Infrastructure: Design, deploy, and operate infrastructure on AWS using Kubernetes to orchestrate containerized services.
Build Tools and Automate Everything: Streamline internal workflows with smart tooling, configuration management, and monitoring systems.
Lead CI/CD Improvements: Define and refine deployment pipelines to support rapid, reliable releases and cross-team agility.
Strengthen Edge Security: Manage WAFs, gateways, and load balancers to balance strong protection with great user experience.
Drive Reliability: Lead incident detection, troubleshooting, and automated recovery-improving uptime and system robustness continuously.
Requirements:
6+ years in DevOps or infrastructure engineering roles.
Deep experience with Kubernetes, Helm, containerization, and cloud.
Strong background in AWS (bonus: experience with GCP or Azure).
Proficiency in CI/CD platforms (GitHub Actions, GitLab, Jenkins, etc.).
Solid understanding of networking, distributed systems, and both SQL and NoSQL databases.
Linux power user with strong scripting skills (e.g., Bash, Python).
Hands-on with infrastructure-as-code tools like Terraform.
Great communicator with a security mindset and a bias for automation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8762093
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
-Cloud Infrastructure Platform Ownership:
Design, implement, and maintain scalable infrastructure in a multi-account AWS organization
Manage and deploy applications using Helm charts, GitOps workflows using ArgoCD
Support CI/CD pipelines and release processes.

-Observability Reliability:
Implement and maintain monitoring and logging solutions.
Define alerts, dashboards, and SLOs to ensure system health and operational excellence.

-MLOps Enablement:
Support and operate ML infrastructure in AWS
Enable reliable model deployment, monitoring, and lifecycle management.

-DevSecOps Security:
Implement security best practices across infrastructure and CI/CD pipelines
Enforce IAM least-privilege policies and secure networking configurations
Integrate security ownership and compliance controls.

-DevFinOps Cost Optimization

-IT Operational Support
Requirements:
-3-5 years of experience working in a DevOps or Platform Engineering role.
-BSc degree in Computer Science, Engineering, or equivalent (required).
-Strong hands-on experience operating in an AWS environment.
Solid understanding of networking fundamentals, including: VPC design, subnets, routing tables, Load- balancers, NAT/Internet gateways.
-Experience maintaining Kubernetes clusters at scale, including managing and deploying Helm charts.
Strong background in observability and monitoring using: Prometheus, Grafana, and Loki and CloudWatch.
-Experience with GitOps workflows and continuous delivery using ArgoCD.
Proficiency with Infrastructure as Code (Terraform / Cloudformation), preferably using AWS CDK ( Python and/or TypeScript).
Core understanding of MLOps principles, including model deployment, monitoring, versioning, and- lifecycle management.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8719575
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
We are looking for a Senior Cloud DevOps Engineer to design, build, and optimize secure, scalable, and efficient cloud infrastructure for a live cybersecurity SaaS product. You will take ownership of infrastructure designs, production reliability, and cloud cost optimization, while working closely with other senior technical teams.
Key Responsibilities:
Design, implement, and optimize cloud infrastructure with a strong focus on scalability, security, and cost-efficiency.
Build and maintain secure cloud networking infrastructure - including VPCs, subnets, routing, VPNs, firewalls, and load balancers.
Develop Infrastructure as Code solutions using Terraform, Kubernetes, and Helm.
Enhance CI/CD pipelines and deployment workflows (Argo CD) to support frequent production releases.
Implement monitoring, alerting, and observability systems using Prometheus, Grafana, and logging tools.
Troubleshoot and resolve complex production issues, ensuring high availability and customer satisfaction.
Collaborate closely with development, security, and operations teams to integrate infrastructure with application needs.
Contribute to cloud cost optimization initiatives by improving efficiency and resource usage.
Maintain cloud security best practices and compliance with standards (e.g., SOC2, ISO27001).
Requirements:
7 - 10 years of experience in Cloud Infrastructure, DevOps, or SRE roles.
Strong hands-on experience with AWS, Azure, and/or GCP, especially in networking (VPC setup, VPNs, routing, firewall configurations).
Expertise in Terraform, Kubernetes, Helm, and scripting languages (Python or Bash).
Proven experience supporting live, production SaaS environments.
Solid understanding of cloud security concepts and practices.
Strong troubleshooting skills in distributed, cloud-based environments.
Excellent communication and collaboration skills.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8713872
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
7 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for an experienced SRE Team Lead to drive the reliability, observability, and automation practices across our private cloud infrastructure and operations. In this role, you will lead a team of site reliability engineers, own the engineering roadmap for monitoring and automation, and act as a key liaison between development, operations, and platform teams. You bring at least 3-4 years of hands-on people management experience and a deep technical background in SRE or DevOps disciplines.


What will you do?

Leadership & Team Management

Lead, mentor, and grow a team of SREs, providing technical direction, career development guidance, and day-to-day management.

Own the team roadmap for reliability, observability, and automation initiatives - prioritizing work, removing blockers, and driving delivery.

Conduct regular 1:1s, performance reviews, and hiring processes to build and sustain a high-performing team.

Foster a culture of operational excellence, blameless post-mortems, and continuous improvement.

Act as an escalation point for complex incidents and reliability issues, leading post-incident reviews and ensuring follow-through on action items.


Automation & Infrastructure

Design, develop, and maintain automation tools to support infrastructure and operations teams at scale.

Manage pipelines and infrastructure workflows using Jenkins, Ansible, Python, and Bash.

Drive the adoption of infrastructure-as-code practices across the organization.

Collaborate with system engineers to improve scalability, performance, and fault tolerance of critical systems.


Monitoring & Observability

Build and extend monitoring and alerting systems using Grafana, the ELK (Elastic) stack, Zabbix, and custom scripts.

Implement and enforce observability best practices to ensure full visibility into systems, applications, and infrastructure.

Define and track SLIs, SLOs, and error budgets across key services.

Partner with development teams to embed observability earlier in the software development lifecycle.


Database & Platform Support

Support monitoring and infrastructure integration for databases including MongoDB and PostgreSQL.

Maintain documentation and champion knowledge sharing around automation, monitoring, and reliability practices.
Requirements:
Experience & Leadership:

3-4+ years of experience in a people management or team lead capacity within SRE, DevOps, or infrastructure engineering.

5-8+ years of overall experience in SRE, DevOps, or infrastructure automation roles.

Proven track record of building, coaching, and retaining high-performing engineering teams.

Experience owning an engineering roadmap and driving cross-functional reliability initiatives.


Technical Skills :

Strong scripting skills in Python and Bash; comfortable building and maintaining production-grade automation.

Hands-on experience with infrastructure automation tools, particularly Ansible.

Solid experience with monitoring and observability platforms - ELK stack, Grafana, and Zabbix.

Good understanding of CI/CD pipelines and related tooling, including Jenkins.

Familiarity with managing and monitoring MongoDB and PostgreSQL in a production environment.

Comfortable working in Linux-based environments.

Excellent problem-solving skills and strong written and verbal communication.


Ability to support the following:

Experience with cloud providers - AWS, GCP, or Azure.

Exposure to containerization technologies such as Docker and Kubernetes.

Familiarity with infrastructure provisioning using Terraform.

Experience introducing SRE practices (SLOs, error budgets, chaos engineering) at an organizational level.

Exposure and experience with migrating/ building AI tools to improve process.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8760168
סגור
שירות זה פתוח ללקוחות VIP בלבד