דרושים » הנדסה » Senior Infrastructure Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 18 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for an Infrastructure Engineer with a strong backend engineering mindset - someone who treats infrastructure as software.
In this role, you won't just operate infrastructure - you'll design systems, write code, and build controllers and automation that power our Kubernetes-native platform.
You'll work closely with platform, DevOps, and backend teams to create scalable, reliable infrastructure abstractions used by hundreds of developers.
What Youll Do:
Design and build backend-driven infrastructure workflows using Kubernetes controllers and declarative patterns.
Develop automation and tooling in Go or Python to manage infrastructure at scale.
Manage AWS resources through Kubernetes-native approaches (ACK or similar controller models).
Develop internal APIs, scripts, or services that improve developer workflows and platform capabilities.
Operate and evolve multi-cluster Kubernetes environments with a focus on automation and reliability.
Implement observability patterns and integrate monitoring signals into infrastructure logic.
Collaborate with backend engineers to design infrastructure that behaves like software - versioned, testable, and automated.
Requirements:
At least 3 years of experience in software engineering, DevOps, or infrastructure roles.
Strong backend development mindset.
Experience with Kubernetes, containers, and distributed systems concepts.
Programming experience in Go or Python.
Understanding of APIs, automation workflows, and system design.
Curiosity to learn and improve how infrastructure is built through code.
Nice to have:
Experience with Kubernetes controllers, CRDs, or operator patterns.
Exposure to GitOps workflows.
Familiarity with monitoring and observability concepts.
Experience with Temporal or workflow engines is a strong advantage.
Interest in AI-driven automation or ChatOps workflows.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8798141
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a Senior Infrastructure Engineer who views "Infrastructure as Software." In 2026, we dont just manage servers; we build high-performance environments that allow multi-agent systems to operate at scale.
You will be a core member of the R&D team, blending deep DevOps expertise with the coding rigor of a Backend Engineer. You arent just "configuring" AWS; you are architecting the distributed systems and data pipelines that power our autonomous security brain. Your mission is to ensure that while our agents are evolving and taking actions, our underlying platform remains immutable, observable, and infinitely scalable.
What You'll Do
Design, build, and operate our company's cloud infrastructure using AWS, Kubernetes, and Infrastructure as Code.
Build internal tools and platform services using Python and Go to improve developer productivity and system reliability.
Own infrastructure automation with Terraform, Pulumi, and modern cloud-native tooling.
Partner closely with Backend, Data Science, and Security Engineering teams to build scalable, reliable platforms.
Improve observability, monitoring, and incident response across distributed production systems.
Design and optimize infrastructure for performance, scalability, security, and cost efficiency.
Help shape engineering best practices, platform architecture, and developer experience as our company continues to grow.
Requirements:
5+ years of experience in Infrastructure, DevOps, Platform Engineering, or Backend Engineering.
Strong software engineering skills with hands-on experience building production systems in Python or Go.
Deep hands-on experience with AWS, including services such as EKS, RDS, VPC, and IAM.
Strong experience designing, operating, and scaling production Kubernetes environments.
Experience with Infrastructure as Code, CI/CD, GitOps, and modern cloud-native development practices.
A systems mindset with the ability to solve architectural challenges across infrastructure and application layers.
Comfortable using modern AI-powered developer tools and agentic workflows to improve engineering productivity.
The company Mindset: You take ownership, act with accountability, collaborate openly, and focus on delivering meaningful impact. You thrive in fast-moving environments, embrace ambiguity, and enjoy solving hard problems together.
Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
Full professional fluency (written and verbal) in both Hebrew and English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
03/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an exceptional Platform Engineer who combines deep software engineering with a robust DevOps approach to help propel development infrastructure to the next level. , SPEED is integral to our DNA. AI transforms the way we develop at speed, and our development infrastructure is key to the scale and hypergrowth.



The DevX team is a multi-disciplinary group of DevOps and development experts focused on building and operating internal tools, infrastructure, and services to make engineering easy and intuitive SPEED.



Your mission is to build a top-tier Continuous Integration and Delivery (CI/CD) platform. You will own both the application services and the infrastructure that supports them, ensuring speed, stability, and reliability for our entire R&D organization.



Key Focus Areas & What You'll Do



Improve CI/CD pipelines and CI workflows (e.g., GitHub Actions) with a focus on speed and reliability.
Reduce CI flakiness and improve overall pipeline stability through systematic triage and root-cause analysis.
Shorten developer feedback loops by optimizing test strategy, pipeline consistency, and local development workflows.
Strengthen CD and release processes for secure, repeatable, and fast deployments.
Promote best practices such as GitOps and progressive delivery where they fit.
Track and improve delivery metrics, with emphasis on Lead Time for Changes and Deployment Frequency.
Partner with engineering teams to identify friction and deliver scalable automation and paved paths.
Requirements:
3+ years of experience as a Platform Engineer or a strong Backend Engineer with a DevOps focus.
Strong knowledge of CI/CD pipelines, versioning, and release management.
Proven experience with modern build tools (e.g., Bazel, SWC) and managing CI workflows (e.g., GitHub Actions, GitLab CI).
Good skills with infrastructure-as-code tools like Terraform and Helm.
Hands-on experience with Kubernetes, including containers (Docker), Kafka and cloud providers (AWS, GCP, or Azure).
Strong programming skills in Python, Go, TypeScript, or another modern backend language.
Experience with monorepo tools (Lerna, NX, PNPM) is a big plus.
Familiarity with GitOps workflows using tools like ArgoCD or Flux.
A strong sense of ownership and a passion for making systems reliable, scalable, and improving the developer experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8765802
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
03/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an exceptional Platform Engineer who combines deep software engineering with a robust DevOps approach to help propel development infrastructure to the next level. , SPEED is integral to our DNA. AI transforms the way we develop at speed, and our development infrastructure is key to the scale and hypergrowth.



The DevX team is a multi-disciplinary group of DevOps and development experts focused on building and operating internal tools, infrastructure, and services to make engineering easy and intuitive SPEED.



Your mission is to build a top-tier Continuous Integration and Delivery (CI/CD) platform. You will own both the application services and the infrastructure that supports them, ensuring speed, stability, and reliability for our entire R&D organization.



Key Focus Areas & What You'll Do



Improve CI/CD pipelines and CI workflows (e.g., GitHub Actions) with a focus on speed and reliability.
Reduce CI flakiness and improve overall pipeline stability through systematic triage and root-cause analysis.
Shorten developer feedback loops by optimizing test strategy, pipeline consistency, and local development workflows.
Strengthen CD and release processes for secure, repeatable, and fast deployments.
Promote best practices such as GitOps and progressive delivery where they fit.
Track and improve delivery metrics, with emphasis on Lead Time for Changes and Deployment Frequency.
Partner with engineering teams to identify friction and deliver scalable automation and paved paths.
Requirements:
What Were Looking For:



3+ years of experience as a Platform Engineer or a strong Backend Engineer with a DevOps focus.
Strong knowledge of CI/CD pipelines, versioning, and release management.
Proven experience with modern build tools (e.g., Bazel, SWC) and managing CI workflows (e.g., GitHub Actions, GitLab CI).
Good skills with infrastructure-as-code tools like Terraform and Helm.
Hands-on experience with Kubernetes, including containers (Docker), Kafka and cloud providers (AWS, GCP, or Azure).
Strong programming skills in Python, Go, TypeScript, or another modern backend language.
Experience with monorepo tools (Lerna, NX, PNPM) is a big plus.
Familiarity with GitOps workflows using tools like ArgoCD or Flux.
A strong sense of ownership and a passion for making systems reliable, scalable, and improving the developer experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8765844
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
04/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a highly skilled Infrastructure Engineer to join our team and own the scaling, management, and automation of our platforms distributed environments. If youre excited about building high-scale distributed systems and solving deep DevOps and infrastructure challenges, lets talk.

What Youll Do:

Own and scale infrastructure - Design, build, and optimize the backbone of our observability platform, ensuring seamless deployment across hundreds of distributed environments.
Solve complex scalability challenges - Tackle unique problems in multi-cluster Kubernetes environments, multi-cloud setups, and high-ingestion observability pipelines.
Manage data at scale - Build and optimize configurable data pipelines, ensuring efficient ingestion, storage, and querying of large volumes of observability data with resilience, consistency, and analytical capabilities.
Automate everything - Develop infrastructure as code, improve CI/CD processes, and automate environment provisioning for reliability and efficiency.
Enhance system reliability - Design robust monitoring, alerting, and self-healing mechanisms for a high-scale production environment.
Collaborate cross-functionally - Work closely with backend engineers, product teams, and customers to design scalable, developer-friendly infrastructure.
Adopt and implement cutting-edge technologies - Continuously evaluate and introduce new tools and frameworks to improve scalability, performance, and cost efficiency.
Improve deployment efficiency - Optimize Helm charts, Kubernetes operators, and Terraform configurations to streamline environment creation and lifecycle management.
Requirements:
5+ years of experience in DevOps, SRE, or Infrastructure Engineering roles.
Strong expertise in Kubernetes, Terraform, Helm, and cloud environments (AWS, GCP, or Azure).
Experience with scalable observability stacks (e.g., ClickHouse, VictoriaMetrics, OpenTelemetry) is a huge plus.
Deep understanding of distributed systems, networking, and containerized workloads.
Proficiency in at least one programming language (Go, Python, or similar) for automation and tooling
Passion for building scalable, reliable, and efficient infrastructure.
A problem-solving mindset with the ability to tackle complex technical challenges independently.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8768205
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
17/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a DevOps Engineer to join our engineering team. The ideal candidate has strong hands-on experience with cloud-native infrastructure, a GitOps mindset, and the ability to independently research and learn new tools and techniques in a fast-moving, large-scale Kubernetes environment.
Responsibilities
Design, operate, and troubleshoot Kubernetes clusters (EKS/AKS) at scale
Manage application delivery and infrastructure using GitOps tools (ArgoCD/Flux)
Build and maintain Helm charts and Kustomize overlays for multi-environment deployments
Provision and manage cloud infrastructure using Crossplane and/or Terraform
Own and optimize CI/CD pipelines (GitHub Actions, GitLab CI) for build, test, and deployment workflows
Maintain and extend observability stacks (Prometheus, Grafana, alerting rules, dashboards)
Write automation scripts and tooling in Python, Go, or Bash to streamline operations
Support AWS infrastructure across multiple accounts/regions (networking, IAM, compute, storage)
Participate in on-call rotation, troubleshoot production incidents, and drive root-cause analysis
Collaborate with platform, security, and application engineering teams on infrastructure design and reliability improvements.
Requirements:
2-3 years of experience in DevOps, SRE, platform engineering, or a related role
Solid working knowledge of Kubernetes and Helm in production environments
Experience with Crossplane and/or Terraform for infrastructure as code
Proficiency with AWS services (EKS, IAM, VPC, networking, compute)
Hands-on experience with GitOps tools such as ArgoCD or Flux
Scripting ability in Python, Go, or Bash for automation and tooling
Experience with Prometheus and Grafana for monitoring and alerting
Practical experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI)
Strong self-learning ability, capable of independently researching unfamiliar technologies, reading documentation, and applying findings without handholding
Solid troubleshooting skills across networking, compute, and distributed systems
Nice to Have
Programming proficiency in Go or Python (beyond scripting)
Experience with service mesh technologies (Istio or similar)
Exposure to multi-cloud environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785730
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 19 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are constantly striving to make our systems reliable, scalable, and simple to operate so our services are available to travelers when they need them most. With our continued growth, we have exciting challenges ahead and we're looking for a Senior Site Reliability Engineer to join our team in Tel Aviv. This role blends classic SRE ownership with pragmatic AI SRE work: you will build and operate the platforms, automation, observability, and incident response practices that keep Navan reliable, while helping teams use AI solutions, AI providers, and their APIs safely and dependably.



This is a hands-on engineering role, not a research role. You will partner with product, platform, data, security, support, and incident response teams to make production systems and AI-powered experiences more resilient. You will use software engineering, infrastructure as code, SLOs, telemetry, provider observability, and automation as your main tools, and you will apply AI where it creates measurable reliability value rather than novelty.



This position is based out of our new Tel Aviv office.



What You'll Do:

Support AI-based application solutions where reliability matters. Partner with the development teams building AI-powered travel experiences to support the development and production operation of their solution.
Work with AI solutions, providers, and APIs. Partner with teams integrating AI capabilities and providers, with attention to API reliability, authentication, quotas, rate limits, latency and provider-specific operational constraints.
Troubleshoot AI tools and provider issues. Diagnose failures across AI-powered workflows, provider APIs, configuration, permission errors, degraded responses and related areas.
Operate reliable production platforms. implement and run cloud infrastructure,and help product teams move quickly without compromising reliability.
Improve observability. Build dashboards, alerts, traces, logs, and runbooks that make service health clear, actionable, and tied to SLOs and customer impact.
Apply AI to SRE workflows. Prototype and productionize AI-assisted systems that create effective and efficient operations
Automate operational toil. Create tools, workflows, and automation that remove repetitive manual work and make operational knowledge easier to use.
Requirements:
5+ years of experience as a Senior SRE, Infrastructure Software Engineer, Production Engineer, or DevOps Engineer.
3+ years of experience operating production, 24x7 customer-facing systems.
Hands-on experience delivering production infrastructure, platform tooling, and automation used by engineering teams.
Strong software engineering skills in Python, Go, Java, or a similar language, with a bias toward production-quality code, tests, monitoring, and documentation.
Experience with cloud infrastructure, container orchestration, Linux systems, networking, CI/CD, and infrastructure as code such as Terraform or CloudFormation.
Experience building, tuning, and automating observability systems such as Grafana, Prometheus, New Relic, Datadog, Splunk, or similar tools.
Familiarity with SLOs, incident response, on-call practices, root cause analysis, and blameless postmortems.
Practical experience or strong interest in AI solutions, AI providers, agents, AI APIs, provider integrations, or AI-assisted internal tools.
Ability to troubleshoot AI tools and provider/API issues, including rate limits, quota, auth, permission errors, latency, SDK or API contract changes, content quality issues, and service degradations.
Excellent communication skills and the ability to work with stakeholders and domain experts across the company.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797911
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
30/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior DevOps Engineer , youll play a critical role in scaling and evolving our infrastructure as we grow. Youll work alongside experienced engineers to drive automation, optimize cloud operations, and ensure our systems are secure, resilient, and high-performing. This role is perfect for a hands-on engineer who thrives in fast-paced environments and wants to shape infrastructure practices at a product-focused security startup.



What Youll Do

Own and Evolve Infrastructure: Design, deploy, and operate infrastructure on AWS using Kubernetes to orchestrate containerized services.
Build Tools and Automate Everything: Streamline internal workflows with smart tooling, configuration management, and monitoring systems.
Lead CI/CD Improvements: Define and refine deployment pipelines to support rapid, reliable releases and cross-team agility.
Strengthen Edge Security: Manage WAFs, gateways, and load balancers to balance strong protection with great user experience.
Drive Reliability: Lead incident detection, troubleshooting, and automated recovery-improving uptime and system robustness continuously.
Requirements:
6+ years in DevOps or infrastructure engineering roles.
Deep experience with Kubernetes, Helm, containerization, and cloud.
Strong background in AWS (bonus: experience with GCP or Azure).
Proficiency in CI/CD platforms (GitHub Actions, GitLab, Jenkins, etc.).
Solid understanding of networking, distributed systems, and both SQL and NoSQL databases.
Linux power user with strong scripting skills (e.g., Bash, Python).
Hands-on with infrastructure-as-code tools like Terraform.
Great communicator with a security mindset and a bias for automation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8762093
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 13 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a DevOps Engineer to join our growing Infrastructure team.
You'll work alongside our DevOps team to design, scale, and optimize the cloud infrastructure powering our company's AI-driven platform. As our product and engineering organization continue to grow rapidly, you'll play a key role in improving reliability, automation, scalability, and developer productivity.
This is a highly hands-on role with significant ownership. You'll collaborate closely with R&D, AI, and Security teams to build reliable infrastructure, streamline development workflows, and help shape the future of our engineering platform.
Responsibilities
Design, build, and maintain scalable cloud infrastructure.
Develop and improve CI/CD pipelines and deployment processes.
Manage and optimize Kubernetes clusters and containerized environments.
Automate infrastructure using Infrastructure as Code (Terraform or similar).
Improve monitoring, observability, and system reliability.
Support production environments and troubleshoot infrastructure issues.
Collaborate closely with R&D, AI, and Security teams to improve developer experience.
Help define infrastructure best practices and drive operational excellence.
Requirements:
4+ years of experience as a DevOps Engineer or Site Reliability Engineer.
Strong experience with AWS (or another major cloud provider).
Hands-on experience with Kubernetes and Docker.
Experience building and maintaining CI/CD pipelines.
Experience with Infrastructure as Code (Terraform preferred).
Strong scripting skills (Python, Bash, or similar).
Experience with Linux environments.
Strong troubleshooting and problem-solving skills.
Excellent communication and collaboration skills.
Bonus Points
Experience supporting AI/ML infrastructure.
Experience with GitHub Actions, ArgoCD, or Helm.
Experience with AI Tools - such as Claude or Cursor.
Experience with monitoring tools such as Prometheus, Grafana, or Datadog.
Experience working in cybersecurity companies.
Startup experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8798830
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד