דרושים » תוכנה » DevOps Engineer (Agentic Search)

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
We're building the infrastructure layer for agentic web interaction at scale. Our API is designed from the ground up to power Retrieval-Augmented Generation (RAG) and real-time reasoning in AI systems. By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that are not only intelligent - but also informed.

We work with some of the most innovative teams in AI - from small startups shaping the ecosystem to the largest enterprises deploying AI at scale. Whether it's powering sales assistants, research copilots, or internal knowledge tools, we're the missing link between LLMs and the real world.

The Role: DevOps Engineer
Managing Kubernetes clusters across multiple environments and regions

Owning infrastructure as code for all resources

Maintaining and improving CI/CD pipelines and GitOps-based deployments

Maintaining and optimize real-time data pipelines that process billions of events per day across distributed queues and stream processors

Building out monitoring, alerting, and observability

Debugging production issues across services

Managing cloud costs and capacity planning

Working closely with a small engineering team - you'd own infra, not a slice of it
Requirements:
3+ years in a DevOps or platform engineering role, working in production environments

Proven experience designing and operating large-scale, distributed systems, with a solid understanding of API design, reliability, and performance at scale

Strong Kubernetes experience in a managed cloud environment

Proficiency with infrastructure as code (Terraform or similar)

Experience with GitOps-based deployment workflows

Built or maintained observability stacks (logging, metrics, alerting)

Experience handling production incidents calmly and methodically
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8818172
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have:
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796409
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
17/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a DevOps Engineer to join our engineering team. The ideal candidate has strong hands-on experience with cloud-native infrastructure, a GitOps mindset, and the ability to independently research and learn new tools and techniques in a fast-moving, large-scale Kubernetes environment.
Responsibilities
Design, operate, and troubleshoot Kubernetes clusters (EKS/AKS) at scale
Manage application delivery and infrastructure using GitOps tools (ArgoCD/Flux)
Build and maintain Helm charts and Kustomize overlays for multi-environment deployments
Provision and manage cloud infrastructure using Crossplane and/or Terraform
Own and optimize CI/CD pipelines (GitHub Actions, GitLab CI) for build, test, and deployment workflows
Maintain and extend observability stacks (Prometheus, Grafana, alerting rules, dashboards)
Write automation scripts and tooling in Python, Go, or Bash to streamline operations
Support AWS infrastructure across multiple accounts/regions (networking, IAM, compute, storage)
Participate in on-call rotation, troubleshoot production incidents, and drive root-cause analysis
Collaborate with platform, security, and application engineering teams on infrastructure design and reliability improvements.
Requirements:
2-3 years of experience in DevOps, SRE, platform engineering, or a related role
Solid working knowledge of Kubernetes and Helm in production environments
Experience with Crossplane and/or Terraform for infrastructure as code
Proficiency with AWS services (EKS, IAM, VPC, networking, compute)
Hands-on experience with GitOps tools such as ArgoCD or Flux
Scripting ability in Python, Go, or Bash for automation and tooling
Experience with Prometheus and Grafana for monitoring and alerting
Practical experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI)
Strong self-learning ability, capable of independently researching unfamiliar technologies, reading documentation, and applying findings without handholding
Solid troubleshooting skills across networking, compute, and distributed systems
Nice to Have
Programming proficiency in Go or Python (beyond scripting)
Experience with service mesh technologies (Istio or similar)
Exposure to multi-cloud environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785730
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
23/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are a well-funded, early-stage startup looking for a talented and motivated Backend Engineer specializing in infrastructure to join our founding team. The focus of this role is to build and scale the infrastructure that powers autonomous AI agents automating complex enterprise workflows. You will own the systems, pipelines, and platforms that let our AI agents run reliably, securely, and at scale in production.

Your Impact
Infrastructure & Platform

Design, build, and own the core infrastructure powering our AI agent platform, from data pipelines to production deployment systems.

Build and scale the backend systems that support high-throughput document processing and data extraction workloads.

Cloud Infrastructure and Scalability

Architect and deploy infrastructure on cloud platforms (AWS, GCP, or Azure) with a focus on scalability, reliability, and cost efficiency.

Own containerization and orchestration (Docker, Kubernetes) for all production workloads.

Build and maintain CI/CD pipelines and DevOps practices that let the team ship fast without breaking things.

Data Infrastructure

Design and manage data pipelines to process and analyze large volumes of documents and unstructured data at scale.

Build the infrastructure layer connecting AI agents to databases, vector stores, and enterprise systems (ERP, CRM).

API & Systems Integration

Build and maintain robust, well-documented APIs connecting AI agents with external systems and enterprise software.

Design for reliability: retries, observability, and graceful degradation across distributed systems.

Security and Compliance

Implement authentication and authorization mechanisms (OAuth2, JWT) to secure AI-driven systems.

Ensure compliance with data privacy standards (e.g. GDPR, HIPAA) and drive best practices for secure data handling across the infrastructure.

Monitoring and Optimization

Build observability and monitoring systems to track infrastructure health, performance, and cost.

Continuously optimize system performance for speed, reliability, and cost-efficiency at scale.

Collaboration

Work closely with AI/ML engineers, product, and the founding team to make sure infrastructure decisions support fast iteration and production-grade reliability.

Participate in code reviews, design discussions, and architecture planning to drive infrastructure strategy.
Requirements:
5+ years of experience in backend or infrastructure engineering, ideally supporting production AI/ML systems or high-throughput data pipelines.

Proven track record of building and scaling infrastructure in production environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793026
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
10/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Senior DevOps Engineer who thrives in high-scale, high-performance environments. In this role, you will design and manage mission-critical infrastructure, develop automation tools, and ensure the reliability and scalability of platform. You will work closely with engineering teams to optimize performance, enhance observability, and prevent incidents before they happen.

If you love solving complex infrastructure challenges and building tools that empower engineers, this role is for you!

What Will You Do?
As a Senior DevOps Engineer , you will:

Design, build, and maintain a highly available, scalable, and resilient production infrastructure handling millions of requests per second.
Manage and optimize tens of AWS accounts in a multi-account cloud environment.
Developing AI-driven automation frameworks and tools that more than 400 R&D engineers use while enhancing system reliability and engineering efficiency.
Design, implement, and support the LLM and agentic workflow infrastructure, ensuring its scalability and reliability.
Manage thousands of servers and containers, ensuring seamless infrastructure operations.
Improve and maintain our logging, monitoring, and alerting stacks for enhanced observability.
Troubleshoot and mitigate production incidents, participating in on-call rotations to ensure system health and stability.
Collaborate with engineering teams to optimize performance, reduce latency, and improve deployment processes.
Requirements:
5+ years of experience in building and maintaining production infrastructure at scale.
5+ years of experience in developing automation tools and server-side applications using Python, Go, Ruby, Java, or Node.js.
Strong cloud experience with AWS, Google Cloud, or similar platforms.
Experience with the deployment, management, and leveraging of Large Language Models (LLMs) and agentic infrastructure.
Deep knowledge of Linux systems, including troubleshooting, architecture, and system internals.
Solid understanding of web servers, load balancers, caching systems, relational databases, and networking.
A proactive, problem-solving mindset with a passion for optimizing infrastructure and AI-driven automation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8775600
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
11/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at Fiverr and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the Fiverr Technology team might be right for you. Fiverr's Engineering team is expanding its DevOps capabilities to manage our rapidly growing cloud infrastructure. We're looking for a mid-level DevOps Engineer to join this high-velocity environment, taking ownership of cloud assets, contributing to production stability, and driving key projects. You'll be instrumental in strengthening our team and supporting our continued innovation.

What am I going to do?:

* Independently deploy and maintain robust cloud infrastructure across AWS, GCP, and CloudFlare, leveraging Kubernetes, Terragrunt, and Ansible.
* Contribute to an on-call rotation, ensuring the stability of critical production services including Kafka, RabbitMQ, and various databases, with a focus on proactive incident resolution.
* Lead small to mid-sized projects from conception through completion, utilizing your expertise in CI/CD pipelines (Jenkins, ArgoCD, Argo Workflows) and scripting languages like Python, NodeJS, Go, or Kotlin.
* Partner closely with security and development teams to embed security best practices throughout the entire software development lifecycle.
* Proactively evaluate and implement new tools and technologies to enhance engineering efficiency, security posture, and operational excellence.
* Manage sensitive information securely using tools like HashiCorp Vault and ensure secure configurations for services like Kong & Nginx.

Equal opportunities:
At Fiverr, we know that talent has no single face. We welcome talent from everywhere and everyone because it makes everything we build better. Need accommodations? Just ask. And if this role excites you but you don't tick every box, apply anyway. The best people rarely fit the mold exactly.
Requirements:
* 4+ years of hands-on DevOps / Platform Engineering experience in production environments within a public cloud environment (AWS preferred)
* Strong, production-grade Kubernetes experience (design, deployment, scaling, and troubleshooting) with solid AWS experience (VPC, IAM, EC2, EKS, Load Balancers, DNS)
* Experience designing and operating highly available, scalable infrastructure systems
* Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis)
* Hands-on experience with Infrastructure as Code and configuration management (Terraform required, Terragrunt & Ansible – advantage)
* Experience with Docker and containerized workloads
* 2+ years of experience building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
* Proficiency in Python for automation and strong Linux administration skills
* Experience with monitoring and observability tools (Prometheus, Grafana)
* Development experience and familiarity with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI) – advantage Working with AI At Fiverr, AI is a powerful partner in our engineering workflows. You'll leverage AI-powered tools like GitHub Copilot for code assistance, intelligent security scanning tools, and automation platforms to streamline deployments and monitoring. While AI handles repetitive tasks and provides insights, your critical thinking, architectural decisions, and complex problem-solving remain at the forefront.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8776958
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
31/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Zipher is building the Autonomous Execution Layer for cloud data and AI workloads. Backed by $50M in funding , we dynamically orchestrate clusters, predict bottlenecks, and auto-heal infrastructure in real time — with zero human intervention . Our platform runs in production at global enterprise customers, including Fortune 500 companies , delivering mission-critical resilience and sub-second optimization We are looking for a DevOps Engineer to build and own the platform our autonomous execution engine runs on. You will own how Zipher ships, scales, and stays up — the delivery pipelines, the Kubernetes infrastructure, and the reliability of systems enterprise customers depend on around the clock.
What You’ll Do

* Architect and own GitOps-based delivery — ArgoCD, Helm, Terraform, GitHub Actions — so the engine ships to production safely, many times a day
* Build and operate the Kubernetes platform , including on-demand environments spun up per pull request and torn down automatically
* Design deployment safety into the platform: canary analysis, blue/green rollouts, automated rollback, and an observability stack (Prometheus, Grafana, OpenTelemetry) that surfaces failure first
* Partner closely with backend and data engineers to make production infrastructure secure, reproducible, and cost-aware across AWS accounts and enterprise deployments
* Drive reliability end to end: define SLOs , own incident response, and raise the operational bar for mission-critical services
What We Offer
* Own the platform behind a new category of autonomous cloud infrastructure High ownership from day one : real architectural influence, direct exposure to founders, and responsibility for mission-critical systems
* A small, technical, high-velocity team that values curiosity, speed, rigor, and engineering craftsmanship Top-of-market compensation and meaningful equity

Ready to own the platform that lets an autonomous execution engine run in production? Hit Apply.
Requirements:
What You’ll Bring 6+ years in DevOps, Platform, or Infrastructure Engineering , including ownership of production environments for a real product at scale
* Deep hands-on experience with Kubernetes, Helm, and Terraform , with GitOps (ArgoCD or equivalent) as your default way to ship
* Strong production experience on AWS — EKS, IAM, VPC networking, and managed services such as Lambda, S3, DynamoDB, or Kinesis — plus scripting in Python, Go, or Bash
* Real depth in observability and operations : metrics, tracing, log aggregation, alerting, and SLOs you defined and defended
* A high-agency, engineering-first mindset : you enjoy ambiguous, high-leverage problems and take responsibility for reliability, performance, and security
Nice to Have
* Experience building ephemeral environments with Crossplane, Terraform, or a home-grown control plane
* Experience operating data and streaming infrastructure at scale , such as Kafka, Spark, EMR, MongoDB Atlas, or Snowflake
* Familiarity with security and compliance in a fast-moving startup: SOC 2, SSO/MFA, secret scanning, IaC policy enforcement, and cloud cost optimization
* Experience in an elite IDF technology unit or another high-performance engineering environment
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8804051
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
13/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a talented DevOps Engineer to join our team. As a DevOps Engineer, you will architect and operate scalable, resilient cloud environments, using advanced automation and AI-based tools to improve system health, reduce manual intervention, and optimize infrastructure costs.
Responsibilities:
Own the full lifecycle of the Coin Master platform. Ensure high availability, performance, and cost-efficiency across our multi-region AWS environment.
Design and maintain robust, high-performance infrastructure using IaC. You will focus on building self-healing environments where AI agents continuously monitor and remediate issues in real-time.
Embed security into the infrastructure layer. Leverage AI-assisted tooling to proactively identify and patch vulnerabilities across our Kubernetes clusters and AWS stack.
Build tools and services that empower our engineering team to ship code faster. You are the architect of the platform, ensuring it is intuitive, scalable, and fully automated.
Requirements:
3+ years of DevOps experience. Proven track record in managing production-ready Kubernetes clusters, AWS infrastructure, and high-traffic distributed systems.
Experience integrating AI tools into daily workflows, including hands-on use of AI agents (Cursor, Claude Code or similar) for infrastructure automation, planning, and debugging.
Extensive management experience of multiple AWS accounts spanning various regions, proficient in services like Lambda, CloudFront, SQS, VPC, and IAM.
Strong understanding of cloud cost optimization at scale through architectural design and automated resource management.
Deep proficiency with modern CI/CD platforms (GitHub Actions, Jenkins, ArgoCD).
Advantages:
Experience with LLM orchestration (LangChain, LangGraph) or building internal tools/agents that interact with cloud APIs to automate operational tasks.
Hands-on experience with Kafka (MSK/Apache Kafka) and high-performance caching layers like Redis/Elasticache.
Experience with CloudFlare management and complex edge-computing configurations.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781328
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Join a team of senior engineers operating in a large-scale, multi-cloud production environment supporting tens of thousands of enterprise customers worldwide. This is not a typical SRE role - youll work at the core of a complex, high-impact system alongside experienced DevOps professionals in a fast-paced, cybersecurity-focused organization.
Your Impact:
Own and operate large-scale, global production environments across multiple cloud providers (GCP, AWS, Azure)
Actively monitor, investigate, and resolve incidents triggered by automated alerting systems (PagerDuty / Incident Response)
Drive end-to-end troubleshooting across complex, distributed systems with high context switching
Design, deploy, and improve monitoring and observability systems (e.g., Prometheus, Grafana) - not just react to alerts
Collaborate closely with internal teams (CX, CS, Engineering) to ensure system reliability and performance
Work hands-on with modern DevOps and infrastructure tools including Kubernetes, Terraform, CI/CD pipelines, and GitOps workflows
Develop and maintain automation and tooling (primarily in Python)
Gain deep understanding of system architecture and interconnected services
Contribute to a culture of operational excellence in a high-scale, high-availability environment
On call responsibilities:
Daytime hours (12:00-20:00)
Occasional weekends and holidays (rotation-based).
Requirements:
Your experience:
5+ years of experience in SRE roles in production environments at scale
Strong hands-on experience with Kubernetes and Terraform
Strong hands-on experience with at least one major cloud platform (GCP or AWS required)
Experience building and configuring monitoring systems (e.g., Prometheus, Grafana)
Familiarity with CI/CD and GitOps tools (GitLab CI, GitHub Actions, Jenkins, Flux)
Proficiency in Python for scripting and automation
Strong troubleshooting and problem-solving skills with a passion for incident handling
Ability to work in fast-paced environments with high context switching
Highly responsive, proactive, and ownership-driven
Strong collaboration and communication skills
Curious mindset and eagerness to learn.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769584
סגור
שירות זה פתוח ללקוחות VIP בלבד