דרושים » תוכנה » DevOps Engineer (Agentic Search)

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 2 שעות
Location: Tel Aviv-Yafo
Job Type: Full Time
We're building the infrastructure layer for agentic web interaction at scale. Our API is designed from the ground up to power Retrieval-Augmented Generation (RAG) and real-time reasoning in AI systems. By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that are not only intelligent - but also informed.

We work with some of the most innovative teams in AI - from small startups shaping the ecosystem to the largest enterprises deploying AI at scale. Whether it's powering sales assistants, research copilots, or internal knowledge tools, we're the missing link between LLMs and the real world.

The Role: DevOps Engineer
Managing Kubernetes clusters across multiple environments and regions

Owning infrastructure as code for all resources

Maintaining and improving CI/CD pipelines and GitOps-based deployments

Maintaining and optimize real-time data pipelines that process billions of events per day across distributed queues and stream processors

Building out monitoring, alerting, and observability

Debugging production issues across services

Managing cloud costs and capacity planning

Working closely with a small engineering team - you'd own infra, not a slice of it
Requirements:
3+ years in a DevOps or platform engineering role, working in production environments

Proven experience designing and operating large-scale, distributed systems, with a solid understanding of API design, reliability, and performance at scale

Strong Kubernetes experience in a managed cloud environment

Proficiency with infrastructure as code (Terraform or similar)

Experience with GitOps-based deployment workflows

Built or maintained observability stacks (logging, metrics, alerting)

Experience handling production incidents calmly and methodically
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8818172
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have:
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796409
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
17/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a DevOps Engineer to join our engineering team. The ideal candidate has strong hands-on experience with cloud-native infrastructure, a GitOps mindset, and the ability to independently research and learn new tools and techniques in a fast-moving, large-scale Kubernetes environment.
Responsibilities
Design, operate, and troubleshoot Kubernetes clusters (EKS/AKS) at scale
Manage application delivery and infrastructure using GitOps tools (ArgoCD/Flux)
Build and maintain Helm charts and Kustomize overlays for multi-environment deployments
Provision and manage cloud infrastructure using Crossplane and/or Terraform
Own and optimize CI/CD pipelines (GitHub Actions, GitLab CI) for build, test, and deployment workflows
Maintain and extend observability stacks (Prometheus, Grafana, alerting rules, dashboards)
Write automation scripts and tooling in Python, Go, or Bash to streamline operations
Support AWS infrastructure across multiple accounts/regions (networking, IAM, compute, storage)
Participate in on-call rotation, troubleshoot production incidents, and drive root-cause analysis
Collaborate with platform, security, and application engineering teams on infrastructure design and reliability improvements.
Requirements:
2-3 years of experience in DevOps, SRE, platform engineering, or a related role
Solid working knowledge of Kubernetes and Helm in production environments
Experience with Crossplane and/or Terraform for infrastructure as code
Proficiency with AWS services (EKS, IAM, VPC, networking, compute)
Hands-on experience with GitOps tools such as ArgoCD or Flux
Scripting ability in Python, Go, or Bash for automation and tooling
Experience with Prometheus and Grafana for monitoring and alerting
Practical experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI)
Strong self-learning ability, capable of independently researching unfamiliar technologies, reading documentation, and applying findings without handholding
Solid troubleshooting skills across networking, compute, and distributed systems
Nice to Have
Programming proficiency in Go or Python (beyond scripting)
Experience with service mesh technologies (Istio or similar)
Exposure to multi-cloud environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785730
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
23/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are a well-funded, early-stage startup looking for a talented and motivated Backend Engineer specializing in infrastructure to join our founding team. The focus of this role is to build and scale the infrastructure that powers autonomous AI agents automating complex enterprise workflows. You will own the systems, pipelines, and platforms that let our AI agents run reliably, securely, and at scale in production.

Your Impact
Infrastructure & Platform

Design, build, and own the core infrastructure powering our AI agent platform, from data pipelines to production deployment systems.

Build and scale the backend systems that support high-throughput document processing and data extraction workloads.

Cloud Infrastructure and Scalability

Architect and deploy infrastructure on cloud platforms (AWS, GCP, or Azure) with a focus on scalability, reliability, and cost efficiency.

Own containerization and orchestration (Docker, Kubernetes) for all production workloads.

Build and maintain CI/CD pipelines and DevOps practices that let the team ship fast without breaking things.

Data Infrastructure

Design and manage data pipelines to process and analyze large volumes of documents and unstructured data at scale.

Build the infrastructure layer connecting AI agents to databases, vector stores, and enterprise systems (ERP, CRM).

API & Systems Integration

Build and maintain robust, well-documented APIs connecting AI agents with external systems and enterprise software.

Design for reliability: retries, observability, and graceful degradation across distributed systems.

Security and Compliance

Implement authentication and authorization mechanisms (OAuth2, JWT) to secure AI-driven systems.

Ensure compliance with data privacy standards (e.g. GDPR, HIPAA) and drive best practices for secure data handling across the infrastructure.

Monitoring and Optimization

Build observability and monitoring systems to track infrastructure health, performance, and cost.

Continuously optimize system performance for speed, reliability, and cost-efficiency at scale.

Collaboration

Work closely with AI/ML engineers, product, and the founding team to make sure infrastructure decisions support fast iteration and production-grade reliability.

Participate in code reviews, design discussions, and architecture planning to drive infrastructure strategy.
Requirements:
5+ years of experience in backend or infrastructure engineering, ideally supporting production AI/ML systems or high-throughput data pipelines.

Proven track record of building and scaling infrastructure in production environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793026
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
11/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at Fiverr and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the Fiverr Technology team might be right for you. Fiverr's Engineering team is expanding its DevOps capabilities to manage our rapidly growing cloud infrastructure. We're looking for a mid-level DevOps Engineer to join this high-velocity environment, taking ownership of cloud assets, contributing to production stability, and driving key projects. You'll be instrumental in strengthening our team and supporting our continued innovation.

What am I going to do?:

* Independently deploy and maintain robust cloud infrastructure across AWS, GCP, and CloudFlare, leveraging Kubernetes, Terragrunt, and Ansible.
* Contribute to an on-call rotation, ensuring the stability of critical production services including Kafka, RabbitMQ, and various databases, with a focus on proactive incident resolution.
* Lead small to mid-sized projects from conception through completion, utilizing your expertise in CI/CD pipelines (Jenkins, ArgoCD, Argo Workflows) and scripting languages like Python, NodeJS, Go, or Kotlin.
* Partner closely with security and development teams to embed security best practices throughout the entire software development lifecycle.
* Proactively evaluate and implement new tools and technologies to enhance engineering efficiency, security posture, and operational excellence.
* Manage sensitive information securely using tools like HashiCorp Vault and ensure secure configurations for services like Kong & Nginx.

Equal opportunities:
At Fiverr, we know that talent has no single face. We welcome talent from everywhere and everyone because it makes everything we build better. Need accommodations? Just ask. And if this role excites you but you don't tick every box, apply anyway. The best people rarely fit the mold exactly.
Requirements:
* 4+ years of hands-on DevOps / Platform Engineering experience in production environments within a public cloud environment (AWS preferred)
* Strong, production-grade Kubernetes experience (design, deployment, scaling, and troubleshooting) with solid AWS experience (VPC, IAM, EC2, EKS, Load Balancers, DNS)
* Experience designing and operating highly available, scalable infrastructure systems
* Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis)
* Hands-on experience with Infrastructure as Code and configuration management (Terraform required, Terragrunt & Ansible – advantage)
* Experience with Docker and containerized workloads
* 2+ years of experience building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
* Proficiency in Python for automation and strong Linux administration skills
* Experience with monitoring and observability tools (Prometheus, Grafana)
* Development experience and familiarity with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI) – advantage Working with AI At Fiverr, AI is a powerful partner in our engineering workflows. You'll leverage AI-powered tools like GitHub Copilot for code assistance, intelligent security scanning tools, and automation platforms to streamline deployments and monitoring. While AI handles repetitive tasks and provides insights, your critical thinking, architectural decisions, and complex problem-solving remain at the forefront.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8776958
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Join a team of senior engineers operating in a large-scale, multi-cloud production environment supporting tens of thousands of enterprise customers worldwide. This is not a typical SRE role - youll work at the core of a complex, high-impact system alongside experienced DevOps professionals in a fast-paced, cybersecurity-focused organization.
Your Impact:
Own and operate large-scale, global production environments across multiple cloud providers (GCP, AWS, Azure)
Actively monitor, investigate, and resolve incidents triggered by automated alerting systems (PagerDuty / Incident Response)
Drive end-to-end troubleshooting across complex, distributed systems with high context switching
Design, deploy, and improve monitoring and observability systems (e.g., Prometheus, Grafana) - not just react to alerts
Collaborate closely with internal teams (CX, CS, Engineering) to ensure system reliability and performance
Work hands-on with modern DevOps and infrastructure tools including Kubernetes, Terraform, CI/CD pipelines, and GitOps workflows
Develop and maintain automation and tooling (primarily in Python)
Gain deep understanding of system architecture and interconnected services
Contribute to a culture of operational excellence in a high-scale, high-availability environment
On call responsibilities:
Daytime hours (12:00-20:00)
Occasional weekends and holidays (rotation-based).
Requirements:
Your experience:
5+ years of experience in SRE roles in production environments at scale
Strong hands-on experience with Kubernetes and Terraform
Strong hands-on experience with at least one major cloud platform (GCP or AWS required)
Experience building and configuring monitoring systems (e.g., Prometheus, Grafana)
Familiarity with CI/CD and GitOps tools (GitLab CI, GitHub Actions, Jenkins, Flux)
Proficiency in Python for scripting and automation
Strong troubleshooting and problem-solving skills with a passion for incident handling
Ability to work in fast-paced environments with high context switching
Highly responsive, proactive, and ownership-driven
Strong collaboration and communication skills
Curious mindset and eagerness to learn.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769584
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
13/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for a talented DevOps Engineer to join our team. As a DevOps Engineer, you will architect and operate scalable, resilient cloud environments, using advanced automation and AI-based tools to improve system health, reduce manual intervention, and optimize infrastructure costs.
Responsibilities:
Own the full lifecycle of the Coin Master platform. Ensure high availability, performance, and cost-efficiency across our multi-region AWS environment.
Design and maintain robust, high-performance infrastructure using IaC. You will focus on building self-healing environments where AI agents continuously monitor and remediate issues in real-time.
Embed security into the infrastructure layer. Leverage AI-assisted tooling to proactively identify and patch vulnerabilities across our Kubernetes clusters and AWS stack.
Build tools and services that empower our engineering team to ship code faster. You are the architect of the platform, ensuring it is intuitive, scalable, and fully automated.
Requirements:
3+ years of DevOps experience. Proven track record in managing production-ready Kubernetes clusters, AWS infrastructure, and high-traffic distributed systems.
Experience integrating AI tools into daily workflows, including hands-on use of AI agents (Cursor, Claude Code or similar) for infrastructure automation, planning, and debugging.
Extensive management experience of multiple AWS accounts spanning various regions, proficient in services like Lambda, CloudFront, SQS, VPC, and IAM.
Strong understanding of cloud cost optimization at scale through architectural design and automated resource management.
Deep proficiency with modern CI/CD platforms (GitHub Actions, Jenkins, ArgoCD).
Advantages:
Experience with LLM orchestration (LangChain, LangGraph) or building internal tools/agents that interact with cloud APIs to automate operational tasks.
Hands-on experience with Kafka (MSK/Apache Kafka) and high-performance caching layers like Redis/Elasticache.
Experience with CloudFlare management and complex edge-computing configurations.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781328
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at our company and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the company Technology team might be right for you.
We're looking for a Senior DevOps Engineer with a strong security orientation to join our company's DevOps team.
our company's platform runs at significant scale, and our DevOps team sits at the core of keeping it fast, reliable, and secure. As we grow, so does the scope of what we own - we need an engineer who can step in as a strong pillar of our company's production ecosystem.
This is a hands-on role for someone who takes security seriously, moves fast, and knows how to get things done in a complex, high-scale environment.
our company's Technology Stack:
AWS, GCP, CLoudFlare ,Kubernetes, Terragrunt, Ansible, Jenkins, ArgoCD, Argo Workflows, Kong & Nginx, HashiCorp Vault, Kafka, RabbitMQ, Mongodb, Aurora Postgresql & Mysql, Prometheus, Grafana, VictoriaMetrics
Programming languages: Python, NodeJS, Go, Kotlin
What am I going to do?
Full Ownership: Drive infrastructure initiatives through their entire lifecycle, taking accountability from initial design to delivery and long-term operations.
Kubernetes Orchestration: Architect, implement, and maintain production-grade Kubernetes clusters, ensuring they remain scalable and resilient under high-scale demand.
Cloud Architecture: Design and manage robust AWS environments, including VPC, IAM, and EKS, to support a highly available platform architecture.
Infrastructure as Code: Utilize Terraform to build and evolve our environment, applying configuration management principles to all IaC workflows.
CI/CD Excellence: Support and improve our deployment pipelines using Jenkins and GitHub Actions to maintain a fast development velocity.
Observability: Implement comprehensive monitoring solutions with Prometheus and Grafana to ensure deep visibility into platform health.
Operational Resilience: Join the DevOps on-call rotation, taking responsibility for mitigating production issues and maintaining site reliability.
Tooling & Innovation: Continuously evaluate and adopt tools - security and otherwise - that raise the bar on engineering efficiency and security posture at our company.
Requirements:
6+ years of hands-on DevOps / Platform Engineering experience in large-scale production environments on a public cloud (AWS preferred).
Proven leadership mindset - able to own projects end to end and be accountable for outcomes.
Strong, production-grade Kubernetes experience across design, deployment, scaling, and troubleshooting.
Solid AWS experience with VPC, IAM, EC2, EKS, Load Balancers, and DNS.
Experience designing and operating highly available, scalable infrastructure systems.
Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis).
Hands-on experience with Infrastructure as Code and configuration management (Terraform required; Terragrunt and Ansible a plus).
Experience with Docker and containerized workloads.
2+ years building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
Experience with monitoring and observability tools (Prometheus, Grafana).
Experience with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8806869
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a talented and motivated Software Engineer with hands-on experience building and operating multi-agent AI systems in production to join our Automation Platform (DAP) team.
The team develops automation tools, orchestration capabilities, and intelligent platforms that simplify the deployment, management, troubleshooting, and optimization of large-scale network and AI infrastructure environments.
You will work on the design and development of AI-powered systems that bridge networking, automation, observability, and distributed infrastructure - running on Kubernetes at scale. Our stack includes LangGraph, Langfuse, RAG pipelines, MCP, and agent-to-agent (A2A) communication patterns.
This role combines strong software engineering with practical AI application development, with a sharp focus on production hardening, tracing, evaluation, and safety of agentic systems - not model training or research prototypes.
Requirements:
5+ years of hands-on software engineering experience building production-grade backend services, APIs, or AI-powered systems.
Proven production experience with multi-agent AI systems: deployment, tracing, guardrails, hardening, and incident management.
Hands-on experience with agentic frameworks such as LangGraph, CrewAI, Google ADK, AutoGen, or equivalent.
Experience building and running evaluation pipelines for agentic solutions - including trajectory tracing, ground truth validation, and harshness/quality scoring.
Strong Python proficiency: comfortable building scalable backend services using gRPC and REST APIs.
Solid understanding of distributed systems: fault tolerance, consistency models, service communication, and operational challenges at scale.
Hands-on Kubernetes experience: deploying and operating containerized services, managing workloads, config, and scaling in production clusters.
Practical experience with embeddings, vector databases, and semantic retrieval systems in production.
Practical experience with RAG pipelines, LLM API integration, structured outputs, and tool calling in production environments including building and serving MCP servers at scale.
Working knowledge of SQL and/or NoSQL databases, schema design, and query optimization.
Strong debugging skills across application logic, APIs, data, and AI agent behavior.
Strong communication skills and a bias toward ownership and delivery.
Nice to Have:
Familiarity with Langfuse/Arize Pheonix.
Familarity with A2A & A2UI protocols.
Experience with network automation, orchestration, or configuration management (Ansible, Terraform, NETCONF, gNMI, or similar).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764207
סגור
שירות זה פתוח ללקוחות VIP בלבד