דרושים » תוכנה » Senior Site Reliability Engineer (DevTools)

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
30/07/2026
משרה זו סומנה ע"י המעסיק כלא אקטואלית יותר
מיקום המשרה: תל אביב יפו
סוג משרה: משרה מלאה
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
10/09/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We're an SRE team within DevTools, looking for someone ready to help maintain and grow our systems
Your responsibilities will include:

Improving services based on user feedback

Building fault-tolerant, self-healing architecture

Finding ways to speed up our systems and reduce user friction

Modifying well-known closed- and open-source solutions Supporting our users
Requirements:
A combination of SRE and SWE experience (for us that's a 50/50 split). Our code is in Java/Kotlin, Go, Python, and Ruby

An understanding of what's happening under the hood in Unix-like systems and the JVM

A passion for improving the user experience

The ability to adapt quickly on the fly in a fast-changing environment
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8817600
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
10/09/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Senior Site Reliability Engineer (SRE)
Your responsibilities will include:

Ensure fault-tolerance, scale, and uninterrupted operations for the service.
Use cutting-edge cloud technology to solve a variety of infrastructure problems.
Implement and improve CI/CD processes.
We expect you to have:

Solid experience with programming languages (like Go, Python, or C++);
Solid understanding of classic algorithms and data structures;
Commercial experience with and deep understanding of Unix systems and network technology;
Experience with systems for containerization and configuration management (Ansible, Salt, Terraform, Docker, K8s, Helm).
Requirements:
Desire to be involved in backend development;
Experience designing, developing, and running high-load distributed systems;
Commercial experience with a variety of cloud platforms.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8817555
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
26/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are constantly striving to make our systems reliable, scalable, and simple to operate so our services are available to travelers when they need them most. With our continued growth, we have exciting challenges ahead and we're looking for a Senior Site Reliability Engineer to join our team in Tel Aviv. This role blends classic SRE ownership with pragmatic AI SRE work: you will build and operate the platforms, automation, observability, and incident response practices that keep Navan reliable, while helping teams use AI solutions, AI providers, and their APIs safely and dependably.



This is a hands-on engineering role, not a research role. You will partner with product, platform, data, security, support, and incident response teams to make production systems and AI-powered experiences more resilient. You will use software engineering, infrastructure as code, SLOs, telemetry, provider observability, and automation as your main tools, and you will apply AI where it creates measurable reliability value rather than novelty.



This position is based out of our new Tel Aviv office.



What You'll Do:

Support AI-based application solutions where reliability matters. Partner with the development teams building AI-powered travel experiences to support the development and production operation of their solution.
Work with AI solutions, providers, and APIs. Partner with teams integrating AI capabilities and providers, with attention to API reliability, authentication, quotas, rate limits, latency and provider-specific operational constraints.
Troubleshoot AI tools and provider issues. Diagnose failures across AI-powered workflows, provider APIs, configuration, permission errors, degraded responses and related areas.
Operate reliable production platforms. implement and run cloud infrastructure,and help product teams move quickly without compromising reliability.
Improve observability. Build dashboards, alerts, traces, logs, and runbooks that make service health clear, actionable, and tied to SLOs and customer impact.
Apply AI to SRE workflows. Prototype and productionize AI-assisted systems that create effective and efficient operations
Automate operational toil. Create tools, workflows, and automation that remove repetitive manual work and make operational knowledge easier to use.
Requirements:
5+ years of experience as a Senior SRE, Infrastructure Software Engineer, Production Engineer, or DevOps Engineer.
3+ years of experience operating production, 24x7 customer-facing systems.
Hands-on experience delivering production infrastructure, platform tooling, and automation used by engineering teams.
Strong software engineering skills in Python, Go, Java, or a similar language, with a bias toward production-quality code, tests, monitoring, and documentation.
Experience with cloud infrastructure, container orchestration, Linux systems, networking, CI/CD, and infrastructure as code such as Terraform or CloudFormation.
Experience building, tuning, and automating observability systems such as Grafana, Prometheus, New Relic, Datadog, Splunk, or similar tools.
Familiarity with SLOs, incident response, on-call practices, root cause analysis, and blameless postmortems.
Practical experience or strong interest in AI solutions, AI providers, agents, AI APIs, provider integrations, or AI-assisted internal tools.
Ability to troubleshoot AI tools and provider/API issues, including rate limits, quota, auth, permission errors, latency, SDK or API contract changes, content quality issues, and service degradations.
Excellent communication skills and the ability to work with stakeholders and domain experts across the company.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797911
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are on an expedition to find an On-Premise Site Reliability Engineer (SRE)- someone who is passionate about building rock-solid, high-performance infrastructure and bringing order to complex environments. In this role, you will own the end-to-end reliability, automation, and deployment of platform across customer sites, working hands-on with cutting-edge AI, bare-metal, and hybrid cloud architectures.
Youll collaborate closely with Product, R&D, and Architecture teams while serving as the ultimate technical authority for our customer deployments. From designing automated Ansible workflows and mastering Kubernetes to troubleshooting complex network topologies, you will eliminate toil, streamline cluster operations, and ensure every deployment is scalable, seamless, and mission-ready.
:Responsibilities
Lead End-to-End On-Prem & Hybrid Deployments: Own the technical delivery and reliability of platform in close collaboration with Product, R&D, and customer technical teams.
Architect, Execute & Improve K8s Deployments: Take a definitive hands-on role in deploying, configuring, operating, and continuously improving our platform using advanced, enterprise-grade Kubernetes architectures.
Helm Chart Management: Design, modify, and manage Helm charts to package, version, and streamline complex application deployments across different environments.
Drive Automation & Simplification: Design, implement, and maintain robust deployment automation using Ansible. You must have a passion for turning complex manual tasks into reliable, repeatable, single-click operations.
Manage Infrastructure as Code: Utilize Git as the single source of truth to manage configurations, manifests, and automation playbooks, enforcing modern engineering best practices.
Bridge On-Prem and Cloud: Leverage AWS resources (specifically EC2 and S3) for hybrid components, staging environments, or cloud-to-on-prem data flows.
Technical Tier-3 Escalation: Serve as the ultimate technical authority for deployment, Linux networking, and Kubernetes orchestration issues.
Continuous Improvement: Constantly refine our delivery pipelines, optimize bootstrap processes, and create rock-solid technical documentation.
דרישות:
SRE / Delivery Mindset: 3-5 years of hands-on experience in enterprise infrastructure deployment, systems engineering, or an on-prem operational reliability role.
Kubernetes & Helm Expert: Deep, production-grade experience with Kubernetes architecture, deployment, advanced troubleshooting, and CNI networking. Proven working experience creating, maintaining, and deploying applications using Helm charts.
Ansible Mastery: Proven experience writing clean, scalable Ansible roles and playbooks for configuration management, automation, and infrastructure provisioning.
Modern Workflows (Git & AWS): Solid working experience using Git for version control and collaborating on code/infrastructure. Practical experience provisioning and managing AWS resources (EC2 and S3).
Core Systems & Linux: Strong Linux background (Ubuntu) with a deep understanding of system internals, containerized runtimes, and troubleshooting distributed applications.
Solid Networking Knowledge: Hands-on experience with routing, firewalls, and switching topology (mainly Cisco)
Storage Foundations: Working knowledge of storage protocols (iSCSI, SAN, local NVMe) and enterprise storage arrays (like DELL) interacting with Kubernetes Persistent Volumes.
GPU & Accelerated Compute: Working knowledge of managing GPU-enabled Kubernetes nodes, including NVIDIA drivers/runtime and basic troubleshooting.
Air-Gapped Deployments: Experience deploying and maintaining software in air-gapped or offline environments, including registry mirroring and artifact staging.
Problem-Solver: Strong debugging and problem-solving skills in complex, distributed environments with an intense ownership and accountability mindset.
Willingness to Travel: Ready to travel to customer sites for physi המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786710
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
26/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a mission-driven BizOps Engineer to join our R&D team!
We specialize in advanced AI predictive models, focusing on LTV, churn, and intent to drive business growth.
As the BizOps Engineer, you will serve as both a platform architect and a hands-on builder. You will own the AI-native platform architecture and governance for all non-core product functions, including Sales, Operations, HR, Legal, and Finance, while also building, maintaining, and scaling custom AI solutions yourself to address specific departmental needs.
Reporting directly to the VP of R&D, this role begins as a high-impact Individual Contributor (IC) with a clear mandate to build and lead a dedicated engineering team as the function scales. Your primary focus will be designing and deploying infrastructure and security protocols while simultaneously delivering the bespoke AI tools that empower teams and drive operational efficiency.
You will have the autonomy to choose the most effective path forward for any given challenge, whether that means leveraging existing SaaS platforms, integrating open-source projects, or building custom solutions from scratch, always guided by the specific needs and projected ROI of the departments you support.
Responsibilities:
Architect and maintain the internal AI-native platform, while personally building and scaling custom AI solutions and automated workflows for cross-functional teams.
Establish and enforce security protocols, governance frameworks, and best practices to enable safe and compliant AI solution development across the organization.
Identify and scale platform enablement opportunities that allow departments to automate their own workflows using approved AI capabilities.
Monitor and manage organizational compliance for internal AI applications, ensuring data privacy and ethical standards are met at scale.
Requirements:
5+ years in software engineering, DevOps, or business systems engineering.
Deep technical proficiency in AI-native tools (LLMs, RAG, agentic workflows) and hands-on experience building complex API-first integrations and data pipelines.
Strong product sense and intuition for user needs to design ROI-driven internal tools that translate operational bottlenecks into technical requirements.
Ability to operate effectively as an individual contributor today while possessing the vision to build and lead an engineering team as the function scales.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797662
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
As the worlds leading vendor of Cyber Security, facing the most sophisticated threats and attacks, weve assembled a global team of the most driven, creative, and innovative people. At our company, our employees are redefining the security landscape by meeting our customers real-time needs and providing our cutting-edge technologies and services to an ever-growing customer base.
our company Software Technologies has been honored by Time Magazine as one of the Worlds Best Companies and recently Gartner rated our company email security as a market leader for product, detection and innovation. We've also earned a spot on the Forbes list of the Worlds Best Places to Work for five consecutive years (2020-2024) and recognized as one of the Worlds Top Female-Friendly Companies. If you're passionate about making the world a safer place and want to be part of an award-winning company culture, we invite you to join us. our company Harmony Email Security and Collaboration (Previously AVANAN) is a unique email solution that fully secures cloud email and cloud platforms using AI.
we are seeking a promising and talented Chaos Engineer (Chaos DevOps) to join our DevOps group. If you thrive in a fast-paced, dynamic environment, can handle multiple requests simultaneously, and enjoy working independently as part of a cutting-edge DevOps team, this is your opportunity to help make the world a safer place!
Requirements:
Hands-on mindset , we all write code daily!
2+ years of relevant DevOps/SRE/Cloud experience working with CI/CD pipelines for both development and production - must.
1+ years of AWS Cloud experience with high-traffic systems and multiple services, and a working understanding of failure modes and high-availability design (multi-AZ) - must.
Solid scripting skills, with fluency in Python - must.
Experience running experiments, load/failure tests, or building automation that validates system behavior under stress - must.
Experience with containers and orchestration tools (Docker, Kubernetes, or ECS) and an understanding of how they fail and self-heal - must.
Experience building and maintaining CI/CD pipelines with GitHub Actions / Jenkins (workflows, runners, reusable actions) - must.
Familiarity with chaos-engineering principles and tooling (Principles of Chaos, steady-state hypothesis, game days; AWS FIS, Chaos Toolkit, Gremlin, LitmusChaos) - an advantage.
Familiarity with AWS CloudFormation and infrastructure-as-code (Terraform) - an advantage.
Exposure to open-source observability and reliability tooling (Grafana, Prometheus, Nagios, Coralogix, etc.), including OpenTelemetry (OTEL) for distributed tracing, metrics, and instrumentation - an advantage.
Awareness of best practices in security, performance, monitoring, and incident response.
Curiosity and willingness to research, evaluate, and try out new technologies, including running proofs of concept.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786701
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
26/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Senior Cloud Engineer to join our Infrastructure team in Tel Aviv. In this role, you will be a key contributor to our AWS CDK projects, taking end-to-end ownership of high-impact infrastructure shifts like our migration to EKS. This is a hands-on position where youll balance long-term architectural projects with day-to-day support for our R&D teams, ensuring our cloud environment remains scalable, secure, and developer-friendly.

Responsibilities:
As a Senior Cloud Engineer, you will be the backbone of our infrastructure, ensuring that our cloud environment is scalable, secure, and developer-friendly. You will:

Serve as a core contributor to our central AWS CDK (TypeScript) project, ensuring every cloud resource-from S3 buckets to complex networking rules-is defined as code.

Act as a consultant for R&D and platform teams, helping microservice owners deploy new services and spin up sandbox environments while maintaining security standards.

Take end-to-end ownership of high-impact infrastructure shifts, including playing a key role in the transition from ECS Fargate to EKS.

Manage the lifecycle and maintenance of critical infrastructure components, such as security/observability agents and hosted open-source solutions.

Collaborate with our Boston-based teams on shared global projects to ensure architectural alignment and cross-region infrastructure support.

Balance long-term project work with active participation in internal help channels, troubleshooting real-time issues and resolving conflicts in central tools.
Requirements:
Bachelor's degree in Computer Science or equivalent experience.

5+ years of experience in DevOps/Software Engineering working in one or more of the big public clouds (AWS / Azure / GCP).

3+ years of experience managing Infrastructure-as-Code (Terraform / AWS CDK / CloudFormation).

Basic understanding of Cloud Networking concepts (VPC peering, Subnetting, Load Balancing, DNS).

Familiarity with Kubernetes components (Control Plane, Worker Nodes, ETCD) and standard maintenance routines (upgrades, scaling, node patching).

Knowledge of AWS common best practices-design, implement, and maintain scalable and reliable cloud infrastructure solutions.

Hands-on coding/scripting skills in TypeScript/Python/Golang and Linux Bash.

Experience with microservices methodology, Git, and CI/CD (preferably GitHub Actions & ArgoCD).

Experience maintaining infrastructure that requires high availability and security.

Strong written and verbal communication skills in both English and Hebrew.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797851
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Site Reliability Engineer at our company 911, you'll own the infrastructure that keeps our platform reliable, scalable, and secure - work that directly supports mission-critical 911 systems used by public safety agencies. You'll drive infrastructure-as-code practices across AWS, lead observability efforts through Datadog, and bring modern AI-assisted engineering approaches into how the team builds and operates.
What You'll Do
Own and evolve AWS infrastructure using Infrastructure-as-Code (Terraform / Terragrunt)
Architect and scale AWS environments
Deploy, scale, and manage containerized workloads using Kubernetes and Docker; contribute to HA/DR architecture and platform strategy
Lead deployment and release processes using Argo (reference JD also names Bitbucket, Jenkins as part of the CI/CD toolset).
Define and enforce SLOs, SLIs, and error budgets; drive toil reduction across the platform
Drive full utilization of Datadog for monitoring, dashboards, and alerting across the platform (reference JD also names Prometheus, Grafana as potential observability tooling)
Build self-service internal developer platforms that empower teams to ship faster.
Take end-to-end ownership of infrastructure projects - define success criteria, execute, and measure outcomes.
Partner cross-functionally with engineering teams (e.g., network engineering, Dev owners) on long-term technical planning.
Bring AI-assisted engineering practices (e.g., Claude, MCP integrations) into daily workflows to improve team efficiency
Document work and provide cross-training to peers.
Resolve JIRA tickets across Cloud, CI/CD, deployments, and monitoring.
Requirements:
At least 6 years of experience as a DevOps/SRE engineer in a cloud environment
Hands-on, production-level AWS experience.
Hands-on production experience with Kubernetes and containerization
Experience with Terraform/Terragrunt (or similar Infrastructure-as-Code tools) - required
Strong Bash scripting skills
Deep understanding of SRE principles: SLOs, SLIs, error budgets, toil reduction, blameless post-mortems
Strong incident management / on-call experience
Solid understanding of APIs, microservices, and distributed systems
Demonstrated experience leading a project end-to-end, from defining success criteria through delivery and measurement
Communicates effectively across teams and can drive long-term technical planning
Practical experience with AI-assisted engineering tools (e.g., Claude, Cursor) and MCP-style integrations is a strong plus
Experience building AI/ML infrastructure (model deployment, inference pipelines)-plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8796929
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
17/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a DevOps Engineer to join our engineering team. The ideal candidate has strong hands-on experience with cloud-native infrastructure, a GitOps mindset, and the ability to independently research and learn new tools and techniques in a fast-moving, large-scale Kubernetes environment.
Responsibilities
Design, operate, and troubleshoot Kubernetes clusters (EKS/AKS) at scale
Manage application delivery and infrastructure using GitOps tools (ArgoCD/Flux)
Build and maintain Helm charts and Kustomize overlays for multi-environment deployments
Provision and manage cloud infrastructure using Crossplane and/or Terraform
Own and optimize CI/CD pipelines (GitHub Actions, GitLab CI) for build, test, and deployment workflows
Maintain and extend observability stacks (Prometheus, Grafana, alerting rules, dashboards)
Write automation scripts and tooling in Python, Go, or Bash to streamline operations
Support AWS infrastructure across multiple accounts/regions (networking, IAM, compute, storage)
Participate in on-call rotation, troubleshoot production incidents, and drive root-cause analysis
Collaborate with platform, security, and application engineering teams on infrastructure design and reliability improvements.
Requirements:
2-3 years of experience in DevOps, SRE, platform engineering, or a related role
Solid working knowledge of Kubernetes and Helm in production environments
Experience with Crossplane and/or Terraform for infrastructure as code
Proficiency with AWS services (EKS, IAM, VPC, networking, compute)
Hands-on experience with GitOps tools such as ArgoCD or Flux
Scripting ability in Python, Go, or Bash for automation and tooling
Experience with Prometheus and Grafana for monitoring and alerting
Practical experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI)
Strong self-learning ability, capable of independently researching unfamiliar technologies, reading documentation, and applying findings without handholding
Solid troubleshooting skills across networking, compute, and distributed systems
Nice to Have
Programming proficiency in Go or Python (beyond scripting)
Experience with service mesh technologies (Istio or similar)
Exposure to multi-cloud environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785730
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Tech is at the center of everything we do at our company and we're looking for people who are builders at their core. From developers to visionaries and everything in between, we want minds who aren't just interested in putting the pieces together but who can find new ways to innovate. Solid communication, creative problem solving and business understanding are all prerequisites. So, if you're tech savvy, inquisitive, and ready to take the road less traveled, the company Technology team might be right for you.
our company's Engineering team is expanding its DevOps capabilities to manage our rapidly growing cloud infrastructure. We're looking for a mid-level DevOps Engineer to join this high-velocity environment, taking ownership of cloud assets, contributing to production stability, and driving key projects. You'll be instrumental in strengthening our team and supporting our continued innovation.
What am I going to do?
Independently deploy and maintain robust cloud infrastructure across AWS, GCP, and CloudFlare, leveraging Kubernetes, Terragrunt, and Ansible.
Contribute to an on-call rotation, ensuring the stability of critical production services including Kafka, RabbitMQ, and various databases, with a focus on proactive incident resolution.
Lead small to mid-sized projects from conception through completion, utilizing your expertise in CI/CD pipelines (Jenkins, ArgoCD, Argo Workflows) and scripting languages like Python, NodeJS, Go, or Kotlin.
Partner closely with security and development teams to embed security best practices throughout the entire software development lifecycle.
Proactively evaluate and implement new tools and technologies to enhance engineering efficiency, security posture, and operational excellence.
Manage sensitive information securely using tools like HashiCorp Vault and ensure secure configurations for services like Kong & Nginx.
Requirements:
4+ years of hands-on DevOps / Platform Engineering experience in production environments within a public cloud environment (AWS preferred)
Strong, production-grade Kubernetes experience (design, deployment, scaling, and troubleshooting) with solid AWS experience (VPC, IAM, EC2, EKS, Load Balancers, DNS)
Experience designing and operating highly available, scalable infrastructure systems
Experience with managed and distributed databases (AWS Aurora, RDS, MongoDB, Redis)
Hands-on experience with Infrastructure as Code and configuration management (Terraform required, Terragrunt & Ansible - advantage)
Experience with Docker and containerized workloads
2+ years of experience building and maintaining CI/CD pipelines (Jenkins, GitHub Actions)
Proficiency in Python for automation and strong Linux administration skills
Experience with monitoring and observability tools (Prometheus, Grafana)
Development experience and familiarity with GenAI platforms (AWS Bedrock, Vertex AI, OpenAI) - advantage.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8806987
סגור
שירות זה פתוח ללקוחות VIP בלבד