דרושים » מחשבים ורשתות » AI Ops Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 16 שעות
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are leading creative technology company on a mission to empower creators and brands to bring their vision to life with video. Offering cutting-edge AI tools and models for image, video, and voiceover creation, alongside high-quality creative assets and powerful editing tools, Artlist enables creators to stay on trend, and achieve their creative goals. Trusted by over 30 million creators worldwide and top brands including Google, Amazon, Microsoft, and Versace, Artlist provides a seamless, subscription-based platform with a global license, giving creators everything they need to produce professional video content efficiently. For more information, visit artlist.io While Artlist builds cutting-edge AI products to empower creators globally, this role is the other side of that coin. You will build the AI the company runs on internally. You will build the layer that enables, guides, and controls AI efforts across the organization- the identity, access, and audit backbone for internal AI - along with the assistants and automations that run safely on top of it. This is a hands-on infrastructure role with a security, fin-ops, and DevOps reflex built in. You will help teams move quickly as you will define the infra they should use and we manage this infra. What you'll own
* Platform ownership. Own the full lifecycle of internal AI and automation platforms - runtime infrastructure, CI/CD, and identity controls, end to end.
* Identity, cost & control. Implement identity, boundaries, technology, cost, efficiency.
* Workflow architecture. Design and service automations and custom AI assistants for internal teams (Support, Marketing, Legal, Finance, HR, and beyond)to help solve real operational pain.
* Research & tech scouting. Continuously evaluate emerging models, agent frameworks, and SaaS tools and onboard and implement them in Artlist.
* Standardization. Establish reusable design patterns, integrations, and prompt libraries so teams can automate their own work safely, efficiently and smoothly, for the best experience..
Requirements:
* 3+ years in DevOps, SRE, or Production Engineering, with a genuine operational and reliability mindset.
* Coding and scripting skills automation, API integrations, and internal tooling.
* Solid DevOps foundation: Infrastructure as Code (Terraform), containerization (Docker), CI/CD pipelines, and secrets management.
* Identity & API fundamentals. Practical, hands-on knowledge of OAuth2, SSO/SCIM, REST APIs, webhooks, and scoped service accounts.
* Practical LLM experience. Hands-on work with LLMs (Claude, ChatGPT APIs, prompt engineering, writing skills, MCPs and Plugins).
* Cross-functional communication. You can sit with a non-technical stakeholder, understand their actual pain, and turn it into a clean, safe automation.
* Cloud Infrastructure: Hands-on experience managing and architecting cloud infrastructure, with specific proficiency in AWS. Nice to have
* Experience with workflow and automation platforms (n8n, Temporal, Workato, make, Hermes).
* Exposure to AI gateway or routing layers (LiteLLM, OpenRouter, Bedrock).
* Familiarity with vector databases, RAG concepts, or LLM observability tools (Langfuse, Helicone).
* A background in IT infrastructure or corporate systems automation.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8777712
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
6 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a Senior Infrastructure Engineer who views "Infrastructure as Software." In 2026, we dont just manage servers; we build high-performance environments that allow multi-agent systems to operate at scale.
You will be a core member of the R&D team, blending deep DevOps expertise with the coding rigor of a Backend Engineer. You arent just "configuring" AWS; you are architecting the distributed systems and data pipelines that power our autonomous security brain. Your mission is to ensure that while our agents are evolving and taking actions, our underlying platform remains immutable, observable, and infinitely scalable.
What You'll Do
Design, build, and operate our company's cloud infrastructure using AWS, Kubernetes, and Infrastructure as Code.
Build internal tools and platform services using Python and Go to improve developer productivity and system reliability.
Own infrastructure automation with Terraform, Pulumi, and modern cloud-native tooling.
Partner closely with Backend, Data Science, and Security Engineering teams to build scalable, reliable platforms.
Improve observability, monitoring, and incident response across distributed production systems.
Design and optimize infrastructure for performance, scalability, security, and cost efficiency.
Help shape engineering best practices, platform architecture, and developer experience as our company continues to grow.
Requirements:
5+ years of experience in Infrastructure, DevOps, Platform Engineering, or Backend Engineering.
Strong software engineering skills with hands-on experience building production systems in Python or Go.
Deep hands-on experience with AWS, including services such as EKS, RDS, VPC, and IAM.
Strong experience designing, operating, and scaling production Kubernetes environments.
Experience with Infrastructure as Code, CI/CD, GitOps, and modern cloud-native development practices.
A systems mindset with the ability to solve architectural challenges across infrastructure and application layers.
Comfortable using modern AI-powered developer tools and agentic workflows to improve engineering productivity.
The company Mindset: You take ownership, act with accountability, collaborate openly, and focus on delivering meaningful impact. You thrive in fast-moving environments, embrace ambiguity, and enjoy solving hard problems together.
Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
Full professional fluency (written and verbal) in both Hebrew and English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Senior DevOps Engineer who thrives in high-scale, high-performance environments. In this role, you will design and manage mission-critical infrastructure, develop automation tools, and ensure the reliability and scalability of platform. You will work closely with engineering teams to optimize performance, enhance observability, and prevent incidents before they happen.

If you love solving complex infrastructure challenges and building tools that empower engineers, this role is for you!

What Will You Do?
As a Senior DevOps Engineer , you will:

Design, build, and maintain a highly available, scalable, and resilient production infrastructure handling millions of requests per second.
Manage and optimize tens of AWS accounts in a multi-account cloud environment.
Developing AI-driven automation frameworks and tools that more than 400 R&D engineers use while enhancing system reliability and engineering efficiency.
Design, implement, and support the LLM and agentic workflow infrastructure, ensuring its scalability and reliability.
Manage thousands of servers and containers, ensuring seamless infrastructure operations.
Improve and maintain our logging, monitoring, and alerting stacks for enhanced observability.
Troubleshoot and mitigate production incidents, participating in on-call rotations to ensure system health and stability.
Collaborate with engineering teams to optimize performance, reduce latency, and improve deployment processes.
Requirements:
5+ years of experience in building and maintaining production infrastructure at scale.
5+ years of experience in developing automation tools and server-side applications using Python, Go, Ruby, Java, or Node.js.
Strong cloud experience with AWS, Google Cloud, or similar platforms.
Experience with the deployment, management, and leveraging of Large Language Models (LLMs) and agentic infrastructure.
Deep knowledge of Linux systems, including troubleshooting, architecture, and system internals.
Solid understanding of web servers, load balancers, caching systems, relational databases, and networking.
A proactive, problem-solving mindset with a passion for optimizing infrastructure and AI-driven automation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8775605
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking a skilled and motivated DevOps engineer with deep familiarity in the streaming ecosystem to join our elite infrastructure team. If you're excited by the challenge of operating mission-critical systems at scale and optimizing the developer experience through automation and tooling, wed love to hear from you.
What you will do:
Automate Deployment and Operation:
Oversee deployment of Kafka and RabbitMQ clusters (including Confluent Cloud & CFK). Build automation pipelines to ensure repeatability and resiliency across environments.
Monitor and Support Production Systems:
Own production stability of global Kafka clusters. Handle on-call rotations, incident management, troubleshooting, and scaling challenges.
Improve Infrastructure Observability
Build and maintain observability systems: dashboards, alerting pipelines, metrics collection (Prometheus, Grafana, etc.).
Optimize System Performance:
Collaborate with peers on benchmarking and optimization initiatives. Work on tuning Kafka brokers, cluster configurations, and runtime parameters.
Provide Developer Support and Training (Infra-focused)
Help developers configure topics, quotas, and consumers appropriately. Train service owners to interpret monitoring data and avoid pitfalls.
Develop and Maintain Infrastructure:
Contribute to building infrastructure tools and scripts (IaC, Helm charts, etc.) that make provisioning and managing clusters reliable and efficient.
Secure Infrastructure Access:
Configure and maintain secure access patterns across streaming infrastructure, ensuring proper authentication and role-based access controls are enforced for both developers and services.
Requirements:
8+ years of experience in DevOps, SRE, or Infrastructure Engineering roles.
Deep hands-on Kafka experience, including deploying, maintaining, scaling, and monitoring clusters.
Experience with RabbitMQ.
Extensive experience with Docker, Kubernetes, Helm, and GitOps-style deployments.
Infrastructure as Code experience (Terraform, Pulumi, etc.).
Strong skills in scripting and automation (Python, Bash, etc.).
Familiarity with Confluent Cloud, Confluent for Kubernetes, and similar tools.
Solid understanding of authentication and authorization mechanisms in distributed systems.
Production support mindset - with proven troubleshooting and incident resolution history.
Collaboration and communication skills - especially with dev teams depending on platform support.
Experience with Istio Service Mesh (bonus).
Experience with GovCloud (bonus).
Bonus Qualities:
Mentorship and leadership experience in infrastructure or SRE teams.
Contributions to automation or monitoring open-source tooling.
Active participant in SRE or DevOps communities.
Conference speaker or internal tech trainer.
Technical writing about infrastructure automation or reliability.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8754250
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
6 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior SRE Engineer, you will be a key player in ensuring the reliability, scalability, and performance of our critical IT infrastructure. You will leverage SRE principles and an automation-first mindset to build and maintain resilient hybrid cloud environments. This role is ideal for a candidate who thrives in a fast-paced, innovative setting and is passionate about solving complex challenges with cutting-edge technology.
Key Responsibilities
Provision, configure, and support resilient hybrid cloud deployment architectures using an Infrastructure-as-Code framework.
Proactively collaborate with development teams to ensure new applications are production-ready, scalable, and reliable from inception.
Develop and maintain tools and frameworks to automate operational tasks, including deployment, monitoring, and recovery.
Conduct thorough root cause analysis of production issues and implement preventative measures to improve system resilience, demonstrating strong problem-solving skills.
Manage CI/CD platforms, Linux infrastructure, and contribute to capacity planning and operational runbooks.
Design and implement proactive service monitoring, alerting, and trend analysis to maintain service availability and performance SLAs.
Participate in an on-call rotation to support critical applications and services, responding to and resolving incidents efficiently.
Contribute to comprehensive documentation related to infrastructure design, deployment, and operational procedures.
Requirements:
Your Expereience:
6+ years of Devops engineering experience on mission-critical, enterprise-level systems in a hybrid (both cloud and on-prem) environment.
3+ years of hands-on experience with cloud environments, preferably Google Cloud Platform (GCP).
Expertise in configuration management and Infrastructure-as-Code using frameworks such as Terraform and Ansible.
Strong programming/scripting knowledge in languages like Python, Bash, or Go for infrastructure automation.
Demonstrated experience with CI/CD pipelines (e.g., GitHub, Jenkins, Artifactory) and a strong foundation in Linux/Unix administration.
Bachelor's degree in Computer Science, Information Technology, or a related field, or equivalent practical experience.
Preferred Qualifications
Experience with containerization and orchestration technologies, particularly Kubernetes.
Hands-on experience with monitoring and observability tools such as Datadog, Grafana, or Prometheus.
Understanding of networking principles including firewalls, load balancers, and complex network designs.
A curious and positive mindset with a passion for applied learning and challenging existing processes for continuous improvement.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769408
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
03/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an exceptional Platform Engineer who combines deep software engineering with a robust DevOps approach to help propel development infrastructure to the next level. , SPEED is integral to our DNA. AI transforms the way we develop at speed, and our development infrastructure is key to the scale and hypergrowth.



The DevX team is a multi-disciplinary group of DevOps and development experts focused on building and operating internal tools, infrastructure, and services to make engineering easy and intuitive SPEED.



Your mission is to build a top-tier Continuous Integration and Delivery (CI/CD) platform. You will own both the application services and the infrastructure that supports them, ensuring speed, stability, and reliability for our entire R&D organization.



Key Focus Areas & What You'll Do



Improve CI/CD pipelines and CI workflows (e.g., GitHub Actions) with a focus on speed and reliability.
Reduce CI flakiness and improve overall pipeline stability through systematic triage and root-cause analysis.
Shorten developer feedback loops by optimizing test strategy, pipeline consistency, and local development workflows.
Strengthen CD and release processes for secure, repeatable, and fast deployments.
Promote best practices such as GitOps and progressive delivery where they fit.
Track and improve delivery metrics, with emphasis on Lead Time for Changes and Deployment Frequency.
Partner with engineering teams to identify friction and deliver scalable automation and paved paths.
Requirements:
What Were Looking For:



3+ years of experience as a Platform Engineer or a strong Backend Engineer with a DevOps focus.
Strong knowledge of CI/CD pipelines, versioning, and release management.
Proven experience with modern build tools (e.g., Bazel, SWC) and managing CI workflows (e.g., GitHub Actions, GitLab CI).
Good skills with infrastructure-as-code tools like Terraform and Helm.
Hands-on experience with Kubernetes, including containers (Docker), Kafka and cloud providers (AWS, GCP, or Azure).
Strong programming skills in Python, Go, TypeScript, or another modern backend language.
Experience with monorepo tools (Lerna, NX, PNPM) is a big plus.
Familiarity with GitOps workflows using tools like ArgoCD or Flux.
A strong sense of ownership and a passion for making systems reliable, scalable, and improving the developer experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8765844
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
06/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for an experienced Senior DevOps Engineer to join our DevOps team in the Posture R&D Group, who is passionate about software design, development and deployment. The role goes beyond traditional DevOps - it focuses on building the infrastructure and platforms that enable AI models and autonomous agents to run in production at scale, across both cloud and on-prem environments. The job involves writing production-grade modern DevOps solutions that will be shipped to the cloud and on-prem solutions, while working with cutting-edge technologies and architectures that push the boundaries of AI-driven cybersecurity systems.
Responsibilities
Build the best solutions for our production platform, enabling high-scale, AI-driven systems and agents to operate reliably in production-scale environments
Everything as a code approach (IaC): Run our infrastructure with a wide range of technologies including Terraform, and Kubernetes
Build and maintain tools for automation, deployment, monitoring, and operations, with a strong focus on scalability, resilience, and observability of distributed system.
Troubleshoot complex issues in our development, production, and test environments, including large-scale, distributed, and AI-integrated systems
Excellent communication and people skills.
Requirements:
8+ of years experience with DevOps technologies.
Extensive background leading the design, build, and evolution of end-to-end DevOps platforms, including infrastructure, tooling, and operational frameworks across the software lifecycle.
Deep expertise with one of the major cloud providers: AWS (preferred), GCP, Azure.
Extensive experience with modern deployment strategies (GitOps, blue/green, canary, Kubernetes-based deployments)
Strong experience designing and optimizing end-to-end CI/CD pipelines, enabling high velocity, reliable software delivery.
Experienced with bootstrapping projects, introducing new technologies and building systems from scratch.
Background in working with AI components and understanding the challenges of bringing AI workloads into production.
Good coding capabilities (Python, Bash, etc.)
Experience mentoring engineers, leading cross-functional initiatives, and influencing technical direction.
Advantages:
Experience with on-prem environments and solutions.
Prior experience with endpoint security products (agents, sensors, collectors).
Tech Stack: AWS, Kubernetes, EKS, Jenkins, IaC, GitHub, Terraform, Python, Docker, ArgoCD, MongoDB, RabbitMQ, Redis, Go, Neo4J, AI, and more.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8725160
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a hands-on Senior DevOps Engineer with a strong cloud-native mindset to build, maintain, and evolve our highly scalable, highly-available cloud infrastructure. This role is pivotal in driving operational excellence, security, and automation across our entire engineering organization. You will promote communication, integration, and collaboration to significantly enhance our software development productivity and reliability. You'll work closely with engineering and product teams to streamline delivery, enforce platform standards, and enable a high-velocity development environment-all while keeping reliability and security top of mind.
Responsibilities:
Design, Automate, and Manage complex cloud infrastructure on AWS using best-in-class Infrastructure as Code (IaC) practices.
Lead the operation and enhancement of our production Kubernetes environments (EKS), focusing on automation, security, observability, and seamless CI/CD integration.
Drive continuous improvement across platform tooling, developer experience, and operational processes to meet our ambitious performance and uptime goals.
Implement and enforce security-first infrastructure patterns, including strong IAM, network segmentation, and secure secrets management.
Actively contribute to high-level technical design discussions and cross-functional architectural decision-making, ensuring solutions align with long-term platform strategy.
Requirements:
7+ years of experience as a DevOps Engineer, Platform Engineer, or in a similar infrastructure-focused role.
Strong hands-on expertise across the AWS Stack (e.g. EC2, EKS, RDS, VPC, IAM, S3, Lambda).
Mastery of Infrastructure as Code - Terraform or equivalent.
Deep operational knowledge of Kubernetes, including architecture, cluster management, networking, and advanced debugging in production environments.
Strong expertise in designing and managing CI/CD methodologies and platforms (e.g. Jenkins, Github Actions).
Experience with monitoring tools such as Prometheus, DataDog, Coralogix (OTEL), Grafana etc.
Proven prior experience building and maintaining highly-available, production-grade, and service-oriented systems.
Strong scripting and automation background in languages such as Python or Bash.
Exceptional communication and collaboration skills with the ability to articulate complex technical needs and influence cross-functional teams.
Strong knowledge of AWS Networking - an advantage.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8748461
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
30/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Who We Are: Incredibuild is the leading platform empowering developers and enterprises to radically accelerate their development cycles. We help world-leading brands like Microsoft, Citi Bank, GM, Amazon, and Adobe shorten build times, allowing for more iterations, faster product releases, and significant resource savings. Our technology streamlines and accelerates everything from compilation to release automation, dramatically expediting time to market for our clients. You aren’t just "fixing CI/CD pipelines" - you are building a robust Internal Developer Platform (IDP) and creating the tools that enable our engineers to ship high-performance code at scale. If you love writing clean code and get equally excited about a perfectly orchestrated Kubernetes cluster or an elegant Terra grunt module, you’ll fit right in.
What you'll do: Architect & Build: Take full ownership of our infrastructure, from backend services to cloud resources. Empower Developers: Develop and maintain our homegrown IDP , ensuring a seamless "golden path" for engineering teams. Scale the Stack: Design and manage scalable, reliable environments using AWS, Argo CD, and Helm. Automate Everything: Build sophisticated CI/CD workflows using GitHub Actions and manage stateful infrastructure with Terra grunt. Innovate: Evaluate and integrate new CNCF tools and internal utilities to keep our platform at the cutting edge.
Requirements:
* 3-5 years of hands-on experience in DevOps, Platform Engineering, Site Reliability Engineering (SRE), or Cloud Infrastructure roles within modern SaaS environments.
* Hands-on experience deploying, operating, and troubleshooting Kubernetes environments in production.
* Strong experience working with AWS services and containerized workloads.
* Proven experience using Terraform or Terragrunt to provision, manage, and maintain cloud infrastructure at scale.
* Experience building and maintaining CI/CD pipelines using tools such as GitHub Actions.
* Familiarity with GitOps practices and technologies including Argo CD and Helm.
* Strong scripting and automation skills using Python, Bash, Go, or similar languages.
* Ability to develop internal tools and operational automation that improve developer productivity and platform reliability.
* Passionate about solving infrastructure challenges related to scalability, reliability, security, and operational efficiency.
* Committed to reducing developer friction and improving the overall engineering experience.
How We Work We are a small, highly collaborative team with a strong sense of ownership. Engineers are encouraged to take initiative, contribute to technical decisions, and have a meaningful impact on the platform's evolution. We move quickly while maintaining high engineering standards, balancing pragmatism with long-term maintainability.
Our Tech Stack AWS | Kubernetes | Terraform | Terragrunt | GitHub Actions | Argo CD | Helm | Homegrown Internal Developer Platform (IDP).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8693323
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Senior DevOps Engineer Platform Engineering.
Responsibilities:
-Design, build, and operate the internal engineering platform powering ' build, test, deployment, and security validation workflows at scale
- Write and maintain production-grade Python and shell tooling that drives platform automation - this is a hands-on coding role, not just pipeline configuration
- Architect and manage hybrid cloud/on-prem execution infrastructure, including large-scale Kubernetes runner pools across multiple AWS regions
- Own and evolve CI/CD pipelines at scale using GitHub Actions, including reusable workflows, ARC-based runner orchestration, and build caching strategies (BuildKit, sccache, Valkey)
-Operate and tune DinD environments (Sysbox, EBS/NVMe, overlay storage, MTU/networking) for build, test, and release workloads
- Connect and manage self-hosted and on-prem runners, routing physical device (wbox) test jobs by site and device type
- Implement DevSecOps controls including least-privilege IAM, OIDC, isolated runner groups, container signing, and automated security scans
- Drive platform observability, cost optimization, and reliability improvements across the engineering infrastructure
- Collaborate cross-functionally with hundreds of engineers to improve engineering velocity and release confidence
- Take end-to-end ownership of complex infrastructure problems and drive them to resolution
Requirements:
- 5+ years of hands-on DevOps experience with a strong software development background - prior development experience is a must
- B.Sc. in Computer Science or equivalent practical experience
- Strong programming skills in Python (or a similar high-level language); ability to write and own production tooling
- Proven experience designing and building scalable systems, automation frameworks, and infrastructure as code using Terraform and Helm
- Solid understanding of Linux, containers (Docker), and Git-basd workflows
- Hands-on experience with CI/CD at scale using GitHub Actions or similar - including reusable actions, workflow design, and automation frameworks
- Deep experience with hybrid cloud infrastructure (AWS and on-prem), including EKS, ARC, Karpenter, ECR, S3, Direct Connect, VPC endpoints, IAM/OIDC, and Secrets Manager
- Experience operating spot and on-demand runner pools for builds, DinD tests, releases, and security scans across multiple AWS regions
- Experience with DinD environments (Sysbox, EBS/NVMe, memory limits, overlay storage, MTU/networking) and build caching (BuildKit, sccache, Valkey)
- Experience connecting on-prem/self-hosted runners and routing physical device (wbox) test jobs by site and device type
- Experience implementing DevSecOps controls and improving platform observability, cost efficiency, and reliability
- Platform & tooling familiarity: Kubernetes (EKS, on-prem) GitHub Actions ARC Karpenter Terraform Helm Docker/DinD Sysbox containerd BuildKit ECR S3 ElastiCache (Valkey) sccache Direct Connect VPC endpoints IAM/OIDC Secrets Manager self-hosted runners
Soft Sills:
- Strong system-level thinking and troubleshooting skills; able to diagnose and resolve complex infrastructure issues independently
- Takes end-to-end ownership and drives problems to resolution without hand-holding
- Excellent communication and cross-team collaboration skills; comfortable working alongside large engineering organizations
Nice to Have / Advantage
-Experience with Jenkis
- Familiarity with GitHub merge queue
- Experience with MinIO or on-prem S3 caching
- Hardware-in-the-loop CI experience
- MTU/VPC networking tuning expertise
- Monorepo CI optimization experience
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764012
סגור
שירות זה פתוח ללקוחות VIP בלבד