דרושים » מחשבים ורשתות » Principal DevOps Engineer - AI Platforms (Cortex)

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 16 שעות
Location: Tel Aviv-Yafo
Job Type: Full Time
We are establishing a new role focused on bridging AI capabilities with our internal DevOps platform. You will build AI-powered tools and services that run on top of the DevOps platform developed by the infrastructure team. This role combines hands-on DevOps/platform engineering work with close collaboration with software development teams, enabling them to build, deploy, and operate AI solutions efficiently, securely, and in alignment with organizational standards. You will act as a technical enabler, helping teams bring AI-driven applications into production while ensuring best practices across architecture, compliance, DevSecOps, FinOps, and AI governance. The role also includes direct contributions to the platform itself.
Your Impact
Enhance DevOps platform capabilities to support AI-based tools and workloads.
Work closely with development teams to enable and support AI application delivery end-to-end.
Assist in designing architecture and production-grade implementation of AI systems.
Lead and enforce DevSecOps practices, compliance requirements, and governance standards for AI solutions.
Develop platform components and internal services as part of the AI infrastructure layer.
Support production readiness, monitoring, performance, reliability, and cost optimization (FinOps).
Serve as a technical interface between multiple engineering teams.
Participate in future on-call rotations.
Requirements:
Your Experience
Strong experience of 5-7 years in DevOps / Platform Engineering roles.
Good understanding of the AI ecosystem and modern AI workflows.
Hands-on familiarity with AI tools such as Claude and other GenAI platforms.
Ability to present a personal AI-related project
Solid understanding of production systems and modern CI/CD practices.
Experience with cloud infrastructure, automation, and deployment pipelines.
Strong communication skills and ability to work across multiple stakeholders.
Nice to have
Exposure to DevSecOps, compliance, and FinOps practices.
Experience building or integrating AI-driven systems in production environments.
Experience building tools or automations using Claude or similar GenAI tools.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781332
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are establishing a new role focused on bridging AI capabilities with our internal DevOps platform. You will build AI-powered tools and services that run on top of the DevOps platform developed by the infrastructure team. This role combines hands-on DevOps/platform engineering work with close collaboration with software development teams, enabling them to build, deploy, and operate AI solutions efficiently, securely, and in alignment with organizational standards. You will act as a technical enabler, helping teams bring AI-driven applications into production while ensuring best practices across architecture, compliance, DevSecOps, FinOps, and AI governance. The role also includes direct contributions to the platform itself.
Your Impact
Enhance DevOps platform capabilities to support AI-based tools and workloads.
Work closely with development teams to enable and support AI application delivery end-to-end.
Assist in designing architecture and production-grade implementation of AI systems.
Lead and enforce DevSecOps practices, compliance requirements, and governance standards for AI solutions.
Develop platform components and internal services as part of the AI infrastructure layer.
Support production readiness, monitoring, performance, reliability, and cost optimization (FinOps).
Serve as a technical interface between multiple engineering teams.
Participate in future on-call rotations.
Requirements:
Your Experience
Strong experience of 5-7 years in DevOps / Platform Engineering roles.
Good understanding of the AI ecosystem and modern AI workflows.
Hands-on familiarity with AI tools such as Claude and other GenAI platforms.
Ability to present a personal AI-related project
Solid understanding of production systems and modern CI/CD practices.
Experience with cloud infrastructure, automation, and deployment pipelines.
Strong communication skills and ability to work across multiple stakeholders.
Nice to have
Exposure to DevSecOps, compliance, and FinOps practices.
Experience building or integrating AI-driven systems in production environments.
Experience building tools or automations using Claude or similar GenAI tools.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8770082
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
15/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
At our company, "It's all about the user. All of them." We're passionate about providing a seamless one-stop experience for business travelers, no matter how they travel, where they stay, or where they're going. we are building cutting-edge solutions at the intersection of travel, expense, payments, and AI. As a leader in the AI for Travel domain, we are using intelligent, practical AI experiences to make business travel simpler, faster, and more reliable for travelers, travel managers, finance teams, and support teams.
We are constantly striving to make our systems reliable, scalable, and simple to operate so our services are available to travelers when they need them most. With our continued growth, we have exciting challenges ahead and we're looking for a Senior Site Reliability Engineer to join our team in Tel Aviv. This role blends classic SRE ownership with pragmatic AI SRE work: you will build and operate the platforms, automation, observability, and incident response practices that keep our company reliable, while helping teams use AI solutions, AI providers, and their APIs safely and dependably.
This is a hands-on engineering role, not a research role. You will partner with product, platform, data, security, support, and incident response teams to make production systems and AI-powered experiences more resilient. You will use software engineering, infrastructure as code, SLOs, telemetry, provider observability, and automation as your main tools, and you will apply AI where it creates measurable reliability value rather than novelty.
This position is based out of our new Tel Aviv office.
What You'll Do:
Support AI-based application solutions where reliability matters. Partner with the development teams building AI-powered travel experiences to support the development and production operation of their solution.
Work with AI solutions, providers, and APIs. Partner with teams integrating AI capabilities and providers, with attention to API reliability, authentication, quotas, rate limits, latency and provider-specific operational constraints.
Troubleshoot AI tools and provider issues. Diagnose failures across AI-powered workflows, provider APIs, configuration, permission errors, degraded responses and related areas.
Operate reliable production platforms. implement and run cloud infrastructure,and help product teams move quickly without compromising reliability.
Improve observability. Build dashboards, alerts, traces, logs, and runbooks that make service health clear, actionable, and tied to SLOs and customer impact.
Apply AI to SRE workflows. Prototype and productionize AI-assisted systems that create effective and efficient operations
Automate operational toil. Create tools, workflows, and automation that remove repetitive manual work and make operational knowledge easier to use.
Requirements:
5+ years of experience as a Senior SRE, Infrastructure Software Engineer, Production Engineer, or DevOps Engineer.
3+ years of experience operating production, 24x7 customer-facing systems.
Hands-on experience delivering production infrastructure, platform tooling, and automation used by engineering teams.
Strong software engineering skills in Python, Go, Java, or a similar language, with a bias toward production-quality code, tests, monitoring, and documentation.
Experience with cloud infrastructure, container orchestration, Linux systems, networking, CI/CD, and infrastructure as code such as Terraform or CloudFormation.
Experience building, tuning, and automating observability systems such as Grafana, Prometheus, New Relic, Datadog, Splunk, or similar tools.
Familiarity with SLOs, incident response, on-call practices, root cause analysis, and blameless postmortems.
Practical experience or strong interest in AI solutions, AI providers, agents, AI APIs, provider integrations, or AI-assisted internal tools.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8739673
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced DevOps Engineer to join our Engineering team and play a key role in building and operating our cloud-native platform. The ideal candidate will have hands-on experience managing production environments at scale, driving cloud transformation initiatives, and supporting the of enterprise systems from on-premises deployments to modern SaaS and cloud-native architectures. You will be responsible for designing, automating, and maintaining scalable infrastructure, CI/CD pipelines, and deployment processes that support both our core products and emerging AI-driven capabilities. Working closely with Engineering, QA, Product, and AI teams, you will help ensure the reliability, security, and performance of our services while driving operational excellence and continuous improvement across our technology stack.
Responsibilities:
Design, implement, and maintain CI/CD pipelines.
Manage and optimize cloud infrastructure across AWS, Azure, and/or GCP.
Develop and maintain Infrastructure as Code using Terraform.
Manage Kubernetes-based environments and GitOps deployment workflows using Argo CD and Kustomize.
Lead and support the migration of enterprise applications and infrastructure from on-premises
environments to scalable SaaS and cloud-native architectures.
Establish, maintain, and continuously improve production environments, ensuring high availability, security, scalability, and operational excellence.
Demonstrate strong production ownership, including incident management, root cause analysis, capacity planning, and performance optimization.
Collaborate with Engineering, QA, Product, and AI teams.
Support the deployment, operation, and monitoring of AI and Generative AI services.
Build and maintain monitoring, logging, and alerting systems.
Troubleshoot and resolve infrastructure, deployment, and production issues.
Requirements:
5+ years of experience as a DevOps Engineer or similar
infrastructure-focused role.
Hands-on experience with Azure, Aws, or GCP.
Experience with CI/CD tools such as Jenkins, GitHub Actions, or similar platforms.
Strong knowledge of Terraform and Infrastructure as Code practices.
Experience with Docker, Kubernetes, Argo CD, and Kustomize.
Experience designing and operating production-grade Kubernetes saas platforms or enterprise environments.
Experience with monitoring and observability tools such as Prometheus, Grafana, and ELK.
Strong troubleshooting, analytical, and communication skills.
B.Sc. in Computer Science, Computer Engineering, Information Systems, or a related field (or equivalent practical experience).
Nice to have
Experience supporting AI, Machine Learning, or Generative AI workloads, including familiarity with MLOps concepts, AI deployment platforms, or cloud-based AI services.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8739903
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
29/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for an experienced SRE Team Lead to drive the reliability, observability, and automation practices across our private cloud infrastructure and operations. In this role, you will lead a team of site reliability engineers, own the engineering roadmap for monitoring and automation, and act as a key liaison between development, operations, and platform teams. You bring at least 3-4 years of hands-on people management experience and a deep technical background in SRE or DevOps disciplines.


What will you do?

Leadership & Team Management

Lead, mentor, and grow a team of SREs, providing technical direction, career development guidance, and day-to-day management.

Own the team roadmap for reliability, observability, and automation initiatives - prioritizing work, removing blockers, and driving delivery.

Conduct regular 1:1s, performance reviews, and hiring processes to build and sustain a high-performing team.

Foster a culture of operational excellence, blameless post-mortems, and continuous improvement.

Act as an escalation point for complex incidents and reliability issues, leading post-incident reviews and ensuring follow-through on action items.


Automation & Infrastructure

Design, develop, and maintain automation tools to support infrastructure and operations teams at scale.

Manage pipelines and infrastructure workflows using Jenkins, Ansible, Python, and Bash.

Drive the adoption of infrastructure-as-code practices across the organization.

Collaborate with system engineers to improve scalability, performance, and fault tolerance of critical systems.


Monitoring & Observability

Build and extend monitoring and alerting systems using Grafana, the ELK (Elastic) stack, Zabbix, and custom scripts.

Implement and enforce observability best practices to ensure full visibility into systems, applications, and infrastructure.

Define and track SLIs, SLOs, and error budgets across key services.

Partner with development teams to embed observability earlier in the software development lifecycle.


Database & Platform Support

Support monitoring and infrastructure integration for databases including MongoDB and PostgreSQL.

Maintain documentation and champion knowledge sharing around automation, monitoring, and reliability practices.
Requirements:
Experience & Leadership:

3-4+ years of experience in a people management or team lead capacity within SRE, DevOps, or infrastructure engineering.

5-8+ years of overall experience in SRE, DevOps, or infrastructure automation roles.

Proven track record of building, coaching, and retaining high-performing engineering teams.

Experience owning an engineering roadmap and driving cross-functional reliability initiatives.


Technical Skills :

Strong scripting skills in Python and Bash; comfortable building and maintaining production-grade automation.

Hands-on experience with infrastructure automation tools, particularly Ansible.

Solid experience with monitoring and observability platforms - ELK stack, Grafana, and Zabbix.

Good understanding of CI/CD pipelines and related tooling, including Jenkins.

Familiarity with managing and monitoring MongoDB and PostgreSQL in a production environment.

Comfortable working in Linux-based environments.

Excellent problem-solving skills and strong written and verbal communication.


Ability to support the following:

Experience with cloud providers - AWS, GCP, or Azure.

Exposure to containerization technologies such as Docker and Kubernetes.

Familiarity with infrastructure provisioning using Terraform.

Experience introducing SRE practices (SLOs, error budgets, chaos engineering) at an organizational level.

Exposure and experience with migrating/ building AI tools to improve process.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8760168
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Were looking for a Senior Infrastructure Engineer who views "Infrastructure as Software." In 2026, we dont just manage servers; we build high-performance environments that allow multi-agent systems to operate at scale.
You will be a core member of the R&D team, blending deep DevOps expertise with the coding rigor of a Backend Engineer. You arent just "configuring" AWS; you are architecting the distributed systems and data pipelines that power our autonomous security brain. Your mission is to ensure that while our agents are evolving and taking actions, our underlying platform remains immutable, observable, and infinitely scalable.
What You'll Do
Design, build, and operate our company's cloud infrastructure using AWS, Kubernetes, and Infrastructure as Code.
Build internal tools and platform services using Python and Go to improve developer productivity and system reliability.
Own infrastructure automation with Terraform, Pulumi, and modern cloud-native tooling.
Partner closely with Backend, Data Science, and Security Engineering teams to build scalable, reliable platforms.
Improve observability, monitoring, and incident response across distributed production systems.
Design and optimize infrastructure for performance, scalability, security, and cost efficiency.
Help shape engineering best practices, platform architecture, and developer experience as our company continues to grow.
Requirements:
5+ years of experience in Infrastructure, DevOps, Platform Engineering, or Backend Engineering.
Strong software engineering skills with hands-on experience building production systems in Python or Go.
Deep hands-on experience with AWS, including services such as EKS, RDS, VPC, and IAM.
Strong experience designing, operating, and scaling production Kubernetes environments.
Experience with Infrastructure as Code, CI/CD, GitOps, and modern cloud-native development practices.
A systems mindset with the ability to solve architectural challenges across infrastructure and application layers.
Comfortable using modern AI-powered developer tools and agentic workflows to improve engineering productivity.
The company Mindset: You take ownership, act with accountability, collaborate openly, and focus on delivering meaningful impact. You thrive in fast-moving environments, embrace ambiguity, and enjoy solving hard problems together.
Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
Full professional fluency (written and verbal) in both Hebrew and English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are seeking a promising and talented Senior DevFinOps Engineer to join our DevOps group. If you thrive in a fast-paced, dynamic environment, can handle multiple requests simultaneously, and enjoy working independently as part of a cutting-edge DevOps team, this is your opportunity to help make the world a safer place!
Key Responsibilities
Act as a DevFinOps Engineer within a highly skilled team, bridging engineering and finance to drive cloud cost efficiency across large-scale operations from development to production.
Design, develop, and maintain Avanan's cloud cost visibility, allocation, and optimization solutions - including tagging strategies, cost dashboards, budgets, and anomaly detection across accounts and services.
Implement tools and procedures for cost monitoring, forecasting, and alerting across our SaaS multi-tenant product family.
Embed FinOps practices into the CI/CD lifecycle - surfacing the cost impact of changes early, and enforcing cost guardrails as part of deployment automation.
build AI-based FinOps agents for cloud services at the infrastructure and application levels
Continuously identify and execute cost-optimization opportunities (right-sizing, reserved capacity/savings plans, spot usage, storage tiering, idle-resource cleanup) without compromising performance, reliability, or security.
Partner with engineering teams to drive cost accountability - providing unit-economics insight (cost per tenant/feature/service) and making cost a first-class engineering metric.
Plan capacity and model the financial impact of scaling decisions, balancing cost efficiency with fault tolerance and growth.
Design and shape our cost reporting, chargeback/showback, and budgeting solutions.
Execute all tasks with cloud infrastructure security and financial governance as guiding principles.
Requirements:
Hands-on mindset - we all write code daily!
3+ years of relevant DevOps/Cloud experience building and operating CI/CD pipelines for both development and production - must.
2+ years of AWS Cloud experience working with high-traffic systems and multiple services, with a strong grasp of AWS pricing models and cost management tooling (Cost Explorer, CUR, Budgets, Compute Optimizer) - must.
Strong scripting skills, with fluency in Python - must.
Experience working with AI tools to achieve cloud or application cost control, identify and investigate the root cause of cost changes (differentiating between organic growth / infrastructure change / config change / application change).
Demonstrated experience driving measurable cloud cost reductions and building cost-optimization tooling or automation - must.
Experience with containers and orchestration tools (Docker, Kubernetes, or ECS) and understanding of their cost drivers - must.
Familiarity with FinOps principles and practices (FinOps Foundation framework, showback/chargeback, unit economics) - an advantage.
Experience with CI integration tools such as Jenkins.
Familiarity with AWS CloudFormation and infrastructure-as-code - an advantage.
Exposure to a wide range of open-source technologies (Redis, Nagios, Grafana, Prometheus, etc.) and cost-analytics tooling.
Knowledge of best practices in security, performance, monitoring, and cost governance.
Proven ability to research, evaluate, and implement new technologies, including running proofs of concept and cost analysis.
Huge advantage: measuring and controlling AI cost.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8723227
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
03/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an exceptional Platform Engineer who combines deep software engineering with a robust DevOps approach to help propel development infrastructure to the next level. , SPEED is integral to our DNA. AI transforms the way we develop at speed, and our development infrastructure is key to the scale and hypergrowth.



The DevX team is a multi-disciplinary group of DevOps and development experts focused on building and operating internal tools, infrastructure, and services to make engineering easy and intuitive SPEED.



Your mission is to build a top-tier Continuous Integration and Delivery (CI/CD) platform. You will own both the application services and the infrastructure that supports them, ensuring speed, stability, and reliability for our entire R&D organization.



Key Focus Areas & What You'll Do



Improve CI/CD pipelines and CI workflows (e.g., GitHub Actions) with a focus on speed and reliability.
Reduce CI flakiness and improve overall pipeline stability through systematic triage and root-cause analysis.
Shorten developer feedback loops by optimizing test strategy, pipeline consistency, and local development workflows.
Strengthen CD and release processes for secure, repeatable, and fast deployments.
Promote best practices such as GitOps and progressive delivery where they fit.
Track and improve delivery metrics, with emphasis on Lead Time for Changes and Deployment Frequency.
Partner with engineering teams to identify friction and deliver scalable automation and paved paths.
Requirements:
3+ years of experience as a Platform Engineer or a strong Backend Engineer with a DevOps focus.
Strong knowledge of CI/CD pipelines, versioning, and release management.
Proven experience with modern build tools (e.g., Bazel, SWC) and managing CI workflows (e.g., GitHub Actions, GitLab CI).
Good skills with infrastructure-as-code tools like Terraform and Helm.
Hands-on experience with Kubernetes, including containers (Docker), Kafka and cloud providers (AWS, GCP, or Azure).
Strong programming skills in Python, Go, TypeScript, or another modern backend language.
Experience with monorepo tools (Lerna, NX, PNPM) is a big plus.
Familiarity with GitOps workflows using tools like ArgoCD or Flux.
A strong sense of ownership and a passion for making systems reliable, scalable, and improving the developer experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8765802
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8779592
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
We are seeking a talented Full Stack Engineer to join our team and help build the Opik product! Opik is an open-source platform designed to streamline the entire lifecycle of LLM applications, helping developers evaluate, test, monitor, and optimize their models and agentic systems.

Ready to shape the future of LLM development? Join us in building the next generation of AI observability tools!
Responsibilities:
Build & ship features at high pace in a fast-moving AI development environment
Develop and maintain comprehensive observability tools for LLM applications and RAG systems
Create intuitive user interfaces for LLM tracing, evaluation dashboards, and production monitoring
Build robust APIs and backend services to support real-time LLM application monitoring
Collaborate with AI researchers and ML engineers to implement cutting-edge evaluation methodologies
Contribute to both frontend (React/TypeScript) and backend (Java) components
Design and implement automated evaluation systems and "LLM as a Judge" capabilities
Work on integrations with popular LLM frameworks and libraries
Participate in open-source community engagement and technical documentation
Requirements:
Experience: 3+ years of full-stack development experience, preferably in AI/ML domains
AI Development Tools: Hands-on experience with AI-assisted coding environments, IDEs, and agent-based workflows (e.g., Cursor, Windsurf, GitHub Copilot, Codex, Claude Code, and similar platforms)
Model Knowledge: Deep understanding of different AI model capabilities, limitations, and appropriate use cases (GPT-5, Claude, Gemini, etc.)
Frontend Development: Expertise in React, TypeScript and modern web development practices
Backend Development: Strong proficiency in Java with frameworks like Dropwizard
Database & Infrastructure: Knowledge of scalable database design and containerization
Observability Tools: Experience with tracing, monitoring, and evaluation systems
Collaboration: Strong communication skills for working in a distributed, global team environment
Problem Solving: Ability to work in ambiguous environments and solve complex technical challenges
Nice to have:
Open Source: Experience contributing to or maintaining open-source projects
AI/ML Experience: Understanding of Large Language Models, RAG systems, and AI application architectures
Preferred Qualifications:
Experience with model evaluation, prompt engineering, and LLM optimization
Knowledge of distributed systems and high-throughput data processing
Familiarity with ML experiment tracking and model monitoring platforms
Experience with DevOps practices and CI/CD pipelines
Understanding of AI safety, model security, and responsible AI practices
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8722649
סגור
שירות זה פתוח ללקוחות VIP בלבד