דרושים » הנדסה » System Architect

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a highly technical, hands-on System Architect to lead the architecture and evolution of Automation Platform (DAP), with focus on:
Infrastructure automation Service lifecycle orchestration Multi-tenant operations Intent-based networking Closed-loop assurance Cloud-native OSS transformation This role is ideal for architects passionate about scalable automation systems, distributed cloud-native platforms, and carrier-grade operational workflows.
The Role As a DAP Infrastructure & Services Automation Architect, you will define and drive the architecture for automation frameworks enabling lifecycle management of Network Cloud deployments at hyperscale.
You will bridge: Network infrastructure NOS internals Kubernetes-native microservices OSS/BSS integration Telemetry and analytics CI/CD operational models AI-assisted operational automation You will work closely with: DAP engineering Infrastructure teams NOS platform groups DevOps/SRE Product management Tier-1 service provider customers
Requirements:
Networking & Infrastructure:
6+ years designing large-scale Service Provider or hyperscale IP/MPLS networks
Strong understanding of:
BGP
ISIS
RSVP-TE
EVPN/VXLAN
Segment Routing
QoS architectures
Deep familiarity with distributed routing systems and disaggregated architectures
Automation & Cloud-Natve
4+ years in network automation and orchestration
Strong knowledge of:
NetConf/YANG
gNMI
REST/gRPC
OpenConfig
Hands-on Kubernetes architecture experience
Strong understanding of:
Docker
Helm
Service meshes
Microservices
Event-driven systems
Observability & Reliability- Experience with:
Prometheus
Grafana
OpenTelemetry
Kafka
Pulsar
RabbitMQ
Strong SRE mindset:
HA
Scalability
Failure domains
Reliability engineering
Operational resiliency
Deep networking expertise
Cloud-native software architecture
Automation-first thinking
Operational scalability mindset
Strong customer-facing communication
Ability to translate hyperscale operational requirements into reusable platform capabilities
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8763828
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are on an expedition to find an On-Premise Site Reliability Engineer (SRE)- someone who is passionate about building rock-solid, high-performance infrastructure and bringing order to complex environments. In this role, you will own the end-to-end reliability, automation, and deployment of platform across customer sites, working hands-on with cutting-edge AI, bare-metal, and hybrid cloud architectures.
Youll collaborate closely with Product, R&D, and Architecture teams while serving as the ultimate technical authority for our customer deployments. From designing automated Ansible workflows and mastering Kubernetes to troubleshooting complex network topologies, you will eliminate toil, streamline cluster operations, and ensure every deployment is scalable, seamless, and mission-ready.
:Responsibilities
Lead End-to-End On-Prem & Hybrid Deployments: Own the technical delivery and reliability of platform in close collaboration with Product, R&D, and customer technical teams.
Architect, Execute & Improve K8s Deployments: Take a definitive hands-on role in deploying, configuring, operating, and continuously improving our platform using advanced, enterprise-grade Kubernetes architectures.
Helm Chart Management: Design, modify, and manage Helm charts to package, version, and streamline complex application deployments across different environments.
Drive Automation & Simplification: Design, implement, and maintain robust deployment automation using Ansible. You must have a passion for turning complex manual tasks into reliable, repeatable, single-click operations.
Manage Infrastructure as Code: Utilize Git as the single source of truth to manage configurations, manifests, and automation playbooks, enforcing modern engineering best practices.
Bridge On-Prem and Cloud: Leverage AWS resources (specifically EC2 and S3) for hybrid components, staging environments, or cloud-to-on-prem data flows.
Technical Tier-3 Escalation: Serve as the ultimate technical authority for deployment, Linux networking, and Kubernetes orchestration issues.
Continuous Improvement: Constantly refine our delivery pipelines, optimize bootstrap processes, and create rock-solid technical documentation.
דרישות:
SRE / Delivery Mindset: 3-5 years of hands-on experience in enterprise infrastructure deployment, systems engineering, or an on-prem operational reliability role.
Kubernetes & Helm Expert: Deep, production-grade experience with Kubernetes architecture, deployment, advanced troubleshooting, and CNI networking. Proven working experience creating, maintaining, and deploying applications using Helm charts.
Ansible Mastery: Proven experience writing clean, scalable Ansible roles and playbooks for configuration management, automation, and infrastructure provisioning.
Modern Workflows (Git & AWS): Solid working experience using Git for version control and collaborating on code/infrastructure. Practical experience provisioning and managing AWS resources (EC2 and S3).
Core Systems & Linux: Strong Linux background (Ubuntu) with a deep understanding of system internals, containerized runtimes, and troubleshooting distributed applications.
Solid Networking Knowledge: Hands-on experience with routing, firewalls, and switching topology (mainly Cisco)
Storage Foundations: Working knowledge of storage protocols (iSCSI, SAN, local NVMe) and enterprise storage arrays (like DELL) interacting with Kubernetes Persistent Volumes.
GPU & Accelerated Compute: Working knowledge of managing GPU-enabled Kubernetes nodes, including NVIDIA drivers/runtime and basic troubleshooting.
Air-Gapped Deployments: Experience deploying and maintaining software in air-gapped or offline environments, including registry mirroring and artifact staging.
Problem-Solver: Strong debugging and problem-solving skills in complex, distributed environments with an intense ownership and accountability mindset.
Willingness to Travel: Ready to travel to customer sites for physi המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786710
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a talented and motivated Software Engineer with hands-on experience building and operating multi-agent AI systems in production to join our Automation Platform (DAP) team.
The team develops automation tools, orchestration capabilities, and intelligent platforms that simplify the deployment, management, troubleshooting, and optimization of large-scale network and AI infrastructure environments.
You will work on the design and development of AI-powered systems that bridge networking, automation, observability, and distributed infrastructure - running on Kubernetes at scale. Our stack includes LangGraph, Langfuse, RAG pipelines, MCP, and agent-to-agent (A2A) communication patterns.
This role combines strong software engineering with practical AI application development, with a sharp focus on production hardening, tracing, evaluation, and safety of agentic systems - not model training or research prototypes.
Requirements:
5+ years of hands-on software engineering experience building production-grade backend services, APIs, or AI-powered systems.
Proven production experience with multi-agent AI systems: deployment, tracing, guardrails, hardening, and incident management.
Hands-on experience with agentic frameworks such as LangGraph, CrewAI, Google ADK, AutoGen, or equivalent.
Experience building and running evaluation pipelines for agentic solutions - including trajectory tracing, ground truth validation, and harshness/quality scoring.
Strong Python proficiency: comfortable building scalable backend services using gRPC and REST APIs.
Solid understanding of distributed systems: fault tolerance, consistency models, service communication, and operational challenges at scale.
Hands-on Kubernetes experience: deploying and operating containerized services, managing workloads, config, and scaling in production clusters.
Practical experience with embeddings, vector databases, and semantic retrieval systems in production.
Practical experience with RAG pipelines, LLM API integration, structured outputs, and tool calling in production environments including building and serving MCP servers at scale.
Working knowledge of SQL and/or NoSQL databases, schema design, and query optimization.
Strong debugging skills across application logic, APIs, data, and AI agent behavior.
Strong communication skills and a bias toward ownership and delivery.
Nice to Have:
Familiarity with Langfuse/Arize Pheonix.
Familarity with A2A & A2UI protocols.
Experience with network automation, orchestration, or configuration management (Ansible, Terraform, NETCONF, gNMI, or similar).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764207
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior SRE Engineer, you will be a key player in ensuring the reliability, scalability, and performance of our critical IT infrastructure. You will leverage SRE principles and an automation-first mindset to build and maintain resilient hybrid cloud environments. This role is ideal for a candidate who thrives in a fast-paced, innovative setting and is passionate about solving complex challenges with cutting-edge technology.
Key Responsibilities
Provision, configure, and support resilient hybrid cloud deployment architectures using an Infrastructure-as-Code framework.
Proactively collaborate with development teams to ensure new applications are production-ready, scalable, and reliable from inception.
Develop and maintain tools and frameworks to automate operational tasks, including deployment, monitoring, and recovery.
Conduct thorough root cause analysis of production issues and implement preventative measures to improve system resilience, demonstrating strong problem-solving skills.
Manage CI/CD platforms, Linux infrastructure, and contribute to capacity planning and operational runbooks.
Design and implement proactive service monitoring, alerting, and trend analysis to maintain service availability and performance SLAs.
Participate in an on-call rotation to support critical applications and services, responding to and resolving incidents efficiently.
Contribute to comprehensive documentation related to infrastructure design, deployment, and operational procedures.
Requirements:
Your Expereience:
6+ years of Devops engineering experience on mission-critical, enterprise-level systems in a hybrid (both cloud and on-prem) environment.
3+ years of hands-on experience with cloud environments, preferably Google Cloud Platform (GCP).
Expertise in configuration management and Infrastructure-as-Code using frameworks such as Terraform and Ansible.
Strong programming/scripting knowledge in languages like Python, Bash, or Go for infrastructure automation.
Demonstrated experience with CI/CD pipelines (e.g., GitHub, Jenkins, Artifactory) and a strong foundation in Linux/Unix administration.
Bachelor's degree in Computer Science, Information Technology, or a related field, or equivalent practical experience.
Preferred Qualifications
Experience with containerization and orchestration technologies, particularly Kubernetes.
Hands-on experience with monitoring and observability tools such as Datadog, Grafana, or Prometheus.
Understanding of networking principles including firewalls, load balancers, and complex network designs.
A curious and positive mindset with a passion for applied learning and challenging existing processes for continuous improvement.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8779502
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
5 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are constantly striving to make our systems reliable, scalable, and simple to operate so our services are available to travelers when they need them most. With our continued growth, we have exciting challenges ahead and we're looking for a Senior Site Reliability Engineer to join our team in Tel Aviv. This role blends classic SRE ownership with pragmatic AI SRE work: you will build and operate the platforms, automation, observability, and incident response practices that keep Navan reliable, while helping teams use AI solutions, AI providers, and their APIs safely and dependably.



This is a hands-on engineering role, not a research role. You will partner with product, platform, data, security, support, and incident response teams to make production systems and AI-powered experiences more resilient. You will use software engineering, infrastructure as code, SLOs, telemetry, provider observability, and automation as your main tools, and you will apply AI where it creates measurable reliability value rather than novelty.



This position is based out of our new Tel Aviv office.



What You'll Do:

Support AI-based application solutions where reliability matters. Partner with the development teams building AI-powered travel experiences to support the development and production operation of their solution.
Work with AI solutions, providers, and APIs. Partner with teams integrating AI capabilities and providers, with attention to API reliability, authentication, quotas, rate limits, latency and provider-specific operational constraints.
Troubleshoot AI tools and provider issues. Diagnose failures across AI-powered workflows, provider APIs, configuration, permission errors, degraded responses and related areas.
Operate reliable production platforms. implement and run cloud infrastructure,and help product teams move quickly without compromising reliability.
Improve observability. Build dashboards, alerts, traces, logs, and runbooks that make service health clear, actionable, and tied to SLOs and customer impact.
Apply AI to SRE workflows. Prototype and productionize AI-assisted systems that create effective and efficient operations
Automate operational toil. Create tools, workflows, and automation that remove repetitive manual work and make operational knowledge easier to use.
Requirements:
5+ years of experience as a Senior SRE, Infrastructure Software Engineer, Production Engineer, or DevOps Engineer.
3+ years of experience operating production, 24x7 customer-facing systems.
Hands-on experience delivering production infrastructure, platform tooling, and automation used by engineering teams.
Strong software engineering skills in Python, Go, Java, or a similar language, with a bias toward production-quality code, tests, monitoring, and documentation.
Experience with cloud infrastructure, container orchestration, Linux systems, networking, CI/CD, and infrastructure as code such as Terraform or CloudFormation.
Experience building, tuning, and automating observability systems such as Grafana, Prometheus, New Relic, Datadog, Splunk, or similar tools.
Familiarity with SLOs, incident response, on-call practices, root cause analysis, and blameless postmortems.
Practical experience or strong interest in AI solutions, AI providers, agents, AI APIs, provider integrations, or AI-assisted internal tools.
Ability to troubleshoot AI tools and provider/API issues, including rate limits, quota, auth, permission errors, latency, SDK or API contract changes, content quality issues, and service degradations.
Excellent communication skills and the ability to work with stakeholders and domain experts across the company.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8797911
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're seeking a Staff engineer - AI Builder to lead the design and implementation of AI-native technical frameworks that redefine how large-scale systems are built, extended, and evolved.

This role is deeply technical and architecture-driven, focused on designing modular, extensible, and AI-first system foundations that enable scalable development with AI as a core engineering collaborator.

You will work across the full stack, building infrastructure where AI actively generates, extends, and maintains software components as part of the native development flow.

This is not about incremental AI integration - this is about architecting the frameworks that let AI scale software engineering 10x faster and smarter.

Responsibilities:
Architect and lead the development of AI-native system frameworks that emphasize modularity, extensibility, and adaptive scalability.
Build platform primitives that enable dynamic AI-driven module and extension generation.
Design and implement developer workflows optimized for AI-assisted software development, embedding AI collaboration as a first-class design principle.
Define and codify AI-native engineering practices, patterns, and guidelines to elevate our software development lifecycle.
Establish and champion best practices for AI-assisted engineering across system design, code quality, testing, observability, and operational scalability.
Drive hands-on prototyping and iterative delivery across frontend, backend, and orchestration layers.
Provide technical leadership and mentorship, fostering an AI-first engineering culture across teams.
Partner closely with Product, AI, and Engineering teams to align system evolution with AI-native goals.
Requirements:
Strong builders mindset - balancing deep system thinking with practical, hands-on delivery.
10+ years of full stack system architecture and complex platform engineering experience.
Deep expertise designing distributed, composable, and extensible systems across frontend and backend environments.
Strong hands-on skills with frontend frameworks (e.g., React, Next.js) and backend systems (e.g., Node.js, Python, Go, microservices, event-driven architectures).
Proven experience integrating and operationalizing AI-assisted development tools (e.g., GitHub Copilot, Cursor) into engineering workflows.
Experience building plugin frameworks, extension systems, or developer platforms that emphasize modularity and scalability.
Track record of defining and promoting engineering best practices - especially in the context of AI-driven development.
Demonstrated technical leadership - mentoring engineers, shaping architecture standards, and driving adoption of new paradigms.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8752189
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a hands-on Tech Lead to join our R&D. Reporting directly to the Chief Architect, you will step into a rare opportunity with a fast-growing company to shape the technical future of our platform. You will serve as a technical anchor for the organization, driving the architectural evolution of our high-scale, cloud-native infrastructure built on AWS and Kubernetes, while pioneering the integration of AI across our product and engineering workflows.
Responsibilities:
Drive the technical roadmap for Firefly's platform, working closely with the Chief Architect: architecting new systems and features and continuously improving existing ones.
Lead architectural decisions across the platform's core surfaces: services, data pipelines, workflow orchestration, and multi-tenant infrastructure.
Stay hands-on in the code: build critical components, prototype new capabilities, and lead by example inside the development teams.
Establish the platform foundations: shared services, libraries, and standards used across teams.
Architect AI-powered capabilities into Firefly's product and drive AI-assisted development workflows that amplify engineering productivity across the org.
Drive end-to-end technical solutions, from API design and service boundaries to data modeling and deployment.
Build proofs-of-concept for critical paths and high-risk decisions, and act as the technical anchor for projects from design through delivery.
Write design documents, RFCs, and architectural decision records that drive cross-team alignment and capture the reasoning behind technical decisions.
Mentor engineers across teams through design and code reviews, and act as a focal point for technical questions across R&D.
Requirements:
8+ years of recent, hands-on experience designing and building large-scale distributed systems, with strong understanding of microservices, event-driven systems, and SaaS architecture patterns.
Expertise in data architecture: schema design, indexing, and data governance.
Strong backend development experience (Go, Java, or similar).
Hands-on expertise across modern data stores: relational, document, and search (e.g., PostgreSQL, MongoDB, Elasticsearch).
Strong API design skills and experience with both synchronous and asynchronous service communication (REST, gRPC, Kafka).
Hands-on experience with cloud-native environments and workload management tools (Kubernetes, AWS/GCP/Azure, or similar).
Experience designing observability for distributed systems (metrics, logs, traces) with tools like OpenTelemetry, Prometheus, and Grafana.
Experience leading architectural decisions across multiple engineering teams, writing design documents, and bridging between product, business, and engineering.
Strong hands-on experience with AI coding agents, with the ability to design, drive, and enhance AI-assisted development workflows across engineering teams.
Advantages:
Strong understanding of LLM-based application development and agent design, including tool execution frameworks and runtime safety.
Experience with workflow orchestration engines such as Temporal or Cadence.
Familiarity with Infrastructure-as-Code tooling (Terraform, OpenTofu) and CI/CD pipelines.
Background in cloud asset management, CSPM, or cloud security domains.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8801952
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
We are seeking a Staff Developer to join our team and take a lead role in shaping the architecture, evolution, and reliability of our systems core services.
You will directly shape Apono's technical direction, keeping our architecture aligned with company strategy while pushing forward what's possible in access management and engineering excellence. This role means owning a roadmap that raises the bar on quality, scalability, and security across our platform, as we bridge the operational security gap for organizations running in the cloud.
What you can expct:
Shape the architecture and long-term technical direction of core services powering Apono's platform.
Own complex, ambiguous initiatives end-to-end, from problem definition through rollout, spanning one or more teams and systems.
Break down large, undefined problems into a coherent execution plan, and drive that plan to completion.
Partner with product, design, and other engineering teams as a technical authority, translating ambiguity into clear architectural decisions.
Balance rapid iteration with long-term system health, scalability, and security.
Raise the bar for engineering standards, code quality, observability, and operational excellence across the org.
Mentor senior and mid-level engineers through design reviews, technical deep-dives, and pairing.
Evaluate new technologies and propose architectural improvements that shape how Apono builds going forward.
Leverage AI-assisted engineering workflows to multiply your own impact and the team's.
See your technical decisions directly shape the platform's ability to scale with the business.
Requirements:
8+ years of experience as a backend software developer, with a track record of owning large-scale production systems.
Proven experience architecting and evolving distributed systems, not just building within them.
Deep experience with cloud-native architectures (microservices, Docker, Kubernetes) at scale.
Strong systems thinking - able to reason about trade-offs across performance, reliability, and security.
Demonstrated ability to lead complex, cross-team technical initiatives from ambiguity to delivery.
A track record of raising technical standards beyond your immediate team.
Ability to balance deep technical ownership with business and product outcomes.
Nice to have:
Hands-on experience in identity and access management.
DevOps or platform engineering technical leadership background.
In-depth experience with cloud service providers, mainly AWS.
Strong security mindset.
Experience mentoring engineers into senior or staff-level roles.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8787464
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Job Type: Full Time
We are looking for an outstanding Senior Networking Software Architect to join the NIC/DPU Software and Firmware Architecture group. In this role, you will help define the next generation of our datacenter and AI networking platforms, with focus on DPU management, QoS, performance, telemetry, and software architecture across stacks. You will work closely with hardware designers, firmware/kernel driver teams, system engineers, validation, product management and customers. The role spans early architecture definition, pre-silicon design, bring-up, and production readiness for large-scale AI and cloud datacenter deployments.

What Youll Be Doing:

Own software and system architecture for next-generation DPU management, QoS, performance, telemetry, and observability features.

Define end-to-end control and management flows across DOCA, host drivers, embedded firmware, BMC, management controllers and external management systems.

Specify telemetry and observability requirements, including counters, logs, traces, events, health monitoring, debug data, and streaming telemetry.

Define management interfaces and APIs for configuration, provisioning, lifecycle operations, diagnostics, and field serviceability.

Write clear architecture specifications, interface definitions, flow diagrams, and design documents for software, firmware, and system teams.

Partner with R&D teams to translate high-level architecture into implementable designs and guide features through development, validation, silicon bring-up, and production.

Analyze system performance bottlenecks, interoperability issues, telemetry gaps, and customer-reported issues, then feed learnings into future architecture.

Collaborate with system and cluster architects to ensure NIC/DPU features fit end-to-end AI datacenter and cloud networking designs.
Requirements:
What We Need To See:

B.Sc. or M.Sc. in Computer Engineering, Computer Science, Electrical Engineering, or equivalent experience.

9+ years of experience in networking, system software, embedded software, firmware, or datacenter infrastructure.

Proven experience in software architecture, or technical leadership roles.

Deep understanding of networking concepts and protocols such as Ethernet, TCP/IP, RDMA/RoCE, congestion control, QoS, virtualization overlays, and traffic management.

Strong background with DPUs, SmartNICs, or other high-performance networking devices.

Experience with system management, provisioning, monitoring, telemetry, diagnostics, or lifecycle-management flows.

Familiarity with management protocols and frameworks such as Redfish, PLDM, MCTP, IPMI, gNMI, SNMP, Netconf, REST, or gRPC-based APIs.

Ability to lead cross-functional architecture discussions across software, firmware, hardware, validation, product, and customer-facing teams.

Excellent written and verbal communication skills, including the ability to create clear architecture documents and present trade-offs.


Ways To Stand Out From The Crowd:

Experience defining software architecture for DPU products, including management, telemetry, QoS, performance, security, virtualization, or offload features.

Hands-on background with Linux networking, device drivers, firmware, embedded Linux, BMC software, DOCA, DPDK, OVS or Kubernetes networking.

Experience with performance counters, profiling tools, eBPF, Prometheus, Grafana, dashboards, heat maps, or large-scale telemetry systems.

Experience in defining and developing GAI-based analysis tools to extract insights from telemetry data and streams.

Background in RAS, diagnosability, serviceability, field failure analysis, production debug, or customer escalation handling.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773378
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Your Career:
Own and continuously improve AWS production infrastructure for scalability, reliability, security, performance, and cost.
Run and evolve Kubernetes environments that support fast, safe product delivery.
Drive developer velocity and production safety through better CI/CD pipelines, release workflows, deployment visibility, and GitOps practices.
Improve observability and incident response - reduce alert noise and raise signal quality.
Design and ship AI-assisted operational agents that change how engineers work - triaging monitoring alerts, summarizing incidents, proposing fixes, onboarding new services, answering questions and requests. This is a core part of the role, not a side project.
Build automation and self-service tooling that removes manual work from provisioning, monitoring, incident response, and developer workflows.
Analyze operational data across incidents, alerts, deployments, infra health, and cost to find reliability gaps, inefficiencies, and automation opportunities.
Partner with engineering, security, product, and leadership to remove bottlenecks and support safe production growth.
Evaluate and introduce new tools and AI-assisted approaches, balancing innovation with reliability, cost, and operational simplicity.
Your Impact:
You'll help scale production systems, improve deployment velocity and reliability, reduce operational overhead, and build automation and AI workflows that help engineering teams move faster and operate more efficiently.
This role is a strong fit for someone who enjoys ownership, collaboration, and operational innovation.
Requirements:
Your Experience:
4+ years operating production infrastructure in AWS.
Deep hands-on experience with Kubernetes, Helm, ArgoCD, Terraform, and CI/CD.
Strong experience with observability and alerting in Datadog or comparable platforms.
Solid grounding in Linux, networking, cloud security, and reliability best practices.
Strong scripting skills in Python and Bash.
Proven ability to own platform projects end-to-end, from design through production operation and ongoing improvement.
Strong troubleshooting across distributed systems, Kubernetes, CI/CD, and live incidents.
Collaborative mindset - comfortable working across engineering, security, product, and leadership.
Comfort in a fast-paced, high-ownership environment where priorities shift but production quality doesn't.
Genuine interest in applying AI, automation, and intelligent workflows to operational work.
Key qualities
Ownership-driven - You take responsibility for the systems you build and operate, from design through production support and continuous improvement.
Collaboration - You work effectively across engineering, security, product, and leadership to align priorities and drive shared outcomes.
Developer experience focus - You are committed to reducing friction for engineering teams through thoughtful automation, self-service workflows, and reliable internal tooling.
Innovation balanced with pragmatism - You actively explore new approaches, particularly in AI-assisted operations, while weighing them against reliability, maintainability, and operational simplicity.
Security mindset - You design and build with least privilege, auditability, and production safety as foundational principles rather than afterthoughts.
Clear communication - You articulate infrastructure, reliability, cost, and security tradeoffs precisely to both technical and non-technical stakeholders.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769987
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
06/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for an experienced Sr. Principal Architect to join the XSIAM engineering group in a highly cross-functional, product-facing role. This position sits at the intersection of engineering, product management, customer success, and product architecture, helping shape scalable and innovative solutions across the XSIAM product.
The ideal candidate combines deep technical expertise with strong product thinking, enabling them to translate business and customer needs into clear architectural direction and execution strategies. This role requires close collaboration with engineering teams, product managers, and strategic customers to drive platform capabilities, integrations, scalability, and long-term architectural vision.
Key Responsibilities
Lead architectural design and technical strategy for XSIAM platform capabilities and integrations.
Define best practices, technical standards, and scalable design patterns across the organization.
Work closely with customers, field teams, and support organizations to understand real-world use cases and pain points.
Evaluate tradeoffs between scalability, performance, security, operational complexity, and delivery timelines.
Support strategic customer engagements and complex deployment scenarios.
Mentor engineers and technical leaders on architecture, system design, and product-oriented engineering thinking.
Act as the bridge between engineering, product, and site reliability (SRE) to ensure that architectural designs translate into operational excellence.
Requirements:
10+ years of hands-on Software Engineering experience, with 2+ years in a Principal or Architect role leading multi-team technical strategies.
Proven experience building SaaS products alongside deep knowledge of message brokers and infrastructure.
Experience architecting systems that handle multi-region, large workloads with extreme throughput requirements.
Expert-level knowledge of cloud-native ecosystems (AWS, GCP, or Azure) and container orchestration at scale (Kubernetes).
Proficiency in Golang and Python, with a deep understanding of low-level performance tuning and concurrency models.
B.Sc. or M.Sc. in Computer Science or Software Engineering, or leadership experience in elite military technology units.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771683
סגור
שירות זה פתוח ללקוחות VIP בלבד