דרושים » הנדסה » Research Team Lead - Distributed AI Systems & Large-Scale Infrastructure

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon and Haifa
Job Type: Full Time
Our team at the Huawei Computing Network Innovation Lab is looking for exceptional talent to join us and lead the development of next-generation data centers. We create cutting-edge technologies that synergize software and hardware in tandem to accelerate compute, storage, and networking at large scale. We aim to drive innovation and deliver software-defined infrastructure and algorithms for HPC, AI/ML, and Big Data applications.
We are looking for an outstanding Research Team Lead with deep hands-on expertise in large-scale distributed systems, AI framework infrastructure, and performance optimization on custom accelerators. If you are a visionary technical leader, a skilled communicator, and a team builder who thrives at the frontier of systems and AI research - you're welcome on board.

What Will You Be Doing?
Lead and grow a world-class team of researchers and engineers working on distributed AI infrastructure and systems software.
Architect and own the software infrastructure enabling distributed training and inference on Huawei's custom accelerator hardware (e.g., Ascend NPU).
Drive research and development on communication libraries, runtime systems, memory management, and graph execution & synchronization at scale.
Optimize end-to-end performance across large-scale clusters, covering both scale-up (multi-device) and scale-out (multi-node) configurations.
Design and implement high-performance communication backends and collective operations (AllReduce, AllGather, broadcast) for distributed training workloads.
Collaborate cross-functionally with hardware architects, compiler teams, and framework engineers to co-design hardware-software solutions.
Publish and present research findings at top international venues (NeurIPS, EuroSys, SC, MLSys, OSDI) and represent the team externally.
Mentor engineers and researchers, conduct performance and growth reviews, and shape team culture and technical direction.
Partner with top academic institutions and open-source communities to advance the state of the art in distributed AI systems.
Requirements:
Requirements
B.Sc. or higher in Computer Science, Computer Engineering, Electrical Engineering, or a closely related field.
8+ years of experience in systems software, distributed computing, or AI infrastructure, with 3+ years in a leadership or team lead role.
Deep expertise in large-scale communication systems: collective communication, RDMA, network topology-aware routing, and bandwidth optimization.
Hands-on experience building software infrastructure for distributed training on custom accelerators or heterogeneous hardware (GPU, NPU, TPU).
Strong knowledge of runtime systems: scheduling, execution graphs, kernel dispatch, synchronization primitives, and pipeline management.
Experience with memory management at scale: activation checkpointing, tensor offloading, rematerialization, KV cache management.
Proficiency in C/C++ and Python, with a focus on high-performance, production-quality code in Linux environments.
Proven ability to define technical vision, lead multi-person projects end-to-end, and deliver results under research and engineering timelines.
Excellent communication skills in English - confident presenting to international audiences, writing technical reports, and driving cross-team alignment.
Strong collaborative mindset and experience working in globally distributed, multicultural teams.
Ways to Stand Out From the Crowd.
M.Sc. or Ph.D. in a relevant field, with a strong publication record at systems or ML venues (EuroSys, OSDI, SC, NeurIPS, MLSys, ISCA).
Hands-on experience with communication frameworks such as NCCL, MPI, HCCL, or UCX.
Experience with compiler and graph optimization for AI workloads (XLA, TVM, Triton, or custom operator fusion).
Background in mixed-precision training, model parallelism (Tensor Parallelism, Pipeline Parallelism, Expert Parallelism), and large model co-design.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792874
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon
Job Type: Full Time
We are looking for a Principal Engineer with deep expertise in database internals, system architecture, and hardware-aware performance optimization.
This role is responsible for defining and driving advanced database performance solutions by understanding and improving the interaction between database engines, operating systems, and modern hardware architectures.
The ideal candidate is an experienced technical leader with a strong background in database systems, performance engineering, and computer architecture. The candidate should be capable of leading complex technical initiatives, influencing architectural decisions, and developing innovative solutions for next-generation database platforms.
This is a hands-on technical leadership role requiring expertise across database kernels, query execution, transaction processing, memory systems, and low-level performance optimization.

Responsibilities
Define and drive technical direction for database performance optimization and hardware-aware database architecture.
Lead the design and implementation of advanced database optimizations focused on scalability, efficiency, and workload performance.
Solve complex performance challenges across the full system stack, including database engines, query execution, transaction processing, concurrency control, storage engines, operating systems, and hardware platforms.
Design and develop optimizations leveraging modern hardware architectures, including multi-core CPUs, large-memory systems, NUMA platforms, CXL memory, RDMA networking, and advanced storage technologies.
Improve database execution engines, caching mechanisms, memory management, storage systems, and workload scheduling.
Analyze system behavior and identify bottlenecks using profiling, tracing, hardware performance counters, and low-level performance analysis.
Develop benchmarks, workload models, and methodologies to evaluate scalability, efficiency, and architectural tradeoffs.
Collaborate with database kernel, infrastructure, and engineering teams to influence architecture decisions and deliver high-impact improvements.
Mentor senior engineers and contribute to building technical expertise across the organization.
Drive innovation through research, prototyping, technical analysis, and architectural improvements.
Requirements:
Education
BSc or MSc in Computer Science, Computer Engineering, or a related technical field.
Strong background in systems research, database technologies, distributed systems, or computer architecture is highly desirable.

Requirements
10+ years of experience in database systems, systems software, performance engineering, or related fields.
Proven experience designing, developing, or optimizing high-performance database systems, storage engines, query execution engines, or large-scale data platforms.
Deep understanding of database internals, including:
o Query optimization and execution engines
o Transaction processing and concurrency control
o Storage engine architecture
o Indexing and caching mechanisms
o Database scalability and workload optimization
Strong understanding of computer architecture and hardware performance optimization.
Hands-on experience with memory systems, including:
o NUMA architectures
o CPU cache hierarchy
o Memory bandwidth and latency analysis
o Large-scale memory management techniques
Strong programming skills in C/C++ and experience developing high-performance software on Linux systems.
Experience using performance analysis and debugging tools, including:
o Linux perf
o Hardware performance counters
o Profiling and tracing frameworks
o Benchmarking tools
Experience optimizing OLTP, OLAP, or mixed transactional/analytical database workloads.
Strong analytical and problem-solving skills with the ability to diagnose and resolve complex performance issues across software and hardware layers.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792760
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Haifa
Job Type: Full Time
Are you passionate about systems programming at the intersection of hardware and software? Do you want your code to directly accelerate the world's largest AI training and inference workloads? The Elastic Fabric Adapter (EFA) Drivers team is looking for Software Development Engineers to design and implement the kernel drivers and userspace libraries that power high-performance networking for AWS's AI/ML and HPC infrastructure.

EFA is the custom network interface that enables thousands of GPUs and accelerators to communicate at near-wire-speed bandwidth and low latency, making it a critical enabler of both foundation model training and real-time inference serving at unprecedented scale. When customers train the next generation of large language models or serve billions of inference requests with tight latency SLAs on our EC2 P5, P6, or Trn instances, it's our driver stack that moves the data.

Key job responsibilities
- Design, develop, and optimize Linux kernel drivers (RDMA/EFA) and userspace provider libraries (rdma-core) that ship to millions of EC2 instances.
- Work directly with custom hardware - collaborate with chip designers to bring new silicon capabilities to life in software.
- Contribute to the upstream Linux kernel RDMA subsystem and rdma-core open-source project.
- Architect solutions for next-generation networking features: GPU-direct RDMA, adaptive routing, collective offloads, and multi-path transport.
- Build monitoring and automation tools to enhance driver testing for functionality, reliability, and performance.
- Own the full lifecycle: from initial hardware bring-up and feature enablement through testing, deployment, and production troubleshooting across the global AWS fleet.
Requirements:
Basic Qualifications
- Bachelor's degree in Computer Science, Engineering, Mathematics, or a related field.
- 3+ years of non-internship professional software development experience in C/C++.
- Experience contributing to the architecture and design (architecture, design patterns, reliability and scaling) of new and current systems.

Preferred Qualifications
- 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
- Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems.
- Experience writing low level drivers.
- Experience with DMA, PCIe, or hardware/software co-design.
- Contributions to open-source projects or upstream kernel work.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774252
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon and Haifa
Job Type: Full Time
In this role, you will be responsible for several teams of architects, engineers and software developers, all working together to conduct state-of-the-art R&D in system and network architecture. As the group lead, you will guide and mentor the individual team leads, and also conduct hands-on work leading architecture, technology innovation and technical planning and of high-performance computing cluster network, which oriented at AI, HPC, and big data.

Responsibilities
You will perform a wide range of duties including:
Architecture Innovation:
Deeply analyzing the advantages and disadvantages of mainstream network systems, to find opportunities for network architecture innovation;
Insight into the technology developing trend of the high-performance computing network field, and leading the corresponding technology planning.
Exploring new architectures of high-performance computing network systems and efficiently integrating communication library, topology, and network protocol to solve performance bottlenecks.
Technical breakthroughs in networking and cluster routing algorithm:
Analyzes computing cluster network performance and leads the development of computing cluster network technologies
Research and optimize the heterogeneous interconnection topology of key computing chips to continuously improve the key competitiveness of Huawei computing heterogeneous chipsets.
Responsible for the research of data center network technologies, and guide network topology design and routing algorithm development
Group leadership:
Lead the development of a comprehensive system architecture for AI Fabric and HPC Fabric solutions.
Manage and mentor highly skilled team leaders, to ensure that the group operates together in pursuit of common goal.
Foster a collaborative and innovative work environment.
Provide technical guidance and support to team members.
Collaborate closely with cross-functional teams internationally, including hardware, software, and ucode design teams, to ensure alignment of architectural decisions with product and platform common objectives.
Initiate and supervise collaborations with top academic researchers in Israel and abroad.
Stay up to date with emerging technologies and industry trends in AI, HPC and big data industries.
Evaluate and recommend technologies and next generation projects.
Occasional travel related to ongoing projects, seminars, conferences etc.
Requirements:
Requirements:
At least 10 years of hands-on experience in system architecture design, or equivalent research experience.
Demonstrated experience in leading R&D team.
Familiarity with high-performance computing cluster services and system architectures, such as AI, HPC and big data.
In-depth understanding of computer networks, communication libraries, and design of AI or HPC cluster networks

Key qualifications you hold include:
Ability to work in a team environment; actively seek out resolution to issues with the team and work co-operatively to build a coherent system; Team-working and excellent inter-personal communication skills.
High level of self-reliance and an autonomous target-oriented work style, can do attitude, eager to learn new things and ready to think outside the box
Ability to work on a schedule, even for un-schedule-able issues such as inventiveness.
Experience with network system of AI, HPC or big data cluster.
Demonstrated capability in several items from the list below:
o Deep understanding of computer network protocol and network topology of computing cluster
o Familiarity with communication library, such as OpenMPI/NCCL
o Extensive experience in leading the network architecture design of computing cluster
o Experience with implementing routing algorithm
o Familiarity with network specifications of CPU/Network Processing Unit/ Neural Processing Unit/Switch chip
Fluent in written and spoken English.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793412
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon
Job Type: Full Time
We are looking for an experienced Software Team Leader to build and lead a new team focused on AI and agent technologies for enterprise data platforms.
This is a hands-on leadership role. The Team Leader will define the technical direction, build the first solutions, recruit engineers, and turn new AI ideas into production systems.
The main focus is AI agents, RAG, natural-language interaction, model integration, and intelligent workflows. Database technologies are an important foundation, but not the center of the role.

Responsibilities
Lead a new software engineering team.
Define the AI architecture, roadmap, and technical priorities.
Build AI agents that reason, use tools, and execute multi-step workflows.
Develop agent orchestration, state management, and workflow execution.
Build agent workflows for data analysis, automation, and decision support.
Integrate LLMs, embedding models, and model-serving platforms.
Develop tool and API integration for AI agents, including MCP where relevant.
Optimize AI systems for latency, throughput, cost, and scale.
Move successful prototypes into production-grade software.
Work closely with data, database, infrastructure, cloud, and product teams.
Requirements:
Requirements
10+ years of software engineering or related technical experience.
Proven experience leading a strong software engineering team.
Strong hands-on software architecture and development skills.
Experience building a new technology area or product from an early stage.
Strong practical knowledge of LLMs, RAG, embeddings, and AI agents.
Experience integrating AI models with applications, APIs, tools, or data systems.
Strong programming skills in Python, C/C++, Java, or similar languages.
Good understanding of distributed systems and modern software architecture.
Experience with production performance, reliability, and scalability.
Ability to evaluate new AI technologies and turn them into a clear engineering plan.

Advantages
LangGraph or similar agent frameworks
Vector databases and semantic-search systems
PyTorch, Hugging Face, or similar AI frameworks
GPU and AI accelerator technologies
Database systems and SQL engines
Research, patents, or strong open-source contributions in AI systems
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792741
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
6 ימים
Location: Hod Hasharon
Job Type: Full Time
As a leading technology innovator, we push the boundaries of what's possible to enable next-generation experiences and drive digital transformation, creating a smarter, connected future for all.

As a company AI Software Engineer, you will develop and implement cutting-edge AI techniques that enable the efficient utilization of state-of-the-art solutions across various technology verticals.

In this position, you will be responsible for leading software development and implementation of the company AI Stack features within our AI Runtime (QAIRT) SDK, specifically Core software including Generative AI Inference Extensions. You will contribute to the efficient execution of advanced deep neural networks (DNNs), large language models (LLMs), and other modern AI architectures. You will also be responsible for supporting AI stack that optimizes runtime efficiency, latency and power consumption of AI applications running Machine Learning workloads on our low-power AI accelerators. The DSP core design work directly supports our NPU and GPU solutions, enabling optimized inference across diverse hardware platforms.

You will have the opportunity to demonstrate your passion for software design and development through your analytical, design, programming, and debugging skills.
Responsibilities
Develop real-time software for our AI DSPs to support the execution of the latest generative AI models on Snapdragon platforms

Develop and implement core components of our AI Stack runtime framework for inference on resource constrained AI systems, operating in low power, small memory footprint platforms

Validate, analyze, and optimize the performance and accuracy of software through detailed testing of machine learning use cases

Debug complex issues, perform root cause analysis, and ensure high system reliability

Collaborate with cross-functional teams to deliver robust, scalable AI software solutions

Contribute to a culture of technical excellence, knowledge sharing, and continuous improvement within the AI Software team

Participate in design and code reviews

Working in global environment and collaborating inside the team and across other AI teams
Requirements:
Minimum Qualifications (Must Have)
Bachelor's degree in Engineering, Information Systems, Computer Science, or related field

7+ years of general software development experience or 4+ years of related work experience on AI DSP systems

Proficiency in software development using C/C++

Experience with development in a Linux environment

Strong software development skills, including data structure and algorithm design, object-oriented or other software design paradigms, software debugging, and testing

Excellent communication skills (verbal, presentation, and writing)


Preferred Qualifications (Desired)
3+ years of experience building embedded software applications

Familiarity with various generative AI model architectures such as LLMs

Strong understanding of hardware acceleration and deployment of generative AI inference on edge devices as mobile, laptop and AR/VR/XR headset.

Experience with development in various platforms such as Linux, Android, or Windows

Experience with agile software development practices and git-based SCM

Ability to collaborate across a globally diverse team and manage multiple interests
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793087
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Hod Hasharon
Job Type: Full Time
We are looking for an experienced Tech Lead to lead a team of 3-5 software engineers developing innovative database security and data protection solutions for enterprise platforms.

This is a hands-on technical leadership role focused on cybersecurity, data security, and advanced software engineering. The successful candidate will combine deep security expertise with strong software architecture skills to define technical direction, lead engineering execution, mentor developers, and build enterprise-grade security solutions.

The role requires strong technical ownership, the ability to solve complex security challenges, and close collaboration with security researchers, product teams, customers, and engineering organizations. Occasional travel may be required for technical discussions and customer engagements.

Responsibilities
Lead and mentor a team of 3-5 software engineers, providing technical guidance and driving engineering excellence.
Define software architecture, technical strategy, and development direction for database security products.
Design and develop advanced cybersecurity and data security solutions for enterprise environments.
Lead security architecture discussions, technical design reviews, and code reviews.
Translate complex security requirements and research concepts into practical, scalable software solutions.
Drive secure software development practices across architecture, design, implementation, and deployment.
Solve complex technical challenges related to security, reliability, scalability, and performance.
Collaborate with security researchers, product management, and cross-functional engineering teams.
Participate in customer-facing technical discussions and occasional international travel.
Remain hands-on and contribute to critical software components and technical decisions.
Requirements:
Requirements
10+ years of software development experience.
Proven experience as a Tech Lead, Senior Software Engineer, or equivalent technical leadership role.
Strong expertise in cybersecurity, data security, or security software development.
Proven experience conducting cybersecurity research and developing security technologies.
Strong software architecture and system design skills.
Extensive hands-on experience with C/C++ development on Linux.
Experience designing and developing complex enterprise software systems.
Strong understanding of secure software development methodologies and security-focused system design.
Ability to analyze complex security problems and develop practical engineering solutions.
Excellent technical leadership, communication, and problem-solving skills.
Ability and willingness to travel occasionally.
B.Sc. in Computer Science, Software Engineering, or a related field (M.Sc. is an advantage).

Preferred Qualifications
Experience with relational databases such as PostgreSQL, Oracle, SQL Server, or MySQL.
Understanding of database security concepts including encryption, auditing, access control, authentication, authorization, and data protection technologies.
Background in security research, vulnerability analysis, threat modeling, or security architecture.
Experience analyzing attack surfaces and designing security mechanisms to protect complex software systems.
Experience with operating system security concepts including process isolation, privilege management, memory protection, and Linux security mechanisms.
Familiarity with cryptographic concepts and secure communication protocols.
Experience with vulnerability research, reverse engineering, binary analysis, fuzzing, or security testing methodologies.
Experience developing security-sensitive enterprise software, infrastructure products, or low-level system components.
Understanding of cloud security concepts and modern enterprise security architectures.
Familiarity with security standards and frameworks such as NIST, ISO 27001, CIS, OWASP, or similar
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792766
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
22/07/2026
Location: Haifa
Job Type: Full Time
We are seeking a Senior Software Development Engineer to design, develop, and optimize mission-critical embedded software for cloud infrastructure. You will join teams focused on networking, machine learning acceleration, and high-performance computing (HPC), impacting millions of AWS services globally.
Requirements:
Basic Qualifications
- Experience as a mentor, tech lead or leading an engineering team
- Experience leading the architecture and design (architecture, design patterns, reliability and scaling) of new and current systems
- Bachelor's degree
- 8+ years of professional experience in embedded software development, with strong proficiency in C/C++
- Hands-on experience developing firmware, device drivers, or user-space applications for embedded systems, including low-level hardware interaction

Preferred Qualifications
- Expertise in networking protocols and performance optimization for high-throughput, low-latency systems
- Ability to work in cross-functional, agile teams and communicate technical concepts effectively to stakeholders
- Experience with AWS cloud infrastructure or other large-scale distributed systems.
- Knowledge of hardware/software co-design.
- Familiarity with storage protocols.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8749497
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon and Haifa
Job Type: Full Time and Part Time
*Part-Time or Full-Time Opportunity

In this job opening, we are seeking a highly motivated and visionary researcher to join our team focused on chiplet-based computing systems. This role involves research and development in chiplet-based architectures, aiming to overcome performance and scalability limits of traditional compute platforms.

Responsibilities:
Conduct cutting-edge research on the co-design of communication and computation for wafer-scale systems. On the communication side, the work focuses on architecting Networks-on-Chip, providing scalable, high-bandwidth, and ultra-low-latency interconnects that efficiently span multiple dies and support fine-grained data movement. On the computation side, the research emphasizes optimizing and mapping training and inference workloads to fully exploit wafer-scale resources, improving performance, scalability, and energy efficiency through intelligent task placement and scheduling.
Investigate various aspects of system design, including:
Network-on-Chip design, further including Topology exploration, routing algorithms, and protocol/flow-control design to support high-throughput collective communication, low-latency synchronization, and efficient model/data parallelism
Expressing and Mapping LLM Parallelism to Wafer-Scale System that simplify LLM parallelism (e.g., tensor, pipeline, and data parallelism) and efficiently map training and inference workloads onto wafer-scale system
Requirements:
Master in Electrical Engineering, Computer Engineering, Computer Science, or a closely related field.
Background in interconnection networks and computer architecture is a must.
Demonstrated research expertise in one or both of the following areas:
Network-on-Chip (NoC) architecture, including routing, flow control, topologies, and performance modeling
LLM parallel execution and co-design, including workload decomposition, communication-computation overlap, and mapping to high-performance interconnects
Excellent analytical, problem-solving, and system-level thinking skills.
Excellent programming skills with hands-on experience in simulator modeling, development and simulator-based evaluation.
Strong verbal and written communication skills, and the ability to work both independently and collaboratively in a cross-disciplinary team.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792813
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon
Job Type: More than one
*Part-Time or Full-Time Opportunity

Responsibilities:
The research focuses on designing next-generation rack and SuperPod architectures that combine elasticity through resource pooling with optimal performance enabled by holistic co-design across applications, parallel programming models, interconnect fabrics, and compute, memory, and switch chip architectures. On the communication side, the work architects scalable, high-bandwidth, and low-latency scale-up fabrics across multiple chips at rack and SuperPod scale. On the memory side, it explores richer protocol semantics beyond traditional load/store operations to reduce unnecessary data movement and rethinks memory hierarchies to expose large-scale capacity with near-local latency. On the compute side, it analyzes the requirements of General Compute and Generative AI workloads and their parallel programming models to fully leverage large-scale system resources.
Investigate and prototype new architectural features, including but not limited to:
o Tiered Memory: Explore advanced memory organizations, including hardware-managed caching, to enable software-transparent, fine-grained promotion and demotion of cache lines across the memory hierarchy.
o Prefetching/Speculation: Evaluate and design existing and next-generation hardware prefetching and speculation mechanisms to effectively hide local and remote memory latencies at rack and SuperPod scale.
o Near-Memory/Network Processing: Develop support for key primitives executed at the memory and network layers to minimize unnecessary data movement across sparse, dense, and pointer-based data structures.
o Workload-Centric Co-Design: Study optimal parallelization strategies at rack and SuperPod scale for both General Compute and Generative AI workloads, and design dedicated hardware support for widely used parallel programming primitives such as RPCs and collective communication.
Write reports and papers on the research results and present them.
Requirements:
MSc in Computer Science, Electrical Engineering, or related field.
Background in Computer Architecture and Computer Fabrics is a must.
Creativity and the ability to think outside the box to develop innovative technologies.
Research experience in at least one of the following areas:
o Computer Architecture: Modern cache hierarchies, cache coherence protocols, memory systems, hardware prefetchers, and the memory-side of the core microarchitecture.
o Scale-up Fabrics: NVLink, UALink, CXL, UPI, IF, or PCIe.
o Parallel programming models: Collective libraires (NCCL or RCCL) and RPCs (gRPC or Thift).
o Workload optimization: General Compute and Generative AI workload composition (internals) and parallelization expertise at the scale of the rack or SuperPoD in both cloud and HPC environments.
Excellent analytical, problem-solving, and system-level thinking skills.
Strong interpersonal skills, with a collaborative spirit and the ability to work independently.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793401
סגור
שירות זה פתוח ללקוחות VIP בלבד