דרושים » תוכנה » Senior Software Advanced Developer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
2 ימים
Location: More than one
Job Type: Full Time
We are seeking a highly skilled and modern software engineer to develop and prototype brand new advancements in distributed training and inference using our Spectrum-X AI fabric. This role offers a rare chance to pioneer AI and networking technology, contributing to ground-breaking projects that will define the landscape of large-scale AI systems. Improve AI app-networking connection by refining communication, crafting congestion control, coding NIC firmware, and expanding switch SDK features for enhanced AI factory efficiency. Your work impacts large AI system development, scaling, and speed.

What youll be doing:

Prototype end-to-end solutions to improve distributed training and disaggregated inference performance.

Analyze and optimize communication flows across application, transport, and network layers.

Develop system software spanning communication libraries, drivers, and firmware integrations.

Collaborate with hardware, firmware, and SDK teams to co-design network features.

Validate and integrate prototypes into our AI infrastructure and products.
Requirements:
What we need to see:

BSc/MSc/PhD in Computer Science or Electrical Engineering.

5+ years of relevant experience and/or knowledge.

Deep understanding of networking and communication internals - NCCL, RDMA/RoCE, congestion control.

Hands-on experience with HW/SW/FW integration and low-level programming (C/C++, kernel, drivers).

Some background in distributed training systems (such as PyTorch DDP, Megatron-LM, DeepSpeed).


Ways to stand out from the crowd:

Demonstrated innovation and leadership turning prototypes into impactful product features.

Experience with programmable data planes (P4, eBPF, DOCA SDK, or switch SDKs).

Familiarity with NIC firmware scheduling, in-network compute, or congestion management.

Contributions to open-source projects, academic papers, or performance benchmarking tools.

Strong background in AI factory architectures, distributed inference, or network telemetry.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771228
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
חברה חסויה
Location: Yokne`am
Job Type: Full Time
We are seeking a highly skilled Senior Performance Engineer to join our Performance and R&D organizations. In this role, you will help build and evolve systems that support performance analysis, telemetry, and optimization for large-scale GPU- and CPU-based clusters used in AI and high-performance computing environments. You will work closely with hardware, networking, firmware, and software teams to collect, analyze, and interpret performance data from live systems. This is a fast-paced R&D environment where system behavior and requirements evolve rapidly, requiring adaptable engineering solutions and strong analytical thinking.

What youll be doing:

Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters.

Explore performance characteristics of high-performance networking and collective communications (e.g., NCCL, RDMA, MPI, RoCE).

Identify performance bottlenecks across networking, compute, memory, and system architecture.

Develop and enhance performance analysis, benchmarking, and diagnostic tools.

Define performance test plans and establish expectations for new technologies and platforms.

Collaborate across hardware, firmware, networking, systems, and software teams to provide actionable performance insights.

Support telemetry collection and data refinement efforts to enable accurate performance analysis.

Maintain high standards for data quality, reproducibility, and traceability of performance results.
Requirements:
What we need to see:

B.Sc. or M.Sc. in Computer Science, Computer Engineering, Software Engineering, or equivalent experience.

5+ years of experience in performance analysis, systems engineering, or HPC/AI infrastructure.

Demonstrated expertise in performance analysis skills and methodologies.

Hands-on experience with high-performance networking (RDMA, MPI, NCCL, congestion control).

Strong understanding of system performance metrics (latency, throughput, resource utilization).

Exposure to hardware, firmware, or embedded telemetry environments.

Strong analytical, problem-solving, and communication skills.

Ability to work effectively in cross-functional, fast-paced R&D teams.


Ways to stand out from the crowd:

Knowledge of CUDA, NCCL internals, and congestion control algorithms.

Deep system-level understanding of CPU architectures, GPUs, HCAs, memory, and PCIe.

Experience with NVIDIA GPUs, CUDA, and deep learning frameworks such as PyTorch or TensorFlow.

Experience with cloud platforms.

Proficiency in Python; experience with Bash and C/C++ is a plus as well as a strong experience working in Linux environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8770037
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are seeking an AI Networking Architect to join the Networking Research Group. This role will help bridge the gap between emerging tasks supported by advanced technologies and the data center infrastructure that powers them. In this role, you will work at the intersection of AI applications, distributed systems, networking hardware, and software architecture.

You will join a focused team of multidisciplinary engineers driving AI workload optimization through deep application understanding, network analysis, and end-to-end systems thinking. Your insights will directly shape our products across the full stack - from applications and software libraries to hardware architecture and physical design.

What Youll Be Doing:

Model the performance of complex AI workloads to identify bottlenecks and recommend system-level optimizations.

Analyze brand-new AI models, distributed training techniques, and inference workloads to understand their infrastructure requirements.

Build Platforms, simulations and HW platforms, execute AI workloads and build analytical tools to evaluate trade-offs across compute, memory, storage, and network behavior.

Translate research insights and workload behavior into actionable software, hardware, and networking architecture requirements.

Partner with architecture, software, and product teams to influence our future networking and AI infrastructure roadmaps.

Drive architectural innovation by applying deep workload analysis to real-world advanced machine learning frameworks.
Requirements:
What we need to see:

B.Sc. Or M.Sc. in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.

3+ years of relevant industry or research experience.

Strong machine learning or data science background, with hands-on experience in LLMs, generative AI, or deep learning systems.

Strong systems-level thinking, capable of estimating end-to-end requirements across the AI stack.

Shown ability to translate research findings and product requirements into clear software and hardware specifications.

Excellent research skills, including the ability to digest academic papers, self-learn new domains, and independently test hypotheses.

Advanced programming skills for performance modeling, data analysis, and prototyping.

Excellent communication skills, demonstrating proficiency in presenting complex technical findings clearly and confidently.


Ways to Stand Out from the crowd:

Experience with distributed training, distributed inference, or large-scale AI serving systems.

Experience in Agentic programming, and AI tools

Familiarity with GPU clusters, collective communication, storage systems, or AI networking bottlenecks.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8767980
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
Location: Ra'anana
Job Type: Full Time
We are looking for a Distinguished Engineer to help shape the future of AI infrastructure. As part of our advanced development and research efforts, you will drive technical vision and innovation across the AI software stack, spanning GPU computing, RDMA, NVLink, storage, and large-scale distributed systems. You will research next-generation system designs, influence product and platform direction, and help define how massive-scale AI environments are built, optimized, and operated through hardware-software co-design.

What youll be doing:

Research, define, and drive co-design solutions across AI software stacks involving GPU computing, RDMA networking, NVLink interconnects, storage, and large-scale distributed infrastructure.

Lead deep technical investigations and proof-of-concept development in areas related to distributed AI, deep learning systems, high-performance computing, virtualization, memory management, and large-scale systems design.

Work closely with multiple teams across us to advance next-generation AI infrastructure technologies and influence cross-stack co-design decisions.

Analyze end-to-end system bottlenecks and opportunities across compute, networking, memory, and storage layers to improve performance, scalability, and efficiency.

Guide the design of environments that support massive-scale AI training and inference workloads in production.
Requirements:
What we need to see:

M.S. or Ph.D. in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience.

20+ years of industry experience in systems design, distributed systems, system software, or related fields.

Chief Scientist, CTO, or equivalent senior technical leadership experience in smaller companies is highly desirable.

Deep background in distributed systems, parallel computation, networking, storage, virtualization, and memory management.

Experience building, scaling, and supporting environments for massive AI workloads.

Strong understanding of large-scale systems operational aspects, including performance, reliability, scalability, and supportability.

Excellent communication, collaboration, and technical leadership skills in a multi-national, multi-time-zone environment.


Ways to stand out from the crowd:

Proven research track record in AI systems, large-scale infrastructure, or advanced systems co-design.

Deep expertise in GPU systems, RDMA, NVLink, and storage technologies for AI and HPC workloads.

Strong record of cross-functional influence spanning hardware, system software, networking, and infrastructure.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8768266
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
2 ימים
Location: Yokne`am
Job Type: Full Time
We are seeking an exceptional Network Engineer to join the Vertical Verification Group. As a senior engineer, you will contribute to the Vertical Verification work, concentrating on validating and improving complex Ethernet and InfiniBand systems within our Data Center and HPC environments.

You will play a key role in testing advanced networking features, building complex customer-like topologies, and ensuring our products meet industry-leading benchmarks of performance, efficiency, and quality. You will collaborate closely with architects, software engineers, and hardware teams across NIC, HCA, switches, CPUs, and GPUs in a fast-paced and highly technical environment. This is an outstanding opportunity to join a highly skilled team and contribute significantly to groundbreaking technology!

What youll be doing:
Review architecture, build, and requirements for new networking features across Ethernet and InfiniBand portfolios.
Compose and build complex testbed topologies that emulate customer environments.
Implement, run and optimize integration test plans including functional, regression, and performance testing.
Identify, reproduce, and debug issues; work closely with R&D to drive root cause analysis and resolution.
Collaborate with automation teams.
Analyze and optimize network performance, latency, and efficiency.
Provide clear status updates, reports, and insights on system quality and performance.
Requirements:
What we need to see:
B.Sc. in Computer Science, Electrical Engineering, or related field (or equivalent experience).
8+ years of experience in networking, system validation, or related engineering roles.
Strong hands-on experience with Linux-based systems.
Deep understanding of networking protocols (TCP/IP, UDP, Ethernet, VLANs, L2/L3).
Experience with routing, switching, and modern data center network architectures.
Strong troubleshooting, debugging, and analytical skills in distributed environments.
Experience with test methodologies (functional, regression, performance, scale).
Scripting or programming experience (Python, Bash, etc.).
Independent, fast learner with strong ownership and communication skills.

Ways to stand out from the crowd:
Experience with RDMA technologies (RoCE / InfiniBand).
Knowledge of congestion control algorithms and performance tuning.
Familiarity with AI workloads and their networking requirements.
Experience with HPC environments and benchmarking tools.
Background with virtualization or container technologies (Kubernetes, KVM, etc.).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771793
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
Location: Yokne`am
Job Type: Full Time
We are looking for a Host and Systems Performance Manager to join ur Networking Performance team!

In this role, you will lead a team of engineers who measure, analyze, and improve NIC, host, DOCA networking stack, and system-level performance across our networking products and platforms. We work across the full product lifecycle, from early architecture and pre-silicon modeling through post-silicon bring-up, validation, characterization, and GA readiness.

We are looking for someone who enjoys building teams, turning complex performance data into clear direction, and partnering across engineering groups to improve products that support AI, HPC, cloud, and accelerated networking workloads.

What Youll Be Doing:
Lead, coach, and develop a team focused on NIC, host, DOCA networking stack, and system-level performance. You will define performance test plans, methodologies, metrics, and success criteria for our new networking technologies.
Guide pre-silicon and post-silicon performance planning, analysis, reporting, and readiness reviews. Your team will benchmark and profile workloads across RDMA, RoCE, InfiniBand, Ethernet, MPI, NCCL, DOCA, storage, security, and host networking stacks.
Identify bottlenecks across NIC, DPU, CPU, memory, PCIe, firmware, drivers, Linux networking, DOCA, and full-system architecture. We will look to you to lead root-cause analysis, coordinate mitigation plans, and help hardware, firmware, software, architecture, validation, and product teams align on performance goals.
Requirements:
What We Need To See:
B.Sc. or M.Sc. in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.
Experience in performance analysis, systems engineering, networking, or HPC/AI infrastructure.
10+ years of software engineering experience, including 4+ years in a leadership or management role.
Experience with high-performance networking technologies such as RDMA, RoCE, InfiniBand, Ethernet, MPI or NCCL.
Hands-on experience with system performance analysis, benchmarking, profiling, and root-cause analysis.
Understanding of host architecture, including CPUs, memory hierarchy, NUMA, PCIe, Linux OS, drivers, firmware, and DPU/NIC interactions.
Experience creating performance test plans for pre-silicon or post-silicon phases.
Programming or scripting experience with Python and Bash; C/C++ experience is helpful.
Ability to communicate technical findings clearly and work across engineering teams.

Ways To Stand Out from the Crowd
Experience with NICs or DPU architecture, networking offloads, host datapath behavior, networking services, or DPU offloads.
Experience with AI/HPC cluster performance, distributed training, inference workloads, collective communication libraries, telemetry pipelines, benchmark automation, CUDA, NCCL internals.
Linux kernel networking, DPDK, OVS, storage acceleration, or security offloads can also help you succeed in this role.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8768234
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
2 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
Our Network Architecture, Modeling and Performance Insights group is seeking a skilled and driven hands-on Network Modeling Architect to conduct advanced network research and optimization while utilizing and enhancing our advanced simulation tools.

In this role, you will be a key contributor to defining the architecture and improving network performance for AI and High Performance Computing (HPC) workloads. You will be responsible for modeling advanced network topologies, configurations, logic and traffic patterns, and analyzing their effect on end-to-end performance. You will develop expertise in our simulation tools and write code daily to contribute directly to their development road map and prioritization, develop core features in the simulator, and expose the capabilities to additional users. You will take part in complex architecture decisions, define and verify the modeling assumptions, and put them to the test vs. real HW and SW performance. If you're passionate about tackling intricate challenges and contributing to comprehensive systems and working on innovative solutions, we want to hear from you.



What you'll be doing:

Analyze and model our AI and High Performance Computing offerings, deep dive into network behavior for training and inference and generate insights to directly impact the architecture of next-generation of AI systems. Conduct advanced network modeling, research and optimization using innovative simulation tools.

Write production-quality C++ and Python code daily. Hands-on modeling of advanced network topologies, configurations and traffic patterns representing real life AI workloads, and analyze their effect on the system's performance. Actively develop the simulator software solution while contributing modular and scalable code.

Actively support the HW and SW development life cycles, map existing and candidate features into the simulation domain to enable their analysis, impact assessment and optimization.

Take full independent ownership of the performance modeling roadmap for a defined set of features. Interface directly with internal clients, communicate the analysis conclusions effectively and iterate on them. Drive projects from concept to completion.

Improve the quality of the simulation as a software product, ensuring robustness and reliability.
Requirements:
What we need to see:

BSc or above in Computer Science, Electrical Engineering, or a related field

5+ years of recent hands-on coding experience with strong proficiency in C++ and Python. You must be comfortable navigating, optimizing, and contributing to a complex, large-scale codebase.

End to end system perspective - capable of bridging the gap between hardware behavior, micro-architecture, and software performance (latency/throughput).

Project ownership capability with proven ability to lead technical initiatives, prioritize features, and manage project lifecycles autonomously with minimal supervision.


Ways to stand out from the crowd:

Strong background in communication networks technology. Deep understanding of Ethernet, NVLink, and/or Infiniband technologies, data center infrastructure, and network protocols.

Familiarity with discrete event simulators (e.g., OMNeT++, SystemC, or proprietary architectural simulators).

 Advanced degree or experience focusing on computer architecture, algorithms, networking, or distributed systems.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771204
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
Location: Ra'anana and Yokne`am
Job Type: Full Time
We are looking for a Senior Software Engineer to join our team developing the core technologies behind our SmartNIC platform. Your work will directly impact next-generation AI infrastructure, cloud networking, and high-performance data centers used by customers around the world.

What you'll be doing:

Architect, design, and develop next-generation networking acceleration technologies.

Build and enhance core software components within the DOCA SDK.

Work closely with architects, customers, and engineering teams to understand requirements and deliver scalable, high-performance solutions.

Collaborate across the software stack, including applications, virtualization, drivers, Linux kernel, firmware, and hardware teams.

Drive technical decisions and contribute to the design of performant, maintainable software.
Requirements:
What we need to see:

B.Sc. in Computer Science, Software Engineering, or equivalent practical experience.

8+ years of experience developing in C/C++.

Experience with Linux development tools and debugging.

Solid understanding of networking concepts and protocols (such as Ethernet, TCP/IP, or similar).

Experience with virtualization technologies.

Strong analytical and problem-solving skills.

Experience optimizing software performance. Good understanding of operating systems and computer architecture.


Ways to stand out from the crowd:

Experience developing high-performance or low-latency software.

Background with DPDK, RDMA, or similar networking technologies.

Experience designing APIs or SDKs.

Contributions to open-source projects such as DPDK, OVS, or the Linux Kernel.

A collaborative mindset, curiosity, and excellent communication skills.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8770044
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
In this role, you will be a key contributor to defining the architecture and improving network performance for AI and High Performance Computing (HPC) workloads. You will be responsible for modeling advanced network topologies, configurations, logic and traffic patterns, and analyzing their effect on end to end performance. You will develop expertise in our simulation tools and contribute directly to their development road map and prioritization, develop core features in the simulator, and expose the capabilities to additional users. You will take part in complex architecture decisions, define and verify the modeling assumptions, and put them to the test vs. real HW and SW performance.

If you're passionate about tackling intricate challenges and contributing to comprehensive systems and working on innovative solutions, we want to hear from you.

What you'll be doing:

Analyze and model our AI and High Performance Computing offerings, deep dive into network behavior for training and inference and generate insights to directly impact the architecture of next-generations of AI systems. Conduct advanced network modeling, research and optimization using innovative simulation tools.

Implement modeling of advanced network topologies, configurations and traffic patterns which represent real life AI workloads, and analyze their effect on the systems performance. Actively develop the simulator software solution while contributing modular and scalable code.

Actively support the HW and SW development life cycles, map existing and candidate features into the simulation domain to enable their analysis, impact assessment and optimization.

Take full independent ownership of the performance modeling roadmap for a defined set of features. Interface directly with internal clients, communicate the analysis conclusions effectively and iterate on them. Drive projects from concept to completion.

Improve the quality of the simulation as a software product, ensuring robustness and reliability.
Requirements:
What we need to see:

BSc or above in Computer Science, Electrical Engineering, or a related field

5+ years of hands-on coding experience with strong proficiency in C++ and Python. You must be comfortable navigating, optimizing, and contributing to a complex, large-scale codebase.

End to end system perspective - capable of bridging the gap between hardware behavior, micro-architecture, and software performance (latency/throughput).

Project ownership capability with proven ability to lead technical initiatives, prioritize features, and manage project lifecycles autonomously with minimal supervision.


Ways to stand out from the crowd:

Strong background in communication networks technology. Deep understanding of Ethernet, NVLink, and/or Infiniband technologies, data center infrastructure, and network protocols.

Familiarity with discrete event simulators (e.g., OMNeT++, SystemC, or proprietary architectural simulators).

Advanced degree or experience focusing on computer architecture, algorithms, networking, or distributed systems.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8768103
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
חברה חסויה
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are looking for an excellent Drivers Team leader to lead the Inbox Linux Drivers team and make a wide impact on the company products and teams worldwide. The role includes leading the development and maintenance of cutting-edge Linux kernel drivers and userspace libraries for our networking technologies, while managing a team of highly skilled Linux engineers responsible for delivering production-quality software across multiple products and distributions.

What youll be doing:

Lead a team of Linux kernel engineers responsible for developing, maintaining, and delivering Linux networking drivers for our NICs across multiple product lines, including Inbox drivers, our Kernel, and DOCA Host.

Drive the design, implementation, maintenance, and integration of Linux kernel drivers and accompanying userspace libraries for state-of-the-art Ethernet and RDMA networking technologies.

Own the teams technical execution, roadmap, prioritization, and delivery while mentoring engineers and fostering technical excellence.

Collaborate closely with architecture, firmware, hardware, validation, DevOps, product, and customer-facing teams across the globe to deliver robust and scalable networking solutions.

Lead complex debugging and performance investigations across the Linux networking stack, working with both internal stakeholders and open-source communities.

Ensure high software quality through strong engineering practices, code reviews, automation, and continuous integration.

Be part of an experienced organization with a collaborative culture, solving challenging problems that power next-generation AI, cloud, and high-performance computing infrastructure.
Requirements:
What we need to see:

Bachelors or Masters degree in Computer Science, Computer Engineering, or equivalent practical experience.

7+ overall years of hands-on experience developing software in C/C++, including Linux systems programming.

3+ years of experience leading engineering teams, driving technical execution, mentoring engineers, and delivering complex software projects.

Strong experience working in Linux environments, development tools, build systems, and debugging methodologies.

Deep understanding of Linux kernel architecture and networking subsystems.

Strong knowledge of networking technologies and protocols, including Ethernet, RDMA, and InfiniBand.

Experience developing, maintaining, or debugging Linux kernel drivers or low-level system software.

Excellent design, coding, troubleshooting, and problem-solving skills.

Ability to balance technical leadership with people management while maintaining a strong focus on execution and delivery.

Strong communication and collaboration skills, with experience working across globally distributed engineering teams.

Ways to stand out from the crowd:

Experience with RDMA technologies, including RoCE and InfiniBand.

Strong Linux kernel development experience, including upstream kernel contributions.

Familiarity with Linux networking internals (Netdev, ethtool, devlink, NAPI, XDP, or related subsystems).

Experience contributing to open-source software and collaborating with upstream Linux communities.

Familiarity with our networking products, DOCA Host, or similar networking software stacks.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8770048
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
3 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
We seek a versatile Senior Software Engineer who is passionate about performance optimization and generative AI. Our team brings the latest research in LLM inference - from novel decoding strategies to quantization schemes - into production across our hardware lineup, from large data center servers to powerful edge devices. We work on the most advanced architectures in the field, with a focus on NVIDIA's own.

What you'll be doing:

Implement and optimize inference algorithms for LLM and omnimodal architectures, including hybrid Mamba-Transformer and mixture-of-experts models.

Profile inference pipelines using NVIDIA's profiling and simulation tools. Correlate simulation predictions against real hardware across data center and edge devices.

Write and tune GPU kernels (CUDA, Triton) for operators like fused MoE layers, SSM state updates, and quantized GEMMs.

Solve distributed inference problems: expert parallelism, communication-compute overlap, collective tuning, multi-node deployment.

Build production-grade software inside major open-source libraries - vLLM, SGLang, Dynamo, FlashInfer.

Own optimization features end-to-end, from scoping through delivery, collaborating with research, product, and engineering teams worldwide.
Requirements:
What we need to see:

B.Sc., M.Sc., or equivalent experience in Computer Science or Computer Engineering.

5+ years of hands-on software engineering experience in performance-critical systems.

Solid understanding of deep learning architectures (Transformers, SSMs, MoE, ).

Experience with systems where hardware constraints matter: GPU programming, memory hierarchy, networking, or distributed computing.

Strong software engineering fundamentals: clean design, extensibility, testability. Good judgment about when complexity is warranted.

Effective communicator who works well across teams and time zones.

Experience optimizing deep learning workloads on our GPUs using roofline models, Nsight/PyTorch profilers and end-to-end traces.


Ways to stand out from the crowd:

Contributions to open-source inference runtimes and libraries - vLLM, SGLang, FlashInfer, Dynamo or similar.

Hands-on work with LLM quantization (FP8, NVFP4, MXFP8, mixed-precision) and practical understanding of numerical precision tradeoffs.

Track record with distributed inference at scale: tensor parallelism, pipeline parallelism, expert parallelism, disaggregation, multi-node orchestration.

Deep knowledge of the latest LLM architectural trends: multi-token predictors, sparse hybrid models, attention and state-space mechanisms.

Experience with performance modeling and simulation-to-silicon correlation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769559
סגור
שירות זה פתוח ללקוחות VIP בלבד