דרושים » הנדסה » Software Engineer - AI Datacenter Networking

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
we are looking for a Software Engineer - AI Datacenter Networking.
Job Summary:
Develop and optimize high-performance software for AI datacenter networking, implementing NOS enhancements, smart NIC integration, and specialized routing features for AI workloads.
Key Responsibilities:
Implement network operating system features for AI-optimized switching scenarios
Integrate smart NIC solutions for end-to-end AI-optimized networking
Create monitoring and telemetry collection mechanisms for datapath analysis
Implement load balancing algorithms optimized for AI workload patterns
Debug and optimize datapath performance bottlenecks
Collaborate with QA teams on feature testing and validation
Requirements:
BSc degree in Computer Science or Engineering
5+ years of software development experience in networking or systems programming
Understanding of networking protocols and packet processing optimization
Preferred Qualifications:
Experience with SONiC development, SAI implementation, or switch software
Rust
Knowledge of Linux kernel networking
Experience with smart NIC programming interfaces and SDK development
Familiarity with DPDK, RDMA, and high-performance networking libraries
Experience with AI/ML framework networking integration
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8763871
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Software Team Lead.
Job Summary
Develop and optimize high-performance software for AI datacenter networking, implementing NOS enhancements, smart NIC integration, and specialized routing features for AI workloads.
Key Responsibilities:
Implement network operating system features for AI-optimized switching scenarios
Integrate smart NIC solutions for end-to-end AI-optimized networking
Create monitoring and telemetry collection mechanisms for datapath analysis
Implement load balancing algorithms optimized for AI workload patterns
Debug and optimize datapath performance bottlenecks
Collaborate with QA teams on feature testing and validation
Requirements:
BSc degree in Computer Science or Engineering
2+ years of team leads.
5+ years of software development experience in networking or systems programming
Understanding of networking protocols and packet processing optimization
Preferred Qualifications
Experience with SONiC development, SAI implementation, or switch software
Rust
Knowledge of Linux kernel networking
Experience with smart NIC programming interfaces and SDK development
Familiarity with DPDK, RDMA, and high-performance networking libraries
Experience with AI/ML framework networking integration
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8763837
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Job Type: Full Time
We are looking for an outstanding Senior Networking Software Architect to join the NIC/DPU Software and Firmware Architecture group. In this role, you will help define the next generation of our datacenter and AI networking platforms, with focus on DPU management, QoS, performance, telemetry, and software architecture across stacks. You will work closely with hardware designers, firmware/kernel driver teams, system engineers, validation, product management and customers. The role spans early architecture definition, pre-silicon design, bring-up, and production readiness for large-scale AI and cloud datacenter deployments.

What Youll Be Doing:

Own software and system architecture for next-generation DPU management, QoS, performance, telemetry, and observability features.

Define end-to-end control and management flows across DOCA, host drivers, embedded firmware, BMC, management controllers and external management systems.

Specify telemetry and observability requirements, including counters, logs, traces, events, health monitoring, debug data, and streaming telemetry.

Define management interfaces and APIs for configuration, provisioning, lifecycle operations, diagnostics, and field serviceability.

Write clear architecture specifications, interface definitions, flow diagrams, and design documents for software, firmware, and system teams.

Partner with R&D teams to translate high-level architecture into implementable designs and guide features through development, validation, silicon bring-up, and production.

Analyze system performance bottlenecks, interoperability issues, telemetry gaps, and customer-reported issues, then feed learnings into future architecture.

Collaborate with system and cluster architects to ensure NIC/DPU features fit end-to-end AI datacenter and cloud networking designs.
Requirements:
What We Need To See:

B.Sc. or M.Sc. in Computer Engineering, Computer Science, Electrical Engineering, or equivalent experience.

9+ years of experience in networking, system software, embedded software, firmware, or datacenter infrastructure.

Proven experience in software architecture, or technical leadership roles.

Deep understanding of networking concepts and protocols such as Ethernet, TCP/IP, RDMA/RoCE, congestion control, QoS, virtualization overlays, and traffic management.

Strong background with DPUs, SmartNICs, or other high-performance networking devices.

Experience with system management, provisioning, monitoring, telemetry, diagnostics, or lifecycle-management flows.

Familiarity with management protocols and frameworks such as Redfish, PLDM, MCTP, IPMI, gNMI, SNMP, Netconf, REST, or gRPC-based APIs.

Ability to lead cross-functional architecture discussions across software, firmware, hardware, validation, product, and customer-facing teams.

Excellent written and verbal communication skills, including the ability to create clear architecture documents and present trade-offs.


Ways To Stand Out From The Crowd:

Experience defining software architecture for DPU products, including management, telemetry, QoS, performance, security, virtualization, or offload features.

Hands-on background with Linux networking, device drivers, firmware, embedded Linux, BMC software, DOCA, DPDK, OVS or Kubernetes networking.

Experience with performance counters, profiling tools, eBPF, Prometheus, Grafana, dashboards, heat maps, or large-scale telemetry systems.

Experience in defining and developing GAI-based analysis tools to extract insights from telemetry data and streams.

Background in RAS, diagnosability, serviceability, field failure analysis, production debug, or customer escalation handling.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773378
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
06/08/2026
Location: More than one
Job Type: Full Time
We are seeking a highly skilled and modern software engineer to develop and prototype brand new advancements in distributed training and inference using our Spectrum-X AI fabric. This role offers a rare chance to pioneer AI and networking technology, contributing to ground-breaking projects that will define the landscape of large-scale AI systems. Improve AI app-networking connection by refining communication, crafting congestion control, coding NIC firmware, and expanding switch SDK features for enhanced AI factory efficiency. Your work impacts large AI system development, scaling, and speed.

What youll be doing:

Prototype end-to-end solutions to improve distributed training and disaggregated inference performance.

Analyze and optimize communication flows across application, transport, and network layers.

Develop system software spanning communication libraries, drivers, and firmware integrations.

Collaborate with hardware, firmware, and SDK teams to co-design network features.

Validate and integrate prototypes into our AI infrastructure and products.
Requirements:
What we need to see:

BSc/MSc/PhD in Computer Science or Electrical Engineering.

5+ years of relevant experience and/or knowledge.

Deep understanding of networking and communication internals - NCCL, RDMA/RoCE, congestion control.

Hands-on experience with HW/SW/FW integration and low-level programming (C/C++, kernel, drivers).

Some background in distributed training systems (such as PyTorch DDP, Megatron-LM, DeepSpeed).


Ways to stand out from the crowd:

Demonstrated innovation and leadership turning prototypes into impactful product features.

Experience with programmable data planes (P4, eBPF, DOCA SDK, or switch SDKs).

Familiarity with NIC firmware scheduling, in-network compute, or congestion management.

Contributions to open-source projects, academic papers, or performance benchmarking tools.

Strong background in AI factory architectures, distributed inference, or network telemetry.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8771228
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are seeking a highly skilled and versatile Performance Research and Analysis Manager to join our Performance Group. This role will drive end-to-end performance strategy and execution for next-generation our data centers and solutions based on GPU systems, NIC, Switch, DPU and Networking technologies. The ideal candidate will oversee, evaluating, and optimizing end-to-end AI GPU cluster-level performance for scaling out large scale distributed training and inference jobs communication. The role will focus heavily on RDMA, Networking Protocols, Collective Communication, Congestion Control, and Load Balancing algorithms. Secondarily, you will lead our DPUs and Storage technologies for N-S use cases to support AI Inference jobs. Third, you will drive our Performance Dashboards and Observability for cluster-level performance analysis from a stream line telemetry across NICs, Switches, GPUs, and NVlink.

What you'll be doing:

Drive end-to-end performance strategy, characterization, test plans, and optimization for next-generation our AI GPU clusters, focusing on large-scale distributed training and inference workloads.

Deeply evaluate and optimize our Networking core technologies performance, including RDMA/PRDMA, networking protocols, collective communication (NCCL), congestion control, and load-balancing algorithms.

Work on performance research and analysis of NVIDIA DPUs and storage technologies in North-South (N-S) use cases and deployment scenarios to maximize performance and efficiency for AI inference jobs.

Drive the strategy for performance observability and dashboards across next-generation NVIDIA data center solutions and supercomputers by leveraging scalable, streamlined telemetry pipelines to build performance dashboards and automated analytics based on real-time performance metrics across NICs, Switches, GPUs, and NVLink boundaries.

Perform deep root-cause analysis (RCA) on complex multi-node performance bottlenecks, driving actionable mitigation plans across hardware, firmware, and software teams.
Requirements:
What we need to see:

B.Sc. or M.Sc. in Computer Science, Computer Engineering, Software Engineering, or equivalent technical experience.

8+ overall years of experience and deep expertise in High Performance Networking, RDMA, and Systems level performance.

3+ years of experience as an engineering team manager leading technical performance or R&D teams.

Hands-on experience analyzing and optimizing collective communication (e.g., NCCL, MPI) and network traffic patterns for large-scale distributed AI workloads (LLM training and inference).

Hands-on experience designing, deploying, and customizing Grafana dashboards for cluster monitoring, alerting, and data visualization.

Exceptional cross-team leadership, analytical thinking, and communication skills to drive alignment across hardware, software, and architecture groups.


Ways to stand out from the crowd:

Proven track record of optimizing NCCL, RDMA/RoCEv2, and custom collective algorithms specifically tailored for multi-thousand GPU deployments running LLMs or Mixture-of-Experts (MoE) architectures.

Deep experience tuning advanced network traffic mechanisms such as adaptive routing, PFC/ECN congestion control, and packet-spraying technologies.

Experience building autonomous performance-driven tools, AI-assisted root cause analysis agents, or automated regression frameworks for continuous cluster-level performance evaluation.

Hands-on experience developing custom Grafana plugins, complex dashboard panels, or integrated alert management workflows using PromQL/LogQL for hyperscale or HPC environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773270
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Software Engineer III, Cloud Networking
About the job
Our software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to our needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.
This Team owns the core data plane infrastructure for all NAT-based products within the Andromeda network stack (e.g., private service connect (PSC) and Cloud network address translation (NAT). These products serve as critical entry and exit gateways that securely connect Cloud customers to their services and networks.
Responsibilities
Design, develop, test, and maintain high-performance software for our Cloud's core networking and secure connectivity platforms.
Manage complex scalability, resource efficiency, and performance optimization challenges to evolve our cloud networking datapath.
Collaborate with cross-functional engineering partners to define and build new network architectures and routing methods.
Manage individual project priorities, deadlines, and deliverables, ensuring high engineering velocity and robust code quality.
Participate in operational rotations, monitor production health, and troubleshoot packet latency or connectivity issues to keep our global systems healthy.
Requirements:
Minimum qualifications:
Bachelors degree or equivalent practical experience.
2 years of experience with software development or 1 year of experience with an advanced degree in an industry setting.
2 years of experience with developing large-scale infrastructure, distributed systems or networks, or experience with compute technologies, storage or hardware architecture.
Preferred qualifications:
Master's degree or PhD in Computer Science or related technical fields.
2 years of experience with data structures and algorithms.
Experience with systems programming (C++ preferred) and developing or debugging software in Unix/Linux user-space or kernel environments.
Experience with software architecture, engineering productivity, C++, C, Python, network architecture, large-scale distributed systems.
Familiarity with building, analyzing, and optimizing networking components and protocols (e.g., Cloud NAT, firewalls, load balancers, TCP/IP, or Software-Defined Networking).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785731
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
04/08/2026
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are seeking an AI Networking Architect to join the Networking Research Group. This role will help bridge the gap between emerging tasks supported by advanced technologies and the data center infrastructure that powers them. In this role, you will work at the intersection of AI applications, distributed systems, networking hardware, and software architecture.

You will join a focused team of multidisciplinary engineers driving AI workload optimization through deep application understanding, network analysis, and end-to-end systems thinking. Your insights will directly shape our products across the full stack - from applications and software libraries to hardware architecture and physical design.

What Youll Be Doing:

Model the performance of complex AI workloads to identify bottlenecks and recommend system-level optimizations.

Analyze brand-new AI models, distributed training techniques, and inference workloads to understand their infrastructure requirements.

Build Platforms, simulations and HW platforms, execute AI workloads and build analytical tools to evaluate trade-offs across compute, memory, storage, and network behavior.

Translate research insights and workload behavior into actionable software, hardware, and networking architecture requirements.

Partner with architecture, software, and product teams to influence our future networking and AI infrastructure roadmaps.

Drive architectural innovation by applying deep workload analysis to real-world advanced machine learning frameworks.
Requirements:
What we need to see:

B.Sc. Or M.Sc. in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.

3+ years of relevant industry or research experience.

Strong machine learning or data science background, with hands-on experience in LLMs, generative AI, or deep learning systems.

Strong systems-level thinking, capable of estimating end-to-end requirements across the AI stack.

Shown ability to translate research findings and product requirements into clear software and hardware specifications.

Excellent research skills, including the ability to digest academic papers, self-learn new domains, and independently test hypotheses.

Advanced programming skills for performance modeling, data analysis, and prototyping.

Excellent communication skills, demonstrating proficiency in presenting complex technical findings clearly and confidently.


Ways to Stand Out from the crowd:

Experience with distributed training, distributed inference, or large-scale AI serving systems.

Experience in Agentic programming, and AI tools

Familiarity with GPU clusters, collective communication, storage systems, or AI networking bottlenecks.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8767980
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
חברה חסויה
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are looking for an outstanding engineer to join our software architecture group and help shape the next generation of our networking products, SmartNICs, and DPUs. In this role, you will work closely with system architects, software/firmware engineers, hardware designers and product stakeholders to define component-level and system-level solutions for advanced datacenter platforms.

You will join a team working at the intersection of AI networking, virtualization, operating systems, security, management, and programmable acceleration. Our mission is to push the boundaries of modern datacenter infrastructure and help define the platforms of tomorrow.

What Youll Be Doing:

Work with architects and engineers to define software architecture for our networking devices, SmartNICs, and DPUs.

Analyze system requirements across networking protocols, virtualization, operating systems, security, management, and services.

Research new ideas, evaluate design alternatives, and build proof-of-concept models, simulations, and prototypes.

Collaborate with software, firmware, and hardware teams on HW/SW co-design, performance, scale, and resource-utilization challenges.

Learn complex datacenter technologies and turn them into clear architecture proposals, design documents, and technical presentations.
Requirements:
What We Need To See:

Engineer holding a B.Sc. or M.Sc. in Computer Science, Electrical Engineering, Computer Engineering, or a related field.

1+ years of relevant experience

Proven strong programming skills, in C/C++, preferably for large and complex embedded systems.

Solid understanding of operating systems, multi-threading, computer architecture, and software fundamentals.

Curiosity and ability to learn new technologies quickly, investigate deeply, and work through ambiguous problems.

Independent problem solver who also works well in a team environment.

Clear written and verbal communication skills.


Ways To Stand Out From The Crowd:

Knowledge of networking protocols such as Ethernet, InfiniBand, RDMA, RoCE, TCP/IP, or congestion control.

Experience with Linux, Linux kernel, device drivers, embedded systems, or system software.

Experience building simulations, performance models, prototypes, or proof-of-concept software.

Exposure to virtualization, security, programmable acceleration, SmartNICs, DPUs, or DOCA.

Open-source contribution, previous internship experience, or hands-on project work in relevant technical domains.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773393
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
חברה חסויה
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
We are looking for an excellent Drivers Team leader to lead the Inbox Linux Drivers team and make a wide impact on the company products and teams worldwide. The role includes leading the development and maintenance of cutting-edge Linux kernel drivers and userspace libraries for our networking technologies, while managing a team of highly skilled Linux engineers responsible for delivering production-quality software across multiple products and distributions.

What youll be doing:

Lead a team of Linux kernel engineers responsible for developing, maintaining, and delivering Linux networking drivers for our NICs across multiple product lines, including Inbox drivers, our Kernel, and DOCA Host.

Drive the design, implementation, maintenance, and integration of Linux kernel drivers and accompanying userspace libraries for state-of-the-art Ethernet and RDMA networking technologies.

Own the teams technical execution, roadmap, prioritization, and delivery while mentoring engineers and fostering technical excellence.

Collaborate closely with architecture, firmware, hardware, validation, DevOps, product, and customer-facing teams across the globe to deliver robust and scalable networking solutions.

Lead complex debugging and performance investigations across the Linux networking stack, working with both internal stakeholders and open-source communities.

Ensure high software quality through strong engineering practices, code reviews, automation, and continuous integration.

Be part of an experienced organization with a collaborative culture, solving challenging problems that power next-generation AI, cloud, and high-performance computing infrastructure.
Requirements:
What we need to see:

Bachelors or Masters degree in Computer Science, Computer Engineering, or equivalent practical experience.

7+ overall years of hands-on experience developing software in C/C++, including Linux systems programming.

3+ years of experience leading engineering teams, driving technical execution, mentoring engineers, and delivering complex software projects.

Strong experience working in Linux environments, development tools, build systems, and debugging methodologies.

Deep understanding of Linux kernel architecture and networking subsystems.

Strong knowledge of networking technologies and protocols, including Ethernet, RDMA, and InfiniBand.

Experience developing, maintaining, or debugging Linux kernel drivers or low-level system software.

Excellent design, coding, troubleshooting, and problem-solving skills.

Ability to balance technical leadership with people management while maintaining a strong focus on execution and delivery.

Strong communication and collaboration skills, with experience working across globally distributed engineering teams.

Ways to stand out from the crowd:

Experience with RDMA technologies, including RoCE and InfiniBand.

Strong Linux kernel development experience, including upstream kernel contributions.

Familiarity with Linux networking internals (Netdev, ethtool, devlink, NAPI, XDP, or related subsystems).

Experience contributing to open-source software and collaborating with upstream Linux communities.

Familiarity with our networking products, DOCA Host, or similar networking software stacks.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8770048
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Software Development Engineer to own the design and implementation of our inference data plane. We build the software that makes large models run efficiently on custom hardware - spanning model execution, memory management, data movement, and serving integration.
Our work covers the full inference path: integrating serving engines with custom hardware, developing high-performance compute kernels, enabling efficient data movement, and driving models from early validation through production. We operate at frontier scale with large distributed models.
This is a ground-up effort with rapidly evolving hardware and software. We need an individual contributor who can write and optimize low-level code for custom hardware, validate model architectures end-to-end, build test and profiling infrastructure, and drive performance across the stack.

Key job responsibilities
- Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference.
- Implement and validate LLM architectures end-to-end - from PyTorch model definition through distributed execution on custom hardware.
- Integrate custom accelerator backends into open-source ML serving frameworks (vLLM, PyTorch), including scheduler extensions, memory management, and model parallelism.
- Build and maintain test infrastructure for model correctness validation across CPU, GPU, simulator, and hardware targets.
- Profile and optimize inference workloads - identify bottlenecks, instrument critical paths, and drive latency and throughput improvements from simulation through hardware bringup.
- Own features end-to-end: from design through implementation, testing, and integration into the broader software stack.
- Contribute to CI/CD pipelines that gate model and kernel changes on correctness and performance regressions.
Requirements:
Basic Qualifications
- Bachelor's degree or equivalent.
- 4+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
- Knowledge of computer architecture, operating systems, and parallel computing.
- Strong proficiency in C/C++.
- Strong Linux systems knowledge.
- Experience developing compute kernels for GPUs, DSPs, or custom accelerators.
- Proven track record of owning and delivering complex software features end-to-end.

Preferred Qualifications
- Knowledge of ML frameworks including JAX, PyTorch, vLLM, SGLang, Dynamo, TorchXLA, and TensorRT.
- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques.
- Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware.
- Familiarity with speculative decoding, KV cache optimization, or other LLM serving optimizations.
- Experience with distributed systems - collective communication, RDMA, or high-speed interconnect programming.
- Demonstrated early adopter of AI-assisted development tools - uses LLMs or code-generation agents as part of daily workflow.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774268
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
we are looking for a Software Engineer Infra.
Responsibilities:
- Design and develop the core infrastructure components that power our products across a large-scale distributed system.
- Own and evolve key infrastructure domains such as high-availability, data-management, security, and telemetry.
- Produce excellent software design, write clean and efficient code, and debug complex issues across the stack.
- Mentor and guide junior engineers, sharing knowledge and raising the technical bar of the team.
- Turn ambitious, out-of-the-box ideas into real, production-grade value for our customers.
Requirements:
- BSc in Computer Science or a related degree, or equivalent practical experience.
- 5+ years of experience working as a Software Engineer.
- Strong proficiency in C++ / C.
- Experience with Rust and Python.
- Fast and self-driven learner of new technologies and programming languages.
Soft Skills
- Genuine passion for software development and craftsmanship.
- Ability to mentor and guide junior engineers.
- Creative, out-of-the-box thinker who can also execute and deliver.
Nice to Have / Advantage
- Experience with networking concepts (IPv4/6, ARP, DAD, routing, neighbors, etc.) - significant advantage.
- Experience with Linux networking, including Netlink and Linux network interfaces - significant advantage.
- Experience developing distributed systems.
- Experience with Linux kernel development or kernel internals.
- Experience working in a Linux environment.
- Experience developing software infrastructures.
- Experience with computer networking software.
- Experience with high-scale systems and performance optimizations.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8763851
סגור
שירות זה פתוח ללקוחות VIP בלבד