דרושים » ניהול ביניים » GPU Senior Software Engineer

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
We are seeking a skilled software engineer to join our NPU software stack development team. This role involves developing high-performance GPU programming frameworks, runtime systems, and libraries for AI/ML workloads. You will be responsible for implementing, optimizing, and maintaining GPU software stack components to support distributed AI training and inference.
Key Responsibilities:
Identify bottlenecks, analysis and optimize in distributed NPU eco-system
Design and develop NPU memory management system
Design and develop optimized NPU development framework, execution path and debugging
Develop compatibility with AI frameworks (Triton, PyTorch, JAX)
Write high-quality, well-tested code with comprehensive documentation
Collaborate with other teams (Hardware, Network, QA, AI Framework Integration)
Participate in code reviews and technical design discussions
Requirements:
5+ years of experience in distributed system programming
3+ years of experience with NPU programming (Triton, CUDA, HIP, OpenCL)
Expert-level C/C++ programming with focus on performance optimization
Expert-level Python programming with focus on DL/ML frameworks (PyTorch/JAX/etc)
Deep understanding of NPU architecture, memory tiering, and programming models
Knowledge of NPU runtime systems
Experience with performance profiling and optimization tools
Strong problem-solving and debugging skills
Experience with version control systems, Ticking system and collaborative development
Team player with excellent communication skills
Fast learner, highly organized, detail-oriented with high motivation
Preferred Qualifications:
Experience with NPU software stack development
Experience with large-scale NPU systems (100+ NPUs)
Experience with DL/ML workloads (oriented AI) and distributed training / inferencing
Familiarity with containerization and orchestration
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764145
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
28/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a Senior Software Engineer.
The group is a high-powered team responsible for implementing algorithms at scales of 100s of PBs. The team also manages the core filesystem components, including blocks and metadata management, snapshots, RAID logic, object-store tiering, unique cloud disaster recovery features, and more. And most importantly, they skillfully handle the most delicate part of the solution - our customers data.
As a Senior Software Engineer, youll:
Design and develop distributed file system components to support data management features such as snapshots, replication, tiering, and advanced data reduction algorithms;
Participate in the design, architecture, and implementation of next-generation storage architecture;
Assist in technically managing initial storage implementations including proofs-of-concept;
Diagnose bottlenecks and implement clean and performant solutions to achieve unbeatable performance;
Design algorithms and data structures to make sure customer data is safe and coherent across our solution in a wide variety of failure modes; and
Constantly revisit the architecture, algorithms, and methodologies to improve productivity, reliability, and maintainability.
Requirements:
Mastery of low-level and performant programming in C or C++/ Rust
A thorough understanding of concurrency, inter-process communication, threading models, and synchronization concepts, including significant experience with complex multithreaded software design
Experience identifying, reproducing, and resolving complex software defects, including root cause isolation, tracing through large source codebases, and implementing long-term fixes as well as short-term workarounds
5+ years of hands-on experience with Linux development and debugging, along with a broad knowledge and understanding of Linux internals
It's nice if you have:
Experience in data-path design and development
Experience with development of highly-distributed systems
Deep familiarity with concepts and features from the storage industry, including snapshots, replication, transparent data migration, and data reduction techniques
Experience with ZFS, XFS, or other file systems or with enterprise storage solutions
Experience working with the Linux filesystem community
Contribution, upstreaming, or maintaining of filesystem code
Experience playing a significant role in the implementation of a concurrent, long-running performant server
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8757553
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Help build an Always-On, low-overhead GPU profiling service that runs in production, scales across cluster environments, and delivers actionable insights for ML workloads. You will be hands-on delivering our profiling solutions across system software, drivers, and CUDA to make profiling continuously available and reliable.

What youll be doing:

Develop low-overhead, high-reliability implementations in C/C++, with bounded CPU/memory budgets.

Lead end-to-end feature delivery spanning user-mode components, driver/platform layers, and performance counter/trace providers.

Establish profiling models that integrate with existing ML/AI workflows (e.g., PyTorch/XLA) to turn low-level signals into actionable insights.
Requirements:
What we need to see:

BS or MS degree or equivalent experience in Computer Engineering, Computer Science, or related degree.

5+ years of system-level C/C++ development, including concurrency, memory management, and performance engineering.

Familiarity with system software design, operating systems fundamentals, computer architectures, performance analysis, and delivering production-quality software.

Strong interpersonal, verbal, and written communication; able to influence across organizations and build trust with external collaborators.

Ways to stand out from the crowd:

Extensive experience with profiling/tracing stacks for CPU/GPU (e.g., CUPTI, Nsight, performance counters, event correlation) and debugging highly concurrent systems.

Deep hands-on knowledge of CUDA and GPU architecture, including runtime/driver APIs, CUDA streams/graphs, and kernel behavior.

Track record building continuous, always-on, or multi-client profiling systems designed for predictable overhead at scale.

Hands-on experience tuning ML training/inference loops based on deep profiling analysis, with familiarity in ML ecosystems (e.g., PyTorch, JAX) and correlating application events with GPU metrics to translate data into actionable performance insights (e.g., bottleneck triage, compute vs. memory bound).

Experience with user-mode driver development and integration within platform security and permissions models.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769934
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Senior Software Engineer, Cloud Storage, AI/ML
About the job
Our software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to our needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.
The Storage AI/ML Solutions team is dedicated to exploring, architecting and developing next-generation storage solutions for the AI/ML ecosystem. By delivering solutions and sharing performance insights through benchmarks, we enable customers to achieve optimal results and solidify Cloud Platform position as the premier platform for AI/ML.
As a Senior Software Engineer, you will design and implement advanced storage solutions tailored for AI/ML workloads on our Cloud Platform. You will address system design issues, develop production-ready code, and contribute to changing incubation projects. This role involves developing benchmarks, resolving performance issues, and collaborating with cross-functional teams to bridge storage infrastructure with evolving AI/ML demands, from training to inference.
Responsibilities
Write and test product or system development code.
Participate in, or lead design reviews with peers and stakeholders to decide amongst available technologies.
Review code developed by other developers and provide feedback to ensure best practices (e.g., style guidelines, checking code in, accuracy, testability, and efficiency).
Contribute to existing documentation or educational content and adapt content based on product/program updates and user feedback.
Triage product or system issues and debug/track/resolve by analyzing the sources of issues and the impact on hardware, network, or service operations and quality.
Requirements:
Minimum qualifications:
Bachelors degree or equivalent practical experience.
5 years of experience with software development in one or more programming languages.
3 years of experience testing, maintaining, or launching software products, and 1 year of experience with software design and architecture.
Preferred qualifications:
Master's degree or PhD in Computer Science, or a related technical field.
5 years of experience with data structures and algorithms.
1 year of experience in a technical leadership role.
Experience with file systems and Linux kernel internals.
Experience with AI/ML workloads.
Familiarity with cloud computing platforms (e.g., Google Cloud Platform, Cloud Computing Platform).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786786
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Senior Software Engineer, CPU, Cloud
About the job
In this role, you will work at the center of our most strategic infrastructure investments. You will join a high-impact team operating on the critical path of our hardware and software evolution, bridging the gap between compute infrastructure and the massive scale of our software stack. You will be responsible for defining the methods and technologies required to model, simulate, and analyze complex hardware systems, with a core focus on security, performance, and correctness analytics, ensuring the efficiency of our future compute platforms. As a software engineer in this group, you will act as an architect, designer, and coder taking ownership of your goal and working in close partnership with cross-functional teams to make it a reality.
You will be at the intersection of hardware design and software optimization for our next-generation high-performance compute cores. You will not just support existing infrastructure; but will be helping build the foundation of our future hardware. Because this work is on the "most critical path" for our future, you will have visibility into and influence over the core technology that powers our most intensive services. You will bridge the gap between silicon-level security research and massive-scale software implementation, working with experimental AI tools that are shaping how we verify and optimize future hardware. You will be building the foundation of our future hardware.
Responsibilities
Drive performance modeling and analysis for next-generation CPUs, while identifying architectural bottlenecks, ensure system correctness, and implement optimization strategies across hardware-software stack to maximize efficiency for compute workloads.
Design, develop, and maintain software modeling tools and simulation environments, leverage C++, Python, and AI-driven methodologies to perform studies and automate analytical workflows.
Conduct security analysis and threat modeling for hardware-software interfaces, while developing and implementing methodologies for detecting, mitigating, and analyzing side-channel and micro-architectural vulnerabilities, ensuring systems meet safety and privacy.
Explore, deploy, and scale experimental AI-driven tooling to automate verification workflows, enhance predictive modeling, and accelerate development cycles, actively contribute to expanding the team's suite of AI-assisted capabilities.
Provide technical leadership in system integration projects, guide cross-functional teams to resolve large-scale compute bottlenecks, influence architectural designs, and facilitate the successful deployment of hardware-software solutions.
Requirements:
Minimum qualifications:
Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field, or equivalent practical experience.
5 years of experience with software development in C++ programming language or 4 years of experience with an advanced degree.
Preferred qualifications:
Masters degree or PhD in Engineering, Computer Science, or a related technical field.
Experience in modern, high-performance CPU/ML architecture and micro-architecture.
Expertise in Security analytics.
Experience with Python programming language.
Demonstrated leadership in complex technical projects, including system integration and performance optimization.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785652
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior or Principal Software Engineer in Cortex Cloud, you will contribute to the development and scaling of cloud-native security solutions for enterprise organizations. This role involves working within an established team to evolve a high-traffic product, with a focus on refining architecture, optimizing the technology stack, and maintaining engineering standards.
Your responsibilities include writing reliable code, influencing product direction, and designing distributed systems. You will be expected to make technical decisions that impact the long-term stability and performance of cloud workload protection services.
AI Integration & Engineering Workflow
A core component of our development process is the use of AI. Rather than basic code completion, we integrate AI assistants as functional components of our workflow. Our team utilizes a multi-agent AI system (IDEX/ProDex) that assists across the development lifecycle: from planning and architecture to code analysis and security reviews.
In this role, you will:
Work with AI Tools:Utilize platforms such asGemini, Claude, and Cursorfor tasks beyond code generation, including root-cause analysis, system design reviews, and architectural assessment.
Develop AI-Augmented Workflows:Help refine how AI is integrated into the SDLC, including the orchestration of agents and the development of internal tools that extend AI capabilities across our codebase.
Maintain Quality Standards:While AI assists in increasing velocity, you are responsible for the technical output. This includes critical review of all generated code and ensuring that AI-assisted work aligns with our architectural requirements and security benchmarks.
Interact with Specialized Agents:Coordinate with AI agents (Product, Architecture, Security) that operate on shared context to assist in managing complex engineering tasks.
We are looking for engineers who are interested in leveraging AI as a technical tool to manage complexity and who want to contribute to the practical application of human-AI collaboration in a cloud environment.
Requirements:
Your Experience
Backend Engineering: 5+ years of experience building and maintaining production-grade distributed systems.
Languages: Proficiency in Go (Golang) is a strong advantage. We are open to engineers with deep expertise in other backend languages (Java, Python, Rust, C#, or Node.js) who are willing to transition to a Go-primary stack and have a focus on clean, well-tested code.
Fundamentals: Strong grasp of system design, data structures, and algorithms in high-scale cloud environments.
Standards: Experience with CI/CD, comprehensive testing (unit, integration, E2E), and rigorous code reviews.
Cloud: Proficiency in AWS, GCP, or Azure, including cloud-native services.
Reliability: Experience with observability (monitoring, logging, tracing) and system profiling.
Education: B.Sc. or M.Sc. in Computer Science, Software Engineering, or equivalent technical/military experience.
Advantages
Advanced Go: Deep experience with concurrency and memory management patterns.
Distributed SaaS: Background in managing multi-tenant, cloud-based SaaS at scale.
Cybersecurity: Familiarity with threat detection or cloud security infrastructure.
AI Systems: Interest in agentic workflows or prompt engineering in production.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8781613
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior or Principal Software Engineer in Cortex Cloud, you will contribute to the development and scaling of cloud-native security solutions for enterprise organizations. This role involves working within an established team to evolve a high-traffic product, with a focus on refining architecture, optimizing the technology stack, and maintaining engineering standards.
Your responsibilities include writing reliable code, influencing product direction, and designing distributed systems. You will be expected to make technical decisions that impact the long-term stability and performance of cloud workload protection services.
AI Integration & Engineering Workflow
A core component of our development process is the use of AI. Rather than basic code completion, we integrate AI assistants as functional components of our workflow. Our team utilizes a multi-agent AI system (IDEX/ProDex) that assists across the development lifecycle: from planning and architecture to code analysis and security reviews.
In this role, you will:
Work with AI Tools:Utilize platforms such asGemini, Claude, and Cursorfor tasks beyond code generation, including root-cause analysis, system design reviews, and architectural assessment.
Develop AI-Augmented Workflows:Help refine how AI is integrated into the SDLC, including the orchestration of agents and the development of internal tools that extend AI capabilities across our codebase.
Maintain Quality Standards:While AI assists in increasing velocity, you are responsible for the technical output. This includes critical review of all generated code and ensuring that AI-assisted work aligns with our architectural requirements and security benchmarks.
Interact with Specialized Agents:Coordinate with AI agents (Product, Architecture, Security) that operate on shared context to assist in managing complex engineering tasks.
We are looking for engineers who are interested in leveraging AI as a technical tool to manage complexity and who want to contribute to the practical application of human-AI collaboration in a cloud environment.
Requirements:
Your Experience
Backend Engineering: 5+ years of experience building and maintaining production-grade distributed systems.
Languages: Proficiency in Go (Golang) is a strong advantage. We are open to engineers with deep expertise in other backend languages (Java, Python, Rust, C#, or Node.js) who are willing to transition to a Go-primary stack and have a focus on clean, well-tested code.
Fundamentals: Strong grasp of system design, data structures, and algorithms in high-scale cloud environments.
Standards: Experience with CI/CD, comprehensive testing (unit, integration, E2E), and rigorous code reviews.
Cloud: Proficiency in AWS, GCP, or Azure, including cloud-native services.
Reliability: Experience with observability (monitoring, logging, tracing) and system profiling.
Education: B.Sc. or M.Sc. in Computer Science, Software Engineering, or equivalent technical/military experience.
Advantages
Advanced Go: Deep experience with concurrency and memory management patterns.
Distributed SaaS: Background in managing multi-tenant, cloud-based SaaS at scale.
Cybersecurity: Familiarity with threat detection or cloud security infrastructure.
AI Systems: Interest in agentic workflows or prompt engineering in production.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8779570
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior or Principal Software Engineer in Cortex Cloud, you will contribute to the development and scaling of cloud-native security solutions for enterprise organizations. This role involves working within an established team to evolve a high-traffic product, with a focus on refining architecture, optimizing the technology stack, and maintaining engineering standards.
Your responsibilities include writing reliable code, influencing product direction, and designing distributed systems. You will be expected to make technical decisions that impact the long-term stability and performance of cloud workload protection services.
AI Integration & Engineering Workflow
A core component of our development process is the use of AI. Rather than basic code completion, we integrate AI assistants as functional components of our workflow. Our team utilizes a multi-agent AI system (IDEX/ProDex) that assists across the development lifecycle: from planning and architecture to code analysis and security reviews.
In this role, you will:
Work with AI Tools:Utilize platforms such asGemini, Claude, and Cursorfor tasks beyond code generation, including root-cause analysis, system design reviews, and architectural assessment.
Develop AI-Augmented Workflows:Help refine how AI is integrated into the SDLC, including the orchestration of agents and the development of internal tools that extend AI capabilities across our codebase.
Maintain Quality Standards:While AI assists in increasing velocity, you are responsible for the technical output. This includes critical review of all generated code and ensuring that AI-assisted work aligns with our architectural requirements and security benchmarks.
Interact with Specialized Agents:Coordinate with AI agents (Product, Architecture, Security) that operate on shared context to assist in managing complex engineering tasks.
We are looking for engineers who are interested in leveraging AI as a technical tool to manage complexity and who want to contribute to the practical application of human-AI collaboration in a cloud environment.
Requirements:
Your Experience
Backend Engineering: 5+ years of experience building and maintaining production-grade distributed systems.
Languages: Proficiency in Go (Golang) is a strong advantage. We are open to engineers with deep expertise in other backend languages (Java, Python, Rust, C#, or Node.js) who are willing to transition to a Go-primary stack and have a focus on clean, well-tested code.
Fundamentals: Strong grasp of system design, data structures, and algorithms in high-scale cloud environments.
Standards: Experience with CI/CD, comprehensive testing (unit, integration, E2E), and rigorous code reviews.
Cloud: Proficiency in AWS, GCP, or Azure, including cloud-native services.
Reliability: Experience with observability (monitoring, logging, tracing) and system profiling.
Education: B.Sc. or M.Sc. in Computer Science, Software Engineering, or equivalent technical/military experience.
Advantages
Advanced Go: Deep experience with concurrency and memory management patterns.
Distributed SaaS: Background in managing multi-tenant, cloud-based SaaS at scale.
Cybersecurity: Familiarity with threat detection or cloud security infrastructure.
AI Systems: Interest in agentic workflows or prompt engineering in production.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773383
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior or Principal Software Engineer in Cortex Cloud, you will contribute to the development and scaling of cloud-native security solutions for enterprise organizations. This role involves working within an established team to evolve a high-traffic product, with a focus on refining architecture, optimizing the technology stack, and maintaining engineering standards.
Your responsibilities include writing reliable code, influencing product direction, and designing distributed systems. You will be expected to make technical decisions that impact the long-term stability and performance of cloud workload protection services.
AI Integration & Engineering Workflow
A core component of our development process is the use of AI. Rather than basic code completion, we integrate AI assistants as functional components of our workflow. Our team utilizes a multi-agent AI system (IDEX/ProDex) that assists across the development lifecycle: from planning and architecture to code analysis and security reviews.
In this role, you will:
Work with AI Tools:Utilize platforms such asGemini, Claude, and Cursorfor tasks beyond code generation, including root-cause analysis, system design reviews, and architectural assessment.
Develop AI-Augmented Workflows:Help refine how AI is integrated into the SDLC, including the orchestration of agents and the development of internal tools that extend AI capabilities across our codebase.
Maintain Quality Standards:While AI assists in increasing velocity, you are responsible for the technical output. This includes critical review of all generated code and ensuring that AI-assisted work aligns with our architectural requirements and security benchmarks.
Interact with Specialized Agents:Coordinate with AI agents (Product, Architecture, Security) that operate on shared context to assist in managing complex engineering tasks.
We are looking for engineers who are interested in leveraging AI as a technical tool to manage complexity and who want to contribute to the practical application of human-AI collaboration in a cloud environment.
Requirements:
Your Experience
Backend Engineering: 5+ years of experience building and maintaining production-grade distributed systems.
Languages: Proficiency in Go (Golang) is a strong advantage. We are open to engineers with deep expertise in other backend languages (Java, Python, Rust, C#, or Node.js) who are willing to transition to a Go-primary stack and have a focus on clean, well-tested code.
Fundamentals: Strong grasp of system design, data structures, and algorithms in high-scale cloud environments.
Standards: Experience with CI/CD, comprehensive testing (unit, integration, E2E), and rigorous code reviews.
Cloud: Proficiency in AWS, GCP, or Azure, including cloud-native services.
Reliability: Experience with observability (monitoring, logging, tracing) and system profiling.
Education: B.Sc. or M.Sc. in Computer Science, Software Engineering, or equivalent technical/military experience.
Advantages
Advanced Go: Deep experience with concurrency and memory management patterns.
Distributed SaaS: Background in managing multi-tenant, cloud-based SaaS at scale.
Cybersecurity: Familiarity with threat detection or cloud security infrastructure.
AI Systems: Interest in agentic workflows or prompt engineering in production.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769910
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We seek a versatile Senior Software Engineer who is passionate about performance optimization and generative AI. Our team brings the latest research in LLM inference - from novel decoding strategies to quantization schemes - into production across our hardware lineup, from large data center servers to powerful edge devices. We work on the most advanced architectures in the field, with a focus on NVIDIA's own.

What you'll be doing:

Implement and optimize inference algorithms for LLM and omnimodal architectures, including hybrid Mamba-Transformer and mixture-of-experts models.

Profile inference pipelines using NVIDIA's profiling and simulation tools. Correlate simulation predictions against real hardware across data center and edge devices.

Write and tune GPU kernels (CUDA, Triton) for operators like fused MoE layers, SSM state updates, and quantized GEMMs.

Solve distributed inference problems: expert parallelism, communication-compute overlap, collective tuning, multi-node deployment.

Build production-grade software inside major open-source libraries - vLLM, SGLang, Dynamo, FlashInfer.

Own optimization features end-to-end, from scoping through delivery, collaborating with research, product, and engineering teams worldwide.
Requirements:
What we need to see:

B.Sc., M.Sc., or equivalent experience in Computer Science or Computer Engineering.

5+ years of hands-on software engineering experience in performance-critical systems.

Solid understanding of deep learning architectures (Transformers, SSMs, MoE, ).

Experience with systems where hardware constraints matter: GPU programming, memory hierarchy, networking, or distributed computing.

Strong software engineering fundamentals: clean design, extensibility, testability. Good judgment about when complexity is warranted.

Effective communicator who works well across teams and time zones.

Experience optimizing deep learning workloads on our GPUs using roofline models, Nsight/PyTorch profilers and end-to-end traces.


Ways to stand out from the crowd:

Contributions to open-source inference runtimes and libraries - vLLM, SGLang, FlashInfer, Dynamo or similar.

Hands-on work with LLM quantization (FP8, NVFP4, MXFP8, mixed-precision) and practical understanding of numerical precision tradeoffs.

Track record with distributed inference at scale: tensor parallelism, pipeline parallelism, expert parallelism, disaggregation, multi-node orchestration.

Deep knowledge of the latest LLM architectural trends: multi-token predictors, sparse hybrid models, attention and state-space mechanisms.

Experience with performance modeling and simulation-to-silicon correlation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769559
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Join our company, backed by Team8 and NewEra, where youll be part of this growing team solving the next AI security challenges at scale. Work in a fast-paced, dynamic environment supporting a product that blends AI, security, and high-performance distributed systems. Collaborate with a tight-knit, passionate team where your ideas matter, your impact is huge, and every day brings something new.
Requirements:
Experience: 5+ years of hands-on software development experience, preferably in Go (Golang). Microservices & Cloud: Strong background in designing, developing, and maintaining microservices architectures and cloud-native applications. Scalability & Performance: Experience working on high-scale, distributed systems with a focus on performance optimization and resilience. Containers & Infrastructure: Hands-on experience with Docker, Kubernetes, and containerized deployments. System Design: Strong understanding of object-oriented programming (OOP) principles and essential design patterns. Databases & Persistence: Proficiency with various data storage technologies, including SQL, NoSQL, and key-value (K/V) stores. Fast Learner & Adaptable: Ability to learn and absorb new technologies quickly in a fast-paced, dynamic environment. Ownership & Dedication: Self-motivated, detail-oriented, and committed to building high-quality software Fun & Positive Attitude: A great team player - someone who makes work enjoyable and contributes to a positive culture.
Advantage:
AI & LLM Expertise: Experience with AI frameworks, LLMs, and their real-world applications, with an understanding of how people use, integrate, and interact with AI models. Familiarity with LLM optimization, deployment, and potential risks is a plus. Debugging & Optimization: Proven ability to debug, profile, and optimize complex distributed systems. Security & Compliance: Strong understanding of security best practices and compliance requirements for cloud-based systems. Event-Driven Architecture: Experience with event-driven and message-based architectures (e.g., Kafka, RabbitMQ, NATS, or similar). Application Communication: Hands-on experience with gRPC, REST APIs, WebSockets, or message queues for efficient service-to-service communication.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8764684
סגור
שירות זה פתוח ללקוחות VIP בלבד