דרושים » AI » Applied scientist, Agentic AI, Agentic AI

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
22/07/2026
משרה זו סומנה ע"י המעסיק כלא אקטואלית יותר
מיקום המשרה: חיפה ותל אביב יפו
סוג משרה: משרה מלאה
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
30/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Principal Engineer who lives at the frontier of AI,
someone who can envision and build autonomous agents that reason about complex infrastructure, make intelligent decisions, and execute large-scale migrations with minimal human intervention. This is a rare opportunity to define the architecture of an AI-native platform from the ground up.

Key job responsibilities

Define and architect the technical vision for an Agentic AI migration platform, where autonomous agents discover, plan, and execute full data center exits.

Design and build multi-agent systems that leverage foundation models, chain-of-thought reasoning, and tool-use patterns to solve complex migration challenges.

Architect AI-powered network transformation capabilities, using generative models to analyze, replicate, and optimize enterprise network topologies.

Provide technical leadership across 4 scrum teams, instilling an AI-first engineering culture and ensuring architectural coherence across agent frameworks.

Drive innovation in agentic orchestration to continuously improve autonomy and accuracy for migration of servers, storage, networks, and applications.

Mentor senior engineers on AI-native development practices and cultivate a culture of rapid experimentation, spec-driven engineering, and high velocity releases on Brownfield projects.
Requirements:
Basic Qualifications

10+ years of professional software development experience, with 3+ years focused on AI/ML systems.

5+ years of designing and building large-scale distributed systems.

Deep hands-on experience with large language models (LLMs), foundation models, or agentic AI frameworks.

Experience designing multi-agent systems, autonomous orchestration engines, or AI-driven workflow platforms.

Proficiency in Python, with experience in ML frameworks (PyTorch, TensorFlow, or similar).


Preferred Qualifications

Experience with prompt engineering and chain-of-thought reasoning patterns.

Track record of shipping AI-powered products or services at production scale.

Experience leading technical strategy and architecture across multiple engineering teams.

Familiarity with reinforcement learning, planning algorithms, or decision-making systems.

Experience with AWS services, cloud-native architectures, and infrastructure-as-code.

Strong demonstrated thought leadership in AI/ML.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8801993
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Hod Hasharon and Haifa
Job Type: Full Time
Our team at the Huawei Computing Network Innovation Lab is looking for exceptional talent to join us and lead the development of next-generation data centers. We create cutting-edge technologies that synergize software and hardware in tandem to accelerate compute, storage, and networking at large scale. We aim to drive innovation and deliver software-defined infrastructure and algorithms for HPC, AI/ML, and Big Data applications.
We are looking for an outstanding Research Team Lead with deep hands-on expertise in large-scale distributed systems, AI framework infrastructure, and performance optimization on custom accelerators. If you are a visionary technical leader, a skilled communicator, and a team builder who thrives at the frontier of systems and AI research - you're welcome on board.

What Will You Be Doing?
Lead and grow a world-class team of researchers and engineers working on distributed AI infrastructure and systems software.
Architect and own the software infrastructure enabling distributed training and inference on Huawei's custom accelerator hardware (e.g., Ascend NPU).
Drive research and development on communication libraries, runtime systems, memory management, and graph execution & synchronization at scale.
Optimize end-to-end performance across large-scale clusters, covering both scale-up (multi-device) and scale-out (multi-node) configurations.
Design and implement high-performance communication backends and collective operations (AllReduce, AllGather, broadcast) for distributed training workloads.
Collaborate cross-functionally with hardware architects, compiler teams, and framework engineers to co-design hardware-software solutions.
Publish and present research findings at top international venues (NeurIPS, EuroSys, SC, MLSys, OSDI) and represent the team externally.
Mentor engineers and researchers, conduct performance and growth reviews, and shape team culture and technical direction.
Partner with top academic institutions and open-source communities to advance the state of the art in distributed AI systems.
Requirements:
Requirements
B.Sc. or higher in Computer Science, Computer Engineering, Electrical Engineering, or a closely related field.
8+ years of experience in systems software, distributed computing, or AI infrastructure, with 3+ years in a leadership or team lead role.
Deep expertise in large-scale communication systems: collective communication, RDMA, network topology-aware routing, and bandwidth optimization.
Hands-on experience building software infrastructure for distributed training on custom accelerators or heterogeneous hardware (GPU, NPU, TPU).
Strong knowledge of runtime systems: scheduling, execution graphs, kernel dispatch, synchronization primitives, and pipeline management.
Experience with memory management at scale: activation checkpointing, tensor offloading, rematerialization, KV cache management.
Proficiency in C/C++ and Python, with a focus on high-performance, production-quality code in Linux environments.
Proven ability to define technical vision, lead multi-person projects end-to-end, and deliver results under research and engineering timelines.
Excellent communication skills in English - confident presenting to international audiences, writing technical reports, and driving cross-team alignment.
Strong collaborative mindset and experience working in globally distributed, multicultural teams.
Ways to Stand Out From the Crowd.
M.Sc. or Ph.D. in a relevant field, with a strong publication record at systems or ML venues (EuroSys, OSDI, SC, NeurIPS, MLSys, ISCA).
Hands-on experience with communication frameworks such as NCCL, MPI, HCCL, or UCX.
Experience with compiler and graph optimization for AI workloads (XLA, TVM, Triton, or custom operator fusion).
Background in mixed-precision training, model parallelism (Tensor Parallelism, Pipeline Parallelism, Expert Parallelism), and large model co-design.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8792874
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Haifa
Job Type: Full Time
Required Research Software Engineer
About the job
Our software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to our needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward. We're reimagining what it means to search for information - any way and anywhere. To do that, we need to solve complex engineering challenges and expand our infrastructure, while maintaining a universally accessible and useful experience that people around the world rely on. In joining the Search team, you'll have an opportunity to make an impact on billions of people globally.
Responsibilities
Write product or system development code.
Collaborate with peers and stakeholders through design and code reviews to ensure best practices amongst available technologies (e.g., style guidelines, checking code in, accuracy, testability, and efficiency).
Contribute to existing documentation or educational content and adapt content based on product/program updates and user feedback.
Triage product or system issues and debug/track/resolve by analyzing the sources of issues and the impact on hardware, network, or service operations and quality.
Implement solutions in one or more specialized ML areas, utilize ML infrastructure, and contribute to model optimization and data processing.
Requirements:
Minimum qualifications:
Bachelors degree or equivalent practical experience.
2 years of experience with software development in one or more programming languages, or 1 year of experience with an advanced degree.
1 year of experience with one or more of the following: Speech/audio (e.g., technology duplicating and responding to the human voice), reinforcement learning (e.g., sequential decision making), ML infrastructure, or specialization in another ML field.
1 year of experience with ML infrastructure (e.g., model deployment, model evaluation, optimization, data processing, debugging).
Preferred qualifications:
Master's degree or PhD in Computer Science or related technical fields.
2 years of experience with data structures and algorithms.
Experience developing accessible technologies.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786813
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are building a new AI venture inside monday.com to redefine how businesses interact with their customers through next-generation conversational AI and agentic systems. Our AI doesnt just react - it plans, reasons, and acts.

Youll join a small, highly autonomous team with the agility, ownership, and pace of a startup, working alongside world-class engineers and researchers to build the next generation of conversational AI. Youll have access to large-scale real-world data, the compute and resources to pursue groundbreaking research, and the freedom to move fast while helping shape Harmonys AI vision.

What youll do

Research novel algorithms and design new deep learning architectures to advance our in-house NLP and conversational models.

Develop methods for translating customer goals into structured representations with controlled states and real-time observations.

Help design the reasoning layer that orchestrates our AI models efficiently alongside a range of LLMs, enabling each agent to adapt its policy throughout a conversation while remaining aligned with customer objectives.

Build analytical tools that trace our AI decision process across a vast corpus of conversations, providing clear visibility into their reasoning and behavior.

Collaborate closely with engineers and product teams to translate cutting-edge research into production at scale.

Proactively lead new ideas and technical initiatives.
Requirements:
MSc or PhD in Computer Science, Mathematics, Physics, Electrical Engineering, or a closely related field.

Proficiency in Python.

Strong theoretical background in machine learning and deep learning principles.

Solid understanding of NLP concepts and relevant linguistic principles (co-reference, natural language inference, grammatical relations, and similar).

Experience using statistics and information theory to evaluate AI models and conduct EDA.

4+ years of experience conducting NLP research with deep learning models and frameworks (PyTorch, TensorFlow, and similar).

Experience shipping AI systems from research to production.

Experience with conversational AI agents is a big plus.

Publications at top-tier conferences (ACL, EMNLP, NeurIPS, ICML, CVPR) are a plus.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8816536
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
Were currently seeking a Senior Developer Technology Engineer, Artificial Intelligence. Would you enjoy researching parallel algorithms to accelerate AI workloads on advanced computer architectures? Do you find it rewarding to identify and eliminate system bottlenecks to achieve the best possible performance on pioneering computer hardware? Could you be thrilled about an opportunity to partner with the developer community, working at the forefront of technology breakthroughs that contribute to the success of an industry leader like us? If so, the Developer Technology Team invites you to consider this role.

What you will be doing:

In this position, you will research and develop techniques to GPU accelerate workloads in deep learning, machine learning or other AI domains.

Work directly with other technical experts in their fields (industry and academia) to perform in-depth analysis and optimization of complex AI and HPC algorithms to ensure optimal AI solutions on modern CPU and GPU architectures.

Publish and/or present discovered optimization techniques in developer blogs or relevant conferences to engage and educate the developer community.

Influence the design of next-generation hardware architectures, software, and programming models in collaboration with research, hardware, system software, libraries, and tools for our teams.
Requirements:
What we need to see:

An advanced degree in Computer Science, Computer Engineering, or related computationally focused science degree (or equivalent experience).

You have 8+ years of relevant experience in software development or research work.

Programming fluency in C/C++ with a deep understanding of algorithms and software development.

A background that includes parallel programming, e.g., CUDA, OpenACC, OpenMP, MPI, pthreads, etc.

Hands on experience doing low-level performance optimizations.

In-depth expertise with CPU and GPU architecture fundamentals.

Effective communication and organization skills, with a logical approach to problem solving, good time management, and prioritization skills.


Ways to stand out from the crowd:

Expertise in parallelization and performance optimization of Deep Learning models arising from Natural Language Processing, Computer Vision, Recommender Systems, etc.

Excellent understanding of linear algebra.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773631
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo and Yokne`am
Job Type: Full Time
The Networking Advanced Development Software team develops new groundbreaking technologies to enable new market shares for the company and tighten customer relationships. These are emerging technologies in networking and distributed computing for the booming AI factories and data centers. They span areas such as AI neural networks, Deep Learning, High Performance Computing (HPC), Storage, Cloud, SW Defined Network, Network Function Virtualization, 5G NR and more. We develop the solutions top-down, all the way from application behavioral analysis, to architecture definition and down to the implementation, using the world-leading our devices. The development traverses any needed component - application SW, middleware SW, OS kernel subsystems, device drivers, embedded SW (Firmware) and CUDA GPU. We collaborate with partners and key customers in the analysis processes and engage with open source communities introducing our leading features.

What youll be doing:

Lead a team of 5 engineers in the advanced technologies development.

Design and implement solutions throughout all layers from high level application, OS and driver subsystem to firmware.

Work on impactful projects involving state-of-the-art high-performance computing hardware and software.

Provide insight and technical guidance and collaborate with peers from across the company - including software architecture, chip architecture, and engineering departments to improve our future technology.

Collaborate with our partners and customers.
Requirements:
What we need to see:

B.Sc. in Computer Science, Electrical Engineering, Computer Engineering, or a related field, or equivalent practical experience.

10+ overall years of industry experience in system programming or related fields and 3+ years of experience leading a team.

Understanding of multi core hardware, operating systems design, concurrency, virtual memory, caching, interrupts, device drivers, real-time.

Excellent programming skills.

Ability to learn complex concepts in a fast pace environment.

A teammate with a can-do attitude, high energy and excellent interpersonal skills.

Ways to stand out from a crowd:

Familiarity with networking protocols.

Experience with open-source projects (coursework, personal, or contributions).

Working in a fast-paced and dynamic environment.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774078
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
05/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We seek a versatile Senior Software Engineer who is passionate about performance optimization and generative AI. Our team brings the latest research in LLM inference - from novel decoding strategies to quantization schemes - into production across our hardware lineup, from large data center servers to powerful edge devices. We work on the most advanced architectures in the field, with a focus on NVIDIA's own.

What you'll be doing:

Implement and optimize inference algorithms for LLM and omnimodal architectures, including hybrid Mamba-Transformer and mixture-of-experts models.

Profile inference pipelines using NVIDIA's profiling and simulation tools. Correlate simulation predictions against real hardware across data center and edge devices.

Write and tune GPU kernels (CUDA, Triton) for operators like fused MoE layers, SSM state updates, and quantized GEMMs.

Solve distributed inference problems: expert parallelism, communication-compute overlap, collective tuning, multi-node deployment.

Build production-grade software inside major open-source libraries - vLLM, SGLang, Dynamo, FlashInfer.

Own optimization features end-to-end, from scoping through delivery, collaborating with research, product, and engineering teams worldwide.
Requirements:
What we need to see:

B.Sc., M.Sc., or equivalent experience in Computer Science or Computer Engineering.

5+ years of hands-on software engineering experience in performance-critical systems.

Solid understanding of deep learning architectures (Transformers, SSMs, MoE, ).

Experience with systems where hardware constraints matter: GPU programming, memory hierarchy, networking, or distributed computing.

Strong software engineering fundamentals: clean design, extensibility, testability. Good judgment about when complexity is warranted.

Effective communicator who works well across teams and time zones.

Experience optimizing deep learning workloads on our GPUs using roofline models, Nsight/PyTorch profilers and end-to-end traces.


Ways to stand out from the crowd:

Contributions to open-source inference runtimes and libraries - vLLM, SGLang, FlashInfer, Dynamo or similar.

Hands-on work with LLM quantization (FP8, NVFP4, MXFP8, mixed-precision) and practical understanding of numerical precision tradeoffs.

Track record with distributed inference at scale: tensor parallelism, pipeline parallelism, expert parallelism, disaggregation, multi-node orchestration.

Deep knowledge of the latest LLM architectural trends: multi-token predictors, sparse hybrid models, attention and state-space mechanisms.

Experience with performance modeling and simulation-to-silicon correlation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769559
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Haifa
Job Type: Full Time
Are you passionate about systems programming at the intersection of hardware and software? Do you want your code to directly accelerate the world's largest AI training and inference workloads? The Elastic Fabric Adapter (EFA) Drivers team is looking for Software Development Engineers to design and implement the kernel drivers and userspace libraries that power high-performance networking for AWS's AI/ML and HPC infrastructure.

EFA is the custom network interface that enables thousands of GPUs and accelerators to communicate at near-wire-speed bandwidth and low latency, making it a critical enabler of both foundation model training and real-time inference serving at unprecedented scale. When customers train the next generation of large language models or serve billions of inference requests with tight latency SLAs on our EC2 P5, P6, or Trn instances, it's our driver stack that moves the data.

Key job responsibilities
- Design, develop, and optimize Linux kernel drivers (RDMA/EFA) and userspace provider libraries (rdma-core) that ship to millions of EC2 instances.
- Work directly with custom hardware - collaborate with chip designers to bring new silicon capabilities to life in software.
- Contribute to the upstream Linux kernel RDMA subsystem and rdma-core open-source project.
- Architect solutions for next-generation networking features: GPU-direct RDMA, adaptive routing, collective offloads, and multi-path transport.
- Build monitoring and automation tools to enhance driver testing for functionality, reliability, and performance.
- Own the full lifecycle: from initial hardware bring-up and feature enablement through testing, deployment, and production troubleshooting across the global AWS fleet.
Requirements:
Basic Qualifications
- Bachelor's degree in Computer Science, Engineering, Mathematics, or a related field.
- 3+ years of non-internship professional software development experience in C/C++.
- Experience contributing to the architecture and design (architecture, design patterns, reliability and scaling) of new and current systems.

Preferred Qualifications
- 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
- Experience in debugging, profiling, and implementing software engineering best practices in large-scale systems.
- Experience writing low level drivers.
- Experience with DMA, PCIe, or hardware/software co-design.
- Contributions to open-source projects or upstream kernel work.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774252
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
23/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are a well-funded, early-stage startup looking for a talented and motivated Backend Engineer to join our founding team. The focus of this role is to design, develop, and deploy autonomous AI agents that can automate and optimize complex enterprise workflows. You will work on building intelligent systems capable of decision-making, data extraction, document processing, and other tasks typically requiring human intervention. This is an opportunity to influence the architecture and strategy of AI-driven automation solutions for large-scale enterprise environments.

Your Impact
AI Agent Development

Design and develop autonomous AI agents using best of breed large language models to automate tasks such as document processing, data extraction, and workflow management.

Implement reinforcement learning techniques to enhance decision-making capabilities of AI agents.

AI Integration

Build and integrate APIs that connect AI agents with external systems and enterprise software (ERP, CRM).

Use frameworks like TensorFlow, PyTorch, or Hugging Face to deploy and optimize AI models for real-time processing.

Data Processing and Pipelines

Design and manage data pipelines to process and analyze large volumes of documents and unstructured data efficiently.

Optimize data handling to improve speed and accuracy of AI agents.

Security and Compliance

Implement authentication and authorization mechanisms to secure AI-driven systems.

Ensure compliance with data privacy standards (e.g., GDPR, HIPAA) and adopt best practices for secure data handling.

Cloud Infrastructure and Scalability

Deploy AI agents on cloud platforms (AWS, GCP, or Azure) ensuring scalability and reliability.

Leverage containerization (Docker, Kubernetes) for efficient deployment and management.

Testing and Optimization

Develop and execute unit, integration, and performance tests for AI-driven systems.

Continuously monitor and optimize system performance for speed, accuracy, and cost-efficiency.

Collaboration

Work closely with AI researchers, front end engineers, and product teams to align AI agent capabilities with business requirements.

Participate in code reviews, design discussions, and architecture planning to drive innovation.
Requirements:
5+ years of experience in software engineering with a focus on AI-driven or autonomous systems.

Proven track record of deploying AI models or agents in production environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793042
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
09/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking a highly motivated Senior Deep Learning Researcher to join our team! This is an outstanding opportunity to conduct impactful research and develop the next generation of large language model (LLM) inference algorithms. You will work on technologies that directly enhance our software, making the latest LLMs more efficient and accessible for users worldwide.

By joining us, you will be part of a strategic effort to establish us as the definitive platform for high-performance LLM inference. You will engage with our skilled problem-solvers and top organizations, crafting AI technology advancements.

What you'll be doing:
Research, invent, and implement groundbreaking algorithms for LLM inference to advance the state of the art in both low-latency and high-throughput scenarios.

Translate research into practical software solutions that directly impact our products and customers.

Collaborate with internal research, engineering, and product teams across the globe to drive the development of advanced inference technologies.

Analyze the performance of new algorithms on our latest hardware, identifying bottlenecks and opportunities for algorithmic optimizations.

Partner with leading scientific organizations and industry pioneers to remain at the forefront of technological advancements and integrate the latest innovations into practical applications.
Requirements:
What we need to see:

MSc/PhD in Computer Science, Electrical Engineering, or a closely related field.

At least 5 years of relevant experience in deep learning research or applied research.

Publications in a top-tier AI/ML conference (e.g., NeurIPS, ICLR, ICML).

Deep understanding of LLM architectures coupled with hands-on experience in training large-scale models.

Excellent programming skills, particularly in Python and deep learning frameworks like PyTorch, and experience with software engineering standards.

A strong problem-solving mentality and a proactive attitude, driven by the ambition to deliver solutions with real-world impact.


Ways to stand out from the crowd:

Hands-on research experience in LLM inference optimization algorithms such as speculative decoding or parallelization strategies.

Proven experience with High-Performance Computing (HPC) environments, including training or running inference on large-scale GPU clusters (tens to hundreds of GPUs).

Deep familiarity and experience with popular LLM inference systems (e.g., vLLM, TensorRT-LLM).

Experience from a world-class industrial research group or a top-tier institution.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773626
סגור
שירות זה פתוח ללקוחות VIP בלבד