דרושים » תוכנה » Software Development Engineer (AWS ML), Machine Learning Israel (MLIL) - FLOW sub-team (Fleet Lifecy

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 12 שעות
Location: Tel Aviv-Yafo
Job Type: Full Time
The MLIL FLOW team is looking for a Software Development Engineer to design and build automation, tooling, and monitoring systems for our next-generation ML accelerator servers. We build production software to validate, initialize, monitor, and qualify these servers - from first silicon through fleet-scale deployment. Our work spans hardware diagnostics, manufacturing test automation, CI/CD pipelines, operational dashboards, and data-driven fleet health monitoring.

Key job responsibilities
Design and develop software infrastructure - automation frameworks, deployment systems, and test orchestration platforms that run at scale across manufacturing and production environments.
Work cross-functionally with Hardware, Manufacturing, and EC2 teams to automate coordinated software delivery and qualification workflows.
Debug and root-cause hardware/software interaction failures using systematic data analysis and automation-assisted triage.
Build and own CI/CD pipelines end-to-end: from code commit through build, test, deploy, and production validation - driving fast, reliable software delivery for hardware teams.
Create data pipelines and analytics systems (ETL, aggregation, real-time reporting) that transform raw hardware test results into actionable engineering insights.
Develop monitoring dashboards, alerting systems, and data visualization tools for fleet health, yield tracking, and performance benchmarking.
Own features end-to-end: from design through implementation, testing, deployment, and operational excellence.
Requirements:
Basic Qualifications
- 3+ years of software development engineer or related occupational experience.
- Bachelor's degree in Computer Science, Electrical Engineering, Computer Engineering or a related discipline or equivalent.
- Experience using Linux, demonstrating proficiency with associated tools or languages.
- Can work proactively and independently, meet deadlines, and deliver on projects and tasks.
- Knowledge of software engineering best practices across the development life cycle, including agile methodologies, coding standards, code reviews, source management, build processes, testing, and operations.

Preferred Qualifications
- Experience building monitoring dashboards and data visualization (Grafana, CloudWatch, QuickSight, or similar).
- Experience with data pipelines, ETL, or analytics (S3, Athena, Spark, or similar).
- Experience with systems programming languages (C, C++, Rust).
- Familiarity with computer architecture concepts (PCIe, memory hierarchy, power management).
- Advantage: experience with hardware bring-up, ASIC/FPGA validation, or manufacturing test development.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774254
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
22/07/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
The MLIL DataPlane team is looking for a Senior Software Development Engineer to own the design and implementation of our inference data plane. We build the software that makes large models run efficiently on custom hardware - spanning model execution, memory management, data movement, and serving integration.
Our work covers the full inference path: integrating serving engines with custom hardware, developing high-performance compute kernels, enabling efficient data movement, and driving models from early validation through production. We operate at frontier scale with large distributed models.
This is a ground-up effort with rapidly evolving hardware and software. We need a senior IC who can write and optimize low-level code for custom hardware, validate model architectures end-to-end, build test and profiling infrastructure, and drive performance across the stack.

Key job responsibilities
- Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference.
- Implement and validate LLM architectures (decoder-only, mixture-of-experts) end-to-end - from PyTorch model definition through distributed execution on custom hardware.
- Integrate custom accelerator backends into open-source ML serving frameworks (vLLM, PyTorch), including scheduler extensions, memory management, and model parallelism.
- Build and maintain test infrastructure for model correctness validation across CPU, GPU, simulator, and hardware targets.
- Profile and optimize inference workloads - identify bottlenecks, instrument critical paths, and drive latency and throughput improvements from simulation through hardware bringup.
- Own features end-to-end: from design through implementation, testing, and integration into the broader software stack.
- Contribute to CI/CD pipelines that gate model and kernel changes on correctness and performance regressions.
- Mentor engineers, drive design reviews, and raise the engineering bar across the team.
Requirements:
Basic Qualifications
- Bachelor's degree in computer science or equivalent
- 7+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques
- Knowledge of computer architecture, operating systems, and parallel computing
- Strong proficiency in C/C++.
- Strong Linux systems knowledge.
- Experience developing compute kernels for GPUs, DSPs, or custom accelerators.
- Proven track record of owning and delivering complex software features end-to-end.

Preferred Qualifications
- Knowledge of ML frameworks including JAX, PyTorch, vLLM, SGLang, Dynamo, TorchXLA, and TensorRT.
- Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware, or experience with CUDA kernels or ML/low-level kernels.
- Familiarity with speculative decoding, KV cache optimization, or other LLM serving optimizations.
- Experience with distributed systems - collective communication, RDMA, or high-speed interconnect programming.
- Experience with hardware simulation environments and model validation workflows.
- Demonstrated early adopter of AI-assisted development tools - uses LLMs or code-generation agents as part of daily workflow.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8749429
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 12 שעות
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Software Development Engineer to own the design and implementation of our inference data plane. We build the software that makes large models run efficiently on custom hardware - spanning model execution, memory management, data movement, and serving integration.
Our work covers the full inference path: integrating serving engines with custom hardware, developing high-performance compute kernels, enabling efficient data movement, and driving models from early validation through production. We operate at frontier scale with large distributed models.
This is a ground-up effort with rapidly evolving hardware and software. We need an individual contributor who can write and optimize low-level code for custom hardware, validate model architectures end-to-end, build test and profiling infrastructure, and drive performance across the stack.

Key job responsibilities
- Develop and optimize compute kernels for a custom ML accelerator architecture, targeting production-level performance for large language model inference.
- Implement and validate LLM architectures end-to-end - from PyTorch model definition through distributed execution on custom hardware.
- Integrate custom accelerator backends into open-source ML serving frameworks (vLLM, PyTorch), including scheduler extensions, memory management, and model parallelism.
- Build and maintain test infrastructure for model correctness validation across CPU, GPU, simulator, and hardware targets.
- Profile and optimize inference workloads - identify bottlenecks, instrument critical paths, and drive latency and throughput improvements from simulation through hardware bringup.
- Own features end-to-end: from design through implementation, testing, and integration into the broader software stack.
- Contribute to CI/CD pipelines that gate model and kernel changes on correctness and performance regressions.
Requirements:
Basic Qualifications
- Bachelor's degree or equivalent.
- 4+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
- Knowledge of computer architecture, operating systems, and parallel computing.
- Strong proficiency in C/C++.
- Strong Linux systems knowledge.
- Experience developing compute kernels for GPUs, DSPs, or custom accelerators.
- Proven track record of owning and delivering complex software features end-to-end.

Preferred Qualifications
- Knowledge of ML frameworks including JAX, PyTorch, vLLM, SGLang, Dynamo, TorchXLA, and TensorRT.
- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/inference lifecycles, and optimization techniques.
- Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware.
- Familiarity with speculative decoding, KV cache optimization, or other LLM serving optimizations.
- Experience with distributed systems - collective communication, RDMA, or high-speed interconnect programming.
- Demonstrated early adopter of AI-assisted development tools - uses LLMs or code-generation agents as part of daily workflow.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8774268
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
21/07/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced engineer to join our team that owns the network stack for EC2 distributed AI/ML systems. The team develops support for a variety of frameworks and communication libraries including NCCL, NVSHMEM, NIXL, NCCL GIN, and Perplexity kernels. Solid knowledge of Linux, networking, and performant coding is important. Experience with embedded systems is valued, and experience with high-speed networking or HPC/RDMA interconnects is highly valued.

Key job responsibilities
Be a senior engineer on a team that builds and maintains the infrastructure that monitors and reports on functionality and performance of massive testing workloads run at scale. Use internal our CI/CD tools, Linux, and public AWS products to automate the delivery of our software to customers, saving developer time. Write Python code that effortlessly spools up large clusters and runs benchmarks and applications for ML and HPC workloads. Use AWS Managed Grafana and Athena to digest the massive amount of performance data generated by these workloads and create dashboards for developers and stakeholders. Invent automatic mechanisms to alert developers to functional and performance regressions so they never reach reach customers. Manage the complexity of infrastructure that covers many instance types, software stacks, Linux operating systems, cutting-edge releases and make it easy to evolve.
Requirements:
Basic Qualifications
- 5+ years of non-internship professional software development experience.
- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience.
- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
- 3+ years as a mentor, tech lead or leading engineering teams.
- 3+years experience in SW/HW Co-Design.

Preferred Qualifications
- Bachelor's degree in computer science or equivalent.
- Experience creating automated dashboards and visualization (such as Grafana).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8748498
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 18 שעות
Job Type: Full Time
We are looking for an outstanding Senior Networking Software Architect to join the NIC/DPU Software and Firmware Architecture group. In this role, you will help define the next generation of our datacenter and AI networking platforms, with focus on DPU management, QoS, performance, telemetry, and software architecture across stacks. You will work closely with hardware designers, firmware/kernel driver teams, system engineers, validation, product management and customers. The role spans early architecture definition, pre-silicon design, bring-up, and production readiness for large-scale AI and cloud datacenter deployments.

What Youll Be Doing:

Own software and system architecture for next-generation DPU management, QoS, performance, telemetry, and observability features.

Define end-to-end control and management flows across DOCA, host drivers, embedded firmware, BMC, management controllers and external management systems.

Specify telemetry and observability requirements, including counters, logs, traces, events, health monitoring, debug data, and streaming telemetry.

Define management interfaces and APIs for configuration, provisioning, lifecycle operations, diagnostics, and field serviceability.

Write clear architecture specifications, interface definitions, flow diagrams, and design documents for software, firmware, and system teams.

Partner with R&D teams to translate high-level architecture into implementable designs and guide features through development, validation, silicon bring-up, and production.

Analyze system performance bottlenecks, interoperability issues, telemetry gaps, and customer-reported issues, then feed learnings into future architecture.

Collaborate with system and cluster architects to ensure NIC/DPU features fit end-to-end AI datacenter and cloud networking designs.
Requirements:
What We Need To See:

B.Sc. or M.Sc. in Computer Engineering, Computer Science, Electrical Engineering, or equivalent experience.

9+ years of experience in networking, system software, embedded software, firmware, or datacenter infrastructure.

Proven experience in software architecture, or technical leadership roles.

Deep understanding of networking concepts and protocols such as Ethernet, TCP/IP, RDMA/RoCE, congestion control, QoS, virtualization overlays, and traffic management.

Strong background with DPUs, SmartNICs, or other high-performance networking devices.

Experience with system management, provisioning, monitoring, telemetry, diagnostics, or lifecycle-management flows.

Familiarity with management protocols and frameworks such as Redfish, PLDM, MCTP, IPMI, gNMI, SNMP, Netconf, REST, or gRPC-based APIs.

Ability to lead cross-functional architecture discussions across software, firmware, hardware, validation, product, and customer-facing teams.

Excellent written and verbal communication skills, including the ability to create clear architecture documents and present trade-offs.


Ways To Stand Out From The Crowd:

Experience defining software architecture for DPU products, including management, telemetry, QoS, performance, security, virtualization, or offload features.

Hands-on background with Linux networking, device drivers, firmware, embedded Linux, BMC software, DOCA, DPDK, OVS or Kubernetes networking.

Experience with performance counters, profiling tools, eBPF, Prometheus, Grafana, dashboards, heat maps, or large-scale telemetry systems.

Experience in defining and developing GAI-based analysis tools to extract insights from telemetry data and streams.

Background in RAS, diagnosability, serviceability, field failure analysis, production debug, or customer escalation handling.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8773378
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
30/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking a detail-oriented and collaborative Senior ML Engineer to help support and maintain our machine learning capabilities. This role is ideal for someone who enjoys working closely with production systems, ensuring reliability, scalability, and explainability of models while enabling research teams to deliver impact faster.



Responsibilities



Collaborate with cross-functional teams to ensure ML systems remain robust, explainable, and aligned with business needs.
Monitor and report on ML model performance, reliability, and explainability metrics.
Participate in model retraining procedures, implement automation and optimization of MLOps pipelines.
Extend and scale monitoring pipelines, including support for new features in development.
Investigate, troubleshoot, and resolve issues in production ML workflows (tiered support from initial triage to root-cause analysis with model owners).
Develop and maintain repositories for feature engineering, inference monitoring pipelines, and artifact monitoring tools.
Perform exploratory data analysis (EDA) on historical datasets to identify quality issues and maintain data health.
Implement and oversee production based adjusters across customer deployments.
Evaluate and track critical ML artifacts such as explainability files, coverage metrics, and alignment of features.
Support development and maintenance of internal tools (e.g., interfaces, registries, and feature monitoring frameworks).
Build and maintain static and temporal features, including seasonality, event-based, and price-related features.
Requirements:
5+ years of hands-on experience in data science, ML operations, or applied ML support.
Proficiency in Python and standard data/ML libraries (Pandas/Polars, NumPy, Scikit-learn, SQL; experience with PyTorch or TensorFlow is a plus).
Strong data visualization and exploratory data analysis skills for monitoring and debugging pipelines.
Experience with time-series data and feature engineering.
Familiarity with explainability tools and model monitoring best practices.
Strong problem-solving skills with the ability to troubleshoot across data, code, and model workflows.
Excellent communication skills to summarize findings for both technical and non-technical audiences.
Experience with cloud-based ML platforms - preferably GCP
Familiarity with containerization (Docker), K8s, CI/CD workflows, or ML observability tools.
Familiarity with orchestration tools such as Airflow, Kedro or Dagster is a plus.
Prior exposure to demand forecasting, pricing, or revenue management.
Bachelor's or Master's in Computer Science, Machine Learning, Statistics, Engineering or a relevant field.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8762148
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
21/07/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced Senior Delivery Consultant - Modernization with deep expertise in Artificial Intelligence to join AWS Professional Services (ProServe). This role combines strategic architectural vision with hands-on technical leadership to deliver innovative AI solutions that drive customer success and business transformation across diverse industries and use cases.

Key job responsibilities
* Architecture & Design: Design and architect end-to-end AI-powered application solutions aligned with customer business objectives and technical requirements.
* Define application architecture patterns, standards, and best practices for AI/ML integration on AWS.
* Create technical roadmaps for customer AI application development and modernization initiatives
* Evaluate and recommend AWS AI/ML services and technologies including our Bedrock, SageMaker, and generative AI solutions
* Design data pipelines and ETL processes to support AI model training and inference using AWS services
* Customer Engagement & Consulting:
Lead customer engagements from discovery through implementation, serving as trusted technical advisor
* Conduct AI readiness assessments and develop adoption strategies tailored to customer maturity levels
* Facilitate architecture workshops and design sessions with customer stakeholders
* Deliver Well-Architected reviews focused on AI/ML workloads
* Build strong relationships with customer technical teams and executive leadership
* Guide customers in constructing AI processes aligned with AWS best practices
* Technical Leadership: Lead cross-functional teams in implementing AI solutions from concept to production
* Provide technical guidance on AI model integration, deployment strategies, and optimization on AWS
* Conduct architecture reviews ensuring solutions meet scalability, performance, security, and cost-efficiency requirements
* Mentor customer teams and junior ProServe consultants on AI best practices and AWS technologies
* Collaborate with data scientists, ML engineers, and software developers to translate AI models into production applications
* AI Solution Development: Design architectures for generative AI applications including RAG (Retrieval-Augmented Generation) systems, chatbots, and intelligent agents using our Bedrock
* Architect real-time and batch AI inference pipelines with appropriate monitoring and observability
* Implement MLOps practices using SageMaker for model versioning, deployment automation, and continuous improvement
* Design solutions for responsible AI including bias detection, explainability, and governance frameworks
* Optimize AI application performance, cost, and resource utilization across AWS services
Knowledge Sharing & Thought Leadership
* Develop reusable assets, reference architectures, and best practice documentation
* Contribute to AWS ProServe knowledge base and customer-facing content.
דרישות:
Basic Qualifications
- 10+ years of experience in application architecture and software development.
- 5+ years of hands-on experience with AI/ML technologies and frameworks (TensorFlow, PyTorch, scikit-learn, Hugging Face).
- Deep expertise in AWS cloud platform with focus on AI/ML services (SageMaker, Bedrock, Comprehend, Rekognition, etc.).
- Proficiency in programming languages such as Python, Java, or similar.
- Strong knowledge of generative AI technologies including LLMs, prompt engineering, fine-tuning, and RAG architectures.
- Understanding of various AI domains: NLP, computer vision, recommendation systems, predictive analytics.
- Willingness to travel to customer sites as needed.

Preferred Qualifications
- AWS Certified Machine Learning Specialty or AI Practitioner or Generative AI - Associate.
- Contributions to open-source AI projects or published research.
- Experience with responsible AI frameworks, governance practices, and compliance requirements.
- Prior experience in ProServe, consulting, or systems integration roles המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8748479
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
28/07/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Technical Lead to drive the architectural direction and engineering excellence of this group. This is a senior, deeply hands-on role for a technology leader who can own the technical roadmap, mentor a team of elite engineers, and build the infrastructure that challenges platform to its theoretical limits.
What You'll Lead:
Define and own the technical architecture of the group's distributed testing and reliability platform - designing for massive scale, real-world workload simulation, and adversarial failure injection
Lead effort involving multiple engineers, setting technical standards, running architecture reviews, driving design decisions, and mentoring engineers to grow
Build the systems that orchestrate millions of concurrent IO operations, inject chaos at the infrastructure layer (latency, packet loss, hardware failures), and expose the hardest-to-find race conditions and consistency bugs
Advance AI-driven approaches to test automation: intelligent scenario generation, LLM-augmented root-cause analysis, and autonomous validation pipelines
Drive observability and reliability engineering across the group - building telemetry pipelines that track P99 latency, jitter, and system health, turning quality into a quantitative discipline
Collaborate deeply with Core R&D, Storage Kernel, and Infrastructure teams - translating architectural knowledge into targeted reliability strategies
Establish engineering practices - design docs, production-grade code reviews, testing philosophy, and cross-team technical alignment
Requirements:
Strong software engineering background with 6+ years of hands-on Python development experience is required. The ability to read, debug, and reason about C++, Rust, or Go is a significant advantage
Deep understanding of distributed systems: concurrency, consistency models, fault tolerance, and large-scale system behavior under stress
Background in one or more of: storage systems, networking (TCP/IP, RDMA), cloud infrastructure, database internals, or high-performance backend systems
Experience building large-scale infrastructure platforms, internal developer platforms, or reliability engineering systems
Leadership:
Proven track record leading complex technical initiatives from architecture through delivery
Experience mentoring and growing engineers - raising the technical bar of a team, not just directing work
Ability to drive technical alignment across teams, communicate tradeoffs clearly, and make high-quality architectural decisions at speed
Comfortable operating at both the strategic and hands-on level - you write code, review designs, and shape roadmaps
Previous experience in people management roles - Advantage
Mindset:
You approach quality through the lens of Site Reliability Engineering: you care about MTTD, observability, and building self-healing systems
You have a "hacker" instinct - you don't just find bugs; you find the architectural flaws that allowed them to exist
You are an early adopter of AI tools and excited about applying LLMs and generative AI to accelerate engineering velocity
Big Advantages
Experience with storage systems, file systems, or high-performance distributed environments
Background in chaos engineering, fault injection, or simulation systems
Familiarity with observability tooling and performance engineering at scale
Experience building testing or reliability platforms as first-class engineering products
Prior experience as a Team Lead in a high-growth infrastructure company
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8757543
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
4 ימים
Location: Tel Aviv-Yafo
Job Type: Full Time
We seek a versatile Senior Software Engineer who is passionate about performance optimization and generative AI. Our team brings the latest research in LLM inference - from novel decoding strategies to quantization schemes - into production across our hardware lineup, from large data center servers to powerful edge devices. We work on the most advanced architectures in the field, with a focus on NVIDIA's own.

What you'll be doing:

Implement and optimize inference algorithms for LLM and omnimodal architectures, including hybrid Mamba-Transformer and mixture-of-experts models.

Profile inference pipelines using NVIDIA's profiling and simulation tools. Correlate simulation predictions against real hardware across data center and edge devices.

Write and tune GPU kernels (CUDA, Triton) for operators like fused MoE layers, SSM state updates, and quantized GEMMs.

Solve distributed inference problems: expert parallelism, communication-compute overlap, collective tuning, multi-node deployment.

Build production-grade software inside major open-source libraries - vLLM, SGLang, Dynamo, FlashInfer.

Own optimization features end-to-end, from scoping through delivery, collaborating with research, product, and engineering teams worldwide.
Requirements:
What we need to see:

B.Sc., M.Sc., or equivalent experience in Computer Science or Computer Engineering.

5+ years of hands-on software engineering experience in performance-critical systems.

Solid understanding of deep learning architectures (Transformers, SSMs, MoE, ).

Experience with systems where hardware constraints matter: GPU programming, memory hierarchy, networking, or distributed computing.

Strong software engineering fundamentals: clean design, extensibility, testability. Good judgment about when complexity is warranted.

Effective communicator who works well across teams and time zones.

Experience optimizing deep learning workloads on our GPUs using roofline models, Nsight/PyTorch profilers and end-to-end traces.


Ways to stand out from the crowd:

Contributions to open-source inference runtimes and libraries - vLLM, SGLang, FlashInfer, Dynamo or similar.

Hands-on work with LLM quantization (FP8, NVFP4, MXFP8, mixed-precision) and practical understanding of numerical precision tradeoffs.

Track record with distributed inference at scale: tensor parallelism, pipeline parallelism, expert parallelism, disaggregation, multi-node orchestration.

Deep knowledge of the latest LLM architectural trends: multi-token predictors, sparse hybrid models, attention and state-space mechanisms.

Experience with performance modeling and simulation-to-silicon correlation.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8769559
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
19/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for an experienced and passionate ML Engineering Team Lead to lead our ML Engineering team and shape the next generation of our AI infrastructure. This is a hands-on leadership role where you'll combine technical leadership, software architecture, and people management to build scalable, production-ready AI systems running on edge devices.
About The Role:
Lead, mentor, recruit, and grow a team of software engineers, fostering a culture of ownership, collaboration, and continuous improvement.
Own the team's technical roadmap, architecture, execution, and project prioritization, aligning delivery with business goals.
Design, build, and maintain scalable software and ML infrastructure across cloud and edge environments.
Partner with AI Researchers to productionize Computer Vision and Deep Learning models into reliable, high-performance systems.
Design and optimize inference pipelines with a focus on scalability, latency, and reliability.
Drive engineering excellence through architecture reviews, code reviews, development best practices, and modern AI-assisted engineering workflows.
Requirements:
6+ years of software development experience, including 3+ years leading software engineering or ML engineering teams.
Strong hands-on experience with Python and C++ or Rust.
Experience building, deploying, and maintaining production-grade Machine Learning systems.
Strong understanding of software architecture, scalable system design, and performance optimization.
Experience collaborating with AI, Machine Learning, or Computer Vision teams.
Excellent leadership, communication, and organizational skills, with a strong ownership mindset.
Experience using modern AI-assisted development tools (such as Cursor, Claude Code, or Codex) while maintaining high engineering quality.
Nice to Have:
Hands-on experience developing and optimizing AI applications on NVIDIA edge platforms, particularly NVIDIA Jetson devices, including GPU acceleration and deployment on resource-constrained systems.
Experience with modern AI and Computer Vision frameworks such as PyTorch, CUDA, TensorRT, NVIDIA DeepStream, and GStreamer.
Experience with containerized and cloud-native development using Docker, Kubernetes, and CI/CD pipelines.
Experience using agentic AI coding tools (such as Cursor, Claude Code, Codex, or similar) as part of the software development lifecycle to improve engineering productivity while maintaining code quality and best practices.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8743461
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
20/07/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
we are looking for a AI Solution Engineer.
The AI Solutions Engineer plays a critical, high-impact role in scaling our clinical AI solutions and ensuring the ability to deploy high quality products, monitor and support them in production.
This position requires strong software and data engineering skills to elevate our operational systems.
You will be responsible for independently driving complex E2E projects, demonstrating flexible thinking and making cost-effective decisions while balancing quality versus speed.
You will serve as a vital technical bridge, forming strong collaborations with both the R&D and Product organizations. Success in this dynamic and fast-driven environment demands a focused "getting things done" and "can-do" attitude to deliver value with speed and quality.
Responsibilities:
Automate and streamline AI operations - develop tools to automate AI deployment, data validation, and reporting for smooth production integration of AI solutions built by our research teams.
Act as the production technical expert - collaborate with Product and R&D teams during feature design to ensure production feasibility, scalability, and operational excellence.
Lead end-to-end tool development - gather requirements, design, implement, and deploy operational tools and pipelines in collaboration with cross-functional teams.
Drive engineering quality and play a key role in team member mentoring - uphold and develop engineering standards while fostering innovation.
Investigate and resolve complex data challenges - analyze and address production data issues, including data quality, availability, performance anomalies, and integration challenges.
Collaborate with customer-facing teams - understand real-world operational challenges, develop scalable solutions, and bridge the gap between customer needs and AI system behavior.
Ensure ongoing AI model reliability - design and implement mechanisms for monitoring data and performance degradation, and maintaining long-term model effectiveness.
Develop monitoring and analytics frameworks - create metrics and dashboards to track production health, data integrity, and AI performance across the deployment lifecycle.
Requirements:
Bachelor's degree in a relevant field (Computer Science, Data Science, Engineering). A Master's degree is an advantage.
At least 3 years of proven experience in software or data engineering in a professional setting.
Familiarity with statistical analysis, machine learning concepts, AI model evaluation and system engineering concepts.
Exceptional problem-solving skills and the ability to troubleshoot complex technical issues.
Strong collaboration and communication skills to work effectively with cross-functional teams.
Flexible and adaptive mindset - we need to deliver solutions with both quality and speed in mind, making decisions based on the dynamic tradeoffs between them along the way.
Experience in working in technically oriented operation environments is an advantage.
Tech stack:
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8745832
סגור
שירות זה פתוח ללקוחות VIP בלבד