דרושים » תוכנה » SRE Intern ( Magshimim )

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
משרה זו סומנה ע"י המעסיק כלא אקטואלית יותר
שם חברה חסוי
מיקום המשרה: תל אביב יפו
סוג משרה: משרה מלאה והתמחות
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
16/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Internship
As a Software Engineer Intern in the MSC IL AI Innovation team, you will contribute to the design, development, testing, and delivery of AI-powered features and services for our Specialized Clouds organization. You will work with mentors and engineering team members to build reliable, secure, and responsible software, while learning modern AI engineering practices across prompts, retrieval, agents, evaluation, observability, and cloud-native development.

This internship is designed for entrepreneurial, creative, and highly motivated students who are passionate about building AI-native products, taking ownership in ambiguous 0 to 1 environments, and using AI throughout the engineering lifecycle to solve complex, high-impact problems in sovereign and regulated cloud environments.


Responsibilities
Develop, test, debug, and maintain software components for AI-powered products, services, or internal tools.

Contribute production-quality code that is maintainable, secure, performant, and aligned with team standards.

Partner with mentors, engineers, product managers, and stakeholders to understand user requirements and translate them into engineering tasks.

Contribute to AI features such as agents, tool/API integrations, prompt flows, retrieval workflows, evaluation harnesses, and monitoring capabilities.

Help define and run tests, evaluations, and quality checks for AI behavior, including accuracy, relevance, groundedness, and reliability.

Investigate issues using logs, traces, test results, and debugging tools, and incorporate feedback into improved solutions.

Learn and apply engineering best practices across code quality, security, responsible AI, CI/CD, documentation, and observability.

Document work clearly and share implementation details with team members to support continuity and knowledge sharing.

Collaborate effectively with team members in a dynamic, ambiguous environment; adapt quickly across technologies, stacks, and engineering tasks; proactively seek and incorporate feedback; and contribute to shared goals while building technical depth.
Requirements:
Required qualifications

Currently pursuing an M.Sc. or M.A. degree in Computer Science, Software Engineering, Artificial Intelligence, Machine Learning, or a related technical field, with at least 3 semesters remaining

Programming experience in one or more general-purpose languages such as Python, C#, Java, JavaScript/TypeScript, C++, or similar.

Understanding of core computer science fundamentals, including data structures, algorithms, software design, and debugging.

Ability to communicate clearly, collaborate effectively, adapt across technologies and engineering tasks, and learn quickly in a team environment.


Preferred qualifications

Currently pursuing a Doctorate in Computer Science, Software Engineering, Artificial Intelligence, Machine Learning, or a related technical field, with at least one semester/term remaining after the internship.

1+ year of programming experience.

6+ months of experience delivering projects in teams, through coursework, internships, open-source work, research, or personal projects.

1+ year of experience developing and applying data structures and algorithms.

Familiarity with AI, machine learning, generative AI, LLMs, agents, retrieval-augmented generation, prompt engineering, or evaluation methods.

Basic familiarity with cloud platforms, REST APIs, SDKs, CI/CD, telemetry, or observability.

Interest in responsible AI, security, reliability, and building software for regulated or sovereign cloud environments.

Demonstrated entrepreneurial mindset, creativity, initiative, ownership, and persistence in identifying meaningful problems, exploring unconventional solutions, and turning ambiguous ideas into tangible outcomes in 0 to 1 product spaces.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8782901
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
21/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking an experienced and highly motivated Senior DevOps Engineer to join our engineering and DevOps team. As a Senior DevOps Engineer, you will be the architect of our infrastructure, ensuring that platform is scalable, resilient, and secure. This is a hands-on role where you will bridge the gap between development and operations, automating our deployment pipelines and managing our cloud-native ecosystem. This is an incredible opportunity to shape the foundational infrastructure of a high-growth data company.

What You'll Do

Design, implement, and manage our cloud infrastructure using tools like Terraform, ensuring environment consistency and scalability.

CI/CD Automation: Take full ownership of our deployment pipelines, optimizing for speed, reliability, and developer productivity.

Cloud Orchestration: Manage and scale our Kubernetes (K8s) clusters on AWS, ensuring high availability and efficient resource utilization.

Observability & Monitoring: Implement and maintain robust monitoring, logging, and alerting systems to ensure platform health and rapid incident response.

Security & Compliance: Drive security best practices across the infrastructure, including IAM management, network security, and vulnerability scanning.

Collaborate: Work closely with software engineers to optimize application performance, containerization strategies, and database reliability.
Requirements:
We are seeking a hands-on builder who can grow into a technical leader for our infrastructure. Someone excited to own hard problems and pick up the specifics of our stack quickly.

Must-Haves

5+ years of experience in DevOps or Site Reliability Engineering (SRE), ideally within a high-growth SaaS or data-heavy environment.

Hands-on experience running production workloads on AWS (e.g., EKS, RDS, S3, IAM, VPC).

Strong experience with Kubernetes and Docker in production environments.

Proficiency with CI/CD tooling (e.g., GitHub Actions, GitLab CI, or Jenkins) and infrastructure-as-code (e.g., Terraform).

Working proficiency in Python, Bash, or Go for automation and internal tooling.

A team player with strong communication skills who can explain infrastructure concepts to cross-functional partners.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8791545
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
27/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
we are looking for a Senior Full Stack Engineer with a backend-oriented mindset to join our growing R&D team.
In this role, youll be responsible for connecting high-scale data infrastructure with our customer-facing portal-turning complex, runtime telemetry into fast, intuitive, and intelligent user experiences. Youll work closely with product, design, and Data pipeline and security engineers to build data pipelines, APIs, and integrations that make platform both powerful and delightful to use.
This is a hands-on role at the intersection of backend systems, high scale data engineering, and user experience, shaping how massive volumes of runtime data flow through data-driven platform.
Youll Be Great For This Role If You Love To:
Design and implement efficient data pipelines and backend integrations that connect our high-scale data to the user-facing portal.
Build and optimize APIs, data services, and caching layers to ensure fast and reliable data access.
Use data modeling, denormalization, and performance optimization to handle complex queries and large datasets effectively.
Collaborate closely with frontend engineers, product managers, and designers to ensure data flows seamlessly into intuitive UI components.
Contribute to the architecture and scalability of hybrid cloud-native platform.
Write clean, testable, production-ready code in Python, Node.js, TypeScript, and React, while keeping performance and maintainability top of mind.
Stay close to users and product feedback loops - helping translate technical insights into better experiences.
Own end-to-end delivery, from backend design through deployment using CI/CD, Temporal, and Kubernetes.
Requirements:
7+ years of fullstack experience, with a strong backend focus.
Deep expertise in Python, Node.js, TypeScript, and PostgreSQL (or similar relational DBs).
Proven experience in high-scale data systems, including data ingestion, aggregation, and performance tuning.
Advantage - Experience in ClickHouse DB / OpenSearch
Strong understanding of data modeling, denormalization, and API performance optimization.
Familiarity with React and modern frontend frameworks for connecting and visualizing backend data.
Experience with CI/CD pipelines, Kubernetes, Docker, and cloud-native architectures.
Strong collaboration, communication, and problem-solving skills, you love working across disciplines to deliver impact.
Experience in startups or high-growth environments, where speed, ownership, and quality go hand-in-hand.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8800214
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
23/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are a well-funded, early-stage startup looking for a talented and motivated Backend Engineer specializing in infrastructure to join our founding team. The focus of this role is to build and scale the infrastructure that powers autonomous AI agents automating complex enterprise workflows. You will own the systems, pipelines, and platforms that let our AI agents run reliably, securely, and at scale in production.

Your Impact
Infrastructure & Platform

Design, build, and own the core infrastructure powering our AI agent platform, from data pipelines to production deployment systems.

Build and scale the backend systems that support high-throughput document processing and data extraction workloads.

Cloud Infrastructure and Scalability

Architect and deploy infrastructure on cloud platforms (AWS, GCP, or Azure) with a focus on scalability, reliability, and cost efficiency.

Own containerization and orchestration (Docker, Kubernetes) for all production workloads.

Build and maintain CI/CD pipelines and DevOps practices that let the team ship fast without breaking things.

Data Infrastructure

Design and manage data pipelines to process and analyze large volumes of documents and unstructured data at scale.

Build the infrastructure layer connecting AI agents to databases, vector stores, and enterprise systems (ERP, CRM).

API & Systems Integration

Build and maintain robust, well-documented APIs connecting AI agents with external systems and enterprise software.

Design for reliability: retries, observability, and graceful degradation across distributed systems.

Security and Compliance

Implement authentication and authorization mechanisms (OAuth2, JWT) to secure AI-driven systems.

Ensure compliance with data privacy standards (e.g. GDPR, HIPAA) and drive best practices for secure data handling across the infrastructure.

Monitoring and Optimization

Build observability and monitoring systems to track infrastructure health, performance, and cost.

Continuously optimize system performance for speed, reliability, and cost-efficiency at scale.

Collaboration

Work closely with AI/ML engineers, product, and the founding team to make sure infrastructure decisions support fast iteration and production-grade reliability.

Participate in code reviews, design discussions, and architecture planning to drive infrastructure strategy.
Requirements:
5+ years of experience in backend or infrastructure engineering, ideally supporting production AI/ML systems or high-throughput data pipelines.

Proven track record of building and scaling infrastructure in production environments.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8793026
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a detail-oriented and technically skilled Senior QA Engineer to join our engineering team and help ensure the quality, reliability, and security of our product.
In this role, you will own the quality of complex cybersecurity features from design through production, working closely with developers, product managers, and automation engineers. You will play a key role in defining test strategies, identifying risks early, and ensuring high-quality releases.
As part of your daily workflow, you will leverage AI-powered tools to improve test planning, discover edge cases, investigate defects, analyze logs, and enhance overall testing efficiency.
The Responsibilities:
Own the automation domain end-to-end: strategy, architecture, implementation, and continuous improvement
Design, build, and maintain scalable test automation frameworks across UI, API, integration, and system layers
Develop hands-on automation code in Python and/or TypeScript
Drive quality standards across engineering teams, partnering closely with developers, QA, DevOps, and product stakeholders
Define and implement automation best practices for reliability, scalability, maintainability, and fast feedback
Integrate and optimize automation within CI/CD pipelines to enable rapid, confident releases
Test complex SaaS, cloud, and cyber-related product flows, including security-sensitive scenarios, permissions, integrations, and system behavior
Use AI tools to support automation development, test generation, debugging, log analysis, documentation, and coverage improvement
Proactively identify coverage gaps and improve system testability, observability, and debuggability
Participate in system design and feature planning to embed quality from the earliest stages
Build and evolve testing environments using scripting, infrastructure-as-code, and cloud-native tools
Mentor and guide engineers, elevating automation standards across the organization
Requirements:
5+ years of hands-on experience in test automation, including ownership of frameworks or large-scale automation systems
Strong programming skills in Python and/or TypeScript, with solid software engineering fundamentals
Proven experience designing, building, and scaling automation frameworks such as Pytest, Playwright, or Selenium
Deep understanding of CI/CD workflows and test integration, using tools such as Jenkins or GitHub Actions
Experience testing distributed, cloud-based SaaS systems on AWS, GCP, or Azure
Strong debugging, problem-solving, and system-level thinking
Hands-on experience with Linux environments and shell scripting
Experience working with logs, monitoring, and technical investigation of system failures
Comfortable using AI tools as part of the daily development and automation workflow
Understanding of cybersecurity concepts, security-sensitive flows, or cloud security environments
Excellent communication skills, with the ability to influence technical decisions and drive quality initiatives across teams
Nice to Have :
Experience working on cybersecurity products, security platforms, or detection/response systems
Experience with AI-driven development tools, agentic workflows, or Model Context Protocol (MCP)
Experience with containerization and infrastructure-as-code tools such as Docker, Terraform, or Ansible
Familiarity with Kubernetes and distributed system architectures
Experience working with databases such as SQL, MongoDB, Redis, or Neo4j
Experience with message queues such as RabbitMQ
Experience in performance, load, reliability, or resilience testing
Experience improving observability, logging, and alerting for test and production-like environments
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786668
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
02/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
Join our company, an innovative startup building a fully-managed LLM-inference platform, that enables data heavy enterprises to perform any AI task at any scale without limits.
We're looking for an Experienced Performance Researcher to join our founding team. Youll be responsible for building and optimizing scalable cloud infrastructure solutions tailored for AI workloads. This role offers a unique opportunity to directly shape our infrastructure strategy, improve system reliability and performance, and contribute to establishing our company as a leader in adaptive AI compute management.
Join us to tackle the magic that make AI tick under the hood and build the backbone powering the AI revolution.
What Youll Do
- Design and build high-performance distributed inference pipelines for LLMs, focused on large-batch, non-real-time scenarios.
- Optimize GPU memory usage, kernel execution, and communication across nodes (NCCL, MPI, etc.).
- Own CUDA kernels, compiler-level tricks, and multi-GPU scheduling logic.
- Lead profiling and performance tuning for throughput, and cost- down to the kernel level.
- Collaborate with infra, product, and research teams to define SLAs, resource allocation logic, and runtime behaviors.
- Help build the core infrastructure that will run LLM workloads across hybrid GPU environments (cloud/on-prem/self-hosted).
Requirements:
- Deep experience with CUDA programming, GPU architecture, and low-level performance engineering.
- Fluency with Python and C++, and a mastery of profiling tools like Nsight, nvprof, perf, etc.
- Experience building systems for large-scale distributed training or inference (PyTorch, DeepSpeed, Ray, Horovod, etc.).
- Hands-on familiarity with cluster and container orchestration tools (Kubernetes, Slurm, Docker).
- Self-motivated and able to operate independently in a fast-moving startup environment.
- Strong analytical skills and a passion for elegant performance wins.
- A collaborative team player with strong interpersonal skills, a positive and easygoing attitude, and the potential to grow into a leadership role.
- Prior experience building inference runtimes or scheduling frameworks.
- Experience with serverless GPU models, model parallelism, tensor slicing, and batching tricks.
- Contributions to open-source HPC or ML infra projects.
- Understanding of AI/ML privacy and compliance concerns in enterprise environments.
- Track record of working on distributed systems at bleeding-edge research labs or infrastructure teams.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8807342
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Software Engineer II, Cloud Network
About the job
Our software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to our needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.
Cloud accelerates every organizations ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage our cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Cloud as their trusted partner to enable growth and solve their most critical business problems.
Responsibilities
Build the tools and data pipelines that provide transparency into our cloud network.
Design, implement, and maintain high-impact observability solutions (e.g., packet sample logs) that enable on-call developers, Site Reliability Engineers (SREs), and support teams to troubleshoot complex networking issues.
Lead the team's AI transformation, contributing to network labs by building and optimizing Agentic tools and diagnostic skills that expose observability data through automated, LLM-driven workflows.
Contribute to the scalability of data pipelines and the reliability of our AI-driven diagnostic suite while participating in the team's on-call rotation to resolve customer escalations.
Requirements:
Minimum qualifications:
Bachelors degree or equivalent practical experience.
1 year of experience with software development in one or more programming languages (e.g., Python, C, C++, Java, JavaScript).
1 year of experience with data structures or algorithms.
Experience integrating generative AI tools or LLM interfaces into workflows.
Preferred qualifications:
Master's degree in Computer Science or a related technical field.
1 year of experience developing or supporting observability tooling (e.g., Metrics, Logs, Traces) using platforms like Prometheus, Grafana, Splunk, ELK, or OpenTelemetry.
Expertise in the networking implementation details of a specific major cloud provider (e.g., VPC, peering, firewalls, GKE/EKS networking).
Proven ability to diagnose and troubleshoot complex network issues in a public cloud environment.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8782527
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
08/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Senior Data Engineer to build and scale the data foundation behind platform and products. You will own complex data end-to-end - from ingestion and transformation through modeling, quality, observability, and production delivery.
This is a hands-on senior IC role for a strong builder who can solve difficult data problems independently, set a high technical bar, and collaborate closely with engineering, DS, and product. You will turn large, fragmented datasets into reliable, reusable capabilities that power every product.
Responsibilities
Build and own scalable data pipelines- Design, implement, and operate robust pipelines for high-volume structured and unstructured data, with validation, monitoring, lineage, and recovery built in.
Scale the platform for growth- A key near-term initiative is re-architecting the system to support a significantly larger customer base. You will own performance and cost-efficiency across pipelines and services, keeping reliability and operating costs under control as the platform scales.
Build across the stack- This is not a pipelines-only role. You will also write backend services and some frontend, including the internal backoffice the team runs on. We hire builders, not narrow specialists.
Own the core data tables- Own schema design and evolution, data contracts, and the modeling standards the team follows - naming, shared dimensions, normalization, documentation. Be accountable when a table is wrong, late, or drifting.
Level up the teams data work- Pair with and advise software engineers and data scientists on Spark, SQL, and modeling, and help turn notebook-grade code into production-grade pipelines.
Partner cross-functionally- Translate product, client, compliance, and business requirements into clear technical designs and dependable production systems.
Requirements:
Spark at scale- You have tuned real Spark jobs for performance and cost - skew, shuffle, partitioning, memory, spill - run pipelines over TB-scale or billions of rows in production, and can reason about the physical execution plan, not just write DataFrame code.
5+ years of professional experience building and owning production systems.
Strong Python and SQL, with maintainable, tested production code.
Strong software engineering fundamentals across the stack. You can own backend services and pick up frontend when the work needs it - not a pipelines-only specialist.
AI-first way of working- You build with AI in your day-to-day development, using it to move faster and raise the quality of what you ship.
Deep experience designing and operating ETL/ELT pipelines, data models, and distributed data-processing systems.
Comfortable advising and pairing with other engineers and data scientists on data work.
Strong AWS experience: S3, Glue, EMR, Athena, and related compute and orchestration services.
Experience with modern data lakehouse or warehouse architectures. Apache Iceberg is a strong advantage.
Experience with workflow orchestration (Airflow or similar), CI/CD, Docker, Git, and infrastructure as code such as AWS CDK and CloudFormation.
Strong understanding of data quality, schema evolution, lineage, observability, privacy, security, and access controls. Experience with regulated or sensitive data, such as healthcare / PHI, is an advantage.
High comfort in a fast-moving environment with incomplete requirements, high ownership, and a strong sense of urgency.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8814966
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
18/08/2026
Location: Tel Aviv-Yafo
Job Type: Full Time
We are on an expedition to find an On-Premise Site Reliability Engineer (SRE)- someone who is passionate about building rock-solid, high-performance infrastructure and bringing order to complex environments. In this role, you will own the end-to-end reliability, automation, and deployment of platform across customer sites, working hands-on with cutting-edge AI, bare-metal, and hybrid cloud architectures.
Youll collaborate closely with Product, R&D, and Architecture teams while serving as the ultimate technical authority for our customer deployments. From designing automated Ansible workflows and mastering Kubernetes to troubleshooting complex network topologies, you will eliminate toil, streamline cluster operations, and ensure every deployment is scalable, seamless, and mission-ready.
:Responsibilities
Lead End-to-End On-Prem & Hybrid Deployments: Own the technical delivery and reliability of platform in close collaboration with Product, R&D, and customer technical teams.
Architect, Execute & Improve K8s Deployments: Take a definitive hands-on role in deploying, configuring, operating, and continuously improving our platform using advanced, enterprise-grade Kubernetes architectures.
Helm Chart Management: Design, modify, and manage Helm charts to package, version, and streamline complex application deployments across different environments.
Drive Automation & Simplification: Design, implement, and maintain robust deployment automation using Ansible. You must have a passion for turning complex manual tasks into reliable, repeatable, single-click operations.
Manage Infrastructure as Code: Utilize Git as the single source of truth to manage configurations, manifests, and automation playbooks, enforcing modern engineering best practices.
Bridge On-Prem and Cloud: Leverage AWS resources (specifically EC2 and S3) for hybrid components, staging environments, or cloud-to-on-prem data flows.
Technical Tier-3 Escalation: Serve as the ultimate technical authority for deployment, Linux networking, and Kubernetes orchestration issues.
Continuous Improvement: Constantly refine our delivery pipelines, optimize bootstrap processes, and create rock-solid technical documentation.
דרישות:
SRE / Delivery Mindset: 3-5 years of hands-on experience in enterprise infrastructure deployment, systems engineering, or an on-prem operational reliability role.
Kubernetes & Helm Expert: Deep, production-grade experience with Kubernetes architecture, deployment, advanced troubleshooting, and CNI networking. Proven working experience creating, maintaining, and deploying applications using Helm charts.
Ansible Mastery: Proven experience writing clean, scalable Ansible roles and playbooks for configuration management, automation, and infrastructure provisioning.
Modern Workflows (Git & AWS): Solid working experience using Git for version control and collaborating on code/infrastructure. Practical experience provisioning and managing AWS resources (EC2 and S3).
Core Systems & Linux: Strong Linux background (Ubuntu) with a deep understanding of system internals, containerized runtimes, and troubleshooting distributed applications.
Solid Networking Knowledge: Hands-on experience with routing, firewalls, and switching topology (mainly Cisco)
Storage Foundations: Working knowledge of storage protocols (iSCSI, SAN, local NVMe) and enterprise storage arrays (like DELL) interacting with Kubernetes Persistent Volumes.
GPU & Accelerated Compute: Working knowledge of managing GPU-enabled Kubernetes nodes, including NVIDIA drivers/runtime and basic troubleshooting.
Air-Gapped Deployments: Experience deploying and maintaining software in air-gapped or offline environments, including registry mirroring and artifact staging.
Problem-Solver: Strong debugging and problem-solving skills in complex, distributed environments with an intense ownership and accountability mindset.
Willingness to Travel: Ready to travel to customer sites for physi המשרה מיועדת לנשים ולגברים כאחד.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8786710
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
Required Software Engineer III, Cloud Networking
About the job
Our software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to our needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.
This Team owns the core data plane infrastructure for all NAT-based products within the Andromeda network stack (e.g., private service connect (PSC) and Cloud network address translation (NAT). These products serve as critical entry and exit gateways that securely connect Cloud customers to their services and networks.
Responsibilities
Design, develop, test, and maintain high-performance software for our Cloud's core networking and secure connectivity platforms.
Manage complex scalability, resource efficiency, and performance optimization challenges to evolve our cloud networking datapath.
Collaborate with cross-functional engineering partners to define and build new network architectures and routing methods.
Manage individual project priorities, deadlines, and deliverables, ensuring high engineering velocity and robust code quality.
Participate in operational rotations, monitor production health, and troubleshoot packet latency or connectivity issues to keep our global systems healthy.
Requirements:
Minimum qualifications:
Bachelors degree or equivalent practical experience.
2 years of experience with software development or 1 year of experience with an advanced degree in an industry setting.
2 years of experience with developing large-scale infrastructure, distributed systems or networks, or experience with compute technologies, storage or hardware architecture.
Preferred qualifications:
Master's degree or PhD in Computer Science or related technical fields.
2 years of experience with data structures and algorithms.
Experience with systems programming (C++ preferred) and developing or debugging software in Unix/Linux user-space or kernel environments.
Experience with software architecture, engineering productivity, C++, C, Python, network architecture, large-scale distributed systems.
Familiarity with building, analyzing, and optimizing networking components and protocols (e.g., Cloud NAT, firewalls, load balancers, TCP/IP, or Software-Defined Networking).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8785731
סגור
שירות זה פתוח ללקוחות VIP בלבד