דרושים » תוכנה » Senior Software Engineer - Spark

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
1 ימים
Location: Merkaz
Job Type: Full Time
Were seeking an outstanding Senior SW Engineer to join our team and help us develop our data frameworks integrations. This is a rare opportunity to influence the future of data processing.

Responsibilities:

Develop our Spark plugin and integrations

Develop in Scala and Java

Contribute to open-source projects

Develop in Python

Design advanced algorithms of processing and optimizing query plans

Define our query abstraction layer

Collaborate with other software teams to optimize software and hardware integration for maximum performance
Requirements:
BSc or equivalent or higher degree in Computer Science, Computer Engineering, or Electrical Engineering.

8+ years of experience in software development with a strong focus on data engineering

Strong proficiency in Java and Scala, Python - an advantage

Experience with Apache Spark / other big data frameworks

Excellent communication skills and ability to explain complex technical concepts to non-technical stakeholders
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8843116
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
We're looking for an experienced Software Engineer to join our Data Platform Group. In this key role, you will build the company data platform: cloud-based microservices and data pipelines that process on the order of 1M records/sec at low latency. Because the platform is the foundation other groups build on, your work has a direct impact on our customers and enables engineering, product, and research teams across the organization.
The Lakehouse team owns the data itself - how it is stored, organized, retained, and served. We own our analytical storage layer, the batch processing built on top of it, and the APIs through which customers and the rest of the company consume data. If you enjoy the problems that only appear at petabyte scale - physical data layout, query performance, storage cost, and retention - this is the role.

Responsibilities:
End-to-end ownership of our large-scale analytical storage layer: data modeling, schema and table design, partitioning, retention, and query performance.
Design and develop the batch processing layer over our data lake using Spark on EMR (Java and PySpark).
Build and evolve Java/Spring Boot services that expose our data through well-defined APIs to customers and to consumers across the company.
Own performance and cost: query optimization, file layout and compaction, cluster sizing, and storage efficiency at scale.
Research new technologies in the lakehouse and analytical-storage space and adapt them for use in our product.
Work closely with product, DevOps, and security teams.
Requirements:
5+ years of hands-on experience designing and developing large-scale distributed data systems in production, with a strong emphasis on performance.
Deep, hands-on expertise in at least one of the following, at a significant scale:
A columnar/analytical database - ClickHouse is a major advantage, including data modeling, query optimization, and operating it in production
Apache Spark at an expert level, including tuning and optimizing large batch jobs.
Experience with open table formats such as Iceberg, Delta Lake, or Hudi
Experience with data lake technologies: Parquet, S3, and SQL query engines such as Athena, Trino, or Presto.
Strong command of analytical data modeling and the design principles behind it: partitioning strategies, denormalization, batch vs. streaming trade-offs, and schema evolution.
Strong Java and solid understanding of object-oriented design and software engineering principles.
Experience building and running microservices on Kubernetes.
Hands-on experience with the AWS platform, particularly EMR, S3, and Glue.
Motivated, fast, independent learner and strong problem solver.
A team player with excellent collaboration and communication skills.
B.Sc. in Computer Science, Software Engineering, or a related field, or equivalent practical experience.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8820063
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
6 ימים
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are looking for a Data Engineer to join our team and play a key role in designing, building, and maintaining scalable, cloud-based data pipelines. You will work with AWS services , Airflow and databricks to integrate, process, and analyze large datasets, ensuring data reliability and efficiency.
Your work will directly impact business intelligence, analytics, and data-driven decision-making across the company.
What Youll Do:
ETL & Data Processing: Develop and maintain ETL processes, integrating data from various sources (APIs, databases, external platforms) using Python, SQL, and cloud technologies.
Utilize LangGraph and other frameworks to create state of the art AI agents that integrate various tools and data sources.
Implement integrations that allow agents to take automated actions on behalf of users, streamlining workflows and reducing overhead.
Data Modeling: Design and maintain logical and physical data models to support business needs.
Optimization & Scalability: Improve process efficiency and optimize runtime performance to handle large-scale data workloads.
Collaboration: Work closely with BI analysts and business stakeholders to define data requirements and functional specifications.
Monitoring & Troubleshooting: Ensure data integrity and reliability by proactively monitoring pipelines and resolving issues.
Requirements:
Education & Experience:
BSc in Computer Science, Engineering, or equivalent practical experience.
5+ years of experience in data engineering or related roles.
Technical Expertise:
Proficiency in Python for data engineering and automation.
Experience with Big Data technologies such as Spark, Databricks, DBT, and Airflow.
Hands-on experience with AWS services (S3, Redshift, Glue, Managed Airflow, Lambda)
Knowledge of Docker, Terraform, Kubernetes, and infrastructure automation.
Strong understanding of data warehouse (DWH) methodologies and best practices.
Soft Skills:
Strong problem-solving abilities and a proactive approach to learning new technologies.
Excellent communication and collaboration skills, with the ability to work independently and in a team.
Nice to Have ( Advantage):

Experience with LangGraph, LangChain, or similar frameworks for building AI agents.
Understanding of LLMs, prompt engineering, and AI agent architectures.
Familiarity with K8s for infrastructure as code.
Experience with JavaScript, React, and Node.js.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8838333
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Petah Tikva
Job Type: Full Time
We are looking for a strong, hands-on Data Engineer to join our team and play a key role in building our data infrastructure from the ground up. In this role, you will design and implement scalable data pipelines and platforms, supporting both batch and real-time use cases. You will work closely with analysts and stakeholders to deliver reliable, high-quality data solutions, and take full ownership of data flows - from ingestion to consumption. This is a great opportunity for an executor who enjoys building, moving fast, and making an impact.
What will your job look like?
Design, build, and maintain robust and scalable data pipelines (batch and real-time) end-to-end.
Design and implement scalable, flexible data architectures to support evolving business needs.
Build and manage data platforms, including data lakes and data warehouses.
Integrate multiple data sources (structured and unstructured) into a unified data platform using batch (ETL) and real-time streaming solutions.
Design and implement efficient data models, schemas, and database structures (SQL / NoSQL).
Develop and implement data quality processes to ensure accuracy, consistency, and reliability.
Monitor, optimize, and troubleshoot data infrastructure to meet performance and SLA requirements.
Requirements:
5+ years of hands-on experience as a Data Engineer, building data systems from scratch in dynamic environments.
Bachelors degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
Strong proficiency in Python and advanced SQL, with solid experience in data modeling.
Proven experience designing and building scalable data pipelines (batch and real-time), including streaming technologies such as Kafka.
Strong experience working with AWS, including services such as S3, Athena and DynamoDB.
Experience working with big data processing frameworks such as Spark, and columnar data formats (e.g., Parquet).
Hands-on experience with workflow orchestration tools such as Airflow.
Strong ownership and execution mindset, with excellent problem-solving skills and high attention to detail, and the ability to collaborate effectively and deliver in ambiguous, fast-paced environments.
Fluent in English.
Nice to have:

Experience with data platform technologies such as Databricks, Snowflake.
Experience building data platforms using modern lakehouse technologies (e.g., Iceberg).
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8827701
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
16/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time
We are seeking talented and passionate Senior Data Engineer to join our Data team. In this pivotal role, you will be instrumental in designing, building, and optimizing the critical data infrastructure that underpins innovative creative intelligence platform. You will tackle complex data challenges, ensuring our systems are robust, scalable, and capable of delivering high-quality data to power our advanced AI models, customer-facing analytics, and internal business intelligence. This is an opportunity to make a significant impact on our product, contribute to a data-driven culture, and help solve fascinating problems at the intersection of data, AI, and marketing technology.

Key Responsibilities
Architect & Develop Data Pipelines: Design, implement, and maintain sophisticated, end-to-end data pipelines for ingesting, processing, validating, and transforming large-scale, diverse datasets.
Manage Data Orchestration: Implement and manage robust workflow orchestration for complex, multi-step data processes, ensuring reliability and visibility.
Advanced Data Transformation & Modeling: Develop and optimize complex data transformations using advanced SQL and other data manipulation techniques. Contribute to the design and implementation of effective data models for analytical and operational use.
Ensure Data Quality & Platform Reliability: Establish and improve processes for data quality assurance, monitoring, alerting, and performance optimization across the data platform. Proactively identify and resolve data integrity and pipeline issues.
Cross-Functional Collaboration: Partner closely with AI engineers, product managers, developers, customer success and other stakeholders to understand data needs, integrate data solutions, and deliver features that provide exceptional value.
Drive Data Platform Excellence: Contribute to the evolution of our data architecture, champion best practices in data engineering (e.g., DataOps principles), and evaluate emerging technologies to enhance platform capabilities, stability, and cost-effectiveness.
Foster a Culture of Learning & Impact: Actively share knowledge, contribute to team growth, and maintain a strong focus on how data engineering efforts translate into tangible product and business outcomes.
Requirements:
What we are looking for:
7+ years of experience as a Data Engineer, building and managing complex data pipelines and data-intensive applications.
Solid understanding and application of software engineering principles and best practices. Proficiency in a relevant programming language (e.g., Python, Scala, Java) is highly desirable.
Deep expertise in writing, optimizing, and troubleshooting complex SQL queries for data transformation, aggregation, and analysis in relational and analytical database environments.
Hands-on experience with distributed data processing systems, cloud-based data platforms, data warehousing concepts, and workflow management tools.
Strong ability to diagnose complex technical issues, identify root causes, and develop effective, scalable solutions.
A genuine enthusiasm for tackling new data challenges, exploring innovative technologies, and continually expanding your skillset.
A keen interest in understanding how data powers product features and drives business value, with a focus on delivering results.
Excellent ability to communicate technical ideas clearly and work effectively within a multi-disciplinary team environment.
Advantages:
Familiarity with the marketing/advertising technology domain and associated datasets.
Experience with data related to creative assets, particularly video or image analysis.
Understanding of MLOps principles or experience supporting machine learning workflows.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8823686
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
28/09/2026
חברה חסויה
Location: Tel Aviv-Yafo
Job Type: Full Time and Hybrid work
Required Senior Data Engineer
Description
We are making the future of Mobility come to life starting today.
We support the worlds largest vehicle fleet operators and transportation providers to optimize existing operations and seamlessly launch new, dynamic business models - driving efficient operations and maximizing utilization.
At the heart of our platform lies the data infrastructure, driving advanced machine learning models and optimization algorithms. As the owner of data pipelines, you'll tackle diverse challenges spanning optimization, prediction, modeling, inference, transportation, and mapping.
As a Senior Data Engineer, you will play a key role in owning and scaling the backend data infrastructure that powers our platform-supporting real-time optimization, advanced analytics, and machine learning applications.
What You'll Do:
Design, implement, and maintain robust, scalable data pipelines for batch and real-time processing using Spark, and other modern tools.
Own the backend data infrastructure, including ingestion, transformation, validation, and orchestration of large-scale datasets.
Leverage Google Cloud Platform (GCP) services to architect and operate scalable, secure, and cost-effective data solutions across the pipeline lifecycle.
Develop and optimize ETL/ELT workflows across multiple environments to support internal applications, analytics, and machine learning workflows.
Build and maintain data marts and data models with a focus on performance, data quality, and long-term maintainability.
Collaborate with cross-functional teams including development teams, product managers, and external stakeholders to understand and translate data requirements into scalable solutions.
Help drive architectural decisions around distributed data processing, pipeline reliability, and scalability.
Requirements:
4+ years in backend data engineering or infrastructure-focused software development.
Proficient in Python, with experience building production-grade data services.
Solid understanding of SQL
Proven track record designing and operating scalable, low-latency data pipelines (batch and streaming).
Experience building and maintaining data platforms, including lakes, pipelines, and developer tooling.
Familiar with orchestration tools like Airflow, and modern CI/CD practices.
Comfortable working in cloud-native environments (AWS, GCP), including containerization (e.g., Docker, Kubernetes).
Bonus: Experience working with GCP
Bonus: Experience with data quality monitoring and alerting
Bonus: Experience with Snowflake, DBT, Flink, Kafka
Bonus: Strong hands-on experience with Spark for distributed data processing at scale.
Degree in Computer Science, Engineering, or related field.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8835914
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
5 ימים
חברה חסויה
Location: Ramat Gan
Job Type: Full Time and Hybrid work
Required Senior Data Engineer
About the role:
You will specialize in designing and building world class, scalable data architectures, ensuring reliable data flow and integration for groundbreaking biotechnological research. Your expertise in big data tools and pipelines will accelerate our ability to derive actionable insights from complex datasets, driving innovations in improving patients outcome and in delivering life savings treatment solutions.
In this role, you will work closely with data scientists, AI/ML engineers, architect and other cross-functional teams to understand their data needs and requirements. You will also be responsible for ensuring that data is easily accessible and can be used to support data-driven decision making.
Location: Ramat Gan, (Hybrid model)
What will you do?
Design, build, and maintain data pipelines that ingest, transform, and load noisy, heterogeneous biological data at scale - from single-cell sequencing, genomic, and clinical sources - across databases, APIs, and flat files
Enhance our data warehouse system to dynamically support multiple analytics use cases
Implement and productize complex scientific and computational algorithms as scalable, production-grade pipeline components
Drive our data infrastructure toward becoming AI-native, enabling AI/ML systems and agents to reliably query, reason over, and act on our data
Implement data governance policies and procedures to ensure data quality, security, and privacy
Collaborate with ML and AI scientists and other cross-functional teams to understand their data needs and requirements
Drive team capability growth by championing an AI-first SDLC, mentoring data engineers on AI-assisted development practices and tooling
Develop and maintain documentation for data pipelines, processes, and systems
Requirements
We will only consider senior data engineers that have demonstrated strong system design skills combined with an extensive background working with data orchestration, data warehousing and ETL tools.
Requirements:
Required qualifications:
Bachelor's or Master's degree in a related field (e.g. computer science, data science, engineering, computational biology)
At least 7 years of experience with programming languages, specifically Python
Must have at least 5+ years of experience as a Data Engineer, ideally with experience in multiple data ecosystems
Proficiency in SQL and experience with database technologies (e.g. MySQL, PostgreSQL, Oracle)
Familiarity with data storage technologies (e.g. HDFS, NoSQL databases)
Experience with ETL tools (e.g. Apache Beam, Apache Spark)
Experience with orchestration tools (e.g. Apache Airflow, Dagster)
Experience with data warehousing technologies (ideally BigQuery)
Experience working with large and complex data sets
Experience working in a cloud environment
Strong problem-solving and communication skills
Familiarity with biotech or healthcare data - an advantage
Desired personal traits:
You want to make an impact on humankind
You prioritize We over I
You enjoy getting things done and striving for excellence
You collaborate effectively with people of diverse backgrounds and cultures
You constantly challenge your own assumptions, pushing for continuous improvement
You have a growth mindset
You make decisions that favor the company, not yourself or your team
You are candid, authentic, and transparent.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8838827
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Ramat Gan
Job Type: Full Time
We are seeking a highly skilled and analytical Senior Data Engineer to join our Data team. In this role, you will design and implement robust data pipelines, while uniquely bridging the gap between engineering and analytics by actively analyzing data to extract actionable insights. You will play a crucial part in architecting our data foundations to support everything from business intelligence to advanced machine learning and agentic AI pipelines.
As a Senior Data Engineer, you will collaborate closely with engineering teams, product managers, and stakeholders across the organization. You will not only build the infrastructure utilizing modern data stack tools but also act as a data analyst when needed, ensuring our systems are fully equipped to operate within and support a cutting-edge agentic AI environment.
Responsibilities
Design, build, and maintain highly scalable ELT/ETL data pipelines.
Architect and manage modern cloud data warehousing solutions.
Develop, maintain, and monitor Python services responsible for robust data collection and ingestion.
Perform hands-on data analysis to interpret complex datasets, identify trends, and deliver business insights, acting in a dual capacity as a Data Analyst.
Develop and optimize data infrastructure specifically designed to support autonomous agentic workflows and LLM integrations.
Collaborate with engineers and analysts to troubleshoot data issues, enforce quality SLAs, and define data requirements.
Document data architecture, flow, and analytics standards for internal team alignment.
Build and maintain dashboards and reports to communicate analytical findings and data health to the organization.
Maintain Kafka consumer applications that process high-volume event streams in real-time, ensuring reliable ingestion into cloud databases.
Requirements:
Must-Have:
5+ years of proven experience in a Data Engineering role, with a strong background in data architecture.
Exceptional proficiency in SQL and Python for data manipulation, scripting, and pipeline automation.
Deep hands-on experience with modern data orchestration and transformation tools, specifically Airflow and dbt.
Extensive experience managing and optimizing cloud data platforms such as BigQuery / Databricks / Snowflake.
Demonstrated experience in data analysis, with the ability to act as a Data Analyst to query data, build reports, and extract actionable insights.
Practical experience designing or supporting data infrastructure for an agentic environment or AI/LLM-driven applications.
Strong attention to detail, analytical mindset, and excellent communication skills.
Experience of one or more of these technologies: Kafka, Kubernetes, ArgoCD, Terraform, Debezium.
Understanding of data modeling principles: dimensional modeling, fact/dimension tables, slowly changing dimensions
Experience with Git workflows: branching, PRs, code reviews, and CI/CD for data pipelines.
Ownership mindset: ability to debug production issues, drive projects to completion independently
Nice-to-Have:
Experience with BI tools (e.g., Looker, Tableau, Power BI) for advanced dashboarding.
Experience working with graph databases or NoSQL databases.
Experience with Python backend APIs (FastAPI/Flask) that serve aggregated analytics data to dashboards.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8810579
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Petah Tikva
Job Type: Full Time
our company seeks an experienced Senior Staff Data Engineer to join our AI and Data Solutions unit and champion an AI-first strategy. You will spearhead the development of Data Lakehouse processes, model core product domains, and define data engineering best practices across the organization. Serving as a technical project lead, you will provide strong direction and collaborate with Data Science, Analytics, R&D, and IT teams to elevate onboarding workflows. This role demands a proactive approach to optimizing system capabilities and driving cross-functional data initiatives.
Key Responsibilities
Own the design, implementation, and optimization of scalable Data Lakehouse and Data Warehouse architectures across diverse cloud ecosystems.
Spearhead the development of robust ETL/ELT pipelines and data ingestion processes utilizing SQL, Python, PySpark, and DBT.
Apply advanced data modelling techniques to structure complex data assets into optimized facts, dimensions, and partitions.
Partner with cross-functional leadership across data science, analytics, R&D, and IT to identify requirements and improve workflows.
Orchestrate complex data workflows combining tools like Airflow and DBT with advanced Gen-AI development workflows using CloudCode.
Define, document, and popularize data engineering best practices, design patterns, and data quality standards across the unit.
Evaluate and implement modern data technologies spanning vendor-specific and open-source landscapes including GCP, AWS, and Apache Iceberg.
Provide technical direction and hands-on guidance to ensure high-performance execution, reliable pipeline delivery, and proactive problem-solving.
Requirements:
5+ years of experience with advanced data modeling techniques and concepts (e.g., Facts, Dimensions, Partitions).
5+ years of experience designing and implementing end-to-end data ingestion and ETL processes within Data Lakehouse architectures.
5+ years of hands-on experience programming in SQL and Python, PySpark, Java, or Scala.
Demonstrated experience leading technical projects, driving stakeholder alignment, and navigating cross-functional dynamics.
Proven experience working within cloud environments (AWS or GCP), with a strong requirement for AWS.
Deep understanding of the modern data landscape and open-source tools (e.g., Iceberg, DBT, Airflow).
Proven capability in using and orchestrating Gen-AI workflows based on CloudCode and Cursor.
Preferred Qualifications
Deep expertise in the AWS ecosystem, including Glue, Athena, SageMaker, and EMR.
Advanced experience with Data Warehouse technologies, especially Apache Iceberg or BigQuery.
Leadership or mentorship experience within a data engineering or analytics unit.
Background or domain knowledge in cybersecurity is a strong advantage.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8834340
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
לפני 7 שעות
חברה חסויה
Location: Ra'anana
Job Type: Full Time
we are a dynamic technology firm dedicated to building exceptional products using modern technologies for the world's most admired companies. We pride ourselves on innovation, collaboration, and delivering excellence in every project.

Job Description
Responsibilities:

Design and maintain scalable data pipelines with Spark Structured Streaming.
Implement Lakehouse architecture with Apache Iceberg.
Develop ETL processes for data transformation.
Ensure data integrity, quality, and governance.
Collaborate with stakeholders and IT teams for seamless solution integration.
Optimize data processing workflows and performance.
Requirements:
5+ years in data engineering.
Expertise in Apache Spark and Spark Structured Streaming.
Hands-on experience with Apache Iceberg or similar data lakes.
Proficiency in Scala, Java, or Python.
Knowledge of big data technologies (Hadoop, Hive, Presto).
Experience with cloud platforms (AWS, Azure, GCP) and SQL.
Strong problem-solving, communication, and collaboration skills.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8845410
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location: Tel Aviv-Yafo
Job Type: Full Time
As a Senior Data Engineer, youll collaborate with top-notch engineers and data scientists to elevate our platform to the next level and deliver exceptional user experiences. Your primary focus will be on the data engineering aspects-ensuring the seamless flow of high-quality, relevant data to train and optimize content models, including GenAI foundation models, supervised fine-tuning, and more.

Youll work closely with teams across the company to ensure the availability of high-quality data from ML platforms, powering decisions across all departments. With access to petabytes of data through MySQL, Snowflake, Cassandra, S3, and other platforms, your challenge will be to ensure that this data is applied even more effectively to support business decisions, train and monitor ML models and improve our products.



Key Job Responsibilities and Duties:

Rapidly developing next-generation scalable, flexible, and high-performance data pipelines.

Dealing with massive textual sources to train GenAI foundation models.

Solving issues with data and data pipelines, prioritizing based on customer impact.

End-to-end ownership of data quality in our core datasets and data pipelines.

Experimenting with new tools and technologies to meet business requirements regarding performance, scaling, and data quality.

Providing tools that improve Data Quality company-wide, specifically for ML scientists.

Providing self-organizing tools that help the analytics community discover data, assess quality, explore usage, and find peers with relevant expertise.

Acting as an intermediary for problems, with both technical and non-technical audiences.

Promote and drive impactful and innovative engineering solutions

Technical, behavioral and interpersonal competence advancement via on-the-job opportunities, experimental projects, hackathons, conferences, and active community participation

Collaborate with multidisciplinary teams: Collaborate with product managers, data scientists, and analysts to understand business requirements and translate them into machine learning solutions. Provide technical guidance and mentorship to junior team members.
Requirements:
Bachelors or masters degree in computer science, Engineering, Statistics, or a related field.

Minimum of 6 years of experience as a Data Engineer or a similar role, with a consistent record of successfully delivering ML/Data solutions.

You have built production data pipelines in the cloud, setting up data-lake and server-less solutions; ‌ you have hands-on experience with schema design and data modeling and working with ML scientists and ML engineers to provide production level ML solutions.

You have experience designing systems E2E and knowledge of basic concepts (lb, db, caching, NoSQL, etc)

Strong programming skills in languages such as Python and Java.

Experience with big data processing frameworks such, Pyspark, Apache Flink, Snowflake or similar frameworks.

Demonstrable experience with MySQL, Cassandra, DynamoDB or similar relational/NoSQL database systems.

Experience with Data Warehousing and ETL/ELT pipelines

Experience in data processing for large-scale language models like GPT, BERT, or similar architectures - an advantage.

Proficiency in data manipulation, analysis, and visualization using tools like NumPy, pandas, and matplotlib - an advantage.

Experience with experimental design, A/B testing, and evaluation metrics for ML models - an advantage.

Experience of working on products that impact a large customer base - an advantage.

Excellent communication in English; written and spoken.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8809572
סגור
שירות זה פתוח ללקוחות VIP בלבד