דרושים » מחשבים ורשתות » Observability and SRE expert

משרות על המפה
 
בדיקת קורות חיים
VIP
הפוך ללקוח VIP
רגע, משהו חסר!
נשאר לך להשלים רק עוד פרט אחד:
 
שירות זה פתוח ללקוחות VIP בלבד
AllJObs VIP
כל החברות >
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location: Merkaz
We are seeking an Observability & SRE Expert with Hands-On experience in On-premise and Cloud environments to join us part-time (100 hours monthly).
עוד על התפקיד
the role involves practical tasks:
Implement monitoring, logging, tracing, and metrics for On-Premise and Cloud
Manage SLOs, SLAs, SLIs, and Error Budgets
Monitor, troubleshoot, and resolve performance issues. Handle incidents and
perform root cause analysis
Develop automation for operational processes. Manage infrastructure and write
maintenance scripts
Build intuitive dashboards and optimize alert management
Cross-Team collaboration: Work with Development, Infrastructure, Security, and
DevOps teams
Apply SRE principles to enhance system reliability and performance.
Requirements:
Experience: 7+ years hands-on in Observability, Monitoring, SRE, On-Premise
&Cloud
Tools: Prometheus, Grafana, ELK, Splunk, Datadog, Terraform, Ansible, etc.
Programming: Python, Java
Big Data: Hadoop, Spark, Kafka
Cloud: AWS/GCP/Azure, CloudWatch, Stackdriver
SLO/SLA/SLI/ Error Budgets: Practical experience
Incident Management: RCA, incident handling
automation: Scripting and process automation
כישורים נדרשים
Degree in Computer Science or a related field
Kubernetes, Docker experience
AI& Machine Learning knowledge.
This position is open to all candidates.
 
Hide
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8835859
סגור
שירות זה פתוח ללקוחות VIP בלבד
משרות דומות שיכולות לעניין אותך
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
Location:
Job Type: Full Time and Public Service / Government Jobs
We are seeking a Site Reliability Engineer (SRE) to join our team responsible for critical infrastructure.
עוד על התפקיד
Operate and manage the Kubernetes platform (RKE2) platform
Set up and manage CI/CD pipelines using tools like Jenkins Argo CD and others
Implement Infrastructure as Code (IaC) and infrastructure automation
Design and maintain monitoring systems using Prometheus, Grafana and Elastic
Develop internal monitoring and quality control tools
Analyze incidents, trends and availability using SQL
Build Self - Service capabilities for development teams
Participate in on - call rotations, handle production incidents, and lead documented post- mortems
Write and Maintain operational documentation
Requirements:
Familiarity with TCP/IP and communication protocols
Experience with Kubernetes in production environment
Experience with Argo CD, Jenkins and other CI/CD tools
Experience with monitoring and logging tools (Prometheus, Grafana, Elastic)
Development experience in Python and Knowledge of SQL
Strong Infrastructure understanding and Troubleshooting skills in complex environments
כישורים נדרשים
Advanced Networking Knowledge: Deep understating of TCP/IP, network protocols (UDP, HTTP, HTTPS, DNS, DHCP) and network security.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8835679
סגור
שירות זה פתוח ללקוחות VIP בלבד
סגור
דיווח על תוכן לא הולם או מפלה
מה השם שלך?
תיאור
שליחה
סגור
v נשלח
תודה על שיתוף הפעולה
מודים לך שלקחת חלק בשיפור התוכן שלנו :)
חברה חסויה
Location:
Job Type: Full Time and Public Service / Government Jobs
We are seeking a talented AI Specialist in DevOPS to lead the integration of artificial intelligence (AI) and machine learning (ML) technologies into our development, research, testing, and overall software systems. The ideal candidate will possess a deep understanding of DevOPS principles, coupled with hands- on experience in implementing AI models for code generation, automation, monitoring, and other critical functions.
עוד על התפקיד
In this role, you will design and establish robust infrastructure to support multiple development teams operating across diverse domains, technologies, and environments. We expect initiative, independent problem-solving, and innovative thinking to drive AI-driven solutions that enhance our software development lifecycle.
Requirements:
Bechelors degree in Computer Science, Software Engineering, Data science, or a related field, with 3 years of experience in DevOPS or similar roles.
5 years of experience in DevOPS or similar roles.
Mandatory AI experience- setting up local AI models, working with harness, using tools, MCP, and more practical experience in agent-based work on large projects: skills, tools, multi-agents.
Proficiency in multiple programming languages (at least python).
Experience with DevOPS tools: Docker, kubernetes, Jenkins, Git.
Deep understanding of language models and agents- context management, multi-agent systems, etc.
This position is open to all candidates.
 
Show more...
הגשת מועמדותהגש מועמדות
עדכון קורות החיים לפני שליחה
עדכון קורות החיים לפני שליחה
8835530
סגור
שירות זה פתוח ללקוחות VIP בלבד