DevOps Engineer (GCP | Kubernetes) – Mid Level
פורסם לפני 5 ימים · 0 מועמדים
התפקיד במילים פשוטות
תפקיד זה כולל תכנון, הקמה ותפעול של תשתיות ענן ב-GCP וסביבות Kubernetes בקנה מידה רחב. העבודה כוללת ניטור ותחזוקה של מערכות, ניהול בסיסי נתונים, פתרון תקלות בזמן אמת ושיפור ביצועי המערכת.
- Strong hands-on experience with Google Cloud Platform (GCP)
- Production experience with Kubernetes (GKE)
- Experience with Prometheus monitoring (PromQL, alerting, HA setups)
- Deep understanding of distributed systems
- Experience with large-scale time-series databases
- Experience with Thanos / Cortex / VictoriaMetrics
- Familiarity with real-time, high-scale systems
- Experience with GitOps workflows
- Scripting skills (Python / Bash)
חולץ מתיאור המשרה · מתעדכן אוטומטית
למי זה מתאים
התפקיד מתאים למהנדסי DevOps בעלי ניסיון מעשי חזק ב-GCP, Kubernetes ו-Prometheus, עם הבנה מעמיקה במערכות מבוזרות. הוא פחות מתאים למי שחסר ניסיון ייצורי בסביבות ענן וסביבות ניטור מתקדמות.
תיאור המשרה המלא
המשרה המקורית · נשמר לעיוןWe are looking for a hands-on DevOps Engineer to join a high-scale infrastructure team.
In this role, you will design, deploy, and operate highly available, large-scale systems supporting real-time data pipelines processing billions of events daily.
This is a production-focused role with real impact, ideal for someone passionate about distributed systems, performance, and reliability at scale.
Responsibilities
Infrastructure & Cloud
• Design, deploy, and manage cloud infrastructure (GCP)
• Own and optimize Kubernetes environments (GKE)
• Implement infrastructure-as-code (Terraform or similar)
Observability & Monitoring
• Build and maintain monitoring systems (Prometheus, Grafana)
• Define SLAs/SLOs and alerting strategies
• Lead incident response and post-mortem processes
Data & Infrastructure
• Design and maintain large-scale time-series databases
• Optimize performance, storage, and query efficiency
• Support real-time data pipelines and observability systems
Performance & Reliability
• Troubleshoot large-scale production systems
• Optimize performance and resource utilization
• Drive reliability practices (load testing, resilience, fault tolerance)
Collaboration
• Work closely with engineering teams on DevOps best practices
• Contribute to documentation, runbooks, and knowledge sharing
Requirements
Must Have
• Strong hands-on experience with Google Cloud Platform (GCP)
• Production experience with Kubernetes (GKE)
• Experience with Prometheus monitoring (PromQL, alerting, HA setups)
• Deep understanding of distributed systems
• Experience with large-scale time-series databases
• Strong troubleshooting and performance optimization skills
Nice to Have
• Experience with Thanos / Cortex / VictoriaMetrics
• Familiarity with real-time, high-scale systems
• Experience with GitOps workflows
• Scripting skills (Python / Bash)
שאלות על המשרה
- המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
- Strong hands-on experience with Google Cloud Platform (GCP), Production experience with Kubernetes (GKE), Experience with Prometheus monitoring (PromQL, alerting, HA setups), Deep understanding of distributed systems, Experience with large-scale time-series databases