דלג לתוכן הראשי

HPC Support Team Lead

Weizmann Institute of Scienceרחובות, מחוז המרכז, ישראללא צויןFull-timeדרגה: ליד

פורסם לפני 4 ימים · 0 מועמדים

שכר לא צוין במשרה זו

שמירה, הגשה או בדיקת התאמה. פתיחת חשבון חינם לוקחת כמה שניות.

תובנת Willbi

התפקיד במילים פשוטות

התפקיד כולל ניהול צוות תמיכה מסדר גודל Tier 2.5, ניטור מדדי שירות ותפעול תשתיות מחשוב עתיר ביצועים (HPC) ובינה מלאכותית במכון. בעל התפקיד מעניק תמיכה טכנולוגית מתקדמת לחוקרים ולצוותים פנימיים, פותר תקלות מורכבות ומסייע בהרצת תוכנות מדעיות וספריות פייתון מתקדמות.

חובה
  • Managing Tier 2.5 team
  • Over 10 years experience managing various Linux environments
  • 5 years team management experience in HPC/AI field
  • At least 5 years experience with DevOps tools such as Jenkins and IaC approach
  • Experience working with compilers, MPI libraries, and scientific computing environments

חולץ מתיאור המשרה · מתעדכן אוטומטית

למי זה מתאים

התפקיד מתאים למנהלים בעלי ניסיון שטח של מעל 10 שנים בסביבות Linux ו-5 שנות ניסיון בניהול צוותים בתחומי ה-HPC וה-AI, לצד ניסיון בכלי DevOps. התפקיד פחות מתאים למי שחסר ניסיון ניהולי או ניסיון מעשי בתשתיות מחשוב מדעי מתקדמות.

תיאור המשרה המלא

המשרה המקורית · נשמר לעיון

Your role will include:

Responsible for managing the Tier 2.5 team, monitoring service metrics, and operating and supporting the institute's core HPC infrastructure, providing advanced technological support to institute researchers and internal teams, including among others:

Supporting High-Performance Computing (HPC) and Artificial Intelligence (AI) infrastructure.

Handling complex service tickets in the institute's core HPC/AI environments and providing advanced solutions for systemic issues.

Providing professional support to internal teams and institute researchers on all matters related to compiling, optimizing, and running complex scientific software in HPC/AI environments, including tools such as WRF, SAM, and Flash.

Working with scientific development environments and advanced Python libraries, including TensorFlow, PyTorch, Dask, mpi4py, NumPy, and SciPy.

Operating and optimizing job scheduling engines such as Slurm, LSF, and PBS.

Skills and abilities:

Significant experience managing various Linux environments, over 10 years -required.

Team management experience in the HPC/AI field, 5 years -required.

Experience working with High-Performance Computing (HPC) and/or AI environments -required.

Experience with DevOps tools such as Jenkins and an IaC approach, at least 5 years -required.

Experience working with compilers, MPI libraries, and scientific computing environments -required.

Experience troubleshooting complex issues in HPC/AI environments, including CPU/GPU performance testing.

Experience debugging scientific applications deployed across multiple servers and GPU accelerators.

Experience working with job scheduling engines such as Slurm, LSF, and PBS.

High-level proficiency in Hebrew and English, written and spoken.

Strong ability to analyze system failures and solve complex problems in production environments.

Ability to work independently, in a team, and in a dynamic, multitasking environment.

Responsibility, initiative, and strong self-learning ability.

Systems thinking and the ability to lead complex technological processes.

Analytical thinking and the ability to quickly learn new technologies.

Strong service orientation and interpersonal communication skills.

אודות Weizmann Institute of Science
פרופיל החברה · בקרוב

ביקורות עובדים · בקרובעוד משרות ב-Weizmann Institute of Science

שאלות על המשרה

  • המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
דומות וקשורות
Weizmann Institute of Science
פורסם לפני 4 ימים · 0 מועמדים
בדקו את ההתאמה