דלג לתוכן הראשי

Research Intern — Human Pose Understanding and Vision-Language Models

Appleהרצליה, מחוז תל אביב, ישראללא צויןInternshipדרגה: ג׳וניור

פורסם לפני 17 ימים · 0 מועמדים

שכר לא צוין במשרה זו

שמירה, הגשה או בדיקת התאמה — כמה שניות להקמת חשבון חינם.

תובנת Willbi

התפקיד במילים פשוטות

המתמחה במחקר יצטרף לפרויקט שמטרתו פרסום בכנס מוביל, ויעסוק בתכנון ופיתוח מערכות המשלבות הבנת תנוחת גוף אנושית עם מודלי ראייה-שפה (VLMs). התפקיד כולל שיתוף פעולה עם חוקרים ומהנדסים בצוות, מימוש שיטות חדשות ב-Python וביצוע הערכות ביצועים מול מדדים אקדמיים מקובלים.

חובה
  • Currently enrolled in a graduate program (M.Sc. or Ph.D.) in Computer Science, Electrical Engineering, or a related field
  • Publications at top-tier venues (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, ACL, EMNLP or similar)
  • Strong programming skills in Python and experience with deep learning frameworks (e.g., PyTorch)
  • Solid foundation in computer vision, natural language processing, or multimodal learning
יתרון
  • Demonstrated expertise working with Vision-Language Models (VLMs) and/or Large Language Models (LLMs)
  • Experience with human pose estimation, motion modeling, or related body-tracking tasks
  • Familiarity with video understanding tasks and temporal modeling
  • Familiarity with multimodal learning and benchmarks that combine language with visual or spatial data
  • Experience with prompt engineering and optimization techniques

חולץ מתיאור המשרה · מתעדכן אוטומטית

למי זה מתאים

תפקיד זה מתאים לסטודנטים לתארים מתקדמים (M.Sc. או Ph.D.) במדעי המחשב או הנדסת חשמל, שיש להם כבר פרסומים בכנסים מובילים ורקע חזק בלמידה עמוקה וראייה ממוחשבת. הוא פחות יתאים למי שאין לו ניסיון מחקרי מוכח או רקע אקדמי מתאים בתחומים אלו.

תיאור המשרה המלא

המשרה המקורית · נשמר לעיון

Summary We are looking for a research intern to join us for a research project aimed at publication at a top-tier venue. The intern will design and develop novel systems that explore the interaction between human pose understanding and vision-language models (VLMs), advancing how these modalities can be combined to reason about human motion, activity, and embodied behavior across images and video.

Description Our group develops hand and body pose tracking algorithms for various apple devices and applications. One such example includes the hand tracking input for the Vision Pro.

Responsibilities

• Design and implement novel methods that integrate pose representations with vision-language models, targeting established academic benchmarks

• Collaborate with researchers and engineers on the team to produce a publication-ready contribution

• Benchmark against established evaluation suites and iterate toward state-of-the-art results

Minimum Qualifications

• Currently enrolled in a graduate program (M.Sc. or Ph.D.) in Computer Science, Electrical Engineering, or a related field

• Publications at top-tier venues (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, ACL, EMNLP or similar)

• Strong programming skills in Python and experience with deep learning frameworks (e.g., PyTorch)

• Solid foundation in computer vision, natural language processing, or multimodal learning

Preferred Qualifications

• Demonstrated expertise working with Vision-Language Models (VLMs) and/or Large Language Models (LLMs)

• Experience with human pose estimation, motion modeling, or related body-tracking tasks

• Familiarity with video understanding tasks and temporal modeling

• Familiarity with multimodal learning and benchmarks that combine language with visual or spatial data

• Experience with prompt engineering and optimization techniques

At Apple, we believe accessibility is a fundamental human right. You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools. By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.

Learn about accessibility in Apple’s workplace

Role Number: 200671694-0865

אודות Apple
פרופיל החברה · בקרוב

ביקורות עובדים · בקרובעוד משרות ב-Apple

שאלות על המשרה

  • המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
Apple
פורסם לפני 17 ימים · 0 מועמדים
בדקו את ההתאמה