דלג לתוכן הראשי

Computer Vision Engineer

Robotic Imaging, Inc.מחוז תל אביב, ישראללא צויןFull-timeדרגה: לא צוין

פורסם לפני 27 ימים · 0 מועמדים

שכר לא צוין במשרה זו

שמירה, הגשה או בדיקת התאמה — כמה שניות להקמת חשבון חינם.

תובנת Willbi

התפקיד במילים פשוטות

התפקיד כולל הפיכת סרטוני וידאו גולמיים שצולמו בחנויות לנתונים מובנים. המהנדס יבנה מנוע תפיסה שינתח את הסרטונים ויענה על שאלות סקר לגבי תכולת החנות, מצבה ועלות החלפת פריטים. העבודה כוללת תרומה לכל שלבי צינור עיבוד הווידאו, אימון מודלים ואיחוד פלט תפיסתי עם רב-מודאליות.

חובה
  • Production-grade computer vision: object detection and segmentation shipped in real systems (not notebooks)
  • Video / multi-frame understanding: tracking, temporal consistency across frames
  • A track record of taking models to production and owning them
  • Comfort with dirty, real-world capture and genuine evaluation rigor — you measure accuracy and reason about failure modes
  • High agency: you spot a perception bottleneck and build the fix without waiting for a ticket
יתרון
  • Modern multimodal models / VLMs for reasoning over perception output
  • Exposure to 3D / depth / point-cloud data

חולץ מתיאור המשרה · מתעדכן אוטומטית

למי זה מתאים

התפקיד מתאים למהנדסי ראייה ממוחשבת עם ניסיון מוכח בהטמעת מודלים למערכות אמיתיות, הבנה בווידאו וריבוי פריימים, ויכולת לפתור בעיות באופן עצמאי. הוא אינו מתאים למי שזקוק להנחיות מפורטות או לתור טיקטים.

תיאור המשרה המלא

המשרה המקורית · נשמר לעיון

Founding-level · in the code · owning perception end to end.

We send technicians into thousands of retail locations (7-Eleven, Kroger, the largest enterprise portfolios in the world) and they walk every store with a camera. Your job is to turn that raw video into structured truth. You'll build the perception engine that watches a walkthrough and answers what a site survey used to need a human for: what's in the store, what condition it's in, and what it costs to replace.

The challenge

• Inputs: raw on-site video from retail walkthroughs — handheld, real-world lighting, occlusion, motion blur.

• Goal: extract structured store attributes (fixtures, equipment, conditions, counts, dimensions) and auto-answer survey questions.

• The catch: an ingestion pipeline that works at scale — not for one client but for hundreds of enterprise clients, each with different site attributes and physical layouts.

What you'll do

• Contribute to the video-to-attributes pipeline end to end: ingestion, detection, segmentation, multi-frame tracking.

• Train and ship models that recognize retail fixtures, equipment, and site conditions from real footage.

• Fuse perception output with multimodal reasoning to populate our canonical attribute registry.

• Keep it fast and cheap at portfolio scale — this runs across thousands of stores, not a demo.

Must have

• Production-grade computer vision: object detection and segmentation shipped in real systems (not notebooks).

• Video / multi-frame understanding: tracking, temporal consistency across frames.

• A track record of taking models to production and owning them.

• Comfort with dirty, real-world capture and genuine evaluation rigor — you measure accuracy and reason about failure modes.

• High agency: you spot a perception bottleneck and build the fix without waiting for a ticket.

Nice to have (or fast to learn)

• Modern multimodal models / VLMs for reasoning over perception output.

• Exposure to 3D / depth / point-cloud data.

• Retail or built-environment domain experience.

Do not apply if you need a ticket queue or detailed specs handed to you. You'll navigate ambiguity and build.

אודות Robotic Imaging, Inc.
פרופיל החברה · בקרוב

ביקורות עובדים · בקרובעוד משרות ב-Robotic Imaging, Inc.

שאלות על המשרה

  • המשרה לא ציינה שכר. אנחנו מציגים שכר רק כשהמעסיק מפרסם אותו.
Robotic Imaging, Inc.
פורסם לפני 27 ימים · 0 מועמדים
בדקו את ההתאמה