AI Research Scientist, Real-Time Video Understanding & Physical AI

Apply Now ↗
📍 Mountain View, California, United StatesFull time💰 USD 230K–270K

About this role

The Problem

As robots powered by learned policies enter real manufacturing environments, making them trustworthy becomes as important as making them capable. Learned policies can be confidently wrong — executing smoothly while doing something other than what was intended — and understanding what a robot is actually doing on a factory floor, in real time and from observation alone, is an open research frontier.

Our lab researches the AI systems that make robot fleets trustworthy in production: real-time perception and reasoning over robot behavior — and, at its core, detecting when a robot is doing something wrong — built on multi-camera video understanding and multimodal signals from the operating environment and the robot itself.

Samsung SDS builds and operates the systems behind Samsung's global manufacturing — thousands of production lines worldwide. This research is being developed with that environment as its destination.

This is not a monitoring dashboard project. It is a frontier problem in video understanding: not just recognizing what a robot is doing, but reliably detecting when it is doing it wrong — on continuous, real-world behavior, in real time, under production latency and cost constraints.

The Team

You would join a small, hands-on lab of PhD-level researchers — no layers between you and the research. Your research happens alongside real robots that our lab operates end-to-end: we collect our own data through teleoperation, train open-source robot foundation models on our GPUs, and deploy them to humanoid robots to test their behavior. Ongoing work extends to dexterous manipulation. The work that proves out in the lab has a path to pilot deployment in real manufacturing settings.

We run on a simple contract: the mission is fixed; the method is yours. The lab's direction is clear and executive-sponsored, and everyone's work compounds toward it — but how you get there (which architectures, which formulations, which experiments) is your call to make and defend.

At this size, a new researcher is not headcount. Your technical judgment shapes how we get there from your first week.

What You Will Work On

  • Detecting anomalous robot behavior from real-time streaming video of humanoid robots at work — fusing external multi-camera views and, potentially, the robot's own egocentric video
  • Efficient VLM research: adapting vision-language models to achieve low-latency, low-cost on-prem deployment without sacrificing reasoning quality
  • Video-language grounding: connecting continuous visual observations of robot behavior with language and structured task knowledge
  • Multimodal fusion beyond vision: combining camera streams with robot state signals and manufacturing context data into a unified representation for judgment
  • World models and video prediction: moving beyond detection — reasoning about the causes and dynamics of robot behavior, and anticipating what happens next

Why This Role

  • Own the technical agenda. This is an executive-sponsored research effort at an early, formative stage. You will shape the architecture, the research questions, and the evaluation standards — not inherit them.
  • Ship into the physical world at scale. Samsung SDS operates the systems behind Samsung's global manufacturing. When this research succeeds, it does not end as a paper or a demo — the deployment path runs onto real production lines, at a scale almost no research organization anywhere can offer. Papers are a milestone here, not the finish line.
  • A data setting few others have. Synchronized multi-camera video of robots at work, paired with rich operational context from real manufacturing environments — a combination that few academic labs or frontier AI labs can match.
  • Real robots, every day. The lab trains and runs learned policies on its own physical platforms — from teleoperation-based learning to dexterous manipulation — so your research has a living testbed: real robots executing real learned behaviors, generating the kind of behavioral data most video researchers never get to touch. Publishing and patenting are part of how we work.

Minimum Qualifications

  • PhD in Computer Vision, Machine Learning, Robotics, or a related field, or equivalent industry research experience
  • First-author publications at top venues (CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, CoRL, RSS, or comparable)
  • Hands-on research experience in video understanding: temporal action detection/segmentation, video anomaly detection, streaming/online perception, or long-form video reasoning

Preferred Qualifications

  • Video-language grounding, instructional/procedural video understanding, or video question answering
  • Multimodal learning combining vision with non-visual signals (sensor time-series, structured/tabular context)
  • World models, video prediction, or causal reasoning over temporal data
  • Exposure to robot learning (VLA models, imitation learning) — useful for understanding how learned policies behave and fail
  • Hands-on experience building perception or ML systems that run on real-world data streams — cameras, sensors, robots, or production data — not only curated benchmarks

Compensation 

This role offers competitive compensation package including base salary, bonus, and benefits.

Expected salary range for this role in Mountain View, CA:

$230,000 – $270,000 base salary, depending on experience, interview assessment results, skills and qualification. On top of base salary this role may participate in performance bonus plan.

Samsung SDSA offers a comprehensive suite of programs to support our employees:

  • Top-notch medical, dental, vision and prescription coverage
  • Wellness program
  • Parental leave
  • 401K match and savings plan
  • Flexible spending accounts
  • Life insurance
  • Paid Holidays
  • Paid Time off
  • Additional benefits

Samsung SDS America, Inc. is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, race, color, religion, sex, sexual orientation, gender identity or expression, national origin, disability, status as a protected veteran, marital status, genetic information, medical condition, or any other characteristic protected by law.

We are committed to providing reasonable accommodations to participate in the job application or interview process for candidates with disabilities. Please let your recruiter know if you need an accommodation at any point during the interview process.

Certain roles are eligible for additional rewards, including annual bonus. U.S.-based employees have access to medical, dental, and vision insurance, a 401(k) plan and company match, short-term and long-term disability coverage, basic life insurance, and wellbeing benefits, among others. U.S.-based employees also receive, per calendar year, up to 10 scheduled paid holidays, and Paid Time Off.

Frequently Asked Questions

What is the salary for the AI Research Scientist, Real-Time Video Understanding & Physical AI role at Samsung SDS America?
The listed salary for this AI Research Scientist, Real-Time Video Understanding & Physical AI position at Samsung SDS America is USD 230K–270K. This is an Full time role.
Where is the AI Research Scientist, Real-Time Video Understanding & Physical AI position at Samsung SDS America located?
This AI Research Scientist, Real-Time Video Understanding & Physical AI role at Samsung SDS America is based in Mountain View, California, United States. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the AI Research Scientist, Real-Time Video Understanding & Physical AI role at Samsung SDS America full-time or part-time?
This is listed as a Full time position. It is posted as a AI Research Scientist, Real-Time Video Understanding & Physical AI role in the SDSRA department at Samsung SDS America.
Which team or department does the AI Research Scientist, Real-Time Video Understanding & Physical AI at Samsung SDS America belong to?
This AI Research Scientist, Real-Time Video Understanding & Physical AI position is part of the SDSRA department at Samsung SDS America. See the full job description for more information about the team structure and responsibilities.
How do I apply for the AI Research Scientist, Real-Time Video Understanding & Physical AI position at Samsung SDS America?
Click the "Apply Now" button on this page. You will be redirected to Samsung SDS America's official application portal hosted on workable where you can submit your application directly.
When was the AI Research Scientist, Real-Time Video Understanding & Physical AI job at Samsung SDS America posted?
This AI Research Scientist, Real-Time Video Understanding & Physical AI position at Samsung SDS America was posted on Aug 19, 2026. Apply as soon as possible — early applications are often reviewed first.
AI Research Scientist, Real-Time Video Understanding & Physical AI
Samsung SDS America · 💰 USD 230K–270K
Apply for this role ↗

You'll be redirected to Samsung SDS America's official application page on workable.