Switch language한국어
Back to the list

Self-Supervised Learning of Structured Dynamics from Videos

TL;DR AI

Key summary

2 min read
  1. Researchers introduced Structured Dynamics Model, a self-supervised approach that disentangles camera motion from object motion in videos.

  2. The model uses frozen features from a pretrained vision transformer and future-feature prediction to learn structured video dynamics.

  3. Training combines unlabeled video with weak supervision from synthetic Kubric scenes.

  4. On the new ProbeMotion benchmark, it outperformed simple backbone features and competed well with some strongly supervised methods.

Read the original