In this study, Zekun asked: When we visually segment observed actions, is perceptual segmentation fully driven by low-level features (e.g., motion dynamics, scene cuts), or also by high-level structure (e.g., distinct phases of a tennis serve).
We find that BOTH matter.