Related Experiment Video
Updated: Feb 14, 2026

Creating Objects and Object Categories for Studying Perception and Perceptual Learning
Published on: November 2, 2012
Slot-BERT: Self-supervised object discovery in surgical video
Guiqiu Liao1, Matjaž Jogan1, Marcel Hussing2
1Department of Surgery, Penn Computer Assisted Surgery and Outcomes Laboratory, University of Pennsylvania, Philadelphia, PA, USA.
None:
Object-centric slot attention is a powerful framework for unsupervised learning of structured and explainable representations that can support reasoning about objects and actions, including in surgical video. However, current object-centric models either fail to reliably capture object dependencies in seconds-long video episodes that encompass surgical actions and tasks or are computationally too expensive for practical implementation. We introduce Slot-BERT, a slot attention model with a temporal slot transformer module to overcome these limitations. Our core innovations are: 1) A bidirectional transformer module that processes object-centric slot representations, enabling longer-range temporal coherence; 2) A slot-contrastive loss that further improves the representation by enforcing slot dissimilarity; 3) We evaluate Slot-BERT on real-world surgical video datasets from abdominal, cholecystectomy, and thoracic procedures, and on real and synthetic videos with everyday objects. Our method surpasses state-of-the-art object-centric approaches under unsupervised training achieving superior performance across these domains. We also demonstrate efficient zero-shot domain adaptation to data from diverse surgical specialties and databases.
Related Concept Videos
Drug Discovery: Overview
Velocity of an Object
Potential Due to a Polarized Object
Potential Due to a Magnetized Object
The vector...
Moment of Inertia of Compound Objects
Consider a child of mass (mc) 25 kg standing at a distance (rc) of 1 m from the axis of a rotating...
Gravitational Potential Energy for Extended Objects

