Challenge Design

The challenge has been designed according to the Equator guideline BIAS. The full challenge design can be found here.

Three Competiton Tracks

SAVE FOCUSForeign Object Contextual Understanding for Safe Surgical AI – is organized in three tracks.

Frame Track

The FRAME Track evaluates a model’s ability to answer clinically relevant questions from a single image. This track targets core surgical scene understanding skills such as foreign object detection, identification, attribute recognition, and spatial localization within a single moment in time.

Segment Track

The SEGMENT Track focuses on short video segments (up to 5 min), requiring models to incorporate local temporal context to answer questions about foreign objects and their interactions with anatomy and instruments.

Procedure Track

The PROCEDURE Track challenges models with long surgical video contexts, ranging from extended segments to full-length laparoscopic procedures, to assess their capacity for long-term memory, persistent tracking, and global reasoning.

Together, these tracks enable a systematic characterization of where current VLMs succeed and fail as task complexity transitions from instantaneous perception to long-context intraoperative reasoning.

Sample Data

When was the first sponge inserted in the abdomen?
Please return the time-point.

This video includes surgical scenes. 
Viewer discretion is advised.