Video Object Segmentation in Cultural Context

Published in In preparation, 2026

Tracking and segmenting culturally significant objects through video, where the object’s identity depends on cultural knowledge rather than appearance alone.

Status

This work is under development — earlier stage than the papers under review, with no preprint and no results to report yet.

If it overlaps something you are working on, get in touch; I would rather compare notes early than collide at submission.

It extends the grounding requirement in Seeing Culture (EMNLP 2025) and RARSeg from single images into video, and shares the video setting of the Cultural Moment benchmark.