This is your work, valued
I am a M.S. student at Beihang University, primarily focused on researching deep learning, and multimodal.
VISA. [ECCV24] VISA: Reasoning Video Object Segmentation via Large Language Model
214PiClick. Official PyTorch implementation of PiClick: Picking the desired mask in click-based interactive segmentation.
26ReVOS-api. [ECCV24] VISA: Reasoning Video Object Segmentation via Large Language Model
22LTCA. The official implementation of "LTCA: Long-range Temporal Context Attention for Referring Video Object Segmentation".
7crossvid-pub.
1