MA-LMM. (2024CVPR) MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
350A2Summ. The official implementation of 'Align and Attend: Multimodal Summarization with Dual Contrastive Losses' (CVPR 2023)
86D-NeRV. The official implementation of 'Towards Scalable Neural Representation for Diverse Videos' (CVPR 2023)
47ASM-Loc. (CVPR2022) ASM-Loc: Action-aware Segment Modeling for Weakly-Supervised Temporal Action Localization
452048. My 2048 game
1