This is your work, valued
Vision-Language Understanding; Visual Relationship Detection
PMFNet. Implementation of "Pose-aware Multi-level Feature Network for Human Object Interaction Detection"(ICCV 2019 Oral)
87cliora. Official codebase for ICLR oral paper Unsupervised Vision-Language Grammar Induction with Shared Structure Modeling
36Weakly-HOI.
5Zeroshot-HOI-with-CLIP. Official Implementation of "Exploiting CLIP for Zero-shot HOI Detection Requires Knowledge Distillation at Multiple Levels", WACVC 2024
1