PhD student in computer vision @HKU
GPT4Scene-and-VLN-R1. GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models
531Tailor3D. This is the official code for the paper Tailor3D
183Gemini-vs-GPT4V.
167gobjaverse-lvis. Jupyter Notebook
9Pointcept. Pointcept: a codebase for point cloud perception research. Latest works: MSC (CVPR'23), CeCo (CVPR'23), PTv2 (NeurIPS'22)
1mmdetection. OpenMMLab Detection Toolbox and Benchmark
1BEVFormer. [ECCV 2022] This is the official implementation of BEVFormer, a camera-only framework for autonomous driving perception, e.g., 3D object detection and semantic map segmentation.
1Multi-Task-Learning-PyTorch. PyTorch implementation of multi-task learning architectures, incl. MTI-Net (ECCV2020).
1