zwangmw@connect.ust.hk · zwangmw@cse.ust.hk
Open to discussion, collaboration, and internships.
zwangmw@connect.ust.hk · zwangmw@cse.ust.hk
Open to discussion, collaboration, and internships.
CAIRI Supervised, Semi- and Self-Supervised Visual Representation Learning Toolbox and Benchmark
OpenSTL: A Comprehensive Benchmark of Spatio-Temporal Predictive Learning
[ICLR 2024] MogaNet: Efficient Multi-order Gated Aggregation Network
[CVPR'25] MergeVQ: A Unified Framework for Visual Generation and Representation with Token Merging and Quantization
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks