-
Taobao, Alibaba
- Hangzhou, China
- https://scholar.google.com/citations?user=KtwTy88AAAAJ
Pinned Loading
-
MME-Benchmarks/Video-MME
MME-Benchmarks/Video-MME Public✨✨[CVPR 2025] Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
-
Video-RAG-master
Video-RAG-master Public✨✨[NeurIPS 2025] This is the official implementation of our paper "Video-RAG: Visually-aligned Retrieval-Augmented Long Video Comprehension"
-
MAC-AutoML/QuoTA
MAC-AutoML/QuoTA Public✨✨[AAAI 2026] This is the official implementation of our paper "QuoTA: Query-oriented Token Assignment via CoT Query Decouple for Long Video Comprehension"
-
MAC-AutoML/OmniScope
MAC-AutoML/OmniScope Public[ACMMM 2026🔥] This is the official implementation of our paper "OmniScope: Modality-decoupled Token Compression for Efficient Omnimodal Video Understanding"
Python 7
-
TaoLiveAIGC/TLive-Omni
TaoLiveAIGC/TLive-Omni PublicTLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming
Python 4
If the problem persists, check the GitHub status page or contact support.


