Skip to content
View xuyang-liu16's full-sized avatar
๐ŸŽฏ
Focusing
๐ŸŽฏ
Focusing

Block or report xuyang-liu16

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
xuyang-liu16/README.md

๐ŸŒˆ I am Xuyang Liu (ๅˆ˜ๆ—ญๆด‹), a first-year PhD student at HK PolyU, where I am a member of the VC Lab under the supervision of Chair Prof. Lei Zhang (IEEE Fellow). I am also currently working as a research intern at OPPO Research Institute. Previously, I earned my M.S. from Sichuan University and spent a wonderful year interning at Alibaba Group and Ant Group. I am fortunate to work closely with Dr. Siteng Huang and Prof. Linfeng Zhang.

๐Ÿ“Œ My research centers on Efficient Multimodal Large Language Models (MLLMs), including:

  • ๐Ÿ–ผ๏ธ Image Understanding: high-resolution understanding via context compression and fast decoding, including GlobalCom2[AAAI'26], V2Drop[CVPR'26], FiCoCo[AAAI'26], and MixKV[ICLR'26].
  • ๐ŸŽฌ Video Understanding: long-video, audio-video, and streaming reasoning via efficient encoding and compression, including VidCom2[EMNLP'25], STC[CVPR'26], V-CAST, and OmniSIFT[ICML'26].
  • ๐ŸŽจ Content Generation: lightweight and efficient AIGC via feature caching, pruning, and fast decoding, including ToCa[ICLR'25], Flash-Unified[CVPR'26 Findings], and STDec.
  • โš™๏ธ Efficiency Toolbox: efficient transfer/fine-tuning and benchmarking for downstream task adaptation, including M2IST[TCSVT'25], V-PETL[NeurIPS'24], and AutoGnothi[ICLR'25].

๐Ÿ“ข Please email me at xu-yang.liu@connect.polyu.hk or liuxuyang@stu.scu.edu.cn for any form of academic collaboration.

Pinned Loading

  1. Awesome-Generation-Acceleration Awesome-Generation-Acceleration Public

    ๐Ÿ“š Collection of awesome generation acceleration resources.

    403 12

  2. Awesome-Token-level-Model-Compression Awesome-Token-level-Model-Compression Public

    ๐Ÿ“š Collection of token-level model compression resources.

    203 8

  3. VidCom2 VidCom2 Public

    [EMNLP 2025 Main] Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models

    Python 131 13

  4. GlobalCom2 GlobalCom2 Public

    [AAAI 2026] Global Compression Commander: Plug-and-Play Inference Acceleration for High-Resolution Large Vision-Language Models

    Python 43 1

  5. Shenyi-Z/ToCa Shenyi-Z/ToCa Public

    [ICLR2025] Accelerating Diffusion Transformers with Token-wise Feature Caching

    Python 223 10

  6. MixKV MixKV Public

    [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models

    Python 31 7