Skip to content
View xuyang-liu16's full-sized avatar
๐ŸŽฏ
Focusing
๐ŸŽฏ
Focusing

Block or report xuyang-liu16

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please donโ€™t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this userโ€™s behavior. Learn more about reporting abuse.

Report abuse
xuyang-liu16/README.md

๐ŸŒˆ I am Xuyang Liu (ๅˆ˜ๆ—ญๆด‹), an incoming PhD student at PolyU, where I will join the VC Lab under the supervision of Chair Prof. Lei Zhang (IEEE Fellow). I am also currently working as a research intern at OPPO Research Institute. Previously, I earned my M.S. from Sichuan University and spent a wonderful year interning at Alibaba Group and Ant Group. I am fortunate to work closely with Dr. Siteng Huang and Prof. Linfeng Zhang.

๐Ÿ“Œ My research centers on Efficient Multimodal Large Language Models (MLLMs), including:

  • ๐Ÿ–ผ๏ธ Image Understanding: high-resolution understanding via context compression and fast decoding, including GlobalCom2[AAAI'26], V2Drop[CVPR'26], FiCoCo[AAAI'26], and MixKV[ICLR'26].
  • ๐ŸŽฌ Video Understanding: long/audio-video, and streaming reasoning via efficient encoding and compression, including VidCom2[EMNLP'25], STC[CVPR'26], V-CAST, and OmniSIFT[ICML'26].
  • ๐ŸŽจ Content Generation: lightweight and efficient AIGC via feature caching, pruning, and fast decoding, including ToCa[ICLR'25], Flash-Unified[CVPR'26 Findings], and STDec.
  • โš™๏ธ Efficiency Toolbox: efficient transfer/fine-tuning and benchmarking for downstream task adaptation, including M2IST[TCSVT'25], V-PETL[NeurIPS'24], and AutoGnothi[ICLR'25].

๐Ÿ“ข If you find these directions interesting, feel free to reach out via email: liuxuyang@stu.scu.edu.cn. I am actively seeking internship opportunities!

Pinned Loading

  1. Awesome-Generation-Acceleration Awesome-Generation-Acceleration Public

    ๐Ÿ“š Collection of awesome generation acceleration resources.

    402 12

  2. Awesome-Token-level-Model-Compression Awesome-Token-level-Model-Compression Public

    ๐Ÿ“š Collection of token-level model compression resources.

    201 8

  3. VidCom2 VidCom2 Public

    [EMNLP 2025 Main] Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models

    Python 129 14

  4. GlobalCom2 GlobalCom2 Public

    [AAAI 2026] Global Compression Commander: Plug-and-Play Inference Acceleration for High-Resolution Large Vision-Language Models

    Python 42 1

  5. Shenyi-Z/ToCa Shenyi-Z/ToCa Public

    [ICLR2025] Accelerating Diffusion Transformers with Token-wise Feature Caching

    Python 221 10

  6. MixKV MixKV Public

    [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models

    Python 29 7