Skip to content
View Xuchen-Li's full-sized avatar
:electron:
Working Hard~
:electron:
Working Hard~
  • CASIA
  • Beijing, China
  • 12:47 (UTC +08:00)

Organizations

@OpenDCAI @bjzgcai

Block or report Xuchen-Li

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Xuchen-Li/README.md

Xuchen Li / 李旭宸

Ph.D. Student of Pattern Recognition and Intelligent System

Institute of Automation, Chinese Academy of Sciences and Zhongguancun Academy

lixuchen2024@ia.ac.cn

Google Scholar / Github

Biography

I am a Ph.D. student at Institute of Automation, Chinese Academy of Sciences (CASIA) and Zhongguancun Academy (ZGCA). Before that, I received my B.E. degree in Computer Science and Technology with overall ranking 1/449 (0.22%) at the School of Computer Science from Beijing University of Posts and Telecommunications (BUPT) in Jun. 2024. I am grateful to be growing up and studying with my twin brother Xuzhao Li (M.S. Student at BIT), which is a truly unique and special experience for me.

Research

  • Multi-modal Adaptive Agentic RL and OPD: Empower agents through reinforcement learning and on-policy distillation to autonomously interact with complex, dynamic environments and carry out multi-turn, long-horizon decision-making tasks, enabling them to better process text-image and video information.
  • Streaming Video Understanding Memory Modeling: Optimize the dense memory compression, memory retrieval triggering mechanism and memory window management capabilities of streaming video understanding models.
  • Fine-grained Evaluation of Multi-modal Reasoning: Build benchmarks targeting the model’s reasoning process to conduct fine-grained evaluation and diagnosis of logical consistency, reasoning effectiveness and the distribution of key information.

Pinned Loading

  1. 983632847/Awesome-Multimodal-Object-Tracking 983632847/Awesome-Multimodal-Object-Tracking Public

    A continuously updated project to track the latest progress in the field of multi-modal object tracking. This project focuses solely on single-object tracking.

    Jupyter Notebook 1.1k 56

  2. llm-arxiv-daily llm-arxiv-daily Public

    Automatically update arXiv papers about LLM Reasoning, LLM Evaluation, LLM & MLLM and Video Understanding using Github Actions.

    Python 146 13

  3. cv-arxiv-daily cv-arxiv-daily Public

    Automatically update arXiv papers about SOT & VLT, Multi-modal Learning, LLM and Video Understanding using Github Actions.

    Python 48 6