Skip to content
View JackXing875's full-sized avatar
🥰
,,ᗜ - ᗜ,,
🥰
,,ᗜ - ᗜ,,

Block or report JackXing875

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
JackXing875/README.md

Hi there, I'm Yuu Waving Hand Emoji

I’m an undergraduate at the Gaoling School of Artificial Intelligence, Renmin University of China.

My current research interests lie in Multimodal Reasoning, Efficient Multimodal Models, and Video Generation. In particular, I am interested in understanding and reducing computational redundancy in multimodal models through KV cache compression, token pruning, sparse attention, and adaptive caching.

I also have experience in Visual SLAM and robotic perception, which has shaped my interest in building efficient and robust intelligent systems that can perceive, reason about, and interact with the real world.

Connect with me:

Google Scholar  X (Twitter)  Email


🚀 Research & Projects

  • Constraint-Guided Prompting and Semantic-Aware Evaluation for LLM-Based ABSA

    — Muzhi Li, Tiancheng Xing, Yuheng Wang, ICIC 2026 · Springer LNAI

    • Introduces Constraint-Guided Prompting (CGP) for more reliable structured extraction with large language models and Sem-F1, a semantic-aware evaluation protocol for Aspect-Based Sentiment Analysis.
  • NeneBot

    • A source-grounded character conversational AI that uses RAG over original game scripts to preserve character knowledge and persona. It combines FAISS-based semantic retrieval, pluggable local/cloud LLM backends, multi-turn session memory, real-time SSE streaming, and an immersive Vue 3 visual-novel interface.
  • SLAMForge

    • A C++20 monocular visual SLAM and dense reconstruction system inspired by ORB-SLAM3. It combines geometric tracking, local bundle adjustment, Sim(3) loop closure, and pose-graph optimization with learned monocular depth, using sparse SLAM landmarks and multi-view consistency to reconstruct a colored dense 3D map.

🧠 Research Interests

  • Multimodal Reasoning: Vision-Language Models · Long-Horizon Reasoning · Visual Information Flow
  • Efficient Multimodal Inference: KV Cache Compression & Eviction · Token Pruning · Sparse Attention · Dynamic Budget Allocation
  • Video Generation: Diffusion Models · Flow Matching · Video DiT · Autoregressive / Chunk-wise Generation · Feature & KV Caching
  • Embodied Intelligence: Vision-Language-Action Models · World Models · Efficient Embodied Reasoning

Pinned Loading

  1. NeneBot NeneBot Public

    綾地寧々は世界一可愛い!

    Python 17 3

  2. SLAMForge SLAMForge Public

    An industrial-grade monocular visual SLAM system implementing the full ORB-SLAM3 pipeline: feature-based tracking with ORB descriptors, local bundle adjustment, Sim(3) loop detection and correction…

    C++ 9 3

  3. AuroraTeX AuroraTeX Public

    A clean, modern, and customizable LaTeX report template for academic reports, course projects, and technical notes.

    TeX 2

  4. bloom-into-you-skill bloom-into-you-skill Public

    An anime-grounded Agent Skill for Yuu Koito and Touko Nanami from *Bloom Into You* — distilled from the TV series to preserve her personality, relationships, and character-consistent behavior.

    Shell 5