🚀 Introducing TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs.
✨TimeLens2-8B achieves SoTA results across 7 benchmarks, outperforming Qwen3.5-397B-A17B by 5.8 points.
Paper: arxiv.org/abs/2607.17423
Model & Data: huggingface.co/collections/MC…
CS Ph.D. student at Nanjing University
Research Intern at Shanghai AI Laboratory
Focus on video-language foundation models

