Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yicheng Xu's picture

Yicheng Xu

linghan199
1 8 14
21world's profile picture SteveSHEN's profile picture Datawitch-Programmer's profile picture
·

AI & ML interests

None yet

Organizations

OpenGVLab's profile picture Nanyang Technological University's profile picture

upvoted a paper 2 months ago

TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs

Paper • 2607.17423 • Published Jul 19 • 98
upvoted a collection 2 months ago

TimeLens2

Collection
Generalist Video Temporal Grounding with Multimodal LLMs • 8 items • Updated Jul 28 • 15
upvoted a paper 2 months ago

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding

Paper • 2607.14935 • Published Jul 16 • 120
upvoted 3 papers 3 months ago

TimeLens: Rethinking Video Temporal Grounding with Multimodal LLMs

Paper • 2512.14698 • Published Dec 16, 2025 • 27

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers

Paper • 2606.13289 • Published Jun 11 • 31

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning

Paper • 2606.12195 • Published Jun 10 • 24
upvoted a paper 6 months ago

Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale

Paper • 2603.25040 • Published Mar 26 • 132
upvoted a paper 11 months ago

ExpVid: A Benchmark for Experiment Video Understanding & Reasoning

Paper • 2510.11606 • Published Oct 13, 2025 • 6
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs