Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
JerExJs's picture

JerExJs

xnnxaxcn93
2 3 24
lee137's profile picture
·
https://jiashenhzju.github.io
  • JerExJs

AI & ML interests

MLLM | multimodal representation, multimodal reasoning, agentic vision

Organizations

None yet

upvoted a paper 3 months ago

Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs

Paper • 2603.02556 • Published Mar 3 • 4
upvoted a collection 3 months ago

VC-STaR

Collection
Official collection of models and datasets for the ICLR 2026 Oral paper "Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs" • 3 items • Updated May 14 • 2
upvoted a collection over 1 year ago

Qwen2.5-Omni

Collection
End-to-End Omni (text, audio, image, video, and natural speech interaction) model based Qwen2.5 • 6 items • Updated Mar 2 • 168
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs