Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

SJTU VisionXLab

community
https://yangxue.site/publication
yangxue0827
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

zpy777  authored a paper 1 day ago
Point2RBox-v2: Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances
zpy777  authored a paper 1 day ago
Beyond Decision Boundaries: Relational Geometry Attacks on Contrastive Embedding Manifolds
zpy777  authored a paper 1 day ago
FIRM-Video: Check Before You Score for Reliable Text-to-Video Reward Modeling
View all activity

Papers

FIRM-Video: Check Before You Score for Reliable Text-to-Video Reward Modeling

Trust Your Critic: Robust Reward Modeling and Reinforcement Learning for Faithful Image Editing and Generation

View all Papers

Xue Yang's profile picture Qingyun Li's profile picture Xuehui Wang's profile picture Junwei Luo's profile picture Shuran Ma's profile picture Xiaoxing Hu's profile picture zpy's profile picture Junming Lin's profile picture zyl's profile picture mingqian's profile picture

VisionXLab 's models 9

VisionXLab/FIRM-Video-8B-InternVL3

Image-Text-to-Text • 8B • Updated 24 days ago • 12

VisionXLab/FIRM-Video-8B-Qwen3VL

Image-Text-to-Text • 9B • Updated 24 days ago • 12

VisionXLab/Point2RBox-v2-jittor

Object Detection • Updated 26 days ago

VisionXLab/Point2RBox-v3-jittor

Object Detection • Updated 30 days ago

VisionXLab/FIRM-Qwen-Edit

Updated Mar 11

VisionXLab/FIRM-SD3.5

Updated Mar 11

VisionXLab/FIRM-Gen-8B

Image-Text-to-Text • 770k • Updated Mar 1 • 13 • 1

VisionXLab/FIRM-Edit-8B

Image-Text-to-Text • 770k • Updated Mar 1 • 13 • 1

VisionXLab/ProCLIP

Image-Text-to-Text • Updated Oct 21, 2025 • 1
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs