Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

SpatialAxiom

https://d2i-ai.github.io/SpatialAxiom/
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

lubinfanĀ  authored a paper about 6 hours ago
PTZ-Calib: Robust Pan-Tilt-Zoom Camera Calibration
lubinfanĀ  authored a paper about 6 hours ago
AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization
lubinfanĀ  authored a paper about 6 hours ago
AddressVLM: Cross-view Alignment Tuning for Image Address Localization using Large Vision-Language Models
View all activity

Papers

CC-VQA: Conflict- and Correlation-Aware Method for Mitigating Knowledge Conflict in Knowledge-Based Visual Question Answering

Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs

View all Papers

Yujing  Lou's profile picture Guo's profile picture chenpingyi's profile picture Lubin Fan's profile picture Jintao Tong's profile picture ShenCao's profile picture Jiaqi Gu's profile picture hongyuyang's profile picture Yunzhuo Hao's profile picture
Organization Card
Community About org cards

šŸ‘‹ Welcome to SpatialAxiom

This is the organization of SpatialAxiom, built by the D2I Lab, Alibaba Group. In this organization, we continuously release spatial intelligence projects. Feel free to enjoy our latest models!

spaces 2

Running on Zero
MCP

SpatialAxiom-9B

🧠

Spatial reasoning VLM for 3D relations and perspective

6 days ago

models 2

SpatialAxiom/SpatialAxiom-9B

Image-Text-to-Text • 9B • Updated 3 days ago • 105 • 9

SpatialAxiom/SpatialAxiom-35B-A3B

Image-Text-to-Text • 35B • Updated 3 days ago • 74 • 7

datasets 0

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs