edatai commited on
Commit
e5d7c45
·
verified ·
1 Parent(s): a8acb72

Add model card

Browse files
Files changed (1) hide show
  1. README.md +12 -3
README.md CHANGED
@@ -18,6 +18,11 @@ SpatioLM adds a plug-and-play spatio-vision module to a frozen vision-language
18
  model and learns physically coherent representations from pseudo depth and
19
  camera-ray supervision. No additional 3D input is required at inference time.
20
 
 
 
 
 
 
21
  ## Installation
22
 
23
  ```bash
@@ -81,8 +86,12 @@ checkpoint interface, see the
81
  ```bibtex
82
  @inproceedings{wu2026spatiolm,
83
  title={SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models},
84
- author={Wu, Jianhua and others},
85
- booktitle={International Conference on Machine Learning},
86
- year={2026}
 
 
 
 
87
  }
88
  ```
 
18
  model and learns physically coherent representations from pseudo depth and
19
  camera-ray supervision. No additional 3D input is required at inference time.
20
 
21
+ ## Resources
22
+
23
+ - GitHub: https://github.com/xiaomi-research/spatio-lm
24
+ - Paper: https://arxiv.org/abs/2608.01899
25
+
26
  ## Installation
27
 
28
  ```bash
 
86
  ```bibtex
87
  @inproceedings{wu2026spatiolm,
88
  title={SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models},
89
+ author={Wu, Jing and Wu, Jianhua and Guan, Jiayi and Chen, Jiahong and Lu, Jinghui and Ye, Hangjun and Gao, Bingzhao and Chen, Long},
90
+ booktitle={International Conference on Machine Learning (ICML)},
91
+ year={2026},
92
+ note={To appear},
93
+ eprint={2608.01899},
94
+ archivePrefix={arXiv},
95
+ url={https://arxiv.org/abs/2608.01899}
96
  }
97
  ```