edatai commited on
Commit
b2491ee
·
verified ·
1 Parent(s): 8aed42b

Add model card

Browse files
Files changed (1) hide show
  1. README.md +12 -3
README.md CHANGED
@@ -19,6 +19,11 @@ SpatioLM adds a plug-and-play spatio-vision module to a frozen vision-language
19
  model and learns physically coherent representations from pseudo depth and
20
  camera-ray supervision. No additional 3D input is required at inference time.
21
 
 
 
 
 
 
22
  ## Installation
23
 
24
  ```bash
@@ -82,8 +87,12 @@ checkpoint interface, see the
82
  ```bibtex
83
  @inproceedings{wu2026spatiolm,
84
  title={SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models},
85
- author={Wu, Jianhua and others},
86
- booktitle={International Conference on Machine Learning},
87
- year={2026}
 
 
 
 
88
  }
89
  ```
 
19
  model and learns physically coherent representations from pseudo depth and
20
  camera-ray supervision. No additional 3D input is required at inference time.
21
 
22
+ ## Resources
23
+
24
+ - GitHub: https://github.com/xiaomi-research/spatio-lm
25
+ - Paper: https://arxiv.org/abs/2608.01899
26
+
27
  ## Installation
28
 
29
  ```bash
 
87
  ```bibtex
88
  @inproceedings{wu2026spatiolm,
89
  title={SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models},
90
+ author={Wu, Jing and Wu, Jianhua and Guan, Jiayi and Chen, Jiahong and Lu, Jinghui and Ye, Hangjun and Gao, Bingzhao and Chen, Long},
91
+ booktitle={International Conference on Machine Learning (ICML)},
92
+ year={2026},
93
+ note={To appear},
94
+ eprint={2608.01899},
95
+ archivePrefix={arXiv},
96
+ url={https://arxiv.org/abs/2608.01899}
97
  }
98
  ```