3 ms·Spatial VLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities1 points by kuter 3y ago