RetoVLA: Reusing Register Tokens for Spatial Reasoning in Vision-Language-Action Models Paper • 2509.21243 • Published Sep 25, 2025 • 1
RetoVLA: Reusing Register Tokens for Spatial Reasoning in Vision-Language-Action Models Paper • 2509.21243 • Published Sep 25, 2025 • 1
SPACE-CLIP: Spatial Perception via Adaptive CLIP Embeddings for Monocular Depth Estimation Paper • 2601.17657 • Published 17 days ago
Inference Optimized Checkpoints (with Model Optimizer) Collection A collection of generative models quantized and optimized for inference with Model Optimizer. • 52 items • Updated 5 days ago • 90