0:37InFlux++: benchmark and synthetic data for dynamic camera intrinsics@ErichLiang · 208 views · 2026-09-01InFluxcamera intrinsicsBenchmarks
4:25DART turns SAM3 into a real-time open-vocabulary detector at 55.8 AP on COCO@rsasaki0109 · 216 views · 2026-08-29open-vocabulary detectionSAM3DART
See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action ModelsThis paper addresses the frame mismatch in VLA models between camera-frame observation and robot-frame action by introducing robot-centric pointmaps—images whose pixels store 3D coordinates in the robot frame. Pointmaps provide robot-frame 3D geometry while preserving image structure, enabling cross-viewpoint generalization across diverse camera setups.Byungkun Lee, Dongyoon Hwang, Dongjin Kim·Jul 13, 2026VLAPointmapcross-viewpointJul 13, 2026
1:25Robot-centric pointmaps align VLA perception with the robot frame@shenzhenfoundry · 54 views · 2026-08-12VLARobot-Centric PointmapsEmbodied AI
Describe Anything, Anywhere, at Any Moment (DAAAM)DAAAM is a spatio-temporal memory framework for real-time 4D scene understanding. It uses an optimization-based frontend to infer detailed semantic descriptions from localized captioning models (DAM) with batch processing, and builds a hierarchical 4D scene graph for large-scale AR and robot autonomy applications.Nicolas Gorlo, Lukas Schmid, Luca Carlone·Nov 29, 2025spatio-temporalDAMMIT-SPARKNov 29, 2025
0:30Insta360 X6: Spatial Capture turns 360 recordings into 3D Gaussian Splat scenes@justinryanio · 75 views · 2026-08-07Insta360Gaussian SplattingSpatial Capture
0:53VidMap: Exploiting temporal structure for video-based Structure-from-Motion@zador_pataki · 95 views · 2026-08-07SLAMSfMECCV2026
0:08RoboSense Ships 719K LiDAR Units in H1 2026@techniahqrobot · 77 views · 2026-08-01LiDARRoboSensePerception