LATIDIA · Robótica
CognitiveReality: Robot-Agnostic Semantic Gaussian Mapping with an LLM Agent for Immersive Collaborative VR Teleoperation
arXiv:2609.31418v1 Tipo de anuncio: nuevo Resumen: Una vista 3D fotorrealista le dice a un teleoperador dónde está un robot, pero no qué contiene la escena, qué tan bien se ha observado cada objeto o cómo girar apuntando y hablando
WhatsApp ↗Telegram ↗
La noticia
arXiv:2609.31418v1 Announce Type: new Abstract: A photorealistic 3D view tells a teleoperator where a robot is, but not what the scene contains, how well each object has been observed, or how to turn pointing and speech into robot action. CognitiveReality turns a robot's RGB-D stream into a live, semantically indexed Gaussian-TSDF map shared by an operator in virtual reality and a tool-using language agent. One mapper binary serves any platform through configuration alone: it ingests poses from robot SLAM, joint kinematics, motion capture or an inline visual tracker, bridges localization outages with a shadow tracker and keyframe-anchored PnP, and maintains open-vocabulary instance identities with per-object quality at 2 Hz. Speech