Fangzhou Zhao, Yao Sun, Xuesong Liu +3 more
A robot or vehicle building a 3D model of its surroundings as it moves has to send a continuous image stream to a server that does the heavy computation. Unlike offline reconstruction, where you can gather everything first and think later, camera poses and geometry are estimated while the platform is still moving, so multi-view consistency becomes a real-time requirement rather than a post-processing concern.
That makes the network link part of the perception system. A dropped or delayed frame is not just missing data, it distorts the geometry being estimated right now, and the authors note how sensitive that estimation is to communication-induced problems.
Semantic communication is the proposed response: transmit what the reconstruction needs rather than the pixels a camera happened to capture. If the server is going to extract geometric structure anyway, sending a faithful image and then discarding most of it is a strange use of scarce bandwidth.
Real-time mobile 3D reconstruction is fundamental to many emerging applications such as autonomous navigation and digital twin construction, where a moving platform continuously captures an image stream and transmit to a computing server for scene understanding. Unlike offline reconstruction, camera poses and scene geometry are estimated on-the-fly during acquisition, making multi-view consistency a real-time requirement and rendering geometric estimation highly sensitive to…
Decoding Market Emotion from Blockchain Activity: A Data-Driven Sentiment Classifier
arXiv (cs.AI) · July 16, 2026SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration
arXiv (cs.CV) · July 16, 2026HoloGeo: Mitigating Landmark Bias in Geo-localization via Evidence-Driven Reasoning
arXiv (cs.AI) · July 16, 2026teLLMe Why (Ain't Nothing but a Jam): Exploratory Causal Analysis of Urban Driving Data