Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
GMT: Goal-Conditioned Multimodal Transformer for 6-DOF Object Trajectory Synthesis in 3D Scenes
Huajian Zeng, Abhishek Saroha, Daniel Cremers +1
Synthesizing controllable 6-DOF object manipulation trajectories in 3D environments is essential for enabling robots to interact with complex scenes, yet remains challenging due to…
cs.CV2026
DrivIng: A Large-Scale Multimodal Driving Dataset with Full Digital Twin Integration
Dominik Rößle, Xujun Xie, Adithya Mohan +3
Perception is a cornerstone of autonomous driving, enabling vehicles to understand their surroundings and make safe, reliable decisions. Developing robust perception algorithms req…
cs.CV2025
UrbanIng-V2X: A Large-Scale Multi-Vehicle, Multi-Infrastructure Dataset Across Multiple Intersections for Cooperative Perception
Karthikeyan Chandra Sekaran, Markus Geisler, Dominik Rößle +6
Recent cooperative perception datasets have played a crucial role in advancing smart mobility applications by enabling information exchange between intelligent agents, helping to o…