2 papers
cs.RO2025
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
Guo Ye, Zexi Zhang, Xu Zhao +4
Vision-Language-Action (VLA) models have shown remarkable generalization by mapping web-scale knowledge to robotic control, yet they remain blind to physical contact. Consequently,…
cs.CV2025
DreamOmni3: Scribble-based Editing and Generation
Bin Xia, Bohao Peng, Jiyang Liu +8
Recently unified generation and editing models have achieved remarkable success with their impressive performance. These models rely mainly on text prompts for instruction-based ed…