From the 1 of 12 linked papers with an AI index.
12 papers
DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention
Xing Lei, Wenyan Yang, Xuetao Zhang +1
The paper proposes DAGR, a method that refines static goal embeddings into state‑conditioned representations using multi‑scale gated cross‑attention, improving performance on navig…
ReMoBot: Retrieval-Based Few-Shot Imitation Learning for Mobile Manipulation with Vision Foundation Models
Yuying Zhang, Wenyan Yang, Francesco Verdoja +2
Imitation learning (IL) algorithms typically distill demonstrations into parametric policies to mimic expert behavior. However, with limited data and partial observability, such as…
An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs
Qian Zhu, Xinnan Guo, Jingjing Huo +5
Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adherence to complex regulations and…
ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration
Yuying Zhang, Francesco Verdoja, Wenyan Yang +1
Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the demonstrated trajectory can…
Smoothing Slot Attention Iterations and Recurrences
Rongzhen Zhao, Wenyan Yang, Juho Kannala +1
Slot Attention (SA) lies at the heart of mainstream Object-Centric Learning (OCL). Image features can be aggregated into object-level representations by SA \textit{iteratively} ref…
Rethinking Temporal Consistency in Video Object-Centric Learning: From Prediction to Correspondence
Zhiyuan Li, Rongzhen Zhao, Wenyan Yang +3
The de facto approach in video object-centric learning maintains temporal consistency through learned dynamics modules that predict future object representations, called slots. We…