works on

From the 2 of 61 linked papers with an AI index.

activity
20242026
collaborators

61 papers

cs.IR2026

A Dual-Expert Strategy Integrating LLMs to Mitigate Negative Transfer in Cross-Domain Sequential Recommendation

Hyeongjun Yun, Kihyuk Song, Jaegul Choo +1

Cross-Domain Sequential Recommendation (CDSR) predicts the next item a user will interact with based on their historical interaction sequences across multiple domains. Recent appro…

cs.LG2026

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control

Donghu Kim, Youngdo Lee, Hojoon Lee +6

Improving sample efficiency remains a core challenge in reinforcement learning (RL), especially in real-world settings like robotics, where data collection is costly. This challeng…

cs.CL2026

EditPPT: Faithful Long-Deck Slide Editing via Structured Tool-Using Multi-Agent with Dual-Modal Validators

Jiheon Kim, Kyudan Jung, Jaegul Choo

Automating slide editing requires simultaneously satisfying modification accuracy, preservation fidelity, and robustness to deck length. Existing LLM-based systems often fail on re…

cs.CV2026

ReDesign: Recovering Editable Design Structures from Images via Agentic Decomposition

Jooyeol Yun, Jintae Park, Hyesu Lim +3

Recovering an editable design file from a raster image is a common and costly bottleneck in modern design workflows, yet remains challenging since editability depends on recovering…

cs.LG2026

Accelerating Masked Diffusion Large Language Models: A Survey of Efficient Inference Techniques

Daehoon Gwak, Minhyung Lee, Junwoo Park +1

The paper surveys methods for speeding up inference of masked diffusion large language models by categorizing algorithmic, architectural, and system-level acceleration techniques a…

cs.RO2026

See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models

Byungkun Lee, Dongyoon Hwang, Dongjin Kim +3

The paper proposes robot-centric pointmaps, which encode 3D scene coordinates in the robot's frame as image pixels, enabling vision‑language‑action models to align visual inputs wi…