NewEvery arXiv paper, its researchers & institutions — mapped.
papers

Publications (110)

cs.CV2024

Masked Video and Body-worn IMU Autoencoder for Egocentric Action Recognition

Mingfang Zhang, Yifei Huang, Ruicong Liu +1

cs.CV2026

GROVE: Growing and Reasoning over Temporally Stratified Memory from Streaming Video Experience

Sitong Gong, Caixin Kang, Tianyu Yan +7

quant-ph2026

TRAM: A Transverse Relaxation Time-Aware Qubit Mapping Algorithm for NISQ Devices

Yifei Huang, Pascal Jahan Elahi, Ugo Varetto +3

cs.CV2026

Unleashing Spatial Reasoning in Multimodal Large Language Models via Textual Representation Guided Reasoning

Jiacheng Hua, Yishu Yin, Yuhang Wu +3

stat.ME2024

Constrained D-optimal Design for Paid Research Study

Yifei Huang, Liping Tong, Jie Yang

cs.GR2025

Sel3DCraft: Interactive Visual Prompts for User-Friendly Text-to-3D Generation

Nan Xiang, Tianyi Liang, Haiwen Huang +6

cs.RO2025

Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers

Jianxin Bi, Kelvin Lim, Kaiqi Chen +2

cs.CV2022

Ego4D: Around the World in 3,000 Hours of Egocentric Video

Kristen Grauman, Andrew Westbury, Eugene Byrne +82

cs.IT2025

Fundamental Limits of Coded Caching with Fixed Subpacketization

Minquan Cheng, Yifei Huang, Youlong Wu +1

cs.LG2021

Rethinking Breiman's Dilemma in Neural Networks: Phase Transitions of Margin Dynamics

Weizhi Zhu, Yifei Huang, Yuan Yao

cs.CV2018

Differentiable Fine-grained Quantization for Deep Neural Network Compression

Hsin-Pai Cheng, Yuanjun Huang, Xuyang Guo +4

cs.CV2025

Learning Streaming Video Representation via Multitask Training

Yibin Yan, Jilan Xu, Shangzhe Di +6

stat.CO2025

CDsampling: An R Package for Constrained D-Optimal Sampling in Paid Research Studies

Yifei Huang, Liping Tong, Jie Yang

cs.CV2024

Vinci: A Real-time Embodied Smart Assistant based on Egocentric Vision-Language Model

Yifei Huang, Jilan Xu, Baoqi Pei +15

cs.CV2025

SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-training

Nie Lin, Takehiko Ohkawa, Yifei Huang +5

cs.CL2019

An Evaluation of Transfer Learning for Classifying Sales Engagement Emails at Large Scale

Yong Liu, Pavel Dmitriev, Yifei Huang +2

cs.CV2025

EgoExoLearn: A Dataset for Bridging Asynchronous Ego- and Exo-centric View of Procedural Activities in Real World

Yifei Huang, Guo Chen, Jilan Xu +8

cs.CV2023

Structural Multiplane Image: Bridging Neural View Synthesis and 3D Reconstruction

Mingfang Zhang, Jinglu Wang, Xiao Li +3

cs.CV2022

Precise Affordance Annotation for Egocentric Action Video Datasets

Zecheng Yu, Yifei Huang, Ryosuke Furuta +3

cs.CY2020

Situation Awareness and Information Fusion in Sales and Customer Engagement: A Paradigm Shift

Yifei Huang

cs.CV2026

CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering

Mingfang Zhang, Jingjing Pan, Ashutosh Kumar +7

cs.CV2021

FACIAL: Synthesizing Dynamic Talking Face with Implicit Attribute Learning

Chenxu Zhang, Yifan Zhao, Yifei Huang +4

cs.CV2025

Joint Image-Instance Spatial-Temporal Attention for Few-shot Action Recognition

Zefeng Qian, Chongyang Zhang, Yifei Huang +2

cs.CV2025

Can MLLMs Read the Room? A Multimodal Benchmark for Verifying Truthfulness in Multi-Party Social Interactions

Caixin Kang, Yifei Huang, Liangyang Ouyang +2

cs.CV2024

Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding

Guo Chen, Yifei Huang, Jilan Xu +7

cs.CV2025

LvBench: A Benchmark for Long-form Video Understanding with Versatile Multi-modal Question Answering

Hongjie Zhang, Lu Dong, Yi Liu +4

cs.SE2026

JAMER: Project-Level Code Framework Dataset and Benchmark on Professional Game Engines

Jianwen Sun, Chuanhao Li, Zizhen Li +5

cs.PL2025

Membership Testing for Semantic Regular Expressions

Yifei Huang, Matin Amini, Alexis Le Glaunec +2

cs.HC2026

AutoBG: A Board Game Design Assistant with Interactive Ideation, Iterative Rulebook Generation, and Individualized Feedback

Zizhen Li, Chuanhao Li, Yibin Wang +6

cs.CV2026

Improving Adversarial Transferability on Vision-Language Pre-training Models via Surrogate-Specific Bias Correction

Lijia Yu, Jiuxin Cao, Yuchen Qiang +3

quant-ph2021

Feynman-path type simulation using stabilizer projector decomposition of unitaries

Yifei Huang, Peter Love

cs.CV2025

Multi-speaker Attention Alignment for Multimodal Social Interaction

Liangyang Ouyang, Yifei Huang, Mingfang Zhang +3

cs.LG2019

Discovery of Bias and Strategic Behavior in Crowdsourced Performance Assessment

Yifei Huang, Matt Shum, Xi Wu +1

quant-ph2018

Approximate stabilizer rank and improved weak simulation of Clifford-dominated circuits for qudits

Yifei Huang, Peter Love

cs.HC2026

Beyond Symbols: Motion Perception Cues Enhance Dual-Task Performance with Wearable Directional Guidance

Qing Zhang, Junyu Chen, Yifei Huang +4

cs.CV2021

Spatio-Temporal Perturbations for Video Attribution

Zhenqiang Li, Weimin Wang, Zuoyue Li +2

q-fin.ST2011

Maximum penalized quasi-likelihood estimation of the diffusion function

Jeff Hamrick, Yifei Huang, Constantinos Kardaras +1

cs.CV2026

SFHand: Learning Embodied Manipulation by Streaming Egocentric 3D Hand Forecasting

Ruicong Liu, Yifei Huang, Liangyang Ouyang +2

cs.CV2024

CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding

Guo Chen, Yicheng Liu, Yifei Huang +6

cs.CV2026

Vinci2: Providing Proactive Assistance in Continuous Egocentric Videos

Gong Sitong, Tianyu Yan, Caixin Kang +6

The paper introduces Vinci2, a proactive on‑device assistant for continuous egocentric video that decides when to intervene by using memory‑augmented reasoning, and presents EgoSer…

#egocentric video#proactive assistance#memory‑augmented agents#temporal reasoning
cs.CV2025

EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos

Jilan Xu, Yifei Huang, Baoqi Pei +6

cs.CV2021

Leveraging Human Selective Attention for Medical Image Analysis with Limited Training Data

Yifei Huang, Xiaoxiao Li, Lijin Yang +6

cs.CV2025

Learning Procedural-aware Video Representations through State-Grounded Hierarchy Unfolding

Jinghan Zhao, Yifei Huang, Feng Lu

cs.CV2026

UniLS: End-to-End Audio-Driven Avatars for Unified Listening and Speaking

Xuangeng Chu, Ruicong Liu, Yifei Huang +3

cs.CV2025

TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation

Tianyi Liang, Jiangqi Liu, Yifei Huang +4

cs.CV2026

The N-Body Problem: Parallel Execution from Single-Person Egocentric Video

Zhifan Zhu, Yifei Huang, Yoichi Sato +1

cs.CV2022

InternVideo-Ego4D: A Pack of Champion Solutions to Ego4D Challenges

Guo Chen, Sen Xing, Zhe Chen +18

cs.CV2020

Towards Visually Explaining Video Understanding Networks with Perturbation

Zhenqiang Li, Weimin Wang, Zuoyue Li +2

quant-ph2017

Semiclassical Formulation of Gottesman-Knill and Universal Quantum Computation

Lucas Kocia, Yifei Huang, Peter Love

quant-ph2022

Efficient quantum imaginary time evolution by drifting real time evolution: an approach with low gate and measurement complexity

Yifei Huang, Yuguo Shao, Weiluo Ren +2

cs.CV2021

Stacked Temporal Attention: Improving First-person Action Recognition by Emphasizing Discriminative Clips

Lijin Yang, Yifei Huang, Yusuke Sugano +1

cs.CV2026

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

Zeyu Wang, Chang Liu, Eduardus Tjitrahardja +22

cs.CV2025

Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision

Yuping He, Yifei Huang, Guo Chen +5

cs.CV2024

Retrieval-Augmented Egocentric Video Captioning

Jilan Xu, Yifei Huang, Junlin Hou +4

cs.IT2016

Mode Selection, Resource Allocation and Power Control for D2D-Enabled Two-Tier Cellular Network

Yifei Huang, Ali A. Nasir, Salman Durrani +1

cs.CV2025

An Egocentric Vision-Language Model based Portable Real-time Smart Assistant

Yifei Huang, Jilan Xu, Baoqi Pei +16

cs.CV2021

Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling

Hong Chen, Yifei Huang, Hiroya Takamura +1

cs.CV2024

Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Kristen Grauman, Andrew Westbury, Lorenzo Torresani +98

cs.CV2023

Memory-and-Anticipation Transformer for Online Action Understanding

Jiahao Wang, Guo Chen, Yifei Huang +2

cs.CV2023

VideoLLM: Modeling Video Sequence with Large Language Models

Guo Chen, Yin-Dong Zheng, Jiahao Wang +8

cs.CV2020

Mutual Context Network for Jointly Estimating Egocentric Gaze and Actions

Yifei Huang, Zhenqiang Li, Minjie Cai +1

quant-ph2025

Quantum computing quantum Monte Carlo algorithm

Yukun Zhang, Yifei Huang, Jinzhao Sun +2

cs.HC2025

Panel-by-Panel Souls: A Performative Workflow for Expressive Faces in AI-Assisted Manga Creation

Qing Zhang, Jing Huang, Yifei Huang +1

cs.CV2024

EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation

Baoqi Pei, Guo Chen, Jilan Xu +8

cs.CV2025

Egocentric Action-aware Inertial Localization in Point Clouds with Vision-Language Guidance

Mingfang Zhang, Ryo Yonetani, Yifei Huang +3

stat.CO2026

ForLion: An R Package for Finding Optimal Experimental Designs with Mixed Factors

Siting Lin, Yifei Huang, Jie Yang

econ.GN2025

Seasonality in the U.S. Housing Market: Post-Pandemic Shifts and Regional Dynamics

Yihan Hu, Yifei Huang, Weizhao Wang

cs.LG2021

Adversarial Robustness of Stabilized NeuralODEs Might be from Obfuscated Gradients

Yifei Huang, Yaodong Yu, Hongyang Zhang +2

quant-ph2017

Discrete Wigner Function Derivation of the Aaronson-Gottesman Tableau Algorithm

Lucas Kocia, Yifei Huang, Peter Love

cs.CV2023

Compound Prototype Matching for Few-shot Action Recognition

Yifei Huang, Lijin Yang, Yoichi Sato

cs.CV2019

Manipulation-skill Assessment from Videos with Spatial Attention Network

Zhenqiang Li, Yifei Huang, Minjie Cai +1

cs.CY2025

OmniScientist: Toward a Co-evolving Ecosystem of Human and AI Scientists

Chenyang Shao, Dehao Huang, Yu Li +18

cs.CV2018

Predicting Gaze in Egocentric Video by Learning Task-dependent Attention Transition

Yifei Huang, Minjie Cai, Zhenqiang Li +1

physics.chem-ph2024

GPU-accelerated Auxiliary-field quantum Monte Carlo with multi-Slater determinant trial states

Yifei Huang, Zhen Guo, Hung Q. Pham +1

cs.HC2026

World Craft: Agentic Framework to Create Visualizable Worlds via Text

Jianwen Sun, Yukang Feng, Kaining Ying +8

cs.CV2025

Can MLLMs Read the Room? A Multimodal Benchmark for Assessing Deception in Multi-Party Social Interactions

Caixin Kang, Yifei Huang, Liangyang Ouyang +3

cs.CV2023

Proposal-based Temporal Action Localization with Point-level Supervision

Yuan Yin, Yifei Huang, Ryosuke Furuta +1

cs.HC2025

Living the Novel: A System for Generating Self-Training Timeline-Aware Conversational Agents from Novels

Yifei Huang, Tianyu Yan, Sitong Gong +5

cs.HC2026

MeepleLM: A Virtual Playtester Simulating Diverse Subjective Experiences

Zizhen Li, Chuanhao Li, Yibin Wang +7

cs.CV2026

Towards Interactive Intelligence for Digital Humans

Yiyi Cai, Xuangeng Chu, Xiwei Gao +16

cs.CV2026

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation

Liangyang Ouyang, Ruicong Liu, Caixin Kang +2

cs.SE2026

HALF: Hollowing Analysis Framework for Binary Programs with Kernel Module Assistance

Zhangbo Long, Letian Sha, Jiaye Pan +4

quant-ph2025

Fault-tolerant quantum algorithms for quantum molecular systems: A survey

Yukun Zhang, Xiaoming Zhang, Jinzhao Sun +4

cs.CV2025

Beyond Label Semantics: Language-Guided Action Anatomy for Few-shot Action Recognition

Zefeng Qian, Xincheng Yao, Yifei Huang +3

quant-ph2025

Digital adiabatic evolution is universally accurate

Yangyu Lu, Yifei Huang, Dong An +3

cs.AI2026

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

Caixin Kang, Tianyu Yan, Sitong Gong +8

cs.CV2025

Weakly Supervised Temporal Sentence Grounding via Positive Sample Mining

Lu Dong, Haiyu Zhang, Hongjie Zhang +5

cs.CV2024

ActionVOS: Actions as Prompts for Video Object Segmentation

Liangyang Ouyang, Ruicong Liu, Yifei Huang +2

stat.CO2024

ForLion: A New Algorithm for D-optimal Designs under General Parametric Statistical Models with Mixed Factors

Yifei Huang, Keren Li, Abhyuday Mandal +1

cs.CV2026

Towards Multimodal Lifelong Understanding: A Dataset and Agentic Baseline

Guo Chen, Lidong Lu, Yicheng Liu +17

cs.CV2024

FineBio: A Fine-Grained Video Dataset of Biological Experiments with Hierarchical Annotation

Takuma Yagi, Misaki Ohashi, Yifei Huang +4

cs.CV2018

Semantic Aware Attention Based Deep Object Co-segmentation

Hong Chen, Yifei Huang, Hideki Nakayama

cs.RO2026

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

Yuhang Huang, Xuan Lv, Junyan Xu +25

cs.CV2025

VideoTG-R1: Boosting Video Temporal Grounding via Curriculum Reinforcement Learning on Reflected Boundary Annotations

Lu Dong, Haiyu Zhang, Han Lin +8

cs.CV2025

EgoExoBench: A Benchmark for First- and Third-person View Video Understanding in MLLMs

Yuping He, Yifei Huang, Guo Chen +4

physics.chem-ph2025

ByteQC: GPU-Accelerated Quantum Chemistry Package for Large-Scale Systems

Zhen Guo, Zigeng Huang, Qiaorui Chen +7

cs.CL2023

Pretraining Language Models with Text-Attributed Heterogeneous Graphs

Tao Zou, Le Yu, Yifei Huang +2

cs.CV2024

InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Yi Wang, Kunchang Li, Xinhao Li +17

cs.IT2026

Placement Delivery Array for Cache-Aided MIMO Systems

Yifei Huang, Kai Wan, Minquan Cheng +2

stat.ME2026

Expected Weighted D-optimal Designs for Experiments with Mixed Factors

Siting Lin, Yifei Huang, Jie Yang