2 citations · 2 across the 1 of their papers we have counts for
4 papers
UniRank: End-to-End Domain-Specific Reranking of Hybrid Text-Image Candidates
Yupei Yang, Lin Yang, Wanxi Deng +3
Reranking is a critical component in many information retrieval pipelines. Despite remarkable progress in text-only settings, multimodal reranking remains challenging, particularly…
Factored Causal Representation Learning for Robust Reward Modeling in RLHF
Yupei Yang, Lin Yang, Wanxi Deng +5
A reliable reward model is essential for aligning large language models with human preferences through reinforcement learning from human feedback. However, standard reward models a…
Hallucination-Resistant Relation Extraction via Dependency-Aware Sentence Simplification and Two-tiered Hierarchical Refinement
Yupei Yang, Fan Feng, Lin Yang +5
Relation extraction (RE) enables the construction of structured knowledge for many downstream applications. While large language models (LLMs) have shown great promise in this task…
IROS 2019 Lifelong Robotic Vision Challenge -- Lifelong Object Recognition Report
Qi She, Fan Feng, Qi Liu +33
This report summarizes IROS 2019-Lifelong Robotic Vision Competition (Lifelong Object Recognition Challenge) with methods and results from the top finalists (out of over~…