Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards
Shiyu Li, Yang Tang, Yifan Wang +2
Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement learning (RL) has emerged as a po…
cs.CL2025
Finetune Once: Decoupling General & Domain Learning with Dynamic Boosted Annealing
Yang Tang, Ruijie Liu, Yifan Wang +2
Large language models (LLMs) fine-tuning shows excellent implications. However, vanilla fine-tuning methods often require intricate data mixture and repeated experiments for optima…
cs.CL2024
Conan-embedding: General Text Embedding with More and Better Negative Samples
Shiyu Li, Yang Tang, Shizhe Chen +1
With the growing popularity of RAG, the capabilities of embedding models are gaining increasing attention. Embedding models are primarily trained through contrastive loss learning,…