3 papers
cs.CL2026
ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues
Dinesh Khandelwal, Gnana Prakash Punnavajhala, GPS Bhargav +4
Multi-turn tool calling is essential for LLMs to function as autonomous agents, yet synthesizing the training data required for these capabilities remains a fundamental challenge.…
cs.CL2025
Systematic Knowledge Injection into Large Language Models via Diverse Augmentation for Domain-Specific RAG
Kushagra Bhushan, Yatin Nandwani, Dinesh Khandelwal +4
Retrieval-Augmented Generation (RAG) has emerged as a prominent method for incorporating domain knowledge into Large Language Models (LLMs). While RAG enhances response relevance b…
cs.CL2025
Selective Self-to-Supervised Fine-Tuning for Generalization in Large Language Models
Sonam Gupta, Yatin Nandwani, Asaf Yehudai +3
Fine-tuning Large Language Models (LLMs) on specific datasets is a common practice to improve performance on target tasks. However, this performance gain often leads to overfitting…