4 papers
CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts
Shanu Kumar, Shubhanshu Khandelwal, Akhila Yesantarao Venkata +3
Prompts tuned for accuracy often grow long, raising inference cost on every model call. The best accuracy-cost trade-off depends on the task and the budget, so prompt optimization…
Read the Trace, Steer the Path: Trajectory-Aware Reinforcement Learning for Diffusion Language Models
Anant Khandelwal, Manish Gupta
Diffusion large language models (dLLMs) generate responses by iteratively unmasking and revising many positions in parallel. This process leaves a rich denoising trace depicting wh…
TripTide: A Benchmark for Adaptive Travel Planning under Disruptions
Priyanshu Karmakar, Soumyabrata Chaudhuri, Shubhojit Mallick +3
Recent efforts like TripCraft and TravelPlanner have advanced the use of Large Language Models ( LLMs) for personalized, constraint aware travel itinerary generation. Yet, real tra…
TripCraft: A Benchmark for Spatio-Temporally Fine Grained Travel Planning
Soumyabrata Chaudhuri, Pranav Purkar, Ritwik Raghav +4
Recent advancements in probing Large Language Models (LLMs) have explored their latent potential as personalized travel planning agents, yet existing benchmarks remain limited in r…