3 citations · 4 across the 3 of their papers we have counts for
3 papers
ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks
Saurabh Jha, Rohan Arora, Yuji Watanabe +40
Realizing the vision of using AI agents to automate critical IT tasks depends on the ability to measure and understand effectiveness of proposed solutions. We introduce ITBench, a…
CodeSift: An LLM-Based Reference-Less Framework for Automatic Code Validation
Pooja Aggarwal, Oishik Chatterjee, Ting Dai +4
The advent of large language models (LLMs) has greatly facilitated code generation, but ensuring the functional correctness of generated code remains a challenge. Traditional valid…
Incorporating Customer Reviews in Size and Fit Recommendation systems for Fashion E-Commerce
Oishik Chatterjee, Jaidam Ram Tej, Narendra Varma Dasaraju
With the huge growth in e-commerce domain, product recommendations have become an increasing field of interest amongst e-commerce companies. One of the more difficult tasks in prod…