2 citations · 2 across the 4 of their papers we have counts for
4 papers
Enhancing Scientific Visual Question Answering via Vision-Caption aware Supervised Fine-Tuning
Janak Kapuriya, Anwar Shaikh, Arnav Goel +8
In this study, we introduce Vision-Caption aware Supervised FineTuning (VCASFT), a novel learning paradigm designed to enhance the performance of smaller Vision Language Models(VLM…
Spiritual-LLM : Gita Inspired Mental Health Therapy In the Era of LLMs
Janak Kapuriya, Aman Singh, Jainendra Shukla +1
Traditional mental health support systems often generate responses based solely on the user's current emotion and situations, resulting in superficial interventions that fail to ad…
Exploring the Role of Diversity in Example Selection for In-Context Learning
Janak Kapuriya, Manit Kaushik, Debasis Ganguly +1
In-Context Learning (ICL) has gained prominence due to its ability to perform tasks without requiring extensive training data and its robustness to noisy labels. A typical ICL work…
MM-PhyQA: Multimodal Physics Question-Answering With Multi-Image CoT Prompting
Avinash Anand, Janak Kapuriya, Apoorv Singh +5
While Large Language Models (LLMs) can achieve human-level performance in various tasks, they continue to face challenges when it comes to effectively tackling multi-step physics r…