3 papers
cs.AI2025
SI-Agent: An Agentic Framework for Feedback-Driven Generation and Tuning of Human-Readable System Instructions for Large Language Models
Jeshwanth Challagundla, Mantek Singh, Siddharth Raina +3
System Instructions (SIs), or system prompts, are pivotal for guiding Large Language Models (LLMs) but manual crafting is resource-intensive and often suboptimal. Existing automate…
cs.LG2024
Making Sigmoid-MSE Great Again: Output Reset Challenges Softmax Cross-Entropy in Neural Network Classification
Kanishka Tyagi, Chinmay Rane, Ketaki Vaidya +3
This study presents a comparative analysis of two objective functions, Mean Squared Error (MSE) and Softmax Cross-Entropy (SCE) for neural network classification tasks. While SCE c…
cs.LG2024
Adaptive multiple optimal learning factors for neural network training
Jeshwanth Challagundla
This thesis presents a novel approach to neural network training that addresses the challenge of determining the optimal number of learning factors. The proposed Adaptive Multiple…