Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Broken Words, Broken Performance: Effect of Tokenization on Performance of LLMs
Sachin Pawar, Manoj Apte, Kshitij Jadhav +2
Tokenization is the first step in training any Large Language Model (LLM), where the text is split into a sequence of tokens as per the model's fixed vocabulary. This tokenization…
cs.CL2025
DRAssist: Dispute Resolution Assistance using Large Language Models
Sachin Pawar, Manoj Apte, Girish K. Palshikar +2
Disputes between two parties occur in almost all domains such as taxation, insurance, banking, healthcare, etc. The disputes are generally resolved in a specific forum (e.g., consu…