Showing 2024Show all
3 papers · 1 filter
cs.CL2024
Efficient Hybrid Inference for LLMs: Reward-Based Token Modelling with Selective Cloud Assistance
Adarsh MS, Jithin VG, Ditto PS
Large language models (LLMs) are known for their exceptional performance across a range of natural language processing tasks, but their deployment comes at a high computational and…
cs.DC2024
Inference Acceleration for Large Language Models on CPUs
Ditto PS, Jithin VG, Adarsh MS
In recent years, large language models have demonstrated remarkable performance across various natural language processing (NLP) tasks. However, deploying these models for real-wor…
cs.CL2024
Intellecta Cognitiva: A Comprehensive Dataset for Advancing Academic Knowledge and Machine Reasoning
Ajmal PS, Ditto PS, Jithin VG
Intellecta dataset emerges as an innovative synthetic dataset, engineered to enhance the cognitive processing capabilities of contemporary language models. With a composition of 11…