1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.CL2025
MossNet: Mixture of State-Space Experts is a Multi-Head Attention
Shikhar Tuli, James Seale Smith, Haris Jeelani +5
Large language models (LLMs) have significantly advanced generative applications in natural language processing (NLP). Recent trends in model architectures revolve around efficient…
cs.LG2024
MoDeGPT: Modular Decomposition for Large Language Model Compression
Chi-Heng Lin, Shangqian Gao, James Seale Smith +5
Large Language Models (LLMs) have reshaped the landscape of artificial intelligence by demonstrating exceptional performance across various tasks. However, substantial computationa…
cs.LG2018★ 1 cited
A New Concept of Deep Reinforcement Learning based Augmented General Sequence Tagging System
Yu Wang, Abhishek Patel, Hongxia Jin
In this paper, a new deep reinforcement learning based augmented general sequence tagging system is proposed. The new system contains two parts: a deep neural network (DNN) based s…