6 papers
RAID-0e: A Resilient Striping Array Architecture for Balanced Performance and Availability
Yanzhao Jia, Zhaobo Wu, Zheyi Cao +3
This paper introduces a novel disk array architecture, designated RAID-0e (Resilient Striping Array), designed to superimpose a low-overhead fault tolerance layer upon traditional…
OpenGrok: Enhancing SNS Data Processing with Distilled Knowledge and Mask-like Mechanisms
Lumen AI, Zaozhuang No. 28 Middle School, Shihao Ji +6
This report details Lumen Labs' novel approach to processing Social Networking Service (SNS) data. We leverage knowledge distillation, specifically a simple distillation method ins…
Enhancing Large Language Model Efficiencyvia Symbolic Compression: A Formal Approach Towards Interpretability
Lumen AI, Tengzhou No. 1 Middle School, Shihao Ji +6
Large language models (LLMs) face significant token efficiency bottlenecks in code generation and logical reasoning tasks, a challenge that directly impacts inference cost and mode…
Chinese Stock Prediction Based on a Multi-Modal Transformer Framework: Macro-Micro Information Fusion
Lumen AI, Tengzhou No. 1 Middle School, Shihao Ji +6
This paper proposes an innovative Multi-Modal Transformer framework (MMF-Trans) designed to significantly improve the prediction accuracy of the Chinese stock market by integrating…
Transformer^-1: Input-Adaptive Computation for Resource-Constrained Deployment
Lumen AI, Tengzhou No. 1 Middle School, Shihao Ji +6
Addressing the resource waste caused by fixed computation paradigms in deep learning models under dynamic scenarios, this paper proposes a Transformer architecture based on…
MyGO Multiplex CoT: A Method for Self-Reflection in Large Language Models via Double Chain of Thought Thinking
Shihao Ji, Zihui Song, Fucheng Zhong +4
Recent advancements in large language models (LLMs) have demonstrated their impressive abilities in various reasoning and decision-making tasks. However, the quality and coherence…