4 papers
To Theoretically Understand Transformer-Based In-Context Learning for Optimizing CSMA
Shugang Hao, Hongbo Li, Lingjie Duan
The binary exponential backoff scheme is widely used in WiFi 7 and still incurs poor throughput performance under dynamic channel environments. Recent model-based approaches (e.g.,…
Online Learning from Strategic Human Feedback in LLM Fine-Tuning
Shugang Hao, Lingjie Duan
Reinforcement learning from human feedback (RLHF) has become an essential step in fine-tuning large language models (LLMs) to align them with human preferences. However, human labe…
Algorithm Design for Continual Learning in IoT Networks
Shugang Hao, Lingjie Duan
Continual learning (CL) is a new online learning technique over sequentially generated streaming data from different tasks, aiming to maintain a small forgetting loss on previously…
Regulating Competition in Age of Information under Network Externalities
Shugang Hao, Lingjie Duan
Online content platforms are concerned about the freshness of their content updates to their end customers, and increasingly more platforms now invite and pay the crowd to sample r…