2 citations · 3 across the 2 of their papers we have counts for
3 papers
cs.CL2025★ 2 cited
MiMo-Audio: Audio Language Models are Few-Shot Learners
Core Team, Dong Zhang, Gang Wang +97
Existing audio language models typically rely on task-specific fine-tuning to accomplish particular audio tasks. In contrast, humans are able to generalize to new audio tasks with…
cs.CL2025
Code Aesthetics with Agentic Reward Feedback
Bang Xiao, Lingjie Jiang, Shaohan Huang +5
Large Language Models (LLMs) have become valuable assistants for developers in code-related tasks. While LLMs excel at traditional programming tasks such as code generation and bug…
cs.CV2024★ 1 cited
Token Pruning for Caching Better: 9 Times Acceleration on Stable Diffusion for Free
Evelyn Zhang, Bang Xiao, Jiayi Tang +5
Stable Diffusion has achieved remarkable success in the field of text-to-image generation, with its powerful generative capabilities and diverse generation results making a lasting…