Publications (14)
Deep Dive into Model-free Reinforcement Learning for Biological and Robotic Systems: Theory and Practice
Yusheng Jiao, Feng Ling, Sina Heydari +3
Animals and robots exist in a physical world and must coordinate their bodies to achieve behavioral objectives. With recent developments in deep reinforcement learning, it is now p…
Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model
Team Seawead, Ceyuan Yang, Zhijie Lin +52
This technical report presents a cost-efficient strategy for training a video generation foundation model. We present a mid-sized research model with approximately 7 billion parame…
Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model
Team Seedance, Heyi Chen, Siyan Chen +194
Recent strides in video generation have paved the way for unified audio-visual generation. In this work, we present Seedance 1.5 pro, a foundational model engineered specifically f…
Spontaneous phase coordination and fluid pumping in model ciliary carpets
Anup Kanale, Feng Ling, Hanliang Guo +2
Ciliated tissues such as in the mammalian lungs, brains, and reproductive tracts, are specialized to pump fluid. They generate flows by the collective activity of hundreds of thous…
Instability-driven Oscillations of Elastic Microfilaments
Feng Ling, Hanliang Guo, Eva Kanso
Cilia and flagella are highly conserved slender organelles that exhibit a variety of rhythmic beating patterns from non-planar cone-like motions to planar wave-like deformations. A…
AffineQuant: Affine Transformation Quantization for Large Language Models
Yuexiao Ma, Huixia Li, Xiawu Zheng +6
The significant resource requirements associated with Large-scale Language Models (LLMs) have generated considerable interest in the development of techniques aimed at compressing…
Utilizing entropy to systematically quantify the resting-condition baroreflex regulation function
Bo-Yuan Li, Xiao-Yang Li, Xia Lu +3
Baroreflex is critical to maintain the blood pressure homeostasis, and the quantification of the baroreflex regulation function (BRF) can provide guidance for disease diagnosis, tr…
PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models
Tianchen Zhao, Ke Hong, Xinhao Yang +8
In visual generation, the quadratic complexity of attention mechanisms results in high memory and computational costs, especially for longer token sequences required in high-resolu…
Flow caching for autoregressive video generation
Yuexiao Ma, Xuzhe Zheng, Jing Xu +9
Autoregressive models, often built on Transformer architectures, represent a powerful paradigm for generating ultra-long videos by synthesizing content in sequential chunks. Howeve…
Seedance 2.0: Advancing Video Generation for World Complexity
Team Seedance, De Chen, Liyang Chen +168
Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro…
Cilia-driven transport in confined ducts: an active porous media model
JP Raimondi, Feng Ling, Eva Kanso
Ciliated organs transport viscous fluids through confined ducts, yet how duct morphology and ciliary activity jointly set the limits of flow rate and sustainable pressure remains u…
Learning to swim in potential flow
Yusheng Jiao, Feng Ling, Sina Heydari +3
Fish swim by undulating their bodies. These propulsive motions require coordinated shape changes of a body that interacts with its fluid environment, but the specific shape coordin…
Cloud detection in Landsat-8 imagery in Google Earth Engine based on a deep neural network
Zhixiang Yin, Feng Ling, Giles M. Foody +2
Google Earth Engine (GEE) provides a convenient platform for applications based on optical satellite imagery of large areas. With such data sets, the detection of cloud is often a…
A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition
Yuan Zhang, Jian Cao, Ling Zhang +4
Learning subtle representation about object parts plays a vital role in fine-grained visual recognition (FGVR) field. The vision transformer (ViT) achieves promising results on com…