papers

Publications (14)

cs.RO2024

Deep Dive into Model-free Reinforcement Learning for Biological and Robotic Systems: Theory and Practice

Yusheng Jiao, Feng Ling, Sina Heydari +3

Animals and robots exist in a physical world and must coordinate their bodies to achieve behavioral objectives. With recent developments in deep reinforcement learning, it is now p…

cs.CV2025

Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model

Team Seawead, Ceyuan Yang, Zhijie Lin +52

This technical report presents a cost-efficient strategy for training a video generation foundation model. We present a mid-sized research model with approximately 7 billion parame…

cs.CV2025

Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model

Team Seedance, Heyi Chen, Siyan Chen +194

Recent strides in video generation have paved the way for unified audio-visual generation. In this work, we present Seedance 1.5 pro, a foundational model engineered specifically f…

cond-mat.soft2022

Spontaneous phase coordination and fluid pumping in model ciliary carpets

Anup Kanale, Feng Ling, Hanliang Guo +2

Ciliated tissues such as in the mammalian lungs, brains, and reproductive tracts, are specialized to pump fluid. They generate flows by the collective activity of hundreds of thous…

physics.flu-dyn2018

Instability-driven Oscillations of Elastic Microfilaments

Feng Ling, Hanliang Guo, Eva Kanso

Cilia and flagella are highly conserved slender organelles that exhibit a variety of rhythmic beating patterns from non-planar cone-like motions to planar wave-like deformations. A…

cs.LG2024

AffineQuant: Affine Transformation Quantization for Large Language Models

Yuexiao Ma, Huixia Li, Xiawu Zheng +6

The significant resource requirements associated with Large-scale Language Models (LLMs) have generated considerable interest in the development of techniques aimed at compressing…

physics.med-ph2024

Utilizing entropy to systematically quantify the resting-condition baroreflex regulation function

Bo-Yuan Li, Xiao-Yang Li, Xia Lu +3

Baroreflex is critical to maintain the blood pressure homeostasis, and the quantification of the baroreflex regulation function (BRF) can provide guidance for disease diagnosis, tr…

cs.CV2025

PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models

Tianchen Zhao, Ke Hong, Xinhao Yang +8

In visual generation, the quadratic complexity of attention mechanisms results in high memory and computational costs, especially for longer token sequences required in high-resolu…

cs.CV2026

Flow caching for autoregressive video generation

Yuexiao Ma, Xuzhe Zheng, Jing Xu +9

Autoregressive models, often built on Transformer architectures, represent a powerful paradigm for generating ultra-long videos by synthesizing content in sequential chunks. Howeve…

cs.CV2026

Seedance 2.0: Advancing Video Generation for World Complexity

Team Seedance, De Chen, Liyang Chen +168

Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro…

physics.flu-dyn2026

Cilia-driven transport in confined ducts: an active porous media model

JP Raimondi, Feng Ling, Eva Kanso

Ciliated organs transport viscous fluids through confined ducts, yet how duct morphology and ciliary activity jointly set the limits of flow rate and sustainable pressure remains u…

q-bio.QM2020

Learning to swim in potential flow

Yusheng Jiao, Feng Ling, Sina Heydari +3

Fish swim by undulating their bodies. These propulsive motions require coordinated shape changes of a body that interacts with its fluid environment, but the specific shape coordin…

eess.IV2020

Cloud detection in Landsat-8 imagery in Google Earth Engine based on a deep neural network

Zhixiang Yin, Feng Ling, Giles M. Foody +2

Google Earth Engine (GEE) provides a convenient platform for applications based on optical satellite imagery of large areas. With such data sets, the detection of cloud is often a…

cs.CV2021

A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition

Yuan Zhang, Jian Cao, Ling Zhang +4

Learning subtle representation about object parts plays a vital role in fine-grained visual recognition (FGVR) field. The vision transformer (ViT) achieves promising results on com…