Publications (9)
Automatic Calibration of Mesoscopic Traffic Simulation Using Vehicle Trajectory Data
Ran Sun, Zihao Wang, Xingmin Wang +1
Traffic simulation models have long been popular in modern traffic planning and operation applications. Efficient calibration of simulation models is usually a crucial step in a si…
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
Guoqing Ma, Haoyang Huang, Kun Yan +112
We present Step-Video-T2V, a state-of-the-art text-to-video pre-trained model with 30B parameters and the ability to generate videos up to 204 frames in length. A deep compression…
OmniMatBench: A Human-Calibrated Multimodal Reasoning Benchmark Across 19 Materials Science Subfields
Wanhao Liu, Jiaqing Xie, Qian Tan +10
As multimodal language models play an increasingly important role in scientific research, materials science offers a critical testbed due to its interdisciplinary, multimodal, and…
Broadband planar electromagnetic hyper-lens with uniform magnification in air
Ran Sun, Fei Sun, Hanchuan Chen +2
A planar hyper-lens, capable of creating sub-wavelength imaging for broadband electromagnetic wave, is designed based on electromagnetic null medium. Subsequently, a scheme for the…
Step-Video-TI2V Technical Report: A State-of-the-Art Text-Driven Image-to-Video Generation Model
Haoyang Huang, Guoqing Ma, Nan Duan +51
We present Step-Video-TI2V, a state-of-the-art text-driven image-to-video generation model with 30B parameters, capable of generating videos up to 102 frames based on both text and…
Layer-Guided UAV Tracking: Enhancing Efficiency and Occlusion Robustness
Yang Zhou, Derui Ding, Ran Sun +2
Visual object tracking (VOT) plays a pivotal role in unmanned aerial vehicle (UAV) applications. Addressing the trade-off between accuracy and efficiency, especially under challeng…