activity
20232026
most citedPrinciple-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision

63 citations · 113 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2026

Better Call CineCrew: Consistent Ultra-Long Narrative-to-Film Generation

Jiaben Chen, Sixun Dong, Qinhong Zhou +4

Long-form narrative-to-film generation requires shot-level controllability and cross-clip consistency in both visual identity and character behavior-requirements that remain diffic…

cs.CV2026

Sentinel: Embodied Cooperative Spatial Reasoning and Planning

Xiangye Lin, Hongxin Zhang, Ruxi Deng +2

In this work, we study Cooperative Spatial Intelligence, the ability of decentralized embodied agents to coordinate effectively under dynamic environmental constraints across city-…

cs.CV2025

Virtual Community: An Open World for Humans, Robots, and Society

Qinhong Zhou, Hongxin Zhang, Xiangye Lin +16

The rapid progress in AI and Robotics may lead to a profound societal transformation, as humans and robots begin to coexist within shared communities, introducing both opportunitie…

cs.CV2025

Ella: Embodied Social Agents with Lifelong Memory

Hongxin Zhang, Zheyuan Zhang, Zeyuan Wang +4

We introduce Ella, an embodied social agent capable of lifelong learning within a community in a 3D open world, where agents accumulate experiences and acquire knowledge through ev…

cs.CV2024★ 2 cited

HAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments

Qinhong Zhou, Sunli Chen, Yisong Wang +6

Recent advances in high-fidelity virtual environments serve as one of the major driving forces for building intelligent embodied agents to perceive, reason and interact with the ph…

cs.CV2023★ 8 cited

See, Think, Confirm: Interactive Prompting Between Vision and Language Models for Knowledge-based Visual Reasoning

Zhenfang Chen, Qinhong Zhou, Yikang Shen +3

Large pre-trained vision and language models have demonstrated remarkable capacities for various tasks. However, solving the knowledge-based visual reasoning tasks remains challeng…