most citedVQ-NeRF: Vector Quantization Enhances Implicit Neural Representations

1 citations · 1 across the 1 of their papers we have counts for

collaborators

6 papers

cs.CV2024

DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Zhende Song, Chenchen Wang, Jiamu Sheng +4

Recent large vision-language models (LVLMs) for video understanding are primarily fine-tuned with various videos scraped from online platforms. Existing datasets, such as ActivityN…

cs.CV2023

M3DBench: Let's Instruct Large Models with Multi-modal 3D Prompts

Mingsheng Li, Xin Chen, Chi Zhang +5

Recently, 3D understanding has become popular to facilitate autonomous agents to perform further decisionmaking. However, existing 3D datasets and methods are often limited to spec…

cs.CV2023

ShapeGPT: 3D Shape Generation with A Unified Multi-modal Language Model

Fukun Yin, Xin Chen, Chi Zhang +6

The advent of large language models, enabling flexibility through instruction-driven approaches, has revolutionized many traditional generative tasks, but large models for 3D data,…

cs.CV2023

LL3DA: Visual Interactive Instruction Tuning for Omni-3D Understanding, Reasoning, and Planning

Sijin Chen, Xin Chen, Chi Zhang +6

Recent advances in Large Multimodal Models (LMM) have made it possible for various applications in human-machine interactions. However, developing LMMs that can comprehend, reason,…

cs.CV2023

PDF: Point Diffusion Implicit Function for Large-scale Scene Neural Representation

Yuhan Ding, Fukun Yin, Jiayuan Fan +6

Recent advances in implicit neural representations have achieved impressive results by sampling and fusing individual points along sampling rays in the sampling space. However, due…

cs.CV20231 cited

VQ-NeRF: Vector Quantization Enhances Implicit Neural Representations

Yiying Yang, Wen Liu, Fukun Yin +4

Recent advancements in implicit neural representations have contributed to high-fidelity surface reconstruction and photorealistic novel view synthesis. However, the computational…