◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Quang Minh Dinh

Simon Fraser University

3 papers hereh-index 237 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CV2
  • cs.CL1
affiliations
  • Simon Fraser University
HomepageORCID 0009-0003-9205-6270

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators

3 papers

cs.CV2026

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting

Quang Minh Dinh, Tuan Kiet Doan

Generative traffic video forecasting aims to synthesize long-horizon, temporally coherent future videos of traffic scenes from a short observation history and textual descriptions.…

cs.CL2025

BERSting at the Screams: A Benchmark for Distanced, Emotional and Shouted Speech Recognition

Paige Tuttösí, Mantaj Dhillon, Luna Sang +6

Some speech recognition tasks, such as automatic speech recognition (ASR), are approaching or have reached human performance in many reported metrics. Yet, they continue to struggl…

cs.CV2024

TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning

Quang Minh Dinh, Minh Khoi Ho, Anh Quan Dang +1

Traffic video description and analysis have received much attention recently due to the growing demand for efficient and reliable urban surveillance systems. Most existing methods…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.