Quang Minh Dinh
minhdinh101202
AI & ML interests
None yet
Recent Activity
authored a paper 2 days ago
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning authored a paper 2 days ago
CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting authored a paper 2 days ago
BERSting at the Screams: A Benchmark for Distanced, Emotional and
Shouted Speech RecognitionOrganizations
None yet