arxiv:2608.07693
Quang Minh Dinh
minhdinh101202
AI & ML interests
None yet
Recent Activity
authored a paper about 19 hours ago
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning authored a paper about 19 hours ago
CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting authored a paper about 19 hours ago
BERSting at the Screams: A Benchmark for Distanced, Emotional and
Shouted Speech RecognitionOrganizations
None yet