arxiv:2510.13251
🤝 Open to Collab
Minji Kim
byminji
AI & ML interests
Video Understanding, Vision and Language Models, Multimodal Large Langauge Models
Recent Activity
updated a model about 15 hours ago
byminji/DBTrimKV-Qwen3-VL-8B-Thinking-full-mixture-frames768 published a model about 15 hours ago
byminji/DBTrimKV-Qwen3-VL-8B-Thinking-full-mixture-frames768 upvoted a paper 5 months ago
Grounding World Simulation Models in a Real-World Metropolis