Buckets:
66 GB
12 files
Updated 6 days ago
Ctrl+K
| Name | Size | Uploaded | Xet hash |
|---|---|---|---|
| images | 1 items | ||
| .gitattributes | 2.6 kB xet | d665d16f | |
| README.md | 1.19 kB xet | 5d64f378 | |
| conversations.json | 3.36 GB xet | 9f63706d | |
| eval.jsonl | 1.43 MB xet | c94c0183 | |
| eval_images.zip | 669 MB xet | a598fd5e | |
| images.z01 | 10.7 GB xet | c988447c | |
| images.z02 | 10.7 GB xet | c0dbe332 | |
| images.z03 | 10.7 GB xet | 13f1307f | |
| images.z04 | 10.7 GB xet | 597a649e | |
| images.z05 | 10.7 GB xet | c473386a | |
| images.zip | 8.33 GB xet | cec08736 |
VideoGameBunny Instruction Following Dataset
[Website]
Overview
We present a comprehensive dataset of 185,259 high-resolution images from 413 video games, sourced from YouTube videos. This dataset addresses the lack of game-specific instruction-following data and aims to improve the ability of open-source models to understand and respond to video game content.
Dataset Composition
Our dataset includes various types of instructions generated for these images using different large multimodal models:
- Short captions
- Long captions
- Image-to-JSON conversions
- Image-based question-answering pairs
Dataset Statistics
| Task | Generator | Samples |
|---|---|---|
| Short Captions | Gemini-1.0-Pro-Vision | 70,673 |
| Long Captions | GPT-4V | 70,799 |
| Image-to-JSON | Gemini-1.5-Pro | 136,974 |
| Question Answering | Llama-3, GPT-4o | 81,122 |
- Total size
- 66 GB
- Files
- 12
- Last updated
- Aug 2
- Pre-warmed CDN
- US EU US EU
