RuleMaze / README.md
nielsr's picture
nielsr HF Staff
Add model card
7e8777a verified
|
Raw
History Blame
629 Bytes
metadata
library_name: transformers
pipeline_tag: image-text-to-text

This repository contains the RuleMaze model checkpoint, a LoRA-fine-tuned version of Qwen2.5-VL-3B for rule-compliant visual spatial planning in multimodal large language models. The model is introduced in the paper RuleMaze: Rule-Compliant Visual Spatial Planning for Multimodal Large Language Models.

For more details about the benchmark, dataset, and training pipeline, please refer to the project page and the GitHub repository.