Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
aryanvibhosale
/
vibe
like
3
Text-to-Audio
Safetensors
4 datasets
English
vibe
music
audio
video
multimodal
music-generation
audio-generation
video-to-music
text-to-music
video-to-audio
text-video-to-audio
flow-matching
diffusion
reinforcement-learning
rlhf
grpo
minicpm
arxiv:
2608.30125
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
vibe
3.52 GB
Ctrl+K
Ctrl+K
1 contributor
History:
10 commits
aryanvibhosale
Add arXiv badge (2608.30125); EMNLP badge now reads Findings
730fc5a
verified
about 24 hours ago
static
Delete files static/vibe_banner_live.gif with huggingface_hub
1 day ago
.gitattributes
1.84 kB
New model card: logo header, EMNLP Findings, accurate architecture and curriculum figures, banner video, dataset links, CMI-RM attribution
1 day ago
README.md
12.3 kB
Add arXiv badge (2608.30125); EMNLP badge now reads Findings
about 24 hours ago
config.json
3.21 kB
VIBE video-to-music checkpoint (Stage-5 RL, LoRA folded)
2 days ago
model.safetensors
3.51 GB
xet
VIBE video-to-music checkpoint (Stage-5 RL, LoRA folded)
2 days ago
special_tokens_map.json
Safe
1.63 kB
VIBE video-to-music checkpoint (Stage-5 RL, LoRA folded)
2 days ago
tokenizer.json
Safe
3.68 MB
VIBE video-to-music checkpoint (Stage-5 RL, LoRA folded)
2 days ago
tokenizer_config.json
Safe
4.95 kB
VIBE video-to-music checkpoint (Stage-5 RL, LoRA folded)
2 days ago