Japanese SFT/DPO data convert to speech via TTS. And audio caption data generated by Qwen3-Omni. All datasets are available for commercial use.
Ayuto Tsutsumi
Atotti
AI & ML interests
None yet
Recent Activity
upvoted an article 2 days ago
YODAS v3: A 1 Million Hour Dataset for the Next Generation of Open Voice AI Research liked a dataset 2 days ago
espnet/yodas3 updated a model 9 days ago
voicist/qwen3-ctd