Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

SaruLab Speech group

Team
university
https://www.sp.ipc.i.u-tokyo.ac.jp/
sarulab-speech
Activity Feed

AI & ML interests

Statistical speech synthesis

Recent Activity

WataruΒ  new activity about 2 months ago
sarulab-speech/DialogueSidon:Question about released DialogueSidon weights and SSL-VAE encoder
WataruΒ  updated a model about 2 months ago
sarulab-speech/DialogueSidon
WataruΒ  published a dataset 3 months ago
sarulab-speech/DuplexChat
View all activity

Wataru-Nakata's profile picture Yuki Saito's profile picture Dong YANG's profile picture Shinnosuke Takamichi's profile picture Joon Park's profile picture Kentaro Seki's profile picture Yuki Okamoto's profile picture Kazuki Yamauchi's profile picture

sarulab-speech 's Spaces 6

Running on Zero
Agents
37

sidon_demo_beta

πŸ‹

Speech restoration demo of Sidon.

May 14
Running on Zero
Agents
14

DialogueSidon Demo

πŸ”₯

Separate two speakers from an audio or video recording

Apr 5
Build error
Agents
11

CoCoCap Beta

🐠

Transcribe Japanese audio to text

Mar 23
Running on Zero
Agents
6

MSR UTMOS

🐒

Multiple sampling rate MOS prediction with SFI conv

Jul 19, 2025
Configuration error
Agents
13

UTMOSv2

πŸŒ–

Generate speech quality score from audio

Aug 1, 2024
Build error
Agents
18

UTMOS Demo

🐒

Evaluate audio quality with MOS score

Dec 19, 2023
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs