File size: 1,864 Bytes
c961de8
644baec
c961de8
 
 
 
4a9f080
c961de8
 
 
 
644baec
26aabf1
0c4a0fc
51d58e3
0c4a0fc
 
 
6d56f33
0c4a0fc
26aabf1
ada1198
f498ca7
ada1198
f498ca7
ada1198
f498ca7
ada1198
26aabf1
0c4a0fc
 
f498ca7
0c4a0fc
26aabf1
 
 
 
 
 
 
0c4a0fc
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
26aabf1
0c4a0fc
6d56f33
 
0c4a0fc
 
 
4a9f080
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
38
39
40
41
42
43
44
45
46
47
48
49
50
51
52
53
54
55
56
57
58
59
60
61
62
63
64
65
---
title: VoiceArena
emoji: πŸŽ™οΈ
colorFrom: indigo
colorTo: pink
sdk: streamlit
sdk_version: 1.45.1
app_file: app.py
pinned: false
---

# VoiceArena πŸŽ™οΈ 
*On a mission to make machines talk like humans.*

Welcome to the official Hugging Face organization for **Josh Talks AI** β€” a data-centric initiative focused on creating high-quality datasets.

🌐 **Learn more**: [https://ai.joshtalks.com](https://ai.joshtalks.com)

---

## 🧭 Problems with ASR
- *Conversational multi speaker speech*

- *Dialect and Accented speech*

- *Child Speech*

- *Small models for real time low latency*

## 🧭 Our Mission

We are building open datasets to power the next generation of Speech AI β€” across languages, domains, and communities.

## 🧭 Our Work

**Datasets**  - Highest quality scientifically designed datasets for training models.

**Benchmarks** - We can only fix what we can measure. Our benchmarks tell researchers exactly where their models fail.


## πŸ“¦ Coming Soon

We are currently preparing a series of open datasets, including:

- **Multilingual speech datasets**  
  Transcribed Indian language audio from real-world conversations and talks.

- **Cultural & regional text corpora**  
  Clean, annotated text datasets in Hindi, Tamil, Bengali, Marathi, and more.

- **Media-rich data from Indian contexts**  
  Video metadata, subtitles, and emotion-rich labels for AI-driven storytelling.

Stay tuned β€” our first releases are just around the corner!

## 🀝 Let's Collaborate

We welcome partnerships with researchers, universities, NGOs, and AI developers working on ASR challenges. If you're interested in contributing or using our datasets, reach out below.

---

## πŸ“¬ Contact

🌐 **Website**: [https://ai.joshtalks.com](https://ai.joshtalks.com)  
πŸ’Ό **LinkedIn**: [Josh Talks](https://www.linkedin.com/company/joshtalks/)