| --- |
| title: README |
| emoji: 🦀 |
| colorFrom: pink |
| colorTo: purple |
| sdk: static |
| pinned: false |
| --- |
| # IngCrowd |
|
|
| **Website:** [日本語](https://www.ingcrowd.jp/) | [English](https://www.ingcrowd.jp/en) |
|
|
| ## About |
|
|
| ### English |
|
|
| IngCrowd provides human-centered data services for AI development and academic research in Japan. |
|
|
| We support Japanese participant recruitment, speech and dialogue recording, image and text data collection, transcription, ELAN annotation, classification, and human evaluation. |
|
|
| From small pilot studies to large-scale projects, our dedicated project directors manage workflow design, participant recruitment, quality control, and final delivery. |
|
|
| We also publish selected open datasets and research resources on Hugging Face to demonstrate our data collection and annotation capabilities. |
|
|
| ### 日本語 |
|
|
| 株式会社イングクラウドは、AI開発や学術研究に必要な、人の参加・判断・作業を伴うデータ支援を提供しています。 |
|
|
| 研究参加者や作業者の募集、日本語の音声・対話収録、画像・テキストデータ収集、文字起こし、ELANアノテーション、分類、人手評価などに対応しています。 |
|
|
| 小規模な試験実施から大規模なプロジェクトまで、実施方法の設計、人材の募集・教育、進捗管理、品質確認、納品を専任ディレクターが一貫して担当します。 |
|
|
| 本ページでは、当社のデータ収集・アノテーション技術の一例として、公開可能なデータセットや研究リソースを公開しています。 |
|
|
| --- |
|
|
| ## What We Publish / 公開内容 |
|
|
| We publish: |
|
|
| - Japanese speech corpora |
| - Spoken dialogue datasets |
| - Time-aligned transcripts and annotations using ELAN |
| - Research resources |
|
|
| 公開している主なコンテンツ |
|
|
| - 日本語音声コーパス |
| - 日本語対話データセット |
| - 人手で作成したELANアノテーション |
| - 研究用リソース |
|
|
| --- |
|
|
| ## Research Areas / 研究分野 |
|
|
| - Speech Recognition (ASR) |
| - Spoken Dialogue Systems |
| - Natural Language Processing (NLP) |
| - Multimodal AI |
| - Corpus Linguistics |
| - Speech Corpus Construction |
|
|
| --- |
|
|
| ## Contact / お問い合わせ |
|
|
| - [English Website](https://www.ingcrowd.jp/en) |
| - [日本語ウェブサイト](https://www.ingcrowd.jp/) |
| - [Email / メール](mailto:aidatasets@ingcrowd.jp) |