pangkaiyu commited on
Commit
ed56617
·
verified ·
1 Parent(s): 0f2d051

Add files using upload-large-folder tool

Browse files
This view is limited to 50 files because it contains too many changes.   See raw diff
Files changed (50) hide show
  1. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl +0 -0
  2. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_default_performance.json +17 -0
  3. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_wer_details.jsonl +0 -0
  4. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank0.log +12 -0
  5. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank1.log +4 -0
  6. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank2.log +4 -0
  7. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank3.log +4 -0
  8. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank4.log +4 -0
  9. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank5.log +4 -0
  10. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank6.log +4 -0
  11. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank7.log +4 -0
  12. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank0.log +6 -0
  13. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank1.log +2 -0
  14. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank2.log +2 -0
  15. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank3.log +2 -0
  16. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank4.log +2 -0
  17. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank5.log +2 -0
  18. different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank7.log +2 -0
  19. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl +0 -0
  20. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_default_performance.json +17 -0
  21. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_wer_details.jsonl +0 -0
  22. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank0.log +9 -0
  23. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank1.log +4 -0
  24. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank2.log +4 -0
  25. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank3.log +4 -0
  26. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank4.log +4 -0
  27. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank5.log +4 -0
  28. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank7.log +4 -0
  29. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus.jsonl +0 -0
  30. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus_default_performance.json +121 -0
  31. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus_wer_details.jsonl +0 -0
  32. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank0.log +4 -0
  33. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank1.log +2 -0
  34. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank2.log +2 -0
  35. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank3.log +2 -0
  36. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank4.log +2 -0
  37. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank5.log +2 -0
  38. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank6.log +2 -0
  39. different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank7.log +2 -0
  40. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/Qwen2.5-Omni-7B-encoder_LibriSpeech.jsonl +0 -0
  41. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank0.log +7 -0
  42. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank1.log +4 -0
  43. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank2.log +4 -0
  44. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank3.log +4 -0
  45. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank4.log +4 -0
  46. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank5.log +4 -0
  47. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank6.log +4 -0
  48. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank7.log +4 -0
  49. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/noizeus/Qwen2.5-Omni-7B-encoder_noizeus.jsonl +0 -0
  50. different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/noizeus/logs/rank0.log +4 -0
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_default_performance.json ADDED
@@ -0,0 +1,17 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "task": "ASR",
3
+ "dataset": "LibriSpeech",
4
+ "model": "Qwen2.5-Omni-7B-all",
5
+ "date": "2025-12-15 12:02:26.247561",
6
+ "performance": {
7
+ "test_clean": {
8
+ "wer": 2.4,
9
+ "total": 2620
10
+ },
11
+ "test_other": {
12
+ "wer": 4.47,
13
+ "total": 2939
14
+ }
15
+ },
16
+ "eval_method": "qwen2-audio-impl"
17
+ }
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_wer_details.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank0.log ADDED
@@ -0,0 +1,12 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ 2025-12-15 11:39:13 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:13 | INFO | Msg example: {'index': 0, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0003.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:13 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
5
+ 2025-12-15 11:48:39 | INFO | waiting for other ranks to finish, time elapsed: 10s
6
+ 2025-12-15 11:48:49 | INFO | waiting for other ranks to finish, time elapsed: 20s
7
+ 2025-12-15 11:48:59 | INFO | waiting for other ranks to finish, time elapsed: 30s
8
+ 2025-12-15 11:49:09 | INFO | waiting for other ranks to finish, time elapsed: 40s
9
+ 2025-12-15 11:49:19 | INFO | waiting for other ranks to finish, time elapsed: 50s
10
+ 2025-12-15 11:49:29 | INFO | waiting for other ranks to finish, time elapsed: 60s
11
+ 2025-12-15 11:49:29 | INFO | model Qwen2.5-Omni-7B-all, data LibriSpeech, all 8 result merged to 7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl.
12
+ 2025-12-15 11:49:29 | INFO | skip eval for LibriSpeech
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank1.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:09 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:09 | INFO | Msg example: {'index': 1, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0012.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:10 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank2.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:11 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:11 | INFO | Msg example: {'index': 2, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0026.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:12 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank3.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:11 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:11 | INFO | Msg example: {'index': 3, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0022.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:12 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank4.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:13 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:13 | INFO | Msg example: {'index': 4, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0002.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:14 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank5.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:13 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:13 | INFO | Msg example: {'index': 5, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0025.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:13 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank6.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:07 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:07 | INFO | Msg example: {'index': 6, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0010.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:08 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank7.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 11:39:11 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 11:39:11 | INFO | Msg example: {'index': 7, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0001.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 11:39:12 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank0.log ADDED
@@ -0,0 +1,6 @@
 
 
 
 
 
 
 
1
+ 2025-12-15 11:50:24 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:50:24 | INFO | Msg example: {'index': 1, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0112/Lab41-SRI-VOiCES-rm1-babb-sp0112-ch123215-sg0025-mc01-stu-clo-dg080.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
3
+ 2025-12-15 11:55:26 | INFO | waiting for other ranks to finish, time elapsed: 10s
4
+ 2025-12-15 11:55:36 | INFO | waiting for other ranks to finish, time elapsed: 20s
5
+ 2025-12-15 11:55:36 | INFO | model Qwen2.5-Omni-7B-all, data voices_dev_clo, all 8 result merged to 7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/Qwen2.5-Omni-7B-all_voices_dev_clo.jsonl.
6
+ 2025-12-15 11:55:36 | INFO | skip eval for voices_dev_clo
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank1.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:49:47 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:49:47 | INFO | Msg example: {'index': 2, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0122/Lab41-SRI-VOiCES-rm1-babb-sp0122-ch121729-sg0002-mc02-lav-clo-dg060.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank2.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:49:50 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:49:50 | INFO | Msg example: {'index': 3, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0122/Lab41-SRI-VOiCES-rm1-babb-sp0122-ch121730-sg0014-mc01-stu-clo-dg000.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank3.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:50:00 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:50:00 | INFO | Msg example: {'index': 4, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0159/Lab41-SRI-VOiCES-rm1-babb-sp0159-ch135897-sg0052-mc01-stu-clo-dg100.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank4.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:50:16 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:50:16 | INFO | Msg example: {'index': 5, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0174/Lab41-SRI-VOiCES-rm1-babb-sp0174-ch084280-sg0013-mc02-lav-clo-dg010.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank5.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:49:44 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:49:44 | INFO | Msg example: {'index': 6, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0188/Lab41-SRI-VOiCES-rm1-babb-sp0188-ch135249-sg0029-mc01-stu-clo-dg170.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1500/Qwen2.5-Omni-7B-all/voices_dev_clo/logs/rank7.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 11:49:28 | INFO | Running Qwen2.5-Omni-7B-all on dataset: voices_dev_clo
2
+ 2025-12-15 11:49:28 | INFO | Msg example: {'index': 8, 'audio': ['/workspace/intern/pangkaiyu/dg/VOiCES_Box_unzip/Development_Data/Automatic_Speech_Recognition/ASR_dev.v2/rm1/babb/sp_0032-1182/sp0208/Lab41-SRI-VOiCES-rm1-babb-sp0208-ch126851-sg0011-mc02-lav-clo-dg070.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'voices', 'dataset_name': 'voices_dev_clo', 'lang': 'en', 'subset': 'babb'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_default_performance.json ADDED
@@ -0,0 +1,17 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "task": "ASR",
3
+ "dataset": "LibriSpeech",
4
+ "model": "Qwen2.5-Omni-7B-all",
5
+ "date": "2025-12-15 12:21:25.471164",
6
+ "performance": {
7
+ "test_clean": {
8
+ "wer": 2.4,
9
+ "total": 2620
10
+ },
11
+ "test_other": {
12
+ "wer": 4.48,
13
+ "total": 2939
14
+ }
15
+ },
16
+ "eval_method": "qwen2-audio-impl"
17
+ }
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech_wer_details.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank0.log ADDED
@@ -0,0 +1,9 @@
 
 
 
 
 
 
 
 
 
 
1
+ 2025-12-15 12:03:21 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:21 | INFO | Msg example: {'index': 0, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0003.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:21 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
5
+ 2025-12-15 12:13:09 | INFO | waiting for other ranks to finish, time elapsed: 10s
6
+ 2025-12-15 12:13:19 | INFO | waiting for other ranks to finish, time elapsed: 20s
7
+ 2025-12-15 12:13:29 | INFO | waiting for other ranks to finish, time elapsed: 30s
8
+ 2025-12-15 12:13:29 | INFO | model Qwen2.5-Omni-7B-all, data LibriSpeech, all 8 result merged to 7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/Qwen2.5-Omni-7B-all_LibriSpeech.jsonl.
9
+ 2025-12-15 12:13:29 | INFO | skip eval for LibriSpeech
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank1.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:22 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:22 | INFO | Msg example: {'index': 1, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0012.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:23 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank2.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:20 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:20 | INFO | Msg example: {'index': 2, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0026.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:20 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank3.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:23 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:23 | INFO | Msg example: {'index': 3, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0022.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:24 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank4.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:22 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:22 | INFO | Msg example: {'index': 4, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0002.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:23 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank5.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:23 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:23 | INFO | Msg example: {'index': 5, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0025.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:23 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/LibriSpeech/logs/rank7.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:03:18 | INFO | Running Qwen2.5-Omni-7B-all on dataset: LibriSpeech
2
+ 2025-12-15 12:03:18 | INFO | Msg example: {'index': 7, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0001.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 12:03:19 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus_default_performance.json ADDED
@@ -0,0 +1,121 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "task": "ASR",
3
+ "dataset": "noizeus",
4
+ "model": "Qwen2.5-Omni-7B-all",
5
+ "date": "2025-12-15 12:21:25.835192",
6
+ "performance": {
7
+ "airport_0dB": {
8
+ "wer": 29.34,
9
+ "total": 30
10
+ },
11
+ "airport_10dB": {
12
+ "wer": 1.65,
13
+ "total": 30
14
+ },
15
+ "airport_15dB": {
16
+ "wer": 1.65,
17
+ "total": 30
18
+ },
19
+ "airport_5dB": {
20
+ "wer": 11.57,
21
+ "total": 30
22
+ },
23
+ "babble_0dB": {
24
+ "wer": 45.87,
25
+ "total": 30
26
+ },
27
+ "babble_10dB": {
28
+ "wer": 2.89,
29
+ "total": 30
30
+ },
31
+ "babble_15dB": {
32
+ "wer": 1.65,
33
+ "total": 30
34
+ },
35
+ "babble_5dB": {
36
+ "wer": 9.92,
37
+ "total": 30
38
+ },
39
+ "car_0dB": {
40
+ "wer": 44.21,
41
+ "total": 30
42
+ },
43
+ "car_10dB": {
44
+ "wer": 1.65,
45
+ "total": 30
46
+ },
47
+ "car_15dB": {
48
+ "wer": 0.83,
49
+ "total": 30
50
+ },
51
+ "car_5dB": {
52
+ "wer": 9.92,
53
+ "total": 30
54
+ },
55
+ "exhibition_0dB": {
56
+ "wer": 28.51,
57
+ "total": 30
58
+ },
59
+ "exhibition_10dB": {
60
+ "wer": 3.31,
61
+ "total": 30
62
+ },
63
+ "exhibition_15dB": {
64
+ "wer": 2.07,
65
+ "total": 30
66
+ },
67
+ "exhibition_5dB": {
68
+ "wer": 10.74,
69
+ "total": 30
70
+ },
71
+ "restaurant_0dB": {
72
+ "wer": 40.91,
73
+ "total": 30
74
+ },
75
+ "restaurant_10dB": {
76
+ "wer": 2.07,
77
+ "total": 30
78
+ },
79
+ "restaurant_15dB": {
80
+ "wer": 2.48,
81
+ "total": 30
82
+ },
83
+ "restaurant_5dB": {
84
+ "wer": 14.88,
85
+ "total": 30
86
+ },
87
+ "station_0dB": {
88
+ "wer": 40.91,
89
+ "total": 30
90
+ },
91
+ "station_10dB": {
92
+ "wer": 2.89,
93
+ "total": 30
94
+ },
95
+ "station_15dB": {
96
+ "wer": 1.24,
97
+ "total": 30
98
+ },
99
+ "station_5dB": {
100
+ "wer": 11.57,
101
+ "total": 30
102
+ },
103
+ "street_0dB": {
104
+ "wer": 43.8,
105
+ "total": 30
106
+ },
107
+ "street_10dB": {
108
+ "wer": 3.72,
109
+ "total": 30
110
+ },
111
+ "street_15dB": {
112
+ "wer": 0.83,
113
+ "total": 30
114
+ },
115
+ "street_5dB": {
116
+ "wer": 17.77,
117
+ "total": 30
118
+ }
119
+ },
120
+ "eval_method": "qwen2-audio-impl"
121
+ }
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus_wer_details.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank0.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 12:13:29 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:29 | INFO | Msg example: {'index': 0, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp01_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
3
+ 2025-12-15 12:14:18 | INFO | model Qwen2.5-Omni-7B-all, data noizeus, all 8 result merged to 7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/Qwen2.5-Omni-7B-all_noizeus.jsonl.
4
+ 2025-12-15 12:14:18 | INFO | skip eval for noizeus
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank1.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:13:16 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:16 | INFO | Msg example: {'index': 1, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp02_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank2.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:12:43 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:12:43 | INFO | Msg example: {'index': 2, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp03_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank3.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:13:21 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:21 | INFO | Msg example: {'index': 3, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp04_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank4.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:13:07 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:07 | INFO | Msg example: {'index': 4, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp05_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank5.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:13:03 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:03 | INFO | Msg example: {'index': 5, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp06_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank6.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:12:41 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:12:41 | INFO | Msg example: {'index': 6, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp07_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_all_noise_ck1688/Qwen2.5-Omni-7B-all/noizeus/logs/rank7.log ADDED
@@ -0,0 +1,2 @@
 
 
 
1
+ 2025-12-15 12:13:05 | INFO | Running Qwen2.5-Omni-7B-all on dataset: noizeus
2
+ 2025-12-15 12:13:05 | INFO | Msg example: {'index': 7, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp08_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/Qwen2.5-Omni-7B-encoder_LibriSpeech.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank0.log ADDED
@@ -0,0 +1,7 @@
 
 
 
 
 
 
 
 
1
+ 2025-12-15 09:41:18 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:18 | INFO | Msg example: {'index': 0, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0003.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:19 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
5
+ 2025-12-15 09:48:06 | INFO | waiting for other ranks to finish, time elapsed: 10s
6
+ 2025-12-15 09:48:06 | INFO | model Qwen2.5-Omni-7B-encoder, data LibriSpeech, all 8 result merged to 7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/Qwen2.5-Omni-7B-encoder_LibriSpeech.jsonl.
7
+ 2025-12-15 09:48:06 | INFO | skip eval for LibriSpeech
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank1.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:18 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:18 | INFO | Msg example: {'index': 1, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0012.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:18 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank2.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:15 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:15 | INFO | Msg example: {'index': 2, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0026.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:16 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank3.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:17 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:17 | INFO | Msg example: {'index': 3, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0022.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:18 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank4.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:21 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:21 | INFO | Msg example: {'index': 4, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0002.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:21 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank5.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:20 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:20 | INFO | Msg example: {'index': 5, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0025.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:21 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank6.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:15 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:15 | INFO | Msg example: {'index': 6, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0010.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:16 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/LibriSpeech/logs/rank7.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:41:16 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: LibriSpeech
2
+ 2025-12-15 09:41:16 | INFO | Msg example: {'index': 7, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/dataset/LibriSpeech/librispeech/LibriSpeech/test-clean/1320/122617/1320-122617-0001.flac'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'LibriSpeech', 'dataset_name': 'LibriSpeech', 'lang': 'en', 'subset': 'test_clean'}}
3
+ 2025-12-15 09:41:17 | INFO | Prompt: You are a speech recognition model.
4
+ Transcribe the English audio into text without any punctuation marks.
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/noizeus/Qwen2.5-Omni-7B-encoder_noizeus.jsonl ADDED
The diff for this file is too large to render. See raw diff
 
different_distribution/noise_score/7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/noizeus/logs/rank0.log ADDED
@@ -0,0 +1,4 @@
 
 
 
 
 
1
+ 2025-12-15 09:48:06 | INFO | Running Qwen2.5-Omni-7B-encoder on dataset: noizeus
2
+ 2025-12-15 09:48:06 | INFO | Msg example: {'index': 0, 'audio': ['/workspace/intern/pangkaiyu/Kimi-Audio/Kimi-Audio-Evalkit/data/downloaded_datasets/noizeus/noizeus/airport/0dB/sp01_airport_sn0.wav'], 'text': 'Please transcribe the spoken content into written text.', 'meta': {'task': 'ASR', 'interactive': 'Audio-analysis', 'audio_type': 'Speech', 'dataset_series': 'noizeus', 'dataset_name': 'noizeus', 'lang': 'en', 'subset': 'airport_0dB'}}
3
+ 2025-12-15 09:48:43 | INFO | model Qwen2.5-Omni-7B-encoder, data noizeus, all 8 result merged to 7B_encoder_noise_ck1500/Qwen2.5-Omni-7B-encoder/noizeus/Qwen2.5-Omni-7B-encoder_noizeus.jsonl.
4
+ 2025-12-15 09:48:43 | INFO | skip eval for noizeus