-
japanese-asr/en2ja.s2t_translation
Viewer • Updated • 32k • 129 • 2 -
japanese-asr/ja2en.s2t_translation
Viewer • Updated • 2.24k • 34 • 1 -
japanese-asr/ja-cascaded-s2t-translation
Automatic Speech Recognition • 0.8B • Updated • 11 • 4 -
japanese-asr/en-cascaded-s2t-translation
Automatic Speech Recognition • 0.8B • Updated • 7 • 1
AI & ML interests
This repo contains models and datasets for Japanese ASR. See our main model https://huggingface.co/kotoba-tech/kotoba-whisper-v1.0.
Japanese ASR Models
-
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-all
0.8B • Updated • 33 • 4 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-large
Automatic Speech Recognition • 0.8B • Updated • 321 • 5 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-medium
Automatic Speech Recognition • 0.8B • Updated • 16 • 2 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-tiny
Automatic Speech Recognition • 0.8B • Updated • 37 • 1
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3 (WER filter applied).
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny.wer_10.0
Viewer • Updated • 1.77k • 7 -
japanese-asr/whisper_transcriptions.reazonspeech.small.wer_10.0
Viewer • Updated • 20.9k • 201 -
japanese-asr/whisper_transcriptions.reazonspeech.medium.wer_10.0
Viewer • Updated • 209k • 48 -
japanese-asr/whisper_transcriptions.reazonspeech.large.wer_10.0
Viewer • Updated • 1.04M • 1.39k
ASR Evaluation Dataset
These are the collection of the Bilingual ASR datasets labelled by the whisper-large-v3. The dataset consists of ASR and S2T translation tasks.
-
japanese-asr/en_asr.mls
Viewer • Updated • 10.4M • 5.9k • 2 -
japanese-asr/whisper_transcriptions.mls
Viewer • Updated • 10.4M • 106 • 1 -
japanese-asr/whisper_transcriptions.mls.wer_10.0
Viewer • Updated • 9.33M • 2.8k • 1 -
japanese-asr/whisper_transcriptions.mls.wer_10.0.vectorized
Viewer • Updated • 7.44M • 14.7k • 1
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3.
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny
Viewer • Updated • 5.32k • 50 -
japanese-asr/whisper_transcriptions.reazonspeech.small
Viewer • Updated • 62k • 130 • 2 -
japanese-asr/whisper_transcriptions.reazonspeech.medium
Viewer • Updated • 619k • 188 -
japanese-asr/whisper_transcriptions.reazonspeech.large
Viewer • Updated • 3.1M • 252
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3 (WER filter applied and transformed into logmel feature).
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny.wer_10.0.vectorized
Viewer • Updated • 1.77k • 9 -
japanese-asr/whisper_transcriptions.reazonspeech.small.wer_10.0.vectorized
Viewer • Updated • 3.22k • 49 -
japanese-asr/whisper_transcriptions.reazonspeech.medium.wer_10.0.vectorized
Viewer • Updated • 3.26k • 47 -
japanese-asr/whisper_transcriptions.reazonspeech.large.wer_10.0.vectorized
Viewer • Updated • 3.26k • 973
-
japanese-asr/en2ja.s2t_translation
Viewer • Updated • 32k • 129 • 2 -
japanese-asr/ja2en.s2t_translation
Viewer • Updated • 2.24k • 34 • 1 -
japanese-asr/ja-cascaded-s2t-translation
Automatic Speech Recognition • 0.8B • Updated • 11 • 4 -
japanese-asr/en-cascaded-s2t-translation
Automatic Speech Recognition • 0.8B • Updated • 7 • 1
These are the collection of the Bilingual ASR datasets labelled by the whisper-large-v3. The dataset consists of ASR and S2T translation tasks.
-
japanese-asr/en_asr.mls
Viewer • Updated • 10.4M • 5.9k • 2 -
japanese-asr/whisper_transcriptions.mls
Viewer • Updated • 10.4M • 106 • 1 -
japanese-asr/whisper_transcriptions.mls.wer_10.0
Viewer • Updated • 9.33M • 2.8k • 1 -
japanese-asr/whisper_transcriptions.mls.wer_10.0.vectorized
Viewer • Updated • 7.44M • 14.7k • 1
Japanese ASR Models
-
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-all
0.8B • Updated • 33 • 4 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-large
Automatic Speech Recognition • 0.8B • Updated • 321 • 5 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-medium
Automatic Speech Recognition • 0.8B • Updated • 16 • 2 -
japanese-asr/distil-whisper-large-v3-ja-reazonspeech-tiny
Automatic Speech Recognition • 0.8B • Updated • 37 • 1
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3.
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny
Viewer • Updated • 5.32k • 50 -
japanese-asr/whisper_transcriptions.reazonspeech.small
Viewer • Updated • 62k • 130 • 2 -
japanese-asr/whisper_transcriptions.reazonspeech.medium
Viewer • Updated • 619k • 188 -
japanese-asr/whisper_transcriptions.reazonspeech.large
Viewer • Updated • 3.1M • 252
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3 (WER filter applied).
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny.wer_10.0
Viewer • Updated • 1.77k • 7 -
japanese-asr/whisper_transcriptions.reazonspeech.small.wer_10.0
Viewer • Updated • 20.9k • 201 -
japanese-asr/whisper_transcriptions.reazonspeech.medium.wer_10.0
Viewer • Updated • 209k • 48 -
japanese-asr/whisper_transcriptions.reazonspeech.large.wer_10.0
Viewer • Updated • 1.04M • 1.39k
These are the collection of the Japanese ASR datasets labelled by the whisper-large-v3 (WER filter applied and transformed into logmel feature).
-
japanese-asr/whisper_transcriptions.reazonspeech.tiny.wer_10.0.vectorized
Viewer • Updated • 1.77k • 9 -
japanese-asr/whisper_transcriptions.reazonspeech.small.wer_10.0.vectorized
Viewer • Updated • 3.22k • 49 -
japanese-asr/whisper_transcriptions.reazonspeech.medium.wer_10.0.vectorized
Viewer • Updated • 3.26k • 47 -
japanese-asr/whisper_transcriptions.reazonspeech.large.wer_10.0.vectorized
Viewer • Updated • 3.26k • 973
ASR Evaluation Dataset