From d9e4b0b04eca1470b32ac41965cfcc4afe9c06e7 Mon Sep 17 00:00:00 2001 From: zhifu gao <zhifu.gzf@alibaba-inc.com> Date: 星期一, 17 四月 2023 14:32:25 +0800 Subject: [PATCH] Merge pull request #365 from alibaba-damo-academy/yufan-aslp-patch-1 --- docs/modelscope_models.md | 7 ++++++- 1 files changed, 6 insertions(+), 1 deletions(-) diff --git a/docs/modelscope_models.md b/docs/modelscope_models.md index 07e590b..646216c 100644 --- a/docs/modelscope_models.md +++ b/docs/modelscope_models.md @@ -40,13 +40,18 @@ | [Conformer](https://modelscope.cn/models/damo/speech_conformer_asr_nat-zh-cn-16k-aishell1-vocab4234-pytorch/summary) | CN | AISHELL (178hours) | 4234 | 44M | Offline | Duration of input wav <= 20s | | [Conformer](https://www.modelscope.cn/models/damo/speech_conformer_asr_nat-zh-cn-16k-aishell2-vocab5212-pytorch/summary) | CN | AISHELL-2 (1000hours) | 5212 | 44M | Offline | Duration of input wav <= 20s | + +#### RNN-T Models + +### Multi-talker Speech Recognition Models + #### MFCCA Models | Model Name | Language | Training Data | Vocab Size | Parameter | Offline/Online | Notes | |:----------------------------------------------------------------------------------------------------------------------:|:--------:|:---------------------:|:----------:|:---------:|:--------------:|:--------------------------------------------------------------------------------------------------------------------------------| | [MFCCA](https://www.modelscope.cn/models/NPU-ASLP/speech_mfcca_asr-zh-cn-16k-alimeeting-vocab4950/summary) | CN | AliMeeting銆丄ISHELL-4銆丼imudata (917hours) | 4950 | 45M | Offline | Duration of input wav <= 20s, channel of input wav <= 8 channel -#### RNN-T Models + ### Voice Activity Detection Models -- Gitblit v1.9.1