python/FunASR-XL.git

			@@ -72,7 +72,7 @@
			- pcm_path, `e.g.`: asr_example.pcm,
			- audio bytes stream, `e.g.`: bytes data from a microphone
			- audio sample point，`e.g.`: `audio, rate = soundfile.read("asr_example_zh.wav")`, the dtype is numpy.ndarray or torch.Tensor
			- wav.scp, kaldi style wav list (`wav_id \t wav_path``), `e.g.`:
			- wav.scp, kaldi style wav list (`wav_id \t wav_path`), `e.g.`:
			```text
			asr_example1 ./audios/asr_example1.wav
			asr_example2 ./audios/asr_example2.wav
			@@ -82,7 +82,7 @@
			- `output_dir`: None (Default), the output path of results if set

			### Inference with multi-thread CPUs or multi GPUs
			FunASR also offer recipes [infer.sh](https://github.com/alibaba-damo-academy/FunASR/blob/main/egs_modelscope/asr/TEMPLATE/infer.sh) to decode with multi-thread CPUs, or multi GPUs.
			FunASR also offer recipes [egs_modelscope/asr/TEMPLATE/infer.sh](https://github.com/alibaba-damo-academy/FunASR/blob/main/egs_modelscope/asr/TEMPLATE/infer.sh) to decode with multi-thread CPUs, or multi GPUs.

			- Setting parameters in `infer.sh`
			- `model`: model name in [model zoo](https://alibaba-damo-academy.github.io/FunASR/en/modelscope_models.html#pretrained-models-on-modelscope), or model path in local disk
			@@ -176,6 +176,21 @@
			- `max_epoch`: number of training epoch
			- `lr`: learning rate

			- Training data formats：
			```sh
			cat ./example_data/text
			BAC009S0002W0122 而对楼市成交抑制作用最大的限购
			BAC009S0002W0123 也成为地方政府的眼中钉
			english_example_1 hello world
			english_example_2 go swim 去游泳

			cat ./example_data/wav.scp
			BAC009S0002W0122 /mnt/data/wav/train/S0002/BAC009S0002W0122.wav
			BAC009S0002W0123 /mnt/data/wav/train/S0002/BAC009S0002W0123.wav
			english_example_1 /mnt/data/wav/train/S0002/english_example_1.wav
			english_example_2 /mnt/data/wav/train/S0002/english_example_2.wav
			```

			- Then you can run the pipeline to finetune with:
			```shell
			python finetune.py
			@@ -186,7 +201,7 @@
			```
			## Inference with your finetuned model

			- Setting parameters in [infer.sh](https://github.com/alibaba-damo-academy/FunASR/blob/main/egs_modelscope/asr/TEMPLATE/infer.sh) is the same with [docs](https://github.com/alibaba-damo-academy/FunASR/tree/main/egs_modelscope/asr/TEMPLATE#inference-with-multi-thread-cpus-or-multi-gpus), `model` is the model name from modelscope, which you finetuned.
			- Setting parameters in [egs_modelscope/asr/TEMPLATE/infer.sh](https://github.com/alibaba-damo-academy/FunASR/blob/main/egs_modelscope/asr/TEMPLATE/infer.sh) is the same with [docs](https://github.com/alibaba-damo-academy/FunASR/tree/main/egs_modelscope/asr/TEMPLATE#inference-with-multi-thread-cpus-or-multi-gpus), `model` is the model name from modelscope, which you finetuned.

			- Decode with multi GPUs:
			```shell
			@@ -210,4 +225,4 @@
			--njob 64 \
			--checkpoint_dir "./checkpoint" \
			--checkpoint_name "valid.cer_ctc.ave.pb"
			```
			```