Retrieval-based-Voice-Conve.../my_utils.py

import ffmpeg
import numpy as np


def load_audio(file, sr):
    try:
        # https://github.com/openai/whisper/blob/main/whisper/audio.py#L26
        # This launches a subprocess to decode audio while down-mixing and resampling as necessary.
        # Requires the ffmpeg CLI and `ffmpeg-python` package to be installed.
        file = (
            file.strip(" ").strip('"').strip("\n").strip('"').strip(" ")
        )  # 防止小白拷路径头尾带了空格和"和回车
        out, _ = (
            ffmpeg.input(file, threads=0)
            .output("-", format="f32le", acodec="pcm_f32le", ac=1, ar=sr)
            .run(cmd=["ffmpeg", "-nostdin"], capture_stdout=True, capture_stderr=True)
        )
    except Exception as e:
        raise RuntimeError(f"Failed to load audio: {e}")

    return np.frombuffer(out, np.float32).flatten()
删除无用文件，增加--colab启动选项 2023-04-01 09:02:53 +02:00			`import ffmpeg`
			`import numpy as np`
Reformat and rewrite _get_name_params (#57) * Reformat * rewrite _get_name_params * Add workflow for automatic formatting * Revert "Add workflow for automatic formatting" This reverts commit 9111c5dbc1830248305fb075587a88be07ad3115. * revert Retrieval_based_Voice_Conversion_WebUI.ipynb --------- Co-authored-by: 源文雨 <41315874+fumiama@users.noreply.github.com> 2023-04-15 13:44:24 +02:00

			`def load_audio(file, sr):`
Add files via upload 2023-03-31 11:54:38 +02:00			`try:`
			`# https://github.com/openai/whisper/blob/main/whisper/audio.py#L26`
			`# This launches a subprocess to decode audio while down-mixing and resampling as necessary.`
			# Requires the ffmpeg CLI and `ffmpeg-python` package to be installed.
Reformat and rewrite _get_name_params (#57) * Reformat * rewrite _get_name_params * Add workflow for automatic formatting * Revert "Add workflow for automatic formatting" This reverts commit 9111c5dbc1830248305fb075587a88be07ad3115. * revert Retrieval_based_Voice_Conversion_WebUI.ipynb --------- Co-authored-by: 源文雨 <41315874+fumiama@users.noreply.github.com> 2023-04-15 13:44:24 +02:00			`file = (`
			`file.strip(" ").strip('"').strip("\n").strip('"').strip(" ")`
			`) # 防止小白拷路径头尾带了空格和"和回车`
Add files via upload 2023-03-31 11:54:38 +02:00			`out, _ = (`
			`ffmpeg.input(file, threads=0)`
some change precision audio processing (#94) * some change precision audio processing * fix clipping problem in resample resample sometimes causes signal clipping, not just librosa.resample * fix error 2023-04-22 13:39:47 +02:00			`.output("-", format="f32le", acodec="pcm_f32le", ac=1, ar=sr)`
fix ffmpeg 2023-04-01 09:10:46 +02:00			`.run(cmd=["ffmpeg", "-nostdin"], capture_stdout=True, capture_stderr=True)`
Add files via upload 2023-03-31 11:54:38 +02:00			`)`
fix: train step2a & add arg --port --pycmd --noparallel 2023-04-01 10:42:19 +02:00			`except Exception as e:`
			`raise RuntimeError(f"Failed to load audio: {e}")`
Add files via upload 2023-03-31 11:54:38 +02:00
Format code (#142) Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com> 2023-04-24 14:35:56 +02:00			`return np.frombuffer(out, np.float32).flatten()`