Skip to content

Latest commit

History

20 Commits

Folders and files

NameName
Last commit message
Last commit date

Repository files navigation

에코 (ECHO)

  • 제24회 충북컴퓨터꿈나무축제 고등학교 공모전(SW제작) 대상 수상작

작품 소개

  • 이 서비스는 Diff-SVC 프로젝트를 기반으로 제작되었습니다.
  • 직접 자신의 목소리를 학습시켜 인공지능 모델을 제작하고, 그 모델로 TTS 기능을 이용할 수 있도록 하는 서비스입니다.
  • Streamlit 패키지로 사용자 친화적 인터페이스를 구현하여 누구나 쉽게 모델 제작 및 음성 생성을 할 수 있는 환경을 제공합니다.

사전 준비

0. 최소사양 및 권장사양

최소사양권장사양
RAM8GB16GB
GPUGeForce GTX 1050 TiGeForce RTX 2070
VRAM4GB8GB

1. Python 3.10.11 설치

2. CUDA Toolkit 11.8 설치

3. FFmpeg 설치

Windows

Debian / Ubuntu

sudo apt update
sudo apt install ffmpeg

4. 현재 리포지토리 다운로드

Windows

Debian / Ubuntu

sudo apt install git
git clone https://github.com/k-yumin/echo.git

5. Pre-trained 모델 다운로드

(Optional) GPU 메모리가 6GB 이상인 경우

6. 필요한 패키지 다운로드

  • setup.py 실행

서비스 실행

Windows

  • run.bat 실행

Debian / Ubuntu

./run.sh

런타임 에러 해결 방법

Traceback (most recent call last):
(...)
File "(...)/torch/functional.py", line 641, in stft
return _VF.stft(input, n_fft, hop_length, win_length, window, # type: ignore[attr-defined]
RuntimeError: stft requires the return_complex parameter be given forreal inputs, and will further require that return_complex=Truein a future PyTorch release.
  • 모델 학습 중 다음과 같은 런타임 에러가 발생했을 경우, 프롬포트에 출력된 경로 (...)/torch/functional.py 파일의 641번째 줄에 다음 코드를 추가한다.
ifnotreturn_complex:
returntorch.view_as_real(_VF.stft(input, n_fft, hop_length, win_length, window, # type: ignore[attr-defined]normalized, onesided, return_complex=True))

About

No description, website, or topics provided.

Resources

Stars

0 stars

Watchers

1 watching

Forks

Releases

Packages

Contributors

Languages