메뉴 건너뛰기
소속 기관 / 학교 인증
인증하면 논문, 학술자료 등을  무료로 열람할 수 있어요.
한국대학교, 누리자동차, 시립도서관 등 나의 기관을 확인해보세요
(국내 대학 90% 이상 구독 중)
고객센터 ENG
주제분류

논문 기본 정보

저자정보
(Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University) (Sungkyunkwan University)
저널정보
한국방송·미디어공학회 한국방송미디어공학회 학술발표대회 논문집 한국방송·미디어공학회 2023 하계학술대회
오류 신고하기

피인용 0

검색

    초록·키워드

    To synthesize natural and intelligible speech with a small amount of data, transfer learning with well-maintained and pre-trained data has been known to be useful. However, little attention has been paid to answer the following research questions with empirically-grounded evidence, "How much pre-trained (source) speech data (e.g., 10 K utterances or 10 hours) used in transfer learning is enough for generating natural and intelligible speech?" and "For generating natural and intelligible speech, how much (target) speech data should at least be provided?", which are essential for the quality of speech synthesis. To answer these questions, this paper conducts extensive experiments on speech synthesis with multiple source and target data with different lengths, speakers, and languages. We show that intelligible and natural speech can be synthesized with only 500 utterances of target data using transfer learning. Our work also reveals that at least 5000 utterances of source pre-trained data are required to synthesize decent speech.

    최근 본 자료 전체보기

      UCI(KEPA) : I410-ECN-0102-2023-567-001939431