The Kyoto Speech-to-Speech Translation System for IWSLT 2023

Zhengdong Yang; Shuichiro Shimizu; Wangjin Zhou; Sheng Li; Chenhui Chu

The Kyoto Speech-to-Speech Translation System for IWSLT 2023

Zhengdong Yang, Shuichiro Shimizu, Wangjin Zhou, Sheng Li, Chenhui Chu

Add to Favorites

The 20th International Conference on Spoken Language Translation Long Paper

TLDR: This paper describes the Kyoto speech-to-speech translation system for IWSLT 2023. Our system is a combination of speech-to-text translation and text-to-speech synthesis. For the speech-to-text translation model, we used the dual-decoderTransformer model. For text-to-speech synthesis model, we took

RocketChat
Abstract

You can open the #paper-IWSLT_40 channel in a separate window.

Abstract: This paper describes the Kyoto speech-to-speech translation system for IWSLT 2023. Our system is a combination of speech-to-text translation and text-to-speech synthesis. For the speech-to-text translation model, we used the dual-decoderTransformer model. For text-to-speech synthesis model, we took a cascade approach of an acoustic model and a vocoder.