Catalogue search • Linguistik portal • Fachinformationsdienst (FID)

1	MeetDot: Videoconferencing with Live Translation Captions ...
	Arkhangorodsky, Arkady; Chu, Christopher; Fang, Scot; Huang, Yiqi; Jiang, Denglin; Nagesh, Ajay; Zhang, Boliang; Knight, Kevin. - : arXiv, 2021
	Abstract: We present MeetDot, a videoconferencing system with live translation captions overlaid on screen. The system aims to facilitate conversation between people who speak different languages, thereby reducing communication barriers between multilingual participants. Currently, our system supports speech and captions in 4 languages and combines automatic speech recognition (ASR) and machine translation (MT) in a cascade. We use the re-translation strategy to translate the streamed speech, resulting in caption flicker. Additionally, our system has very strict latency requirements to have acceptable call quality. We implement several features to enhance user experience and reduce their cognitive load, such as smooth scrolling captions and reducing caption flicker. The modular architecture allows us to integrate different ASR and MT services in our backend. Our system provides an integrated evaluation suite to optimize key intrinsic evaluation metrics such as accuracy, latency and erasure. Finally, we present an ... : 7 pages, 4 figures, Accepted as EMNLP 2021 demo paper ...
	Keyword: Artificial Intelligence cs.AI; Computation and Language cs.CL; FOS Computer and information sciences
	URL: https://dx.doi.org/10.48550/arxiv.2109.09577 https://arxiv.org/abs/2109.09577
	BASE
	Hide details

2	MeetDot: Videoconferencing with Live Translation Captions ...
	The 2021 Conference on Empirical Methods in Natural Language Processing 2021; Arkhangorodsky, Arkady; Chu, Christopher. - : Underline Science Inc., 2021
	BASE
	Show details

3	Learning to Pronounce Chinese Without a Pronunciation Dictionary ...
	Chu, Christopher; Fang, Scot; Knight, Kevin. - : arXiv, 2020
	BASE
	Show details

Search in the Catalogues and Directories