Segmentation-Free Streaming Machine Translation

Javier Iranzo-Sánchez; Jorge Iranzo-Sánchez; Adrià Giménez; Jorge Civera; Alfons Juan

Vol. 12 (2024)

TACL approved

Segmentation-Free Streaming Machine Translation

Published 2024-09-07

Javier Iranzo-Sánchez
Jorge Iranzo-Sánchez
Adrià Giménez
Jorge Civera
Alfons Juan

Javier Iranzo-Sánchez
Valencian Research Institute for Artificial Intelligence, Universitat Politècnica de València, València, Spain; Applications Technology (AppTek), València, Spain

Jorge Iranzo-Sánchez
Machine Learning and Language Processing, VRAIN, Universitat Politècnica de València

Adrià Giménez
Departament d’Informàtica, Escola Tècnica Superior d’Enginyeria, Universitat de València

Jorge Civera
Machine Learning and Language Processing, VRAIN, Universitat Politècnica de València

Alfons Juan
Machine Learning and Language Processing, VRAIN, Universitat Politècnica de València

Abstract

Streaming Machine Translation (MT) is the task of translating an unbounded input text stream in real-time. The traditional cascade approach, which combines an Automatic Speech Recognition (ASR) and an MT system, relies on an intermediate segmentation step which splits the transcription stream into sentence-like units. However, the incorporation of a hard segmentation constrains the MT system and is a source of errors. This paper proposes a Segmentation-Free framework that enables the model to translate an unsegmented source stream by delaying the segmentation decision until after the translation has been generated. Extensive experiments show how the proposed Segmentation-Free framework has better quality-latency trade-off than competing approaches that use an independent segmentation model.

Article at MIT Press Presented at ACL 2024