arXiv · 2412.18707
Learning from Many Voices: Literary MT Using Multi-Reference Human and Synthetic Data
Abstract
Multiple valid translations of a single literary work naturally exist. We investigate strategies for leveraging these multi-reference datasets to improve literary machine translation. We propose a filtering framework based on semantic similarity to identify source texts whose references display meaningful variation while remaining faithful. We find that fine-tuning with medium to high semantic similarity data substantially outperforms low semantic similarity data. Moreover, using medium and high semantic similarity data achieves comparable or better performance than using the full unfiltered data. Synthetic translations generated by LLMs are economical and convenient alternatives to human expert translations; however, we find fine-tuning on human expert translations outperforms fine-tuning on synthetically augmented data in automatic metrics and human evaluations, demonstrating the indispensable value of human expert translations for fine-tuning literary machine translation models.
Explore related subjects
Keep this discovery
Si Wu, John Wieting, David A. Smith. 2026-08-29. Learning from Many Voices: Literary MT Using Multi-Reference Human and Synthetic Data. https://arxiv.org/abs/2412.18707
Cite the original work for its findings. Save a collection to share your selection of sources.
Discover connections
Connections use source metadata and explicit phrase matches, not verified experimental comparisons.