Gender-specific Machine Translation with Large Language Models

Sánchez, Eduardo; Andrews, Pierre; Stenetorp, Pontus; Artetxe, Mikel; Costa-jussà, Marta R.

Computer Science > Computation and Language

arXiv:2309.03175v1 (cs)

[Submitted on 6 Sep 2023 (this version), latest version 16 Apr 2024 (v2)]

Title:Gender-specific Machine Translation with Large Language Models

Authors:Eduardo Sánchez, Pierre Andrews, Pontus Stenetorp, Mikel Artetxe, Marta R. Costa-jussà

View PDF

Abstract:Decoder-only Large Language Models (LLMs) have demonstrated potential in machine translation (MT), albeit with performance slightly lagging behind traditional encoder-decoder Neural Machine Translation (NMT) systems. However, LLMs offer a unique advantage: the ability to control the properties of the output through prompts. In this study, we harness this flexibility to explore LLaMa's capability to produce gender-specific translations for languages with grammatical gender. Our results indicate that LLaMa can generate gender-specific translations with competitive accuracy and gender bias mitigation when compared to NLLB, a state-of-the-art multilingual NMT system. Furthermore, our experiments reveal that LLaMa's translations are robust, showing significant performance drops when evaluated against opposite-gender references in gender-ambiguous datasets but maintaining consistency in less ambiguous contexts. This research provides insights into the potential and challenges of using LLMs for gender-specific translations and highlights the importance of in-context learning to elicit new tasks in LLMs.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2309.03175 [cs.CL]
	(or arXiv:2309.03175v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2309.03175

Submission history

From: Eduardo Sánchez [view email]
[v1] Wed, 6 Sep 2023 17:24:06 UTC (7,614 KB)
[v2] Tue, 16 Apr 2024 19:16:46 UTC (8,716 KB)

Computer Science > Computation and Language

Title:Gender-specific Machine Translation with Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Gender-specific Machine Translation with Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators