PixT3: Pixel-based Table-To-Text Generation

Alonso, Iñigo; Agirre, Eneko; Lapata, Mirella

Computer Science > Computation and Language

arXiv:2311.09808 (cs)

[Submitted on 16 Nov 2023 (v1), last revised 3 Jun 2024 (this version, v3)]

Title:PixT3: Pixel-based Table-To-Text Generation

Authors:Iñigo Alonso, Eneko Agirre, Mirella Lapata

View PDF HTML (experimental)

Abstract:Table-to-text generation involves generating appropriate textual descriptions given structured tabular data. It has attracted increasing attention in recent years thanks to the popularity of neural network models and the availability of large-scale datasets. A common feature across existing methods is their treatment of the input as a string, i.e., by employing linearization techniques that do not always preserve information in the table, are verbose, and lack space efficiency. We propose to rethink data-to-text generation as a visual recognition task, removing the need for rendering the input in a string format. We present PixT3, a multimodal table-to-text model that overcomes the challenges of linearization and input size limitations encountered by existing models. PixT3 is trained with a new self-supervised learning objective to reinforce table structure awareness and is applicable to open-ended and controlled generation settings. Experiments on the ToTTo and Logic2Text benchmarks show that PixT3 is competitive and, in some settings, superior to generators that operate solely on text.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2311.09808 [cs.CL]
	(or arXiv:2311.09808v3 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2311.09808

Submission history

From: Iñigo Alonso [view email]
[v1] Thu, 16 Nov 2023 11:32:47 UTC (8,547 KB)
[v2] Thu, 22 Feb 2024 16:59:55 UTC (2,442 KB)
[v3] Mon, 3 Jun 2024 17:43:23 UTC (8,328 KB)

Computer Science > Computation and Language

Title:PixT3: Pixel-based Table-To-Text Generation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:PixT3: Pixel-based Table-To-Text Generation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators