Benchmarking General-Purpose In-Context Learning

Wang, Fan; Lin, Chuan; Cao, Yang; Kang, Yu

Computer Science > Artificial Intelligence

arXiv:2405.17234 (cs)

[Submitted on 27 May 2024 (v1), last revised 26 Jun 2024 (this version, v5)]

Title:Benchmarking General-Purpose In-Context Learning

Authors:Fan Wang, Chuan Lin, Yang Cao, Yu Kang

View PDF HTML (experimental)

Abstract:In-context learning (ICL) empowers generative models to address new tasks effectively and efficiently on the fly, without relying on any artificially crafted optimization techniques. In this paper, we study extending ICL to address a broader range of tasks with an extended learning horizon and higher improvement potential, namely General-Purpose In-Context Learning (GPICL). To this end, we introduce two lightweight benchmarks specifically crafted to train and evaluate GPICL functionalities. Each benchmark encompasses a vast number of tasks characterized by significant task variance, facilitating meta-training that minimizes inductive bias. These tasks are also crafted to promote long-horizon in-context learning through continuous generation and interaction. These characteristics necessitate the models to leverage contexts and history interactions to enhance their capabilities, across domains such as language modeling, decision-making, and world modeling. Our experiments on the baseline models demonstrate that meta-training with minimal inductive bias and ICL from the ground up is feasible across all the domains we've discussed. Additionally, our findings indicate that the scale of parameters alone may not be crucial for ICL or GPICL, suggesting alternative approaches such as increasing the scale of contexts and memory states.

Subjects:	Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2405.17234 [cs.AI]
	(or arXiv:2405.17234v5 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2405.17234

Submission history

From: Fan Wang [view email]
[v1] Mon, 27 May 2024 14:50:42 UTC (15,747 KB)
[v2] Wed, 29 May 2024 13:35:01 UTC (17,517 KB)
[v3] Thu, 6 Jun 2024 11:18:17 UTC (38,564 KB)
[v4] Mon, 17 Jun 2024 10:12:59 UTC (38,555 KB)
[v5] Wed, 26 Jun 2024 07:59:40 UTC (35,370 KB)

Computer Science > Artificial Intelligence

Title:Benchmarking General-Purpose In-Context Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Benchmarking General-Purpose In-Context Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators