Search-Adaptor: Embedding Customization for Information Retrieval

Yoon, **sung; Arik, Sercan O; Chen, Yanfei; Pfister, Tomas

Computer Science > Machine Learning

arXiv:2310.08750 (cs)

[Submitted on 12 Oct 2023 (v1), last revised 12 Mar 2024 (this version, v2)]

Title:Search-Adaptor: Embedding Customization for Information Retrieval

Authors:**sung Yoon, Sercan O Arik, Yanfei Chen, Tomas Pfister

View PDF HTML (experimental)

Abstract:Embeddings extracted by pre-trained Large Language Models (LLMs) have significant potential to improve information retrieval and search. Beyond the zero-shot setup in which they are being conventionally used, being able to take advantage of the information from the relevant query-corpus paired data can further boost the LLM capabilities. In this paper, we propose a novel method, Search-Adaptor, for customizing LLMs for information retrieval in an efficient and robust way. Search-Adaptor modifies the embeddings generated by pre-trained LLMs, and can be integrated with any LLM, including those only available via prediction APIs. On multiple English, multilingual, and multimodal retrieval datasets, we show consistent and significant performance benefits for Search-Adaptor -- e.g., more than 5% improvements for Google Embedding APIs in nDCG@10 averaged over 14 BEIR datasets.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2310.08750 [cs.LG]
	(or arXiv:2310.08750v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2310.08750

Submission history

From: **sung Yoon [view email]
[v1] Thu, 12 Oct 2023 22:30:15 UTC (145 KB)
[v2] Tue, 12 Mar 2024 22:09:41 UTC (14,706 KB)

Computer Science > Machine Learning

Title:Search-Adaptor: Embedding Customization for Information Retrieval

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Search-Adaptor: Embedding Customization for Information Retrieval

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators