Score: 1

DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing

Published: April 30, 2025 | arXiv ID: 2504.21589v1

By: Lisa Kluge, Maximilian Kähler

Potential Business Impact:

Tags library books automatically for better searching.

Business Areas:

Natural Language Processing Artificial Intelligence, Data and Analytics, Software

This paper presents our system developed for the SemEval-2025 Task 5: LLMs4Subjects: LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog. Our system relies on prompting a selection of LLMs with varying examples of intellectually annotated records and asking the LLMs to similarly suggest keywords for new records. This few-shot prompting technique is combined with a series of post-processing steps that map the generated keywords to the target vocabulary, aggregate the resulting subject terms to an ensemble vote and, finally, rank them as to their relevance to the record. Our system is fourth in the quantitative ranking in the all-subjects track, but achieves the best result in the qualitative ranking conducted by subject indexing experts.

SemEval-2025 Task 5: LLMs4Subjects -- LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog

Computation and Language

Helps libraries sort science papers automatically.

9 Apr 2025 1

89%

Annif at SemEval-2025 Task 5: Traditional XMTC augmented by LLMs

Computation and Language

Helps libraries automatically sort books by topic.

28 Apr 2025 1

89%

NBF at SemEval-2025 Task 5: Light-Burst Attention Enhanced System for Multilingual Subject Recommendation

Computation and Language

Helps computers sort academic papers by topic.

6 May 2025 0

View PDF Login to Bookmark

Repos / Data Links

github.com github.com

Page Count

11 pages

DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing

Tags library books automatically for better searching.

Technical Abstract

SemEval-2025 Task 5: LLMs4Subjects -- LLM-based Automated Subject Tagging for a National Technical Library's Open-Access Catalog

Annif at SemEval-2025 Task 5: Traditional XMTC augmented by LLMs

NBF at SemEval-2025 Task 5: Light-Burst Attention Enhanced System for Multilingual Subject Recommendation