A 10.5K-question benchmark for scholarly QA forces systems to combine DBLP and SemOpenAlex knowledge graph facts with Wikipedia text.
Integrating SPARQL and LLMs for Question Answering over Scholarly Data Sources
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
The Scholarly Hybrid Question Answering over Linked Data (QALD) Challenge at the International Semantic Web Conference (ISWC) 2024 focuses on Question Answering (QA) over diverse scholarly sources: DBLP, SemOpenAlex, and Wikipedia-based texts. This paper describes a methodology that combines SPARQL queries, divide and conquer algorithms, and a pre-trained extractive question answering model. It starts with SPARQL queries to gather data, then applies divide and conquer to manage various question types and sources, and uses the model to handle personal author questions. The approach, evaluated with Exact Match and F-score metrics, shows promise for improving QA accuracy and efficiency in scholarly contexts.
fields
cs.CL 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Hybrid-SQuAD: Hybrid Scholarly Question Answering Dataset
A 10.5K-question benchmark for scholarly QA forces systems to combine DBLP and SemOpenAlex knowledge graph facts with Wikipedia text.