Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research?

Longquan Jiang; Ricardo Usbeck

doi:10.1145/3477495.3531751

Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research?

Research output: Contributions to collected editions/works › Article in conference proceedings › Research › peer-review

Standard

Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research? / Jiang, Longquan; Usbeck, Ricardo.
SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval. ed. / Enrique Amigo; Pablo Castells; Julio Gonzalo. New York: Association for Computing Machinery, Inc, 2022. p. 3209-3218 (Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval; Vol. 2022).

Research output: Contributions to collected editions/works › Article in conference proceedings › Research › peer-review

Harvard

Jiang, L & Usbeck, R 2022, Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research? in E Amigo, P Castells & J Gonzalo (eds), SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval. Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval, vol. 2022, Association for Computing Machinery, Inc, New York, pp. 3209-3218, 45th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval - SIGIR 2022, Madrid, Spain, 11.07.22. https://doi.org/10.1145/3477495.3531751, https://doi.org/10.48550/arxiv.2205.06573

APA

Jiang, L., & Usbeck, R. (2022). Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research? In E. Amigo, P. Castells, & J. Gonzalo (Eds.), SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval (pp. 3209-3218). (Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval; Vol. 2022). Association for Computing Machinery, Inc. https://doi.org/10.1145/3477495.3531751, https://doi.org/10.48550/arxiv.2205.06573

Vancouver

Jiang L, Usbeck R. Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research? In Amigo E, Castells P, Gonzalo J, editors, SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval. New York: Association for Computing Machinery, Inc. 2022. p. 3209-3218. (Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval). doi: 10.1145/3477495.3531751, 10.48550/arxiv.2205.06573

Bibtex

@inbook{1747a699866c4ae3adc29801090597ba,

title = "Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research?",

abstract = "Existing approaches on Question Answering over Knowledge Graphs (KGQA) have weak generalizability. That is often due to the standard i.i.d. assumption on the underlying dataset. Recently, three levels of generalization for KGQA were defined, namely i.i.d., compositional, zero-shot. We analyze 25 well-known KGQA datasets for 5 different Knowledge Graphs (KGs). We show that according to this definition many existing and online available KGQA datasets are either not suited to train a generalizable KGQA system or that the datasets are based on discontinued and out-dated KGs. Generating new datasets is a costly process and, thus, is not an alternative to smaller research groups and companies. In this work, we propose a mitigation method for re-splitting available KGQA datasets to enable their applicability to evaluate generalization, without any cost and manual effort. We test our hypothesis on three KGQA datasets, i.e., LC-QuAD, LC-QuAD 2.0 and QALD-9). Experiments on re-splitted KGQA datasets demonstrate its effectiveness towards generalizability. The code and a unified way to access 18 available datasets is online at https: //github.com/semantic-systems/KGQA-datasets as well as https: //github.com/semantic-systems/KGQA-datasets-generalization.",

keywords = "benchmark, evaluation, generalizability, generalization, kgqa, question answering, Informatics, Business informatics",

author = "Longquan Jiang and Ricardo Usbeck",

note = "Publisher Copyright: {\textcopyright} 2022 ACM.; 45th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval - SIGIR 2022, ACM SIGIR 2022 ; Conference date: 11-07-2022 Through 15-07-2022",

year = "2022",

month = jul,

day = "7",

doi = "10.1145/3477495.3531751",

language = "English",

series = "Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval",

publisher = "Association for Computing Machinery, Inc",

pages = "3209--3218",

editor = "Enrique Amigo and Pablo Castells and Julio Gonzalo",

booktitle = "SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval",

address = "United States",

url = "https://sigir.org/sigir2022/",

}

RIS

TY - CHAP

T1 - Knowledge Graph Question Answering Datasets and Their Generalizability

T2 - 45th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval - SIGIR 2022

AU - Jiang, Longquan

AU - Usbeck, Ricardo

N1 - Conference code: 45

PY - 2022/7/7

Y1 - 2022/7/7

N2 - Existing approaches on Question Answering over Knowledge Graphs (KGQA) have weak generalizability. That is often due to the standard i.i.d. assumption on the underlying dataset. Recently, three levels of generalization for KGQA were defined, namely i.i.d., compositional, zero-shot. We analyze 25 well-known KGQA datasets for 5 different Knowledge Graphs (KGs). We show that according to this definition many existing and online available KGQA datasets are either not suited to train a generalizable KGQA system or that the datasets are based on discontinued and out-dated KGs. Generating new datasets is a costly process and, thus, is not an alternative to smaller research groups and companies. In this work, we propose a mitigation method for re-splitting available KGQA datasets to enable their applicability to evaluate generalization, without any cost and manual effort. We test our hypothesis on three KGQA datasets, i.e., LC-QuAD, LC-QuAD 2.0 and QALD-9). Experiments on re-splitted KGQA datasets demonstrate its effectiveness towards generalizability. The code and a unified way to access 18 available datasets is online at https: //github.com/semantic-systems/KGQA-datasets as well as https: //github.com/semantic-systems/KGQA-datasets-generalization.

AB - Existing approaches on Question Answering over Knowledge Graphs (KGQA) have weak generalizability. That is often due to the standard i.i.d. assumption on the underlying dataset. Recently, three levels of generalization for KGQA were defined, namely i.i.d., compositional, zero-shot. We analyze 25 well-known KGQA datasets for 5 different Knowledge Graphs (KGs). We show that according to this definition many existing and online available KGQA datasets are either not suited to train a generalizable KGQA system or that the datasets are based on discontinued and out-dated KGs. Generating new datasets is a costly process and, thus, is not an alternative to smaller research groups and companies. In this work, we propose a mitigation method for re-splitting available KGQA datasets to enable their applicability to evaluate generalization, without any cost and manual effort. We test our hypothesis on three KGQA datasets, i.e., LC-QuAD, LC-QuAD 2.0 and QALD-9). Experiments on re-splitted KGQA datasets demonstrate its effectiveness towards generalizability. The code and a unified way to access 18 available datasets is online at https: //github.com/semantic-systems/KGQA-datasets as well as https: //github.com/semantic-systems/KGQA-datasets-generalization.

KW - benchmark

KW - evaluation

KW - generalizability

KW - generalization

KW - kgqa

KW - question answering

KW - Informatics

KW - Business informatics

UR - http://www.scopus.com/inward/record.url?scp=85135050347&partnerID=8YFLogxK

UR - https://www.mendeley.com/catalogue/b6989f93-815e-39f3-8194-cc48d3490cd6/

U2 - 10.1145/3477495.3531751

DO - 10.1145/3477495.3531751

M3 - Article in conference proceedings

AN - SCOPUS:85135050347

T3 - Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval

SP - 3209

EP - 3218

BT - SIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval

A2 - Amigo, Enrique

A2 - Castells, Pablo

A2 - Gonzalo, Julio

PB - Association for Computing Machinery, Inc

CY - New York

Y2 - 11 July 2022 through 15 July 2022

ER -

Other publications by the same author(s)

ShortPathQA: A Dataset for Controllable Fusion of Large Language Models with Knowledge Graphs

Salnikov, M., Sakhovskiy, A., Nikishina, I., Usmanova, A., Kraft, A., Möller, C., Banerjee, D., Huang, J., Jiang, L., Abdullah, R., Yan, X., Tutubalina, E., Usbeck, R. & Panchenko, A., 2026, Natural Language Processing and Information Systems: 30th International Conference on Applications of Natural Language to Information Systems, NLDB 2025, Proceedings. Ichise, R. (ed.). Springer Science and Business Media Deutschland, p. 95-110 16 p. (Lecture Notes in Computer Science; vol. 15836 LNCS).

Research output: Contributions to collected editions/works › Article in conference proceedings › Research › peer-review

Analyzing the Influence of Knowledge Graph Information on Relation Extraction

Möller, C. & Usbeck, R., 2025, The Semantic Web: 22nd European Semantic Web Conference, ESWC 2025 Portoroz, Slovenia, June 1–5, 2025 Proceedings, Part I. Curry, E., Acosta, M., Poveda-Villalón, M., van Erp, M., Ojo, A., Hose, K., Shimizu, C. & Lisena, P. (eds.). Cham: Springer Nature Switzerland AG, Vol. 1. p. 460-480 21 p. (Lecture Notes in Computer Science ; vol. 15718).

Research output: Contributions to collected editions/works › Article in conference proceedings › Research › peer-review

Automating SPARQL Query Translations between DBpedia and Wikidata

Bartels, M. C., Banerjee, D. & Usbeck, R., 14.07.2025, SEMANTiCS Conference 2025.

Research output: Contributions to collected editions/works › Article in conference proceedings › Research

Bridge-Generate: Scholarly Hybrid Question Answering

Taffa, T. A. & Usbeck, R., 23.05.2025, WWW Companion 2025 - Companion Proceedings of the ACM Web Conference 2025: Companion Proceedings of the ACM Web Conference 2025, April 28-May 2, 2025 Sydney, NSW, Australia. Long, G., Blumestein, M., Chang, Y., Lewin-Eytan, L., Huang, H. & Yom-Tov, E. (eds.). New York: Association for Computing Machinery, Inc, p. 1321-1325 5 p.

Research output: Contributions to collected editions/works › Article in conference proceedings › Research › peer-review

Junior fellows and distinguished dissertation of the GI and AI for crisis

Usbeck, R., Kraft, A. & Westphal, P., 01.02.2025, In: IT - Information Technology. 67, 1, p. 1-2 2 p.

Research output: Journal contributions › Other (editorial matter etc.) › Research

DOI

https://doi.org/10.1145/3477495.3531751
Final published version
https://doi.org/10.48550/arxiv.2205.06573
Other version

Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research?

Standard

Harvard

APA

Vancouver

Bibtex

RIS

Other publications by the same author(s)

ShortPathQA: A Dataset for Controllable Fusion of Large Language Models with Knowledge Graphs

Analyzing the Influence of Knowledge Graph Information on Relation Extraction

Automating SPARQL Query Translations between DBpedia and Wikidata

Bridge-Generate: Scholarly Hybrid Question Answering

Junior fellows and distinguished dissertation of the GI and AI for crisis

DOI

Recently viewed

Researchers

Projects

Activities

Prizes

Publications

Press / Media