Knowledge Graph Question Answering Datasets and Their Generalizability: Are They Enough for Future Research?

Publikation: Beiträge in SammelwerkenAufsätze in KonferenzbändenForschungbegutachtet

Authors

Existing approaches on Question Answering over Knowledge Graphs (KGQA) have weak generalizability. That is often due to the standard i.i.d. assumption on the underlying dataset. Recently, three levels of generalization for KGQA were defined, namely i.i.d., compositional, zero-shot. We analyze 25 well-known KGQA datasets for 5 different Knowledge Graphs (KGs). We show that according to this definition many existing and online available KGQA datasets are either not suited to train a generalizable KGQA system or that the datasets are based on discontinued and out-dated KGs. Generating new datasets is a costly process and, thus, is not an alternative to smaller research groups and companies. In this work, we propose a mitigation method for re-splitting available KGQA datasets to enable their applicability to evaluate generalization, without any cost and manual effort. We test our hypothesis on three KGQA datasets, i.e., LC-QuAD, LC-QuAD 2.0 and QALD-9). Experiments on re-splitted KGQA datasets demonstrate its effectiveness towards generalizability. The code and a unified way to access 18 available datasets is online at https: //github.com/semantic-systems/KGQA-datasets as well as https: //github.com/semantic-systems/KGQA-datasets-generalization.

OriginalspracheEnglisch
TitelSIGIR 2022 - Proceedings of the 45th International ACM SIGIR Conference on Research and Development in Information Retrieval
HerausgeberEnrique Amigo, Pablo Castells, Julio Gonzalo
Anzahl der Seiten10
ErscheinungsortNew York
VerlagAssociation for Computing Machinery, Inc
Erscheinungsdatum07.07.2022
Seiten3209-3218
ISBN (elektronisch)9781450387323
DOIs
PublikationsstatusErschienen - 07.07.2022
Extern publiziertJa
Veranstaltung45th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval - SIGIR 2022 - Online + Círculo de Bellas Artes (Circle of Beaux Arts), Madrid, Spanien
Dauer: 11.07.202215.07.2022
Konferenznummer: 45
https://sigir.org/sigir2022/

Bibliographische Notiz

Funding Information:
The authors acknowledge the financial support by the Federal Ministry for Economic Affairs and Climate Action of Germany in the project CoyPu (project number 01MK21007G).

Publisher Copyright:
© 2022 ACM.

Zuletzt angesehen

Forschende

  1. Jana Hüttmann

Publikationen

  1. Frame-based Optimal Design
  2. Mechanics of sheet-bulk indentation
  3. Auditors' Perceptions of Client Firms
  4. Team Ambidexterity and its Prerequisites: An Exploratory Study of an IT Service Management Team
  5. Nitrate Pollution of Groundwater Long Exceeding Trigger Value
  6. Integration of expertise or collaborative practice?
  7. Finding the Best Match — a Case Study on the (Text‑) Feature and Model Choice in Digital Mental Health Interventions
  8. An Integrative and Comprehensive Methodology for Studying Aesthetic Experience in the Field
  9. Approaches and Lessons in Political Career Research
  10. Set oriented computation of transport rates in 3-degree of freedom systems
  11. Reconfiguring Desecuritization
  12. Competence-Oriented Teaching
  13. Collaborative business in supply chains - a system dynamics approach
  14. Organizational Practices for the Aging Workforce
  15. Backward Extended Kalman Filter to Estimate and Adaptively Control a PMSM in Saturation Conditions
  16. Control of Permanent Magnet Synchronous Motors for Track Applications
  17. How to specify the structure of substituted blade-like zigzag diamondoids
  18. State-wide university implementation of an online platform for eating disorders screening and intervention.
  19. A Model Based Feedforward Regulator Improving PI Control of an Ice-Clamping Device Activated by Thermoelectric Cooler
  20. Elution of monomers from provisional composite materials
  21. Learning Strategies of First Year University Students
  22. Ownership Patterns and Enterprise Groups in German Structural Business Statistics
  23. Exploring the Capacity of Water Framework Directive Indices to Assess Ecosystem Services in Fluvial and Riparian Systems
  24. Towards a Sustainable Use of Phosphorus
  25. If you call for frameworks in sustainability management... editorial to the special issue
  26. Effects of strategy instructions on learning from text and pictures
  27. Is fairness intuitive? An experiment accounting for subjective utility differences under time pressure