Real-time RDF extraction from unstructured data streams

Research output: Contributions to collected editions/worksArticle in conference proceedingsResearchpeer-review

Authors

  • Daniel Gerber
  • Sebastian Hellmann
  • Lorenz Bühmann
  • Tommaso Soru
  • Ricardo Usbeck
  • Axel Cyrille Ngonga Ngomo

The vision behind the Web of Data is to extend the current document-oriented Web with machine-readable facts and structured data, thus creating a representation of general knowledge. However, most of the Web of Data is limited to being a large compendium of encyclopedic knowledge describing entities. A huge challenge, the timely and massive extraction of RDF facts from unstructured data, has remained open so far. The availability of such knowledge on the Web of Data would provide significant benefits to manifold applications including news retrieval, sentiment analysis and business intelligence. In this paper, we address the problem of the actuality of the Web of Data by presenting an approach that allows extracting RDF triples from unstructured data streams. We employ statistical methods in combination with deduplication, disambiguation and unsupervised as well as supervised machine learning techniques to create a knowledge base that reflects the content of the input streams. We evaluate a sample of the RDF we generate against a large corpus of news streams and show that we achieve a precision of more than 85%.

Original languageEnglish
Title of host publicationThe Semantic Web, ISWC 2013 : 12th International Semantic Web Conference, Proceedings
EditorsHarith Alani, Lalana Kagal, Achille Fokoue, Paul Groth, Chris Biemann, Josiane Xavier Parreira, Lora Aroyo, Natasha Noy, Chris Welty, Krzyztof Janowicz
Number of pages16
PublisherSpringer
Publication date2013
Pages135-150
ISBN (print)9783642413346
DOIs
Publication statusPublished - 2013
Externally publishedYes
Event12th International Semantic Web Conference, ISWC 2013 - Sydney Convention Centre , Sydney, NSW, Australia
Duration: 21.10.201325.10.2013
http://iswc2013.semanticweb.org

Recently viewed

Publications

  1. Modeling items for text comprehension assessment using confirmatory factor analysis
  2. Effectiveness of a guided multicomponent internet and mobile gratitude training program - A pragmatic randomized controlled trial
  3. A statistical study of the spatial evolution of shock acceleration efficiency for 5 MeV protons and subsequent particle propagation
  4. Four Methods to Distinguish between Fractal Dimensions in Time Series through Recurrence Quantification Analysis
  5. Comparing the performance of computational estimation methods for physicochemical properties of dimethylsiloxanes and selected siloxanols
  6. Tree diversity increases forest temperature buffering via enhancing canopy density and structural diversity
  7. Stepwise-based optimizing approaches for arrangements of loudspeaker in multi-zone sound field reproduction
  8. A Review of the Application of Machine Learning and Data Mining Approaches in Continuum Materials Mechanics
  9. Formative Perspectives on the Relation Between CSR Communication and CSR Practices
  10. Sensor Fusion for Power Line Sensitive Monitoring and Load State Estimation
  11. Experimentally established correlation of friction surfacing process temperature and deposit geometry
  12. Changes in the Complexity of Limb Movements during the First Year of Life across Different Tasks
  13. Neural network-based estimation and compensation of friction for enhanced deep drawing process control
  14. Does thinking-aloud affect learning, visual information processing and cognitive load when learning with seductive details as expected from self-regulation perspective?
  15. Privatizing the commons
  16. Tree diversity and mycorrhizal type co-determine multitrophic ecosystem functions
  17. The temporal and spatial development of MeV proton acceleration at interplanetary shocks
  18. Exploring the limits of graph invariant- and spectrum-based discrimination of (sub)structures.
  19. Effects of diversity versus segregation on automatic approach and avoidance behavior towards own and other ethnic groups