Real-time RDF extraction from unstructured data streams

Publikation: Beiträge in SammelwerkenAufsätze in KonferenzbändenForschungbegutachtet

Authors

  • Daniel Gerber
  • Sebastian Hellmann
  • Lorenz Bühmann
  • Tommaso Soru
  • Ricardo Usbeck
  • Axel Cyrille Ngonga Ngomo

The vision behind the Web of Data is to extend the current document-oriented Web with machine-readable facts and structured data, thus creating a representation of general knowledge. However, most of the Web of Data is limited to being a large compendium of encyclopedic knowledge describing entities. A huge challenge, the timely and massive extraction of RDF facts from unstructured data, has remained open so far. The availability of such knowledge on the Web of Data would provide significant benefits to manifold applications including news retrieval, sentiment analysis and business intelligence. In this paper, we address the problem of the actuality of the Web of Data by presenting an approach that allows extracting RDF triples from unstructured data streams. We employ statistical methods in combination with deduplication, disambiguation and unsupervised as well as supervised machine learning techniques to create a knowledge base that reflects the content of the input streams. We evaluate a sample of the RDF we generate against a large corpus of news streams and show that we achieve a precision of more than 85%.

OriginalspracheEnglisch
TitelThe Semantic Web, ISWC 2013 : 12th International Semantic Web Conference, Proceedings
HerausgeberHarith Alani, Lalana Kagal, Achille Fokoue, Paul Groth, Chris Biemann, Josiane Xavier Parreira, Lora Aroyo, Natasha Noy, Chris Welty, Krzyztof Janowicz
Anzahl der Seiten16
VerlagSpringer Verlag
Erscheinungsdatum2013
Seiten135-150
ISBN (Print)9783642413346
DOIs
PublikationsstatusErschienen - 2013
Extern publiziertJa
Veranstaltung12th International Semantic Web Conference, ISWC 2013 - Sydney Convention Centre , Sydney, NSW, Australien
Dauer: 21.10.201325.10.2013
http://iswc2013.semanticweb.org

DOI

Zuletzt angesehen

Publikationen

  1. What does it mean to be sensitive for the complexity of (problem oriented) teaching?
  2. Age effects on controlling tools with sensorimotor transformations
  3. Considerations on efficient touch interfaces - How display size influences the performance in an applied pointing task
  4. An analytical approach to evaluating bivariate functions of fuzzy numbers with one local extremum
  5. Explaining and controlling for the psychometric properties of computer-generated figural matrix items
  6. Foundations and applications of computer based material flow networks for einvironmental management
  7. Robust feedback linearization control of a throttle plate by using an approximated pd regulator
  8. On the Decoupling and Output Functional Controllability of Robotic Manipulation
  9. Integration of laser scanning and projection speckle pattern for advanced pipeline monitoring
  10. Artificial Intelligence Algorithms for Collaborative Book Recommender Systems
  11. Partitioned beta diversity patterns of plants across sharp and distinct boundaries of quartz habitat islands
  12. Using Fuzzy PD Controllers for Soft Motions in a Car-like Robot
  13. Switching from a Managing to a Monitoring Function on the Board
  14. The fuzzy relationship of intelligence and problem solving in computer simulations
  15. Performance concepts and performance theory
  16. A Structure and Content Prompt-based Method for Knowledge Graph Question Answering over Scholarly Data
  17. Changes of Perception
  18. Digging into the roots
  19. For a return to the forgotten formula: 'Data 1 + Data 2 > Data 1'
  20. Errors in Training Computer Skills
  21. Using augmented video to test in-car user experiences of context analog HUDs
  22. GENESIS - A generic RDF data access interface
  23. Factor structure and measurement invariance of the Students’ Self-report Checklist of Social and Learning Behaviour (SSL)
  24. Model predictive control for switching gain adaptation in a sliding mode controller of a DC drive with nonlinear friction
  25. Semantic Evaluation Services for Web-Based Exercises
  26. More input, better output
  27. How Much Home Office is Ideal? A Multi-Perspective Algorithm
  28. Optimizing price levels in e-commerce applications with respect to customer lifetime values
  29. Correlation of Microstructure and Local Mechanical Properties Along Build Direction for Multi-layer Friction Surfacing of Aluminum Alloys
  30. Emergency detection based on probabilistic modeling in AAL-environments
  31. Sliding-Mode-Based Input-Output Linearization of a Peltier Element for Ice Clamping Using a State and Disturbance Observer
  32. A general structural property in wavelet packets for detecting oscillation and noise components in signal analysis
  33. Eighth Workshop on Mining and Learning with Graphs
  34. Applied quality assurance methods under the open source development model