Populating a Linked Data Entity Name System

Populating a Linked Data Entity Name System
Author :
Publisher : IOS Press
Total Pages : 190
Release :
ISBN-10 : 9781614996927
ISBN-13 : 161499692X
Rating : 4/5 (27 Downloads)

Book Synopsis Populating a Linked Data Entity Name System by : M. Kejriwal

Download or read book Populating a Linked Data Entity Name System written by M. Kejriwal and published by IOS Press. This book was released on 2016-12-09 with total page 190 pages. Available in PDF, EPUB and Kindle. Book excerpt: Resource Description Framework (RDF) is a graph-based data model used to publish data as a Web of Linked Data. RDF is an emergent foundation for large-scale data integration, the problem of providing a unified view over multiple data sources. An Entity Name System (ENS) is a thesaurus for entities, and is a crucial component in a data integration architecture. Populating a Linked Data ENS is equivalent to solving an Artificial Intelligence problem called instance matching, which concerns identifying pairs of entities referring to the same underlying entity. This publication presents an instance matcher with 4 properties, namely automation, heterogeneity, scalability and domain independence. Automation is addressed by employing inexpensive but well-performing heuristics to automatically generate a training set, which is employed by other machine learning algorithms in the pipeline. Data-driven alignment algorithms are adapted to deal with structural heterogeneity in RDF graphs. Domain independence is established by actively avoiding prior assumptions about input domains, and through evaluations on 10 RDF test cases. The full system is scaled by implementing it on cloud infrastructure using MapReduce algorithms. Resource Description Framework (RDF) is a graph-based data model used to publish data as a Web of Linked Data. RDF is an emergent foundation for large-scale data integration, the problem of providing a unified view over multiple data sources. An Entity Name System (ENS) is a thesaurus for entities, and is a crucial component in a data integration architecture. Populating a Linked Data ENS is equivalent to solving an Artificial Intelligence problem called instance matching, which concerns identifying pairs of entities referring to the same underlying entity. This publication presents an instance matcher with 4 properties, namely automation, heterogeneity, scalability and domain independence. Automation is addressed by employing inexpensive but well-performing heuristics to automatically generate a training set, which is employed by other machine learning algorithms in the pipeline. Data-driven alignment algorithms are adapted to deal with structural heterogeneity in RDF graphs. Domain independence is established by actively avoiding prior assumptions about input domains, and through evaluations on 10 RDF test cases. The full system is scaled by implementing it on cloud infrastructure using MapReduce algorithms.


Populating a Linked Data Entity Name System Related Books

Populating a Linked Data Entity Name System
Language: en
Pages: 190
Authors: M. Kejriwal
Categories: Computers
Type: BOOK - Published: 2016-12-09 - Publisher: IOS Press

DOWNLOAD EBOOK

Resource Description Framework (RDF) is a graph-based data model used to publish data as a Web of Linked Data. RDF is an emergent foundation for large-scale dat
Populating a Linked Data Entity Name System
Language: en
Pages:
Authors: Mayank Kejriwal
Categories:
Type: BOOK - Published: 2016 - Publisher:

DOWNLOAD EBOOK

Knowledge Graphs
Language: en
Pages: 559
Authors: Mayank Kejriwal
Categories: Computers
Type: BOOK - Published: 2021-03-30 - Publisher: MIT Press

DOWNLOAD EBOOK

A rigorous and comprehensive textbook covering the major approaches to knowledge graphs, an active and interdisciplinary area within artificial intelligence. Th
Populating a Linked Data Entity Name System
Language: en
Pages: 0
Authors: Mayank Kejriwal
Categories:
Type: BOOK - Published: 2016 - Publisher:

DOWNLOAD EBOOK

Resource Description Framework (RDF) is a graph-based data model used to publish data as a Web of Linked Data. RDF is an emergent foundation for large-scale dat
Identity of Long-tail Entities in Text
Language: en
Pages: 229
Authors: F. Ilievski
Categories: Computers
Type: BOOK - Published: 2019-11-29 - Publisher: IOS Press

DOWNLOAD EBOOK

The digital era has generated a huge amount of data on the identities (profiles) of people, organizations and other entities in a digital format, largely consis