EconPapers    
Economics at your fingertips  
 

Improving Curated Web-Data Quality with Structured Harvesting and Assessment

Kevin Chekov Feeney, Declan O'Sullivan, Wei Tai and Rob Brennan
Additional contact information
Kevin Chekov Feeney: Trinity College Dublin, Dublin, Ireland
Declan O'Sullivan: Trinity College Dublin, Dublin, Ireland
Wei Tai: Trinity College Dublin, Dublin, Ireland
Rob Brennan: Trinity College Dublin, Dublin, Ireland

International Journal on Semantic Web and Information Systems (IJSWIS), 2014, vol. 10, issue 2, 35-62

Abstract: This paper describes a semi-automated process, framework and tools for harvesting, assessing, improving and maintaining high-quality linked-data. The framework, known as DaCura1, provides dataset curators, who may not be knowledge engineers, with tools to collect and curate evolving linked data datasets that maintain quality over time. The framework encompasses a novel process, workflow and architecture. A working implementation has been produced and applied firstly to the publication of an existing social-sciences dataset, then to the harvesting and curation of a related dataset from an unstructured data-source. The framework's performance is evaluated using data quality measures that have been developed to measure existing published datasets. An analysis of the framework against these dimensions demonstrates that it addresses a broad range of real-world data quality concerns. Experimental results quantify the impact of the DaCura process and tools on data quality through an assessment framework and methodology which combines automated and human data quality controls.

Date: 2014
References: Add references at CitEc
Citations:

Downloads: (external link)
http://services.igi-global.com/resolvedoi/resolve. ... 18/ijswis.2014040103 (application/pdf)

Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.

Export reference: BibTeX RIS (EndNote, ProCite, RefMan) HTML/Text

Persistent link: https://EconPapers.repec.org/RePEc:igg:jswis0:v:10:y:2014:i:2:p:35-62

Access Statistics for this article

International Journal on Semantic Web and Information Systems (IJSWIS) is currently edited by Brij Gupta

More articles in International Journal on Semantic Web and Information Systems (IJSWIS) from IGI Global
Bibliographic data for series maintained by Journal Editor ().

 
Page updated 2025-03-19
Handle: RePEc:igg:jswis0:v:10:y:2014:i:2:p:35-62