International Journal of Science and Research (IJSR)

International Journal of Science and Research (IJSR)
Call for Papers | Fully Refereed | Open Access | Double Blind Peer Reviewed

ISSN: 2319-7064


Downloads: 121 | Views: 247

Research Paper | Computer Science & Engineering | India | Volume 4 Issue 7, July 2015 | Popularity: 6.9 / 10


     

Data Extraction and Annotation Methods Using Tag Value Structure

Tushar Jadhav, Santosh Chobe


Abstract: The world wide web generates search result pages which is based. On the users input query. It is very crucial for many applications like data integration which requires combining more databases to automatically extract the data from the search results. A unique method for extracting the data and then aligning is implemented which uses Unsupervised duplicate detection algorithm which identifies and segments the result records first and then aligns the segmented results in a table, in which data values of similar attributes are put in same column. The new technique is implemented so as to handle the case when the search results are not adjoining which might happen because of auxiliary data such as advertisements, comments etc. and also to handle nested tag structure which might be present in the search results. The results shows that the implemented algorithm performs well than existing methods.


Keywords: data extraction, data annotation, data alignment, wrapper generation


Edition: Volume 4 Issue 7, July 2015


Pages: 1968 - 1972




Text copied to Clipboard!
Tushar Jadhav, Santosh Chobe, "Data Extraction and Annotation Methods Using Tag Value Structure", International Journal of Science and Research (IJSR), Volume 4 Issue 7, July 2015, pp. 1968-1972, https://www.ijsr.net/getabstract.php?paperid=SUB156670

Similar Articles

Downloads: 97

Review Papers, Computer Science & Engineering, India, Volume 3 Issue 11, November 2014

Pages: 1191 - 1194

Web Data Extraction by Using Trinity

Sayali Khodade, Nilav Mukharjee

Share this Article

Downloads: 101

Research Paper, Computer Science & Engineering, India, Volume 4 Issue 11, November 2015

Pages: 1579 - 1582

Data Hiding in H.264/AVC Video Encryption with XOR-ed User Information and Data in File Format

Neenu Shereef

Share this Article

Downloads: 107

Survey Paper, Computer Science & Engineering, India, Volume 4 Issue 10, October 2015

Pages: 1434 - 1436

Method for Repossession of Content Based Video using Speech and Text Information

Manasi A. Kabade, U.A. Jogalekar

Share this Article

Downloads: 108

Survey Paper, Computer Science & Engineering, India, Volume 3 Issue 11, November 2014

Pages: 1152 - 1154

A Survey on Content based Video Retrieval Using Speech and Text information

Laxmikant S. Kate, M. M. Waghmare

Share this Article

Downloads: 109

Survey Paper, Computer Science & Engineering, India, Volume 3 Issue 11, November 2014

Pages: 2425 - 2428

Survey of Various Techniques on Cheating Prevention in Visual Cryptography with Steganography Scheme

Sneha A.Deshmukh, P.B.Sambhare

Share this Article
Top