Downloads: 121 | Views: 278
Research Paper | Computer Science & Engineering | India | Volume 4 Issue 7, July 2015 | Popularity: 6.9 / 10
Data Extraction and Annotation Methods Using Tag Value Structure
Tushar Jadhav, Santosh Chobe
Abstract: The world wide web generates search result pages which is based. On the users input query. It is very crucial for many applications like data integration which requires combining more databases to automatically extract the data from the search results. A unique method for extracting the data and then aligning is implemented which uses Unsupervised duplicate detection algorithm which identifies and segments the result records first and then aligns the segmented results in a table, in which data values of similar attributes are put in same column. The new technique is implemented so as to handle the case when the search results are not adjoining which might happen because of auxiliary data such as advertisements, comments etc. and also to handle nested tag structure which might be present in the search results. The results shows that the implemented algorithm performs well than existing methods.
Keywords: data extraction, data annotation, data alignment, wrapper generation
Edition: Volume 4 Issue 7, July 2015
Pages: 1968 - 1972
Make Sure to Disable the Pop-Up Blocker of Web Browser
Similar Articles
Downloads: 97
Review Papers, Computer Science & Engineering, India, Volume 3 Issue 11, November 2014
Pages: 1191 - 1194Web Data Extraction by Using Trinity
Sayali Khodade, Nilav Mukharjee
Downloads: 101
Research Paper, Computer Science & Engineering, India, Volume 4 Issue 11, November 2015
Pages: 1579 - 1582Data Hiding in H.264/AVC Video Encryption with XOR-ed User Information and Data in File Format
Neenu Shereef
Downloads: 107 | Weekly Hits: ⮙1 | Monthly Hits: ⮙1
M.Tech / M.E / PhD Thesis, Computer Science & Engineering, India, Volume 4 Issue 2, February 2015
Pages: 1282 - 1284Document Annotation Based on Query Workload, Content-Value and User Expectation Tracking Form
Alfia A P, Chashu Mol R
Downloads: 107
Survey Paper, Computer Science & Engineering, India, Volume 4 Issue 10, October 2015
Pages: 1434 - 1436Method for Repossession of Content Based Video using Speech and Text Information
Manasi A. Kabade, U.A. Jogalekar
Downloads: 108
Survey Paper, Computer Science & Engineering, India, Volume 3 Issue 11, November 2014
Pages: 1152 - 1154A Survey on Content based Video Retrieval Using Speech and Text information
Laxmikant S. Kate, M. M. Waghmare