A DOM-based Web Information Extraction
Xiong Xuan-dong
Abstract
Xiong Xuan-dong
Abstract
This paper proposes a DOM-based Web Information Extraction solution,to get the location path of extracted infor- mation byinduction study,to edit the extraction pattern through characteristics of XPath and XSLT on data location and data transform,and according to the mapping relation between Web page's elements and DOM's nodes,to judge if the gained infor- mation sources are the same as that of the generated extraction pattern.
A significance statement is not available in the OpenAlex record.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
This paper proposes a DOM-based Web Information Extraction solution,to get the location path of extracted infor- mation byinduction study,to edit the extraction pattern through characteristics of XPath and XSLT on data location and data transform,and according to the mapping relation between Web page's elements and DOM's nodes,to judge if the gained infor- mation sources are the same as that of the generated extraction pattern.
Key concepts: XPath, Computer science, Document Object Model, XSLT, Relationship extraction, Path (computing), Information extraction, Data extraction