2019Unpublished venueRequires access

Office Document Search Engine

Diki Ardian Wirasandi, Saiful Akbar, Fitra Arifiansyah

Open publisher page 0 citations

Abstract

Search engine have been widely used to find some documents for many reasons. One of frequently used kind of document is office document. Office document is classified as semi-structured document because sometimes they have consistent structure in a document category. Office document also has various categories and formats. To build a search engine, there are two main processes that must be implemented. Those processes are indexing process and query process. Every process consists of some methods that has some function for each of them. Not all kind of methods can be used and implemented for that processes. A suitable method needs to be selected in order to produce an optimal search engine for a specific defined domain. This paper will explain how to recognize office document's pattern that will be used to build a search engine. It will also explain about selection of methods that were used to build an optimal search engine in office document domain. This search engine will be evaluated with some testing scenario to calculate its precision for some queries and to know how optimal it is. This proposed search engine more focus on having effective result than efficiency of processing. However, the evaluation still covers both of effectiveness and efficiency of the system.

About this research paper

What this paper is about

Search engine have been widely used to find some documents for many reasons. One of frequently used kind of document is office document. Office document is classified as semi-structured document because sometimes they have consistent structure in a document category. Office document also has various categories and formats. To build a search engine, there are two main processes that must be implemented. Those processes are indexing process and query process. Every process consists of some methods that has some function for each of them. Not all kind of methods can be used and implemented for that processes. A suitable method needs to be selected in order to produce an optimal search engine for a specific defined domain. This paper will explain how to recognize office document's pattern that will be used to build a search engine. It will also explain about selection of methods that were used to build an optimal search engine in office document domain. This search engine will be evaluated with some testing scenario to calculate its precision for some queries and to know how optimal it is. This proposed search engine more focus on having effective result than efficiency of processing. However, the evaluation still covers both of effectiveness and efficiency of the system.

Why it matters

A significance statement is not available in the OpenAlex record.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Search engine have been widely used to find some documents for many reasons. One of frequently used kind of document is office document. Office document is classified as semi-structured document because sometimes they have consistent structure in a document category. Office document also has various categories and formats. To build a search engine, there are two main processes that must be implemented. Those processes are indexing process and query process. Every process consists of some methods that has some function for each of them. Not all kind of methods can be used and implemented for that processes. A suitable method needs to be selected in order to produce an optimal search engine for a specific defined domain. This paper will explain how to recognize office document's pattern that will be used to build a search engine. It will also explain about selection of methods that were used to build an optimal search engine in office document domain. This search engine will be evaluated with some testing scenario to calculate its precision for some queries and to know how optimal it is. This proposed search engine more focus on having effective result than efficiency of processing. However, the evaluation still covers both of effectiveness and efficiency of the system.

Key concepts: Search engine indexing, Computer science, Search engine, Information retrieval, Search analytics, Process (computing), Spamdexing, Metasearch engine

Related papers

Back to paper searchBrowse research topicsOriginal source
Office Document Search Engine — Research Paper | ScholarLens