• Title/Summary/Keyword: Web search

Search Result 1,666, Processing Time 0.026 seconds

Clustering of Web Objects with Similar Popularity Trends (유사한 인기도 추세를 갖는 웹 객체들의 클러스터링)

  • Loh, Woong-Kee
    • The KIPS Transactions:PartD
    • /
    • v.15D no.4
    • /
    • pp.485-494
    • /
    • 2008
  • Huge amounts of various web items such as keywords, images, and web pages are being made widely available on the Web. The popularities of such web items continuously change over time, and mining temporal patterns in popularities of web items is an important problem that is useful for several web applications. For example, the temporal patterns in popularities of search keywords help web search enterprises predict future popular keywords, enabling them to make price decisions when marketing search keywords to advertisers. However, presence of millions of web items makes it difficult to scale up previous techniques for this problem. This paper proposes an efficient method for mining temporal patterns in popularities of web items. We treat the popularities of web items as time-series, and propose gapmeasure to quantify the similarity between the popularities of two web items. To reduce the computation overhead for this measure, an efficient method using the Fast Fourier Transform (FFT) is presented. We assume that the popularities of web items are not necessarily following any probabilistic distribution or periodic. For finding clusters of web items with similar popularity trends, we propose to use a density-based clustering algorithm based on the gap measure. Our experiments using the popularity trends of search keywords obtained from the Google Trends web site illustrate the scalability and usefulness of the proposed approach in real-world applications.

PSR: Pre-Computing Solutions in RDBMS for Efficient Web Services Composition Search (PSR : 효율적인 웹 서비스 컴포지션 검색을 위한 RDBMS 기반의 선 계산 기법)

  • Kwon, Joon-Ho;Park, Kyu-Ho;Lee, Dae-Wook;Lee, Suk-Ho
    • Journal of KIISE:Databases
    • /
    • v.35 no.4
    • /
    • pp.333-344
    • /
    • 2008
  • In recent years, the web services composition has received much attention. By web services composition, we mean providing a new service that does not exist on the repository. In this paper, we propose a new system called PSR for web services composition search using a relational database. We also propose algorithms for pre-computing web services composition using joins and indices. We store ontologies from web services in RDBMS, so that the PSR system returns web services composition in order of similarity with user query through the degree of the ontology matching. We demonstrated that our pre-computing web services composition approach in RDBMS yields lower execution time and good scalability when handling a large number of web services and user queries.

A Retrieval Technique of Personal Information in a Web Environment (웹 환경에서의 개인정보 검색기법)

  • Seo, Young-Duk;Chang, Jae-Young
    • The Journal of the Institute of Internet, Broadcasting and Communication
    • /
    • v.15 no.4
    • /
    • pp.145-151
    • /
    • 2015
  • Since we use internet every day, the internet privacy has become important. We need to find out what kinds of personal information is exposed to the internet and to eliminate the exposed information. However, it is not efficient to search the personal information using only fragmentary clues in web search engines because the ranking results are not relevant to the exposure degree of personal information. In this paper, we introduced a personal information retrieval system and proposed a process to remove private data from the web easily. We also compared our proposed method with previous methods by evaluating the search performance.

Design and Implementation of Customer Information Retrieval System based on Semantic Web (시맨틱 웹 기반의 고객 정보 검색 시스템의 설계 및 구현)

  • Hwang Jeong-Hee;Gu Mi-Sug;Lee Hyun-Ah;Ryu Keun-Ho
    • The KIPS Transactions:PartD
    • /
    • v.13D no.4 s.107
    • /
    • pp.525-534
    • /
    • 2006
  • Ontology specifies the knowledge in a specific domain and defines the concepts of knowledge and the relationships between concepts. It is possible to provide the service based on the semantic web through the ontology. Therefore, to specify and define the knowledge in a specific domain, it is required to generate the ontology which conceptualizes the knowledge. Accordingly, to search the information of potential customers for home-delivery marketing of post office, we design the specific domain to generate the ontology based on the semantic web in this paper. And we propose how to retrieve the information, using the generated ontology. We implement the data search robot which collects the information based on the generated ontology. Also, we confirm that the ontology and the search robot perform the information retrieval exactly.

Trends of Search Behavior of Korean Web Users (국내 웹 이용자의 검색 행태 추이 분석)

  • Park Soyeon;Lee Joon Ho
    • Journal of the Korean Society for Library and Information Science
    • /
    • v.39 no.2
    • /
    • pp.147-160
    • /
    • 2005
  • This study examines trends of web query types and topics submitted to NAVER during one year period by analyzing query logs and click logs. There was a seasonal difference in the distribution of query types. Query type distribution was also different between weekdays and weekends, and between different days of the week. The log data show seasonal changes in terms of the topics of queries. Search topics seem to change between weekdays and weekends, and between different days of the week. However, there was little change in overall patterns of search behavior across one year. The implications for system designers and web content providers are discussed.

Development of A Plagiarism Detection System Using Web Search and Morpheme Analysis (인터넷 검색과 형태소분석을 이용한 표절검사시스템의 개발에 관한 연구)

  • Hwang, In-Soo
    • Journal of Information Technology Applications and Management
    • /
    • v.16 no.1
    • /
    • pp.21-36
    • /
    • 2009
  • As the World Wide Web (WWW) has become a major channel for information delivery, the data accumulated in the Internet increases at an incredible speed, and it derives the advances of information search technologies. It is the search engine that solves the problem of information overloading and helps people to identify relevant information. However, as search engines become a powerful tool for finding information, the opportunities of plagiarizing have increased significantly in e-Learning. In this paper, we developed an online plagiarism detection system for detecting plagiarized documents that incorporates the functions of search engines and acts in exactly the same way of plagiarizing. The plagiarism detection system uses morpheme analysis to improve the performance and sentence-based comparison to investigate document comes from multiple sources. As a result of applying this system in e-Learning, the performance of plagiarism detection was improved.

  • PDF

Search and Visualization Method on the Semantic Web Portal (시맨틱 웹 포털에서의 검색과 시각화 방법 연구)

  • Lee, Myung-Jin;Lee, Ki-Jun;Park, Sang-Un;Hong, June-Seok;Kim, Woo-Ju
    • Proceedings of the Korea Database Society Conference
    • /
    • 2008.05a
    • /
    • pp.389-403
    • /
    • 2008
  • As the information of web dramatically increase, the existing web reveals more and more limitations in information search because web pages are designed only for human consumption by mixing content with presentation. In order to improve this situation, the Semantic Web comes on the stage by W3C. Semantic web is based on ontology that defines relationships between resources and it is enough to bring a significant advancement in web search. But to do this, the Semantic Web must provide a novel search and visualization methods which can make users instantly and intuitively understand why and how the results are retrieved because ontology has formal explicit descriptions of meaning. In this paper, we propose a semantic association-based search methodology that consists of how to find relevant information for a given user's query in the ontology, that is, a semantic network of resources and properties and how to provide proper visualization and navigation methods on the results. From this work, users can search the semantically associated resources for their query and also navigate such associations between resources.

  • PDF

Layout Analysis for Calculation of Web Page Similarity as Image

  • Mitsuhashi, Noriaki;Yamaguchi, Toru;Takama, Yasufumi
    • Proceedings of the Korean Institute of Intelligent Systems Conference
    • /
    • 2003.09a
    • /
    • pp.142-145
    • /
    • 2003
  • When we search information on the Web using search engines, they only analyze the text information collected from the source files of Web pages. However, there is a limit to analyze the layout of a Web page only from its source file, although Web page design is the most important factor for a user to estimate a page. In particular it often happens on the Web that the pages of similar design ofter similar information. We propose a method to analyze layout for comparing the design of pages by treating the displayed page as image.

  • PDF

A Study on LibraryLookup Services Using Bookmarklets (북마크릿을 활용한 LibraryLookup 서비스 제공방안에 관한 연구)

  • Gu, Jung-Eok;Lee, Eung-Bong
    • Journal of the Korean Society for information Management
    • /
    • v.23 no.3 s.61
    • /
    • pp.49-68
    • /
    • 2006
  • It is required to enhance the value of ISBN as a tool for book search, identification, browsing, and improve the accessability and search capability of library OPAC. Bookmarklet is a small size javascript which can be saved as URL in a web browser bookmark or web page hyperlink. Open source bookmarklet can extract ISBN from web pages and search a book from library OPAC using the ISBN, so it is recognized as a simple but powerful search tool. In foreign countries, commercial library system vendors, libraries, OCLC, etc. are providing bookmarklets which allow a user to search for library holdings and loan information in a real time while he/she is travelling in an online bookshop web page. Therefore, this paper compared and analyzed international bookmarklets application examples and proposed LibraryLookup service in which library OPAC and online bookshop can make use of the bookmarklets.

Modified Spreading Activation Network for Intelligent Profile Construction in Research Agent System (리서치 에이전트시스템에서의 지능적 프로파일 구축을 위한 개선된 확산 활성화 네트워크)

  • 조영임;김유신
    • Journal of Korea Multimedia Society
    • /
    • v.6 no.6
    • /
    • pp.1111-1119
    • /
    • 2003
  • The research of science and engineering needs the latest information from internet resources. But it is a complex and repeated procedure to search and filter web documents from the huge Internet resources. In this paper, we propose the PREA system, which can organize the research paper databases and search World Wide Web documents that the user is interested in. It observes the usage of the local Paper databases and presented web documents and then constructs a profile intelligently. However, to make a profile, we used the modified spreading activation network(MSAN) so that the PREA can search and filter web documents by semantic meaning of user's interest in realtime. The system constructed in multi-agents manner that can cooperate together effectively. The results show the effectiveness of our system to search web documents compared with a commercial search engine.

  • PDF