년 - 년
Improving Web Query Processing Through an Intelligent Algorithm for Heterogeneous Databases
보안공학연구지원센터(IJDTA) International Journal of Database Theory and Application vol.4 no.2 2011.06 pp.13-22
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
Performance of web query processing becomes slow caused by increasing number of data. In this paper, intelligent algorithm was created to fix this problem. Four main components involved in this intelligent algorithm are assigning initial query, exploit query, assign to any possible query and query matching. Firstly, web user is needed to enter a keyword then this keyword automatically will assign as an initial query. After that, this initial query must be exploited before assigning process to any possible queries. These possible queries will match to existing query form temporary file that contains data schema and map to only select data sources. In this methodology, XML will use for mapping to heterogeneous databases. XML is really effective for mapping to heterogeneous databases. This methodology has been implemented in a prototype and applied to web queries. Another issue in with web query is difficult to search for information that best reflects the user’s need information. According this problem, intelligent algorithm was created and tested by developing simple application based on system architecture. This intelligent algorithm that was good for fixed problems by improving web query processing for heterogeneous databases. The intelligent algorithm was implemented and tested for heterogeneous database environment.
A New Distributed Caching Technique for Accelerating the Web Query Processing SCOPUS
보안공학연구지원센터(IJSEIA) International Journal of Software Engineering and Its Applications Vol.7 No.3 2013.05 pp.109-116
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
Because of the fast growing volume of web documents during the past decades, the efficiency of the web search engine has become more crucial than ever. Such efficiency can be estimated with both factors of the query relevance of search results answered and the financial cost for query processing. Between them, the ways for improving query relevance of web searches have been intensively studied in the research topics like hyperlink-based ranking, topic-sensitive document classifications, and semantic-awareness in rank evaluations. However, there have been not studies that provide an efficient solution to cut the financial cost of query processing, while retaining high query relevance. In this light, we propose a distributed cache scheme and a server-clustering technique that can be used to reduce the query processing cost. With the help of such techniques for accelerating the web query processing, we saved around 70% of the server cost of a commercial web search engine implemented in South Korea. We believe that our experiences can give a valuable insight to anyone who wants to develop a large-scale search engine.
Pre-Processing of Query Logs in Web Usage Mining
[Kisti 연계] 대한산업공학회 Industrial engineering & management systems Vol.11 No.1 2012 pp.82-86
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
In For the past few years, query log data has been collected to find user's behavior in using the site. Many researches have studied on the usage of query logs to extract user's preference, recommend personalization, improve caching and pre-fetching of Web objects, build better adaptive user interfaces, and also to improve Web search for a search engine application. A query log contain data such as the client's IP address, time and date of request, the resources or page requested, status of request HTTP method used and the type of browser and operating system. A query log can offer valuable insight into web site usage. A proper compilation and interpretation of query log can provide a baseline of statistics that indicate the usage levels of website and can be used as tool to assist decision making in management activities. In this paper we want to discuss on the tasks performed of query logs in pre-processing of web usage mining. We will use query logs from an online newspaper company. The query logs will undergo pre-processing stage, in which the clickstream data is cleaned and partitioned into a set of user interactions which will represent the activities of each user during their visits to the site. The query logs will undergo essential task in pre-processing which are data cleaning and user identification.
웹 온톨로지 저장소의 질의 처리 성능에 대한 비교 평가
[Kisti 연계] 한국정보과학회 한국정보과학회 학술대회논문집 2007 pp.17-22
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
이 논문에서는 관계형 데이터베이스 모델 기반의 OWL 웹 온톨로지 모델을 보이고 이에 대한 실험 및 비교 평가 결과에 대하여 기술한다. OWL은 W3C에 의해 2004년 12월에 권고안으로 채택된 이후에 많은 연구가 진행되고 있다. 편집 도구 개발, 저장소 개발로 OWL 기반의 추론 엔진까지 이와 관련된 다양한 연구가 진행 중이다. 특히 OWL 온톨로지의 영구적인 저장 및 관리를 위해 관계형 데이터베이스 모델을 이용한 많은 연구 결과들이 발표되고 있다. 이 논문에서는 널리 알려진 관계형 모델 기반의 저장소 보다 나은 성능을 제공하기 위해 제안한 모델에 대한 평가 결과에 대하여 기술한다. 기존 유사 연구의 경우, 비교 평가를 위한 평가 항목으로 온톨로지 로드 시간을 고려하기도 하지만 이 논문에서는 질의응답 시간에만 초점을 둔다. 이는 매우 특수한 상황을 제외한 대부분의 상황에서 질의 처리 시간이 가장 중요한 요소이며 실질적인 활용성 측면에서 핵심적으로 다루어야 하는 평가 항목이기 때문이다. 실험을 위한 데이터로서는 많은 연구에서 활용하고 있는 LUBM 데이터 셋을 이용하며 실험 대상으로는 오픈 소스이며 널리 알려진 시스템인 Jena의 저장소와 Sesame를 이용한다. 실험 및 비교 평가 결과, 제안 시스템이 비교 대상 시스템들에 비해 나은 성능을 보임을 알 수 있다.
[Kisti 연계] 한국정보과학회 한국정보과학회 학술대회논문집 2006 pp.51-55
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
시맨틱 웹은 현재 웹의 확장된 개념으로 사람뿐만 아니라 컴퓨터 스스로가 데이터를 이해하고 처리할 수 있도록 정보에 의미를 부여하는 것이다. 시맨틱 웹 데이터를 기술하는 RDF를 통해 메타데이터를 표현하고 의미론적 추론이 가능하게 되었다. 따라서 기존에 일반 사용자가 쉽게 사용할 수 있는 키워드 검색 방법을 시맨틱 웹 데이터인 RDF/RDF 스키마에 적용함으로써 차세대 웹으로 인식되고 있는 시맨틱 웹을 일반 사용자도 쉽게 활용할 수 있도록 한다. 본 논문에서는 RDF 문서의 효율적인 검색을 위해 RDF 인스턴스와 RDF 스키마 정보를 저장하고, 키워드, 속성, 클래스 타입의 복합 조건 검색을 만족시키는 키워드 인덱스와 스키마 테이블 구조를 제안한다. 본 논문에서 제안한 구조는 다양한 조건들을 만족하는 리소스 정보의 빠르고 정확한 검색이 가능하도록 한다.
시맨틱 웹 문서를 위한 관계형 저장 스키마 설계 및 질의 처리 기법
[Kisti 연계] 한국컴퓨터정보학회 Journal of the Korea society of computer and information Vol.14 No.1 2009 pp.35-45
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
최근 들어 온톨로지 문서의 활용이 증가하고 있는 추세 속에서 시맨틱 정보를 효율적으로 검색하기 위해서는 온톨로지 데이터를 효과적으로 저장 및 질의 처리를 할 수 있는 관리 시스템이 필요하다. 본 논문에서는 W3C에서 제안한 온톨로지 언어인 RDF/RDFS를 기반으로 하는 시맨틱 웹 문서를 관계형 데이터베이스에 저장하고 효율적으로 검색하기 위한 저장 스키마를 제안한다. 특별히 제안한 저장스키마는 계층 정보를 효과적으로 검색할 수 있도록 설계하여 질의 처리의 효율성을 증가시킨다. 또한 본 논문에서는 RQL 시맨틱 질의를 SQL로 변환하여 질의를 처리하는 메카니즘을 기술하며 MS-ACCESS를 사용하여 데이터베이스를 구축 및 구현한다. 구현 결과를 통하여 트리플 모델에 기반한 데이터 질의 뿐 만 아니라 스키마나 계층정보에 대한 질의도 간단하게 SQL로 변환됨을 알 수 있다.
According to the widespread use of ontology documents, a management system which store ontology data and process queries is needed for retrieving semantic information efficiently. In this paper I propose a storage schema that stores and retrieves semantic web documents based on RDF/RDFS ontology language developed by W3C in a relational databases. Specially, the proposed storage schema is designed to retrieve efficiently hierarchy information and to increase efficiency of query processing. Also, I describe a mechanism to transform RQL semantic queries to SQL relational queries and build up database using MS-ACCESS and implement in this paper. According to the result of implementation, we can blow that not only data query based on triple model but also query for schema and hierarchy information are transformed simply to SQL.
시맨틱 웹 데이터의 키워드 질의 처리를 위한 인덱싱 및 저장 기법
[Kisti 연계] 한국컴퓨터정보학회 Journal of the Korea society of computer and information Vol.12 No.5 2007 pp.93-102
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
시맨틱 웹에서는 메타데이터와 온톨로지를 이용하여 질의를 처리하기 때문에 보다 정확한 검색 결과를 얻을 수 있을 뿐만 아니라 추론을 통하여 얻어진 새로운 지식도 검색 결과에 포함시킬 수 있다. 메타데이터와 온톨로지를 기술하기 위한 시맨틱 웹 언어 중 RDF와 RDF 스키마가 보편적으로 많이 활용되고 있다. 따라서 RDF와 RDF 스키마로 기술된 시맨틱 웹 언어에 대한 효과적인 검색 기법이 요구된다. 본 논문에서는 키워드 질의 처리 결과의 기본 단위를 전체 웹 문서나 부분이 아닌 정보 리소스로 정의하였다. 그리고 메타데이터와 온톨로지 정보를 모두 고려한 시맨틱 웹 환경의 키워드 질의를 3가지 유형으로 분류하고 다양한 관련 질의에 대한 처리를 효과적으로 지원하기 위하여 키워드 인덱스와 저장 구조를 제안하였다. 본 논문에서 제안한 키워드 인덱스는 질의 조건으로 주어진 키워드를 직접 포함하고 있는 리소스는 물론 의미적 관계에 의해 간접적으로 포함하고 있는 리소스에 관련된 정보를 쉽게 제공할 수 있다. 그리고 본 논문에서는 클래스와 속성의 일반적인 정보와 계층 정보를 단순한 레이블링 기법을 이용하여 표현한 후 제안된 저장 구조를 이용해 정보를 유지하여 시맨틱 웹 환경에 적합한 키위드 질의 처리를 지원하고자 한다.
Metadata and ontology can be used to retrieve related information through the inference mure accurately and simply on the Semantic Web. RDF and RDF Schema are general languages for representing metadata and ontology. An enormous number of keywords on the Semantic Web are very important to make practical applications of the Semantic Web because most users prefer to search with keywords. In this paper, we consider a resource as a unit of query results. And we classily queries with keyword conditions into three patterns and propose indexing techniques for keyword-search considering both metadata and ontology. Our index maintains resources that contain keywords indirectly using conceptual relationships between resources as well as resources that contain keywords directly. So, if user wants to search resources that contain a certain keyword, all resources are retrieved using our keyword index. We propose a structure of table for storing RDF Schema information that is labeled using some simple methods.
0개의 논문이 장바구니에 담겼습니다.
선택하신 파일을 압축중입니다.
잠시만 기다려 주십시오.