정보관리학회지, 한국정보관리학회

31

웹 검색어 선택과정에서의 이용자 불확실성의 유형 : 자연과학연구자들의 정보탐색환경에 대한 고찰

김양우(한성대학교) 2006, Vol.23, No.2, pp.287-309 https://doi.org/10.3743/KOSIM.2006.23.2.287

초록보기

초록

다수의 연구에서 정보추구 과정상 불 확신성(Uncertainty) 의 중요성이 지적되었지만, 실제 정보검색시스템을 이용한 탐색과정에서 이용자들의 불 확신성에 대한 연구는 많지 않았다. 본 연구는 실제로 정보를 추구하는 이용자들의 웹 검색어 선정과정에서의 불 확신성 인식을 조사하여, 정보탐색 과정에서의 다양한 불 확신성 유형을 식별하였다. 불 확신성 유형에 입각하여 발견된 불 확신성의 주요 원인(Origins)은 정보검색시스템 및 서비스 발전을 위한 시사점을 제시하여준다.

Abstract

While numerous studies have suggested the significance of uncertainty during the process of information-seeking, less research has investigated user uncertainty in the actual search process using a real system. This study investigated user perceptions of uncertainty in the process of the selection of Web search terms in the real information-seeking process. The subjects at the doctoral or post-doctoral level were limited to the discipline of science in order to understand user perceptions in this field. The findings revealed various dimensions, types, and incidents of uncertainty. The typology of uncertainty facilitated an understanding of the subjects' information-seeking context by identifying various aspects of the context that constituted the subjects’ uncertainty. The identification of two principal origins of uncertainty based on the different types of uncertainty generated implications to improve information systems and services.

32

데이터 융합을 이용한 내용기반 이미지 검색에 관한 연구

백우진(건국대학교) ; Sun-Eun Jung(Konkuk U) ; Euigun Ahn(Yonsei U) ; 김기용(건국대학교) ; 신문선(건국대학교) 2008, Vol.25, No.2, pp.49-68 https://doi.org/10.3743/KOSIM.2008.25.2.049

초록보기

초록

Abstract

In many information retrieval experiments, the data fusion techniques have been used to achieve higher effectiveness in comparison to the single evidence-based retrieval. However, there had not been many image retrieval studies using the data fusion techniques especially in combining retrieval results based on multiple retrieval methods. In this paper, we describe how the image retrieval effectiveness can be improved by combining two sets of the retrieval results using the Sobel operator-based edge detection and the Self Organizing Map(SOM) algorithms. We used the clip art images from a commercial collection to develop a test data set. The main advantage of using this type of the data set was the clear cut relevance judgment, which did not require any human interven- tion.

33

온톨로지를 이용한 인터넷웹 검색에 관한 실험적 연구

김현희(명지대학교) ; 안태경(대외경제정책연구원) 2003, Vol.20, No.1, pp.417-455 https://doi.org/10.3743/KOSIM.2003.20.1.417

초록보기

초록

온톨로지는 웹자원을 지식화함으로써 정보의 효율적 검색, 통합, 재사용을 도모할 수 있는 새로운 기술인 시맨틱 웹의 구현을 위한 가장 핵심적인 요소 기술로 알려지고 있다. 온톨로지는 사람간에 그리고 서로 다른 응용 시스템간에 지식을 공유하고 재이용하는 방법을 제공하는 기술로서 특정 주제에 관한 지식 용어들의 집합으로서 이들 용어뿐만 아니라 용어간의 의미적 연결 관계와 간단한 추론 규칙을 포함한다. 본 연구에서는 인터넷 웹상에서 국제기구에 관한 정보를 체계적으로 관리하고 검색하기 위해서 국제기구 온톨로지를 설계하고 이 온톨로지에 기반 하여 검색 시스템을 구현해 보고, 이 시스템을 20개의 탐색 질문들을 이용하여 기존의 인터넷 검색엔진과 적합성과 탐색 시간이라는 두 가지 요인을 통해서 비교해 보았다. 실험 결과에 의하면 적합성 측정은 온톨로지 기반 시스템은 평균 4.53, 인터넷 검색엔진은 평균 2.51로 온톨로지 기반 시스템의 적합도가 1.80배 높은 것으로 나타났다. 또한 탐색시간은 온톨로지 기반 시스템은 평균 1.96분, 인터넷 검색엔진은 평균 4.74분으로 인터넷 검색엔진이 온톨로지 기반 시스템 보다 2.42배 정도 더 많은 탐색시간이 필요한 것으로 나타났다.

Abstract

Ontologies are formal theories that are suitable for implementing the semantic web, which is a new technology that attempts to achieve effective retrieval, integration, and reuse of web resources. Ontologies provide a way of sharing and reusing knowledge among people and heterogeneous applications systems. The role of ontologies is that of making explicit specified conceptualizations. In this context, domain and generic ontologies can be shared, reused, and integrated in the analysis and design stage of information and knowledge systems. This study aims to design an ontology for international organizations, and build an Internet web retrieval system based on the proposed ontology, and finally conduct an experiment to compare the system performance of the proposed system with that of Internet search engines focusing relevance and searching time. This study found that average relevance of ontology- based searching and Internet search engines are 4.53 and 2.51, and average searching time of ontology-based searching and Internet search engines are 1.96 minutes and 4.74 minutes.

34

학과분류체계의 학위논문검색 적용에 관한 연구

심원식(성균관대학교) ; 김성환(성균관대학교) 2007, Vol.24, No.4, pp.153-171 https://doi.org/10.3743/KOSIM.2007.24.4.153

초록보기

초록

본 연구는 이용이 매우 활발한 국내 학위논문의 원문 검색 서비스를 개선하기 위한 하나의 방법으로 한국직업능력개발원이 개발한 커리어넷의 학과정보 분류체계를 한국교육학술정보원이 운영하는 RISS에 포함된 학위논문 정보에 적용한 것이다. 연구 결과 커리어넷의 학과정보 분류체계는 최근 3년간 국내에서 생산되거나 이용된 학위논문을 분류하는데 비교적 적합한 것으로 나타났다. 최근 3개년 동안의 학과분류별 논문생산량과 논문이용량을 분석하였으며, 이를 바탕으로 학과분류 적용 가능성을 검토하고 활용방안을 모색하였다.

Abstract

This study suggests that improvement of theses and dissertations retrieval can be made by applying appropriate academic department classification. We applied the Korea Research Institute for Vocational Education and Training(KRIVET)'s department classification to these and dissertations being serviced by Korea Education & Research Information Service(KERIS). The results show that the chosen classification appropriately represents diverse academic department information contained in the theses and dissertations either published or used within the recent three year period. The study also makes a number of suggestions that will facilitate the application of an academic department classification to a live system.

35

피벗 역문헌빈도 가중치 기법에 대한 연구

이재윤(경기대학교) 2003, Vol.20, No.4, pp.233-248 https://doi.org/10.3743/KOSIM.2003.20.4.233

초록보기

초록

역문헌빈도 가중치 기법은 문헌 집단에서 출현빈도가 낮을수록 색인어의 중요도가 높다는 가정에 근거하고 있다. 그런데 이는 중간빈도어를 중요하게 여기는 여타 이론과는 일치하지 않는 것이다. 이 연구에서는 저빈도어보다 중간빈도어가 더 중요하다는 가정에 근거하여 역문헌빈도 가중치 공식을 수정한 피벗 역문헌빈도 가중치 기법을 제안하였다. 제안된 기법을 검증하기 위해서 세 실험집단을 대상으로 검색실험을 수행한 결과. 피벗 역문헌빈도 가중치기법이 역문헌빈도 가중치 기법에 비해서 특히 검색결과 상위에서의 성능을 향상시키는 것으로 나타났다.

Abstract

The Inverse Document Frequency (IDF) weighting method is based on the hypothesis that in the document collection the lower the frequency of a term is, the more important the term is as a subject word. This well-known hypothesis is, however, somewhat questionable because some low frequency terms turn out to be insufficient subject words. This study suggests the pivoted IDF weighting method for better retrieval effectiveness, on the assumption that medium frequency terms are more important than low frequency terms. We thoroughly evaluated this method on three test collections and it showed performance improvements especially at high ranks.

36

검색 성능 향상을 위한 약품 온톨로지 기반 연관 피드백

임수연(경북대학교) 2005, Vol.22, No.2, pp.41-56 https://doi.org/10.3743/KOSIM.2005.22.2.041

초록보기

초록

기계가 정보의 의미를 이해하고 처리할 수 있도록 기존의 웹을 확장하는 것을 목적으로 하는 시멘틱 웹은 온톨로지를 이용하여 지식을 공유하게 된다. 본 논문에서는 정교한 질의의 처리를 위하여 온톨로지 내에 존재하는 의미 관계들을 질의의 확장을 위한 연관피드백 정보로 이용하는 방안을 제안한다. 실험은 도메인 온톨로지인 Medicine 온톨로지를 대상으로 하였으며, 출현 용어들의 빈도정보만을 이용한 키워드기반 문서검색과 제안한 온톨로지기반 문서검색의 성능을 비교하였다. 이 때, 두 시스템의 정확률과 재현율을 성능 평가의 기준으로 삼았다. 그 결과, 검색 엔진은 온톨로지에 정의된 개념들과 규칙들을 활용하면서 검색의 정확률을 향상시키는데 도움이 되었고 검색 성능을 향상시키기 위한 추론의 기반으로도 사용될 수 있었다.

Abstract

For the purpose of extending the Web that is able to understand and process information by machine, Semantic Web shared knowledge in the ontology form. For exquisite query processing, this paper proposes a method to use semantic relations in the ontology as relevance feedback information to query expansion. We made experiment on pharmacy domain. And in order to verify the effectiveness of the semantic relation in the ontology, we compared a keyword based document retrieval system that gives weights by using the frequency information compared with an ontology based document retrieval system that uses relevant information existed in the ontology to a relevant feedback. From the evaluation of the retrieval performance, we knew that search engine used the concepts and relations in ontology for improving precision effectively. Also it used them for the basis of the inference for improvement the retrieval performance.

37

통합 검색 환경에서 이용자 적합성 판단 기준에 관한 탐색적 연구

박정아(다음커뮤니케이션) 2012, Vol.29, No.2, pp.113-133 https://doi.org/10.3743/KOSIM.2012.29.2.113

초록보기

초록

본 연구는 한국 통합 검색 환경에서의 이용자 적합성 판단 기준에 관한 탐색적 연구이다. 이를 위해 10명의 참가자들을 대상으로 반구조화(semi-structured) 인터뷰를 수행하여 데이터를 수집하였다. 참가자들은 네이버, 다음 등과 같은 통합 검색 환경에서 본인들이 관심 있거나 필요로 하는 다양한 검색을 수행하고, 그 과정에서 문서가 적합한지와 그 판단 기준에 대해 기술하였다. 연구 결과 8개의 적합성 판단 기준과 비적합성 판단 기준, 그리고 검색 환경이 변화하여도 이용자가 적합성을 판단하는 기준들이 크게 변화하지는 않지만 데이터 증가와 이용자 요구의 고도화로 특수성과 구체성이 중요한 적합성 판단 기준으로 부각되는 점을 발견하였다.

Abstract

This study is an exploratory research on the user relevance criteria in Korean search service environments that provide integrated search results. Data were collected from 10 participants using a semi-structured interview technique. The participants conducted a web search using integrated search services, such as Naver or Daum on a self-selected topic. They were asked to judge the relevance of retrieved documents and to report their relevance criteria. As a result, the research indicated 8 user-defined relevance and non-relevance criteria. The research shows that specificity and richness are the two most important criteria yet, the user’s relevance criteria have not changed much despite the change in search environment.

38

객체-관계형 데이터베이스에 의한 XML문헌의 검색성능 평가

김희섭(경북대학교) 2004, Vol.21, No.2, pp.189-210 https://doi.org/10.3743/KOSIM.2004.21.2.189

초록보기

초록

본 연구의 목적은 객체-관계형 데이터베이스 접근에 의한 XML 문헌의 검색 성능을 평가하는 것이다. 본 논문에서는 INEX(Initiative for the Evaluation of XML retrieval)에서의 XML 문헌의 색인 및 검색 방법에 대하여, 그리고 실험 방법론들에 대하여 기술하고 있다. 대부분의 전통적인 정보검색 성능평가 실험에서와 같이 본 연구에서 사용된 테스트 콜렉션(test collection)은 문헌(즉, XML 문헌), 토픽, ad hoc 검색, 적합성 판단, 평가로 이루어졌다. 그리고 ORDBMS 기술들을 기반으로 개발된 전용 XML 데이터베이스의 일종인 EXIMATM Supply을 사용하여 INEX에서 제공한 대규모 XML 문헌들을 저장하고 검색하였다. 본 논문에서는 실험에서 사용한 시스템에 대한 개략적인 기능들과 색인 및 검색 과정 그리고 INEX 2002에서의 성능평가 결과에 대하여, 앞으로 개선되어야 할 기능에 대하여 논하고 있다.

Abstract

The purpose of this study is to evaluate the performance of XML retrieval based on ORDBMSs(Object-Relational Database Management Systems) approach. This paper describes indexing and retrieval methods for XML documents and the methodologies of experiments at INEX(Initiative for the Evaluation of XML retrieval). Like any other traditional information retrieval experiment, the test collection was consists of documents, topics/queries, task, relevance assessments and evaluation. EXIMATM Supply, a kind of native XML DB based on ORDBMS technologies, is used for this experiment. Although this approach has many benefits, for example, no delay in storing and searching XML documents, but it showed relatively disappointed retrieval performance at INEX 2002. This result may caused since the given topics had to be decomposed and modified to be processed by the XPath processor, and during this modification the original meaning of topics can be changed inevitably and some important information may pass over.

39

한글로마자표기에 대한 국제기관의 규정과 표기의 실제에 관한 연구

오경묵(숙명여자대학교) 2007, Vol.24, No.4, pp.33-51 https://doi.org/10.3743/KOSIM.2007.24.4.033

초록보기

초록

인터넷 환경에서 정보검색의 기본적인 사안은 선택된 언어의 문자와 긴밀한 연관을 갖고 있다. 매큔-라이샤워시스템은 학술적 및 비학술적 적용을 위한 국제표준으로서, 목록 및 검색시 이용되고 있을 뿐만 아니라 대부분의 한국자료 이용자들에게서 널리 사용되고 있다. 현재 ISO, UNGEGN, LC, ALA, BL, 영국지명위원회와 유럽, 호주, 캐나다 등의 유관기관들은 모두 매큔-라이샤워시스템을 채택하여 사용하고 있다. 따라서 현재 도서관 일각에서 진행하려고 시도하는 2000년식 새한글로마자시스템으로의 표기방식 전환은 도서관 목록과 온라인DB 등에서 많은 혼란을 일으키게 할 것이다. 본 논문에서는 국제기관에서의 이 분야에 대한 노력을 소개하고, 현재 사용하고 있는 상세한 규정을 통하여 로마자시스템을 심층적으로 분석, 소개하여 향후 이 문제를 둘러싼 한국 도서관계가 현명한 판단과 대처를 할 수 있도록 연구결과를 제시하였다.

Abstract

The fundamental issue of information retrieval in the Internet-based society is closely interrelated with the characteristics of language selected. The McCune-Reischauer Romanization system is not only considered as the international standard for romanizing Korean language, it is also familiar to the majority of the Korean material users internationally. McCune-Reischauer system is adopted by the ISO, UNGEGN, ALA, LC, British PCGN, BL, and the relevant agencies in Europe, Canada and Australia etc. Encouraging for switching to the new Romanization system(2000) would result in complications among the library's catalogs and online databases, causing confusion for both staffs and readers. This paper analysed that the international efforts and rules for Romanizing Korean language materials and recommended direction for bibliographical issues.

40

노드정보를 이용한 문서검색의 성능에 관한 연구

윤소영(국사편찬위원회) 2007, Vol.24, No.1, pp.103-120 https://doi.org/10.3743/KOSIM.2007.24.1.103

초록보기

초록

통신기술과 정보기기의 발달로 대학에서 교육과정에 정보를 활용하는 방식이 급격히 변화하고 있어 저작권이 있는 정보를 윤리적, 합법적으로 교육 자료로 사용하게 될 경우 인지해야할 사항들이 점점 늘어나고 있다. 이 연구에서는 대학에 필요한 교육적 목적의 정보 공정사용에 관련한 저작권 법과 각종 지침을 분석한 후 대학도서관에서 교육자와 학생에게 인지시켜야 할 주요 개념을 도출하여 대학 도서관이 정보 공정사용 지침에 포함해야 할 주요 영역을 제시하고자 하였다. 또한 영역별로 도출된 주요 개념들이 국내 대학 도서관 사이트에서 적절하게 교육자와 학생들에게 제공되고 있는지를 조사하였다.

Abstract

Due to the radical changes of information technology, it becomes indispensable for educators and students of university to learn how to use copyrighted works ethically and legally without violating the copyright law. As a result, academic libraries need to take responsibilities to inform them fair use criteria and to provide proper fair use guidelines. This study analysed various fair use guidelines for them and copyright law to identify key areas of fair use guideline for the academic libraries. It also investigated 10 university libraries' web sites to find that the identified key areas are delivered to the educators and students.

바로가기메뉴

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

초록

Abstract

정보관리학회지