[논문]강화학습을 이용한 주제별 웹 탐색

임수연

doi:10.5391/jkiis.2005.15.4.395

강화학습을 이용한 주제별 웹 탐색
Topic directed Web Spidering using Reinforcement Learning 원문보기

퍼지 및 지능시스템학회 논문지 = Journal of fuzzy logic and intelligent systems, v.15 no.4, 2005년, pp.395 - 399

초록
AI-Helper

본 논문에서는 특정 주제에 관한 웹 문서들을 더욱 빠르고 정확하게 탐색하기 위하여 강화학습을 이용한 HIGH-Q 학습 알고리즘을 제안한다. 강화학습의 목적은 환경으로부터 주어지는 보상(reward)을 최대화하는 것이며 강화학습 에이전트는 외부에 존재하는 환경과 시행착오를 통하여 상호작용하면서 학습한다. 제안한 알고리즘이 주어진 환경에서 빠르고 효율적임을 보이기 위하여 넓이 우선 탐색과 비교하는 실험을 수행하고 이를 평가하였다. 실험한 결과로부터 우리는 미래의 할인된 보상을 이용하는 강화학습 방법이 정답을 찾기 위한 탐색 페이지의 수를 줄여줌으로써 더욱 정확하고 빠른 검색을 수행할 수 있음을 알 수 있었다.

Abstract ▼ AI-Helper

In this paper, we presents HIGH-Q learning algorithm with reinforcement learning for more fast and exact topic-directed web spidering. The purpose of reinforcement learning is to maximize rewards from environment, an reinforcement learning agents learn by interacting with external environment through trial and error. We performed experiments that compared the proposed method using reinforcement learning with breath first search method for searching the web pages. In result, reinforcement learning method using future discounted rewards searched a small number of pages to find result pages.

주제어

참고문헌 (15)

박찬건, 양성봉, '강화 학습에서의 탐색과 이용의 균 형을 통한 범용적 온라인 Q-학습이 적용된 에이전트 의 구현,' 정보과학회 논문지(B), Vol. 30, No. 7, pp. 672-680, 2003
정태진, 장병탁, '강화 학습을 이용한 웹 정보 검색,' 정보과학회 제 28회 추계학술대회, Vol. 28, No. 2, pp. 94-96, 2001
C. J. Watkins and P. Dayan, 'Technical note : QLearning,' Machine Learning, 8, pp .279-292, 1992

상세보기
F. Menczer, 'ARACHNID: Adaptive retrieval agents choosing heuristic neighborhoods for information discovery,' In proceedings of 14th International Conference on Machine Learning, pp. 227-235, 1997
H. Lieberman, 'Letizia: An agent that assists web browsing,' In Proocedings of the International Joint Conference on Arti cial Intelligence (IJCAI95), pp. 924-929, 1995
J. Boyan, D. Freitag, and T. Joachimas, 'A machine learning architecture for optimizing web search engines,' In proceedings of AAAI workshop on Internet-Based Information Systems, pp. 1-8, 1996
J. Peng, and R. Williams, 'Incremental multi-step Q-learning,' Machine Learning, vol. 22, pp. 283- 290, 1996

상세보기
J. Rennie and A. McCallum, 'Using Reinforcement Learning to Spider the Web Efficiently,' In proceedings of the 16th International Conference on Machine Learning(ICML-99), pp. 335-343, 1999
L. P. Kaelbling, 'Learning in Embedded System,' PhD thesis, Departmenr of Computer Science, Stanford University, 1990
R. Dearden, N. Friedman and S. Russell, 'Bayesian Q-Learning,' In proceedings of AAA-98, 1989
R. S. Sutton and A. G. Barto, Reinforcement Learning : An Introduction. The MIT Press, 1998
S. B. Thrun, 'The role of exploration in learning control,' Handbook of Intelligent Control:Neural, Fussy and Adaptive Approaches. 1992
T. Joachims, D. Freitag, and T. M. Mitchell. 'A WebWatcher: A Tour Guide for the World Wide Web,' In Proceedings of the Fifteenth International Joint Conference on Artificial Intelligence (IJCAI'97), pp. 770-777, 1997
T. M. Mitchell, Machine Learning, McGraw-Hill, 1997
M. Tan, Multi-agent reinforcement learning: Independent vs. cooperative agents. In Proc. of the Tenth International Conf. on Machine Learning, pp. 330.337, 1993

내보내기 구분	파일저장 인쇄 메일전송
구성항목	기본정보 상세정보 관리번호, 논문명, 저널/프로시딩명, 저자 , 발행년, 권, 호, 시작페이지, 끝페이지, 발행기관 관리번호, 논문명, 대등논문명, 저자 , 저널/프로시딩명, 발행기관, 발행년, 발행언어, 권, 호, 시작페이지, 끝페이지, ISBN, ISSN, 주제분야, 키워드, 초록(한글), 초록(영문), 저자(소속기관)
저장형식	Text(ASCII format) Excel format RefWorks Direct Export RIS format (for Reference Manager, ProCite, EndNote), Scholar's Aids, Mendeley
메일정보	받는사람 (필수) @ 보내는사람 (선택) @ 제목 내용 KISTI 검색결과 이메일 서비스
안내	총 건의 자료가 검색되었습니다. 다운받으실 자료의 인덱스를 입력하세요. (1-10,000) 검색결과의 순서대로 최대 10,000건 까지 다운로드가 가능합니다. 데이타가 많을 경우 속도가 느려질 수 있습니다.(최대 2~3분 소요) 다운로드 파일은 UTF-8 형태로 저장됩니다. 파일의 내용이 제대로 보이지 않을실 때는 웹브라우저 상단의 보기 -> 인코딩 -> 자동선택 여부를 확인하십시오. ~ Text(ASCII format) Excel format

연합인증

강화학습을 이용한 주제별 웹 탐색
Topic directed Web Spidering using Reinforcement Learning 원문보기

초록
AI-Helper

Abstract ▼ AI-Helper

주제어

참고문헌 (15)

이 논문을 인용한 문헌

저자의 다른 논문 :

관련 콘텐츠

원문 보기

원문 URL 링크

오픈액세스(OA) 유형

연관된 기능

이 논문과 함께 이용한 콘텐츠

AI-Helper ※ AI-Helper는 오픈소스 모델을 사용합니다.

선택된 텍스트

연합인증

강화학습을 이용한 주제별 웹 탐색 Topic directed Web Spidering using Reinforcement Learning 원문보기

초록 용어보기논문에서 용어와 풀이말을 자동 추출한 결과로, 시범 서비스 중입니다. AI-Helper

Abstract ▼ AI-Helper

주제어

참고문헌 (15)

이 논문을 인용한 문헌

저자의 다른 논문 :

임수연 (11)

관련 콘텐츠

원문 보기

원문 URL 링크

오픈액세스(OA) 유형

연관된 기능

이 논문과 함께 이용한 콘텐츠

AI-Helper ※ AI-Helper는 오픈소스 모델을 사용합니다.

선택된 텍스트

강화학습을 이용한 주제별 웹 탐색
Topic directed Web Spidering using Reinforcement Learning 원문보기

초록
AI-Helper