Detecting High Obfuscation Plagiarism : Exploring Multi-Features Fusion via Machine Learning

Leilei Kong; Zhimao Lu; Haoliang Qi; Zhongyuan Han

216.73.217.110

개인회원 가입

개인회원
기관회원

개인회원 로그인

개인회원 가입으로 더욱 편리하게 이용하세요. 개인회원 가입

아이디/비밀번호를 잊으셨나요? 아이디/비밀번호 찾기

기관회원 로그인

소속기관에서 검색되지 않는 기관은 무료원문다운이 불가능합니다. 개인회원 가입 후 유료구매를 하시거나 소속기관 도서관에 이용문의해 주세요.

Home

Detecting High Obfuscation Plagiarism : Exploring Multi-Features Fusion via Machine Learning

발행기관

보안공학연구지원센터(IJUNESST) 바로가기
간행물

International Journal of u- and e- Service, Science and Technology 바로가기
통권

Vol.7 No.4 (2014.08)바로가기
페이지

pp.385-396
저자

Leilei Kong, Zhimao Lu, Haoliang Qi, Zhongyuan Han
언어

영어(ENG)
URL

https://www.earticle.net/Article/A231858

※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.

원문정보

초록

영어: Providing effective methods of identification of high-obfuscation plagiarism seeds presents a significant research problem in the field of plagiarism detection. The conventional methods of plagiarism detection are based on single type of features to capture plagiarism seeds. But for high-obfuscation plagiarism detection, these single type features are not sufficient for identifying the plagiarism seeds effectively because of the varied plagiarism methods used in high-obfuscation plagiarism. This paper presents a multi-features fusion method for the high-obfuscation plagiarism seeds identification. This method exploits Logical Regression model to integrate lexicon features, syntax features, semantics features and structure features which extracted from suspicious document and source document. A multi-feature fusion classifier based on Logical Regression model is proposed to decide whether a text fragment pair can be regarded as plagiarism seeds or not. Experimental results on the PAN@CLEF2013 summary-obfuscation corpus show that the fusion of different types of features produces more accurate results.

Abstract
1. Introduction
2. Related Work
3. Multi-features fusion via Logical Regression model
  3.1. Problem
  3.2. Multi-Features Fusion
  3.3. Features set for high-obfuscation plagiarism detection
4. Result
  4.1. Corpus
  4.2. Measures
  4.3. Baseline
  4.4. Experimental results
5. Conclusion
Acknowledgements
References

키워드

plagiarism detection high-obfuscation plagiarism plagiarism seeds multi-feature fusion logical regression model

저자

Leilei Kong [ Harbin Engineering University, Heilongjiang Institute of Technology ]
Zhimao Lu [ Harbin Engineering University ]
Haoliang Qi [ Heilongjiang Institute of Technology ]
Zhongyuan Han [ Heilongjiang Institute of Technology ]

참고문헌

자료제공 : 네이버학술정보

간행물 정보

발행기관

발행기관명

보안공학연구지원센터(IJUNESST) [Science & Engineering Research Support Center, Republic of Korea(IJUNESST)]
설립연도
2006
분야
공학>컴퓨터학
소개
1. 보안공학에 대한 각종 조사 및 연구 2. 보안공학에 대한 응용기술 연구 및 발표 3. 보안공학에 관한 각종 학술 발표회 및 전시회 개최 4. 보안공학 기술의 상호 협조 및 정보교환 5. 보안공학에 관한 표준화 사업 및 규격의 제정 6. 보안공학에 관한 산학연 협동의 증진 7. 국제적 학술 교류 및 기술 협력 8. 보안공학에 관한 논문지 발간 9. 기타 본 회 목적 달성에 필요한 사업

간행물

간행물명

International Journal of u- and e- Service, Science and Technology
간기
격월간
pISSN
2005-4246
수록기간
2008~2016
십진분류
KDC 505 DDC 605

이 권호 내 다른 논문 / International Journal of u- and e- Service, Science and Technology Vol.7 No.4

피인용수 : 0건 (자료제공 : 네이버학술정보)

함께 이용한 논문 이 논문을 다운로드한 분들이 이용한 다른 논문입니다.

출처 : 네이버학술정보

0개의 논문이 장바구니에 담겼습니다.

페이지 저장

소속기관 조회

이용자님의 소속기관(단체)이 서비스에 가입되어 있는지 확인해 보십시오.
기관회원에 소속되어 있는 이용자는 원문을 무료로 이용할 수 있습니다.

상호: 주식회사 학술교육원 I 대표: 노방용 I 사업자등록번호: 122-81-88227 I 통신판매업신고번호: 제2008-인천부평-00176호 I 정보보호책임자: 이두영
주소: (21319)인천광역시 부평구 영성중로 50 미래타워 701호 I 전화: 0505-555-0740 I 팩스: 0505-555-0741 I 이메일: earticle@earticle.net

음성지원 및 돋보기 서비스

Earticle

Detecting High Obfuscation Plagiarism : Exploring Multi-Features Fusion via Machine Learning

원문정보

초록

목차

키워드

저자

참고문헌

간행물 정보

발행기관

간행물

이 권호 내 다른 논문 / International Journal of u- and e- Service, Science and Technology Vol.7 No.4

피인용수 : 0건 (자료제공 : 네이버학술정보)

함께 이용한 논문 이 논문을 다운로드한 분들이 이용한 다른 논문입니다.