Earticle

현재 위치 Home 검색결과

결과 내 검색

발행연도

-

학문분야

자료유형

간행물

검색결과

검색조건
검색결과 : 200
No
1

4,000원

In this paper, in order to maximize the input process efficiency of the building energy simulation field, the authors developed the automatic extraction module of spatial information based BIM geometry information. Existing research or software extracts geometry information based on object information, but it can not be used in the field of energy simulation because it is inconsistent with the geometry information of the object constituting the thermal zone of the actual building model. Especially, IFC-based geometry information extraction module is needed to link with other architectural fields from the viewpoint of reuse of building information. The study method is as follows. (1) Grasp the category and attribute information to be extracted for energy simulation and Analyze the IFC structure based on spatial information (2) Design the algorithm for extracting and reprocessing information for energy simulation from IFC file (use programming language Phython) (3) Develop the module that generates a geometry information database based on spatial information using reprocessed information (4) Verify the accuracy of the development module. In this paper, the reprocessed information can be directly used for energy simulation and it can be widely used regardless of the kind of energy simulation software because it is provided in database format. Therefore, it is expected that the energy simulation process efficiency in actual practice can be maximized.

2

Automatic Extraction of Dependencies between Web Components and Database Resources in Java Web Applications

Oh, Jaewon, Ahn, Woo Hyun, Kim, Taegong

[Kisti 연계] 한국정보통신학회 Journal of information and communication convergence engineering Vol.17 No.2 2019 pp.149-160

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Web applications typically interact with databases. Therefore, it is very crucial to understand which web components access which database resources when maintaining web apps. Existing research identifies interactions between Java web components, such as JavaServer Pages and servlets but does not extract dependencies between the web components and database resources, such as tables and attributes. This paper proposes a dynamic analysis of Java web apps, which extracts such dependencies from a Java web app and represents them as a graph. The key responsibility of our analysis method is to identify when web components access database resources. To fulfill this responsibility, our method dynamically observes the database-related objects provided in the Java standard library using the proxy pattern, which can be applied to control access to a desired object. This study also experiments with open source web apps to verify the feasibility of the proposed method.

3

Automatic Extraction of Blood Flow Area in Brachial Artery for Suspicious Hypertension Patients from Color Doppler Sonography with Fuzzy C-Means Clustering

Kim, Kwang Baek, Song, Doo Heon, Yun, Sang-Seok

[Kisti 연계] 한국정보통신학회 Journal of information and communication convergence engineering Vol.16 No.4 2018 pp.258-263

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Color Doppler sonography is a useful tool for examining blood flow and related indices. However, it should be done by well-trained operator, that is, operator subjectivity exists. In this paper, we propose an automatic blood flow area extraction method from brachial artery that would be an essential building block of computer aided color Doppler analyzer. Specifically, our concern is to examine hypertension suspicious (prehypertension) patients who might develop their symptoms to established hypertension in the future. The proposed method uses fuzzy C-means clustering as quantization engine with careful seeding of the number of clusters from histogram analysis. The experiment verifies that the proposed method is feasible in that the successful extraction rates are 96% (successful in 48 out of 50 test cases) and demonstrated better performance than K-means based method in specificity and sensitivity analysis but the proposed method should be further refined as the retrospective analysis pointed out.

4

Applying Lexical Semantics to Automatic Extraction of Temporal Expressions in Uyghur

Murat, Alim, Yusup, Azharjan, Iskandar, Zulkar, Yusup, Azragul, Abaydulla, Yusup

[Kisti 연계] 한국정보처리학회 Journal of information processing systems Vol.14 No.4 2018 pp.824-836

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

The automatic extraction of temporal information from written texts is a key component of question answering and summarization systems and its efficacy in those systems is very decisive if a temporal expression (TE) is successfully extracted. In this paper, three different approaches for TE extraction in Uyghur are developed and analyzed. A novel approach which uses lexical semantics as an additional information is also presented to extend classical approaches which are mainly based on morphology and syntax. We used a manually annotated news dataset labeled with TIMEX3 tags and generated three models with different feature combinations. The experimental results show that the best run achieved 0.87 for Precision, 0.89 for Recall, and 0.88 for F1-Measure in Uyghur TE extraction. From the analysis of the results, we concluded that the application of semantic knowledge resolves ambiguity problem at shallower language analysis and significantly aids the development of more efficient Uyghur TE extraction system.

5

FPGA-Based Hardware Accelerator for Feature Extraction in Automatic Speech Recognition

Choo, Chang, Chang, Young-Uk, Moon, Il-Young

[Kisti 연계] 한국정보통신학회 Journal of information and communication convergence engineering Vol.13 No.3 2015 pp.145-151

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

We describe in this paper a hardware-based improvement scheme of a real-time automatic speech recognition (ASR) system with respect to speed by designing a parallel feature extraction algorithm on a Field-Programmable Gate Array (FPGA). A computationally intensive block in the algorithm is identified implemented in hardware logic on the FPGA. One such block is mel-frequency cepstrum coefficient (MFCC) algorithm used for feature extraction process. We demonstrate that the FPGA platform may perform efficient feature extraction computation in the speech recognition system as compared to the generalpurpose CPU including the ARM processor. The Xilinx Zynq-7000 System on Chip (SoC) platform is used for the MFCC implementation. From this implementation described in this paper, we confirmed that the FPGA platform is approximately 500× faster than a sequential CPU implementation and 60× faster than a sequential ARM implementation. We thus verified that a parallelized and optimized MFCC architecture on the FPGA platform may significantly improve the execution time of an ASR system, compared to the CPU and ARM platforms.

6

Similar Image Retrieval Technique based on Semantics through Automatic Labeling Extraction of Personalized Images

Jung-Hee, Seo

[Kisti 연계] 한국정보통신학회 Journal of information and communication convergence engineering Vol.22 No.1 2024 pp.56-63

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Despite the rapid strides in content-based image retrieval, a notable disparity persists between the visual features of images and the semantic features discerned by humans. Hence, image retrieval based on the association of semantic similarities recognized by humans with visual similarities is a difficult task for most image-retrieval systems. Our study endeavors to bridge this gap by refining image semantics, aligning them more closely with human perception. Deep learning techniques are used to semantically classify images and retrieve those that are semantically similar to personalized images. Moreover, we introduce a keyword-based image retrieval, enabling automatic labeling of images in mobile environments. The proposed approach can improve the performance of a mobile device with limited resources and bandwidth by performing retrieval based on the visual features and keywords of the image on the mobile device.

7

Feature Extraction of Non-proliferative Diabetic Retinopathy Using Faster R-CNN and Automatic Severity Classification System Using Random Forest Method

Jung, Younghoon, Kim, Daewon

[Kisti 연계] 한국정보처리학회 Journal of information processing systems Vol.18 No.5 2022 pp.599-613

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Non-proliferative diabetic retinopathy is a representative complication of diabetic patients and is known to be a major cause of impaired vision and blindness. There has been ongoing research on automatic detection of diabetic retinopathy, however, there is also a growing need for research on an automatic severity classification system. This study proposes an automatic detection system for pathological symptoms of diabetic retinopathy such as microaneurysms, retinal hemorrhage, and hard exudate by applying the Faster R-CNN technique. An automatic severity classification system was devised by training and testing a Random Forest classifier based on the data obtained through preprocessing of detected features. An experiment of classifying 228 test fundus images with the proposed classification system showed 97.8% accuracy.

8

딥러닝 언어 모델을 이용한 연구보고서의 참고문헌 자동추출 연구

한유경, 최원석, 이민철

[Kisti 연계] 한국정보관리학회 정보관리학회지 Vol.40 No.2 2023 pp.115-135

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

본 연구는 단행본, 학술지, 보고서 등 다양한 종류의 발간물로 구성된 연구보고서의 참고문헌 데이터베이스를 효율적으로 구축하기 위한 것으로 딥러닝 언어 모델을 이용하여 참고문헌의 자동추출 성능을 비교 분석하고자 한다. 연구보고서는 학술지와는 다르게 기관마다 양식이 상이하여 참고문헌 자동추출에 어려움이 있다. 본 연구에서는 참고문헌 자동추출에 널리 사용되는 연구인 메타데이터 추출과 더불어 참고문헌과 참고문헌이 아닌 문구가 섞여 있는 환경에서 참고문헌만을 분리해내는 원문 분리 연구를 통해 이 문제를 해결하였다. 자동 추출 모델을 구축하기 위해 특정 연구기관의 연구보고서 내 참고문헌셋, 학술지 유형의 참고문헌셋, 학술지 참고문헌과 비참고문헌 문구를 병합한 데이터셋을 구성했고, 딥러닝 언어 모델인 RoBERTa+CRF와 ChatGPT를 학습시켜 메타데이터 추출과 자료유형 구분 및 원문 분리 성능을 측정하였다. 그 결과 F1-score 기준 메타데이터 추출 최대 95.41%, 자료유형 구분 및 원문 분리 최대 98.91% 성능을 달성하는 등 유의미한 결과를 얻었다. 이를 통해 비참고문헌 문구가 포함된 연구보고서의 참고문헌 추출에 대한 딥러닝 언어 모델과 데이터셋 유형별 참고문헌 구축 방향을 제안하였다.

The purpose of this study is to assess the effectiveness of using deep learning language models to extract references automatically and create a reference database for research reports in an efficient manner. Unlike academic journals, research reports present difficulties in automatically extracting references due to variations in formatting across institutions. In this study, we addressed this issue by introducing the task of separating references from non-reference phrases, in addition to the commonly used metadata extraction task for reference extraction. The study employed datasets that included various types of references, such as those from research reports of a particular institution, academic journals, and a combination of academic journal references and non-reference texts. Two deep learning language models, namely RoBERTa+CRF and ChatGPT, were compared to evaluate their performance in automatic extraction. They were used to extract metadata, categorize data types, and separate original text. The research findings showed that the deep learning language models were highly effective, achieving maximum F1-scores of 95.41% for metadata extraction and 98.91% for categorization of data types and separation of the original text. These results provide valuable insights into the use of deep learning language models and different types of datasets for constructing reference databases for research reports including both reference and non-reference texts.

9

기술문서 정의문 패턴을 이용한 전문용어사전 자동추출 및 활용방안

한희정, 김태영, 두효철, 오효정

[Kisti 연계] 한국정보관리학회 정보관리학회지 Vol.34 No.4 2017 pp.81-99

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

기술문서는 지식정보사회에서 생성되는 중요 연구 성과물로, 이를 제대로 활용하기 위해서는 정보 요약 및 정보추출과 같은 개선된 정보 처리 방법을 토대로 기술문서 활용의 편의성을 높여줄 필요가 있다. 이에 본 연구는 기술문서의 핵심 정보를 추출하기 위한 방안으로, 기술문서의 구조와 정의문 패턴을 기반으로 전문용어 및 정의문을 자동 추출하고, 이를 기반으로 전문용어사전을 구축할 수 있는 시스템을 제안하였다. 나아가 전문용어사전을 지식메모리로서 보다 다양하게 활용할 수 있도록 전문용어사전에 기반한 개인화서비스 제공방안을 제안하였다. 이처럼 전문용어 및 정의문 자동추출을 기반으로 전문용어사전을 구축하게 되면 새롭게 등장하는 전문용어를 빠르게 수용할 수 있어 이용자들이 최신정보를 보다 손쉽게 찾을 수 있다. 더불어 개인화된 전문용어사전을 이용자에게 제공한다면 전문용어사전의 가치와 활용성, 검색의 효율성을 극대화할 수 있다.

Technical documents are important research outputs generated by knowledge and information society. In order to properly use the technical documents properly, it is necessary to utilize advanced information processing techniques, such as summarization and information extraction. In this paper, to extract core information, we automatically extracted the terminologies and their definition based on definitional sentences patterns and the structure of technical documents. Based on this, we proposed the system to build a specialized terminology dictionary. And further we suggested the personalized services so that users can utilize the terminology dictionary in various ways as an knowledge memory. The results of this study will allow users to find up-to-date information faster and easier. In addition, providing a personalized terminology dictionary to users can maximize the value, usability, and retrieval efficiency of the dictionary.

10

다중 관심영역의 자동 추출 및 부호화 방법

서영건, 홍도순, 박재흥

[Kisti 연계] 한국디지털콘텐츠학회 디지털콘텐츠학회 논문지 Vol.12 No.1 2011 pp.1-9

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

JPEG2000에서는 영상에서 원하는 영역을 타 영역(배경)보다 고화질로 압축하는 기법인, 관심영역 부호화 방법을 제공하고 있는데, 본 연구에서는 얼굴이 포함된 영상을 이용하여, 얼굴 영역이 가장 우선적으로 처리되고 높은 품질로 압축되도록 부가 서비스를 제공한다. 제안 기법은 크게 두 단계로 구성된다. 첫 번째는 얼굴 추출 단계이고, 두 번째는 관심영역 부호화 단계이다. 얼굴 추출은 영상의 모든 화소에 대해 $20{\times}20$ 윈도우 화소 크기로 자르거나 축소하여 전처리 과정을 거친 후 신경망을 이용하여 인식한다. 추출된 각 영역은 관심영역 마스크로 표시되고, Maxshift 방식을 이용하여 부호화된다. 이후에 EBCOT 과정을 거처 압축 및 저장된다. 기존의 방법은 고주파 성분의 분포에 의해 관심영역을 찾은 후 부호화하는 방법이 많이 연구되었다. 반면에 본 연구는 인간의 인지 능력을 이용하여, 여러 개의 얼굴이 포함된 영상에서 충분히 유용한 기법임을 보인다.

JPEG2000 offers the technique which compresses the interested regions with higher quality than the background. It is called by an ROI(Region-of-Interest) coding method. In this paper, we use images including the human faces, which are processed uppermost and compressed with high quality. The proposed method consists of 2 steps. The first step extracts some faces and the second one is ROI coding. To extract the faces, the method cuts or scale-downs some regions with $20{\times}20$ window pixels for all the pixels of the image, and after preprocessing, recognizes the faces using neural networks. Each extracted region is identified by ROI mask and then ROI-coded using Maxshift method. After then, the image is compressed and saved using EBCOT. The existing methods searched the ROI by edge distributions. On the contrary, the proposed method uses human intellect. And the experiment shows that the method is sufficiently useful with images having several human faces.

11

뇌파측정기술(EEG)과 판별분석을 이용한 영상물의 키프레임 자동 분류 방안 연구

김현희

[Kisti 연계] 한국정보관리학회 정보관리학회지 Vol.32 No.3 2015 pp.377-396

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

본 연구는 뇌파측정기술(EEG)의 ERP와 판별분석을 이용하여 이용자 기반의 비디오 키프레임들을 자동으로 추출할 수 있는 방안을 제안하였다. 구체적으로, 20명의 피험자를 대상으로 한 실험을 통해서 이미지 처리 과정을 다섯 개의 정보처리단계들(자극주목, 자극지각, 기억인출, 자극/기억 대조, 적합 판정)로 구분하여 각 단계별로 적합한 뇌파측정기술의 ERP 유형을 제안하여 검증해 보았다. 검증 결과, 각 단계별로 서로 다른 ERP 유형(N100, P200, N400, P3b, P600)을 나타냈다. 또한 세 그룹(적합, 부분적합 및 비적합 프레임)간을 구별할 수 있는 중요한 변수들로는 P3b에서 P7의 양전위 최고값과 FP2의 음전위 최저값의 잠재기로 나타났고, 이러한 변수들을 이용해 판별분석을 수행한 후 적합 및 비적합 프레임들을 분류할 수 있었다.

This study proposed a key-frame automatic extraction method for video storyboard surrogates based on users' cognitive responses, EEG signals and discriminant analysis. Using twenty participants, we examined which ERP pattern is suitable for each step, assuming that there are five image recognition and process steps (stimuli attention, stimuli perception, memory retrieval, stimuli/memory comparison, relevance judgement). As a result, we found that each step has a suitable ERP pattern, such as N100, P200, N400, P3b, and P600. Moreover, we also found that the peak amplitude of left parietal lobe (P7) and the latency of FP2 are important variables in distinguishing among relevant, partial, and non-relevant frames. Using these variables, we conducted a discriminant analysis to classify between relevant and non-relevant frames.

12

JPEG2000 이미지의 에지 분포를 이용한 ROI 마스크 생성과 자동 관심영역 추출

서영건, 김희민, 김상복

[Kisti 연계] 한국디지털콘텐츠학회 디지털콘텐츠학회 논문지 Vol.16 No.4 2015 pp.583-593

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

오늘날 컴퓨터와 통신 기술의 발달로 멀티미디어(이미지 데이터)는 다양한 응용 분야에서 사용되고 있다. 여기에 가장 널리 사용되고 있는 JPEG2000는 관심영역(ROI) 기술을 제공한다. ROI의 추출은 사용자에게 우선적으로 보여져야 하기 때문에 빠르게 수행되어야 하고 큰 이미지에서 자동적으로 추출되어야 한다. 이를 위해, 본 연구는 JPEG2000의 코드 블록 내에 있는 에지 분포를 이용하여 ROI의 자동 추출과 우선적 처리에 관한 방법을 제안한다. 먼저 에지 추출을 위한 처리와, 다음으로 에지 정보를 이용해 ROI를 자동적으로 추출한다. 그리고, 추출된 ROI 블록을 이용하여 ROI를 그룹핑 하고, ROI 블록의 마스크를 생성한다. 이후에는 양자화를 하고 우선적 처리를 하는 ROI 코딩을 하고 EBCOT를 실행한다. 제안 방법의 유효성을 보이기 위하여 JPEG2000에서 사용되는 다른 ROI 추출 기법들과 비교하고 ROI 코딩을 하지 않는 기법과 ROI 코딩이 포함된 기법 간의 PSNR을 평가하여 품질을 비교한다.

Today, caused by the growth of computer and communication technology, multimedia, especially image data are being used in different application divisions. JPEG2000 that is widely used these days provides a Region-of-Interest(ROI) technique. The extraction of ROI has to be rapidly executed and automatically extracted in a huge amount of image because of being seen preferentially to the users. For this purpose, this paper proposes a method about preferential processing and automatic extraction of ROI using the distribution of edge in the code block of JPEG2000. The steps are the extracting edges, automatical extracting of a practical ROI, grouping the ROI using the ROI blocks, generating the mask blocks and then quantization, ROI coding which is the preferential processing, and EBCOT. In this paper, to show usefulness of the method, we experiment its performance using other methods, and executes the quality evaluation with PSNR between the images not coding an ROI and coding it.

13

전 세계 항공 기업과 국가 기관들은 수십 년간 안전 보고서를 작성하고 이를 분석하여 항공 사고 예방을 위해 지속 적으로 노력해왔다. 그러나 보고서의 규모가 방대해지고 내용이 복잡해짐에 따라 수동 분석만으로는 한계가 있다. 또한, 보안상의 이유로 웹에서 서비스하는 대형 언어 모델의 사용이 어려운 경우가 많다. 이러한 문제를 해결하기 위해 본 논문에서는 항공 안전사고 보고서에서 사고 원인을 추출하기 위해 항공 도메인에 특화된 자연어 처리 모델 인 AirGemma를 제안한다. AirGemma는 Gemma2-2B 모델을 기반으로 항공 도메인 데이터를 활용한 DAPT (Domain Adaptive Pre-Training) 기법을 적용해 항공 도메인 이해도를 향상시켰다. 이후 PEFT(Parameter Efficient Fine-Tuning) 기법을 활용한 미세조정을 통해 사고 원인 추출 성능을 높였다. 실험 결과, AirGemma 가 사전학습과 미세조정을 적용하지 않은 모델 대비 F1-score, ROUGE, BLEU 지표에서 우수한 성능을 기록했 다. 또한 GPT-4를 평가자로 사용한 쌍대비교 결과, AirGemma는 GPT-3.5 Turbo보다 높은 승률을 기록했고 단 일 답변 평가 결과 LLaMA3-70B와 GPT-3.5 Turbo에 비해 사고 원인 분석에 있어 더 높은 사실성 점수를 보였 다. 이러한 결과는 항공 도메인에 특화된 모델이 사고 원인 식별에 효과적임을 입증한다. AirGemma는 항공 산업 데이터의 보안 및 제한 조건을 고려하여 로컬 환경에서 안전하게 동작할 수 있도록 설계되었으며, 항공 안전사고 분 석 및 예방을 위한 새로운 접근 방안을 제시한다.

Analysis of safety accident reports is crucial for global aviation companies and national agencies to prevent aviation accidents. However, with increasing volume and complexity of these reports, manual analysis has its limitations. Moreover, due to security concerns, using large-scale language models served through the web is often not inapplicable. To address these challenges, this paper proposes a domain-specific natural language processing model called AirGemma, which is specifically designed to extract accident causes from aviation safety reports. AirGemma is built upon the Gemma2-2B model and enhances its domain understanding through Domain Adaptive Pre-Training(DAPT) using aviation-specific data. The performance of the proposed model is further improved by applying Parameter Efficient Fine-Tuning(PEFT). Experimental results show that AirGemma outperforms models without pre-training and fine-tuning in terms of F1-score, ROUGE, and BLEU metrics. Additionally, comparative evaluations using GPT-4 as a judge reveal that AirGemma achieves a higher win rate than GPT-3.5 Turbo, and in single-answer assessments, it demonstrated greater accuracy in accident cause analysis conpared to both LLaMA3-70B and GPT-3.5 Turbo. These findings demonstrate that AirGemma is effective in identifying accident causes within the aviation domain. Designed to operate securely in a local environment, AirGemma offers a new approach to aviation safety accident analysis and prevention.

14

병열 코퍼스에서 이중언어 용어사전의 자동 추출

Sun Le, Jin Youbing, Du Lin, Sun Yufang

한국어정보학회 한국어정보학 제5ㆍ6집 2002.01 pp.117-123

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

15

비정형 보안 인텔리전스 보고서 기반 토픽 자동 추출 모델 KCI 등재

허윤아, 이찬희, 김경민, 임희석

한국융합학회 한국융합학회논문지 제10권 제6호 2019.06 pp.33-39

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

지능형 사이버 공격 기법이 다양화됨에 따라 보안 침해 사건, 글로벌 범죄 등의 사건 발생이 증가하고 있다. 지능형 공격을 예측하고 대응하기 위해서는 공격 기법의 특성, 수법, 유형을 파악해야 한다. 이를 위해 수많은 보안 기업 회사에서는 다양한 공격 기법을 빠르게 파악하고 더 큰 피해를 막기 위해 보안 인텔리전스 보고서를 배포한다. 하지만 각 기업에서 배포하는 보고서에 대한 형식이 맞춰져 있지 않으며, 대량의 비정형 보안 인텔리전스 보고서가 배포되고 있다. 본 논문은 비정형한 보안 인텔리전스 보고서에 대한 문제점을 고려하여 정형화된 데이터로 추출하는 방안을 제안 한다. 또한, 대량의 보안 인텔리전스 보고서를 파악하기 위해 소요되는 시간을 줄이고자 대량의 보고서를 주제별로 분류 할 수 있는 보안 인텔리전스 보고서 토픽 자동 추출 모델을 제안한다.

As cyber attack methods are becoming more intelligent, incidents such as security breaches and international crimes are increasing. In order to predict and respond to these cyber attacks, the characteristics, methods, and types of attack techniques should be identified. To this end, many security companies are publishing security intelligence reports to quickly identify various attack patterns and prevent further damage. However, the reports that each company distributes are not structured, yet, the number of published intelligence reports are ever-increasing. In this paper, we propose a method to extract structured data from unstructured security intelligence reports. We also propose an automatic intelligence report analysis system that divides a large volume of reports into sub-groups based on their topics, making the report analysis process more effective and efficient.

16

4,000원

본 연구는 교과서 콘텐츠의 하이라이팅 대상인 ‘후보 핵심 용어’를 식별하기 위한 ‘텍스트 하이라이팅 모델(Text Highlighting Model)’ 관련 선행연구를 확장한 연구이 다. 개선된 ‘핵심 용어 추출 모델’은 1~2어절의 ’후보 핵심 용어‘를 교과목의 ’핵심 용 어‘로 자동 분류하여, 학습자에게 미리 주요 ’핵심 용어‘를 제시하고, 강조하는 방안을 제안한다.

17

합성곱 신경망(CNN)은 다수의 2차원 필터가 계층적으로 중첩된 구조를 가지므로, 그 내부 구조와 데이터 흐름을 직관적으로 시각화하는 데 본질적인 어려움이 있다. 기존의 시각화 도구들은 각기 다른 표현 방식과 추상화 수준을 사용함으로써 서로 다른 모델 간 구조를 비교하거나 해석하는 데 한계를 가지며, 이러한 비표준화된 시각화 환경은 연구자가 모델을 이해하고 비교하는 과정에서 인지적 부담을 초래한다. 본 연구에서는 이러한 문제를 개선하기 위해 CNN 아키텍처 설명으로부터 구조 정보를 추출하고, 이를 구조화된 표현으로 변환한 뒤, 상호작용 가능한 웹 기반 3차원 시각화 프레임워크를 제안한다. 제안하는 시스템은 대형 언어 모델(LLM)을 활용하여 논문 텍스트를 JSON 형태의 중간 표현으로 변환하고, 사용자가 해당 프레임워크 상에서 CNN 아키텍처를 확대·축소 및 회전 등의 상호작용을 통해 탐색하고 비교할 수 있도록 한다. 이러한 접근은 모델 구현 파일에 의존하지 않고 논문에 기술된 아키텍처 구조를 직관적으로 분석할 수 있는 환경을 제공하며 복잡한 모델에 대한 이해를 지원하고 연구의 접근성을 제고하는 데 기여한다.

Convolutional Neural Networks (CNNs) consist of hierarchically stacked two-dimensional filters, which makes their internal structures and data flows inherently difficult to visualize intuitively. Existing visualization tools employ diverse representation styles and abstraction levels, limiting consistent comparison and interpretation across different models and increasing cognitive load for researchers. To address this issue, we propose an interactive web-based 3D visualization framework that extracts architectural information from CNN descriptions in academic papers and converts it into a structured representation. The proposed system utilizes a Large Language Model (LLM) to transform unstructured paper text into a JSON-based intermediate format, enabling users to explore and compare CNN architectures through interactions such as zooming and rotation within the framework. This approach supports intuitive analysis of architectures described in papers without relying on model implementation files and contributes to improving the accessibility of deep learning research.

18

온라인 상품비교문 추출을 위한 형용사술어 비교 구문 연구 KCI 등재

남지순

한국언어과학회 언어과학 제16권 3호 2009.10 pp.63-95

※ 기관로그인 시 무료 이용이 가능합니다.

7,500원

In this paper, we describe comparative sentences based on adjectival predicates and classify them into 6 syntactic classes: there are 3 types of complex sentences obtained by combination of simple sentences and 3 types of simple sentences lexically containing comparative meanings in their adjectival predicates. These 3 types in each case correspond to the adjectival sentences in English like "as Adj as N", "more/less Adj than N", "the most Adj among N". In this study, these classes are formally defined by syntactic structures like ‘Na Nb-mankeum Nc-ga Adj', ‘Na Nb-boda Nc-ga deu Adj', ‘Na Nb-jungeso Nc-ga gajang Adj', ‘Na Nb-wa Nc-ga Adj', ‘Na Nb-boda Nc-ga Adj', ‘Na Nb-jungeso Nc-ga Adj', and named <ACB>, <ACR>, <ACT>, <ASB>, <ASR>, and <AST> respectively. The other types of comparative sentences which are not treated in this study should be analyzed as well to complete a global description of comparative sentences in order that we extract automatically comparative opinions from on-line documents.

19

영상에서 객체와 배경의 색상 특징을 이용한 자동 객체 추출 기법 KCI 등재

이승갑, 박영수, 이강성, 이종용, 이상훈

한국디지털정책학회 디지털융복합연구 제11권 제12호 2013.12 pp.459-465

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

본 논문은 영상 속 객체와 배경의 컬러 특징을 이용한 주요 객체의 자동 추출 방법에 관한 연구이다. 인간 이 객체를 판단할 때에는 배경과 객체의 색상 차이를 이용하는데 이러한 요소를 객체 추출 방법에 적용시키기 위해 서는 배경과 객체의 색차를 강조하여야 한다. 따라서 본 논문에서는 원 RGB 영상을 인간의 시각 시스템과 유사한 HSV 색 공간으로 변환하고 각기 다른 분포도의 메디안 필터를 적용한 두 개의 영상을 생성한 뒤 두 개의 메디안 필 터가 적용된 영상들을 합산하였고 데이터 군집화 방법인 Mean Shift 알고리즘을 적용하여 색상 특징을 그룹화 하였 다. 마지막으로 이진화 작업을 위하여 영상의 채널 수를 3 채널에서 1 채널로 정규화 한 뒤 영상 내 픽셀들의 평균 값을 임계값으로 이용하는 이진화 방법으로 객체 지도 영상을 생성하였고 주요 객체를 추출하였다.

This paper is a study on an object extraction method which using color features of an object and background in the image. A human recognizes an object through the color difference of object and background in the image. So we must to emphasize the color's difference that apply to extraction result in this image. Therefore, we have converted to HSV color images which similar to human visual system from original RGB images, and have created two each other images that applied Median Filter and we merged two Median filtered images. And we have applied the Mean Shift algorithm which a data clustering method for clustering color features. Finally, we have normalized 3 image channels to 1 image channel for binarization process. And we have created object map through the binarization which using average value of whole pixels as a threshold. Then, have extracted major object from original image use that object map.

20

GrabCut의 자동 객체 추출을 위한 저주파 영역 탐지 기반의 윈도우 생성 기법 KCI 등재

유태훈, 이강성, 이상훈

한국디지털정책학회 디지털융복합연구 제10권 제8호 2012.09 pp.211-217

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

기존의 GrabCut 알고리즘은 자동 객체 추출이 아닌 사용자가 객체 영역에 사각형 윈도우를 설정해야하는 알고리즘이다. 본 논문에서는 자동 시스템으로 변환하기 위해 인간의 시각 시스템을 기반으로 영상에서 가장 눈에 띄는 영역을 탐지하는 방법을 연구하였다. 주의 시각 영역인 Saliency Map을 생성하기 위해서 인간이 색채를 감지하는 ‘적/녹’ ‘황/청’의 대립색설을 기반으로 하는 Lab 색공간을 이용하여 생성한다. 생성된 Saliency Map을 주파수 공간으로 변환하여 저주파 영역에 국부적인 경계를 나타내고 경계를 탐지해내어 Saliency Point를 생성한다. 이렇게 생성된 Saliency Point의 좌표 값을 이용하여 윈도우를 자동으로 생성한 후 GrabCut 알고리즘을 기반으로 객체를 추출하였다. 다양한 영상에 제안한 알고리즘을 적용한 결과 객체 영역에 자동으로 윈도우가 생성되었고 객체가 추출되었다.

Conventional GrabCut algorithm is semi-automatic algorithm that user must be set rectangle window surrounds the object. This paper studied automatic object detection to solve these problem by detecting salient region based on Human Visual System. Saliency map is computed using Lab color space which is based on color opposing theory of 'red-green' and 'blue-yellow'. Then Saliency Points are computed from the boundaries of Low-Frequency region that are extracted from Saliency Map. Finally, Rectangle windows are obtained from coordinate value of Saliency Points and these windows are used in GrabCut algorithm to extract objects. Through various experiments, the proposed algorithm computing rectangle windows of salient region and extracting objects has been proved

 
1 2 3 4 5
페이지 저장