Earticle

현재 위치 Home 검색결과

결과 내 검색

발행연도

-

학문분야

자료유형

간행물

검색결과

검색조건
검색결과 : 563
No
1

시간흐름을 고려한 특징 추출과 군집 분석을 이용한 헬스리스크 관리 KCI 등재

강지수, 정경용, 정호일

한국융합학회 한국융합학회논문지 제12권 제1호 2021.01 pp.99-104

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

본 논문에서는 시간 흐름을 고려한 특징추출과 군집분석을 이용한 헬스 리스크 관리를 제안한다. 제안하는 방법은 세단계로 진행한다. 첫 번째는 전처리 및 특징추출 단계이다. 이는 웨어러블 디바이스를 이용하여 라이프로그를 수집하여 불완전데이터, 에러, 잡음, 모순된 데이터를 제거하며 결측 값을 처리한다. 그 다음 특징추출을 위해 주성분 분석을 통해 중요 변수를 선택하고, 상관계수와 공분산을 통해 데이터 간의 관계와 유사한 데이터들의 분류를 진행한 다. 또한 라이프로그에서 추출한 특징을 분석하기 위해 시간의 흐름을 고려하여 K-means 알고리즘을 통해 동적 군집 을 진행한다. 새로운 데이터는 오차 제곱합의 증가분을 기반으로 유사성 거리 측정 방법을 통해 군집을 진행하고, 시간 의 흐름을 고려하여 군집에 대한 정보를 추출한다. 따라서 특징 군집을 통해 헬스 의사결정 시스템을 이용하여 신체적 특성, 생활습관, 질병여부, 헬스케어 이벤트 발생위험, 예상 정도 등의 요소를 통해 리스크를 관리할 수 있다. 성능평가 는 Precision, Recall, F-measure을 사용하여 제안하는 방법과 퍼지방법, 커널기반 방법을 비교한다. 평가결과 제안 하는 방법이 우수하게 평가된다. 따라서 제안하는 방법을 통해 유병자와의 유사도를 이용하여 정확한 사용자의 잠재적 건강 위험을 예측 및 적절한 관리가 가능하다.

In this paper, we propose health risk management using feature extraction and cluster analysis considering time flow. The proposed method proceeds in three steps. The first is the pre-processing and feature extraction step. It collects user’s lifelog using a wearable device, removes incomplete data, errors, noise, and contradictory data, and processes missing values. Then, for feature extraction, important variables are selected through principal component analysis, and data similar to the relationship between the data are classified through correlation coefficient and covariance. In order to analyze the features extracted from the lifelog, dynamic clustering is performed through the K-means algorithm in consideration of the passage of time. The new data is clustered through the similarity distance measurement method based on the increment of the sum of squared errors. Next is to extract information about the cluster by considering the passage of time. Therefore, using the health decision-making system through feature clusters, risks able to managed through factors such as physical characteristics, lifestyle habits, disease status, health care event occurrence risk, and predictability. The performance evaluation compares the proposed method using Precision, Recall, and F-measure with the fuzzy and kernel-based clustering. As a result of the evaluation, the proposed method is excellently evaluated. Therefore, through the proposed method, it is possible to accurately predict and appropriately manage the user’s potential health risk by using the similarity with the patient.

2

감마톤 특징 추출 음향 모델을 이용한 음성 인식 성능 향상 KCI 등재

안찬식, 최기호

한국디지털정책학회 디지털융복합연구 제11권 제7호 2013.07 pp.209-214

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

음성 인식 시스템에서는 인식 성능 향상을 위한 방법으로 인간의 청취 능력을 인식 시스템에 접목하였으며 잡음 환경에서 음성 신호와 잡음을 분리하여 원하는 음성 신호만을 선택할 수 있도록 구성되었다. 하지만 실용적 측면에서 음성 인식 시스템의 성능 저하 요인으로 인식 환경 변화에 따른 잡음으로 인한 음성 검출이 정확하지 못하여 일어나는 것과 학습 모델이 일치하지 않는 것을 들 수 있다. 따라서 본 논문에서는 음성 인식 향상을 위해 감마톤을 이용하여 특징을 추출하고 음향 모델을 이용한 학습 모델을 제안하였다. 제안한 방법은 청각 장면 분석을 이용한 특징을 추출을 통해 인간의 청각 인지 능력을 반영하였으며 인식을 위한 학습 모델 과정에서 음향 모델을 이용하여 인식 성능을 향상시켰다. 성능 평가를 위해 잡음 환경의 -10dB, -5dB 신호에서 잡음 제거를 수행하여 SNR을 측정한 결과 3.12dB, 2.04dB의 성능이 향상됨을 확인하였다.

Improve the recognition performance of speech recognition systems as a method for recognizing human listening skills were incorporated into the system. In noisy environments by separating the speech signal and noise, select the desired speech signal. but In terms of practical performance of speech recognition systems are factors. According to recognized environmental changes due to noise speech detection is not accurate and learning model does not match. In this paper, to improve the speech recognition feature extraction using gamma tone and learning model using acoustic model was proposed. The proposed method the feature extraction using auditory scene analysis for human auditory perception was reflected In the process of learning models for recognition. For performance evaluation in noisy environments, -10dB, -5dB noise in the signal was performed to remove 3.12dB, 2.04dB SNR improvement in performance was confirmed.

3

4,000원

음성 인식 시스템은 정확하지 않게 입력된 음성으로부터 학습 모델을 구성하고 유사한 음소 모델로 인식하 기 때문에 인식률 저하를 가져온다. 따라서 본 논문에서는 바타차랴 알고리즘을 이용한 음성 인식 최적 학습 모델 구 성 방법을 제안하였다. 음소가 갖는 특징을 기반으로 학습 데이터의 음소에 HMM 특징 추출 방법을 이용하였으며 유사한 학습 모델은 바타챠랴 알고리즘을 이용하여 정확한 학습 모델로 인식할 수 있도록 하였다. 바타챠랴 알고리즘 을 이용하여 최적의 학습 모델을 구성하여 인식 성능을 평가하였다. 본 논문에서 제안한 시스템을 적용한 결과 음성 인식률에서 98.7%의 인식률을 나타내었다.

Speech recognition system is shall be composed model of learning from the inaccurate input speech. Similar phoneme models to recognize, because it leads to the recognition rate decreases. Therefore, in this paper, we propose a method of speech recognition optimal learning model configuration using the Bhattacharyya algorithm. Based on feature of the phonemes, HMM feature extraction method was used for the phonemes in the training data. Similar learning model was recognized as a model of exact learning using the Bhattacharyya algorithm. Optimal learning model configuration using the Bhattacharyya algorithm. Recognition performance was evaluated. In this paper, the result of applying the proposed system showed a recognition rate of 98.7% in the speech recognition.

4

MFCC와 LPC 특징 추출 방법을 이용한 음성 인식 오류 보정 KCI 등재

오상엽

한국디지털정책학회 디지털융복합연구 제11권 제6호 2013.06 pp.137-142

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

음성 인식 시스템은 부정확한 음성 신호의 입력으로 특징을 추출하여 인식할 경우 오인식의 결과가 나타나 거나 유사한 음소로 인식된다. 따라서 본 논문에서는 음소가 갖는 특징을 기반으로 음소 유사율과 신뢰도 측정을 이 용한 음성 인식 오류 보정 방법을 제안하였다. 음소 유사율은 학습 모델의 음소에 MFCC와 LPC 특징 추출 방법을 이용하여 구하였으며 신뢰도로 측정하였다. 음소 유사율과 신뢰도를 측정하여 오인식되는 오류를 최소화하였으며 음 성 인식 과정에서 오류로 판명된 음성에 대하여 오류 보정을 수행하였다. 본 논문에서 제안한 시스템을 적용한 결과 98.3%의 인식률과 95.5%의 오류 보정율을 나타내었다.

Speech recognition system is input of inaccurate vocabulary by feature extraction case of recognition by appear result of unrecognized or similar phoneme recognized. Therefore, in this paper, we propose a speech recognition error correction method using phoneme similarity rate and reliability measures based on the characteristics of the phonemes. Phonemes similarity rate was phoneme of learning model obtained used MFCC and LPC feature extraction method, measured with reliability rate. Minimize the error to be unrecognized by measuring the rate of similar phonemes and reliability. Turned out to error speech in the process of speech recognition was error compensation performed. In this paper, the result of applying the proposed system showed a recognition rate of 98.3%, error compensation rate 95.5% in the speech recognition.

5

다중센서 영상융합을 위한 FACET기반의 특징점 추출 KCI 등재후보

이완재, 박장한

한국방위산업학회 한국방위산업학회지 제17권 제1호 2010.06 pp.103-126

※ 기관로그인 시 무료 이용이 가능합니다.

6,100원

In this paper, we propose a FACET-based feature point extraction method and an image registration method for image fusion based on multi-sensors. The most important part of image registration in image fusion is the extraction of feature points in IR and VIS images. The proposed method of extracting feature points uses the FACET-based filter in an IR image and uses Harris corner detector in a VIS image because it is difficult to find common characteristics and correlations in the IR and VIS images. The proposed method of image registration consists of 5 stages: 1) extraction of the feature points, 2) detection of correspondence points by using mutual information (MI), 3) a 2D projective transformation, 4) interpolation to integers of a pixel coordination, 5) a similar measurement by using normalized mutual information (NMI). The final outcome of the fusion image expresses a pseudo color based on hue (H), saturation (S), and intensity (I). Experimental results showed that the proposed method provides robust and accurate registration results. The expression of a pseudo color image after image fusion showed methods for enhanced target detection.

6

An Infant Audio Classification Using Deep Learning Technology

Won Gyeong Hong, Eunjee Lee, Jinhwa Kim

대한산업경영학회 International Journal of Intelligent Technologies and Innovative Practices Vol. 1 No. 1 2026.01 pp.25-31

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

The integration of deep learning techniques in the field of audio signal processing has marked a significant leap forward in the capability to analyze and classify complex sounds, including the nuanced and information-rich cries of infants. Deep learning's promise in this domain lies in its potential to decipher the subtle cues contained within these cries, offering insights into an infant's health, emotional state, and developmental needs. This potential application stands at the intersection of technology and healthcare, promising to enhance our understanding and response to the needs of the youngest members of society. This study experimentally demonstrates that a convolutional neural network–based audio classification model effectively learns discriminative spectral and temporal features from audio signals. Experimental results show that the proposed convolutional neural networks architecture achieves significantly higher classification accuracy than traditional machine-learning baselines, particularly when trained on spectrogram-based representations. The findings confirm that deep learning models not only improve overall performance but also provide robust generalization across different audio classes and noisy conditions.

7

4,000원

Wireless sensor systems are primarily used for monitoring natural environments or industrial automation. The physical environment where these systems are installed is often unstable, making it difficult to replenish sensor energy immediately. Complex and harsh conditions can impact the network's structure, affecting monitoring performance. Wireless sensor systems consist of hundreds of sensors that collect data from hazardous environments and transmit information to a central system. However, due to the system's physical structure, information delays or losses may occur. This paper proposes a distance-based tree structure to address these issues in wireless sensor systems, and experimental results confirm its superior performance.

8

DTW 거리 기반 kNN을 활용한 시계열 데이터 정보 추출 및 회귀 예측 KCI 등재

양현준, 임채국, 정우혁, 우지환

한국경영정보학회 경영정보학연구 제26권 제2호 2024.05 pp.83-93

※ 기관로그인 시 무료 이용이 가능합니다.

4,200원

본 연구는 도금욕 공정의 완성도 예측을 위한 시계열 데이터의 효과적인 표현을 목표로, Dynamic Time Warping(DTW) 및 k-Nearest Neighbors(kNN) 기반의 전처리 방법론을 제안한다. 제안된 DTW 기반 kNN 전처리 방법을 다양한 회귀 모델에 적용하여 비교한 결과, 기존 결정 나무(Decision tree) 대비 최대 RMSE에서 43%과 MAE에서 24% 개선된 성능 향상을 보였으며, 신경망 구조를 갖는 회귀 모델과 결합했을 때 성능 향상이 두드러졌다. 본 논문에서 제안하는 전처리 방법과 회귀 모델을 결합한 구조는 길이가 긴 시계열 데이터와 제한된 데이터 샘플이 있는 상황에서 적합할 것으로 사료되며, 데이터가 부족한 상황에서도 과적합의 위험을 감소시키며, 합리적인 예측을 가능하게 함을 시사한다. 그러나 DTW 및 kNN 알고리즘은 데이터 샘플이 많아질수록 연산량이 늘어난다는 한계가 존재하며, 향후 연구를 통해 이러한 계산 효율성의 문제를 개선할 수 있는 연구가 필요할 것으로 보인다.

This study proposes a preprocessing methodology based on Dynamic Time Warping (DTW) and k-Nearest Neighbors (kNN) to effectively represent time series data for predicting the completion quality of electroplating baths. The proposed DTW-based kNN preprocessing approach was applied to various regression models and compared. The results demonstrated a performance improvement of up to 43% in maximum RMSE and 24% in MAE compared to traditional decision tree models. Notably, when integrated with neural network-based regression models, the performance improvements were pronounced. The combined structure of the proposed preprocessing method and regression models appears suitable for situations with long time series data and limited data samples, reducing the risk of overfitting and enabling reasonable predictions even with scarce data. However, as the number of data samples increases, the computational load of the DTW and kNN algorithms also increases, indicating a need for future research to improve computational efficiency.

9

외곽 검출로 보강한 Pyramid ViT 기반 Image Segmentation 방법

배주원, 서동환

한국ITS학회 한국ITS학회 학술대회 Inclusive ITS Technologies 2024.04 pp.295-297

※ 기관로그인 시 무료 이용이 가능합니다.

3,000원

10

4,000원

모바일 악성 앱이 급증하고 있으며, 전 세계 모바일 OS 시장의 대부분을 차지하고 있는 안드로이드가 모바일 사이버 보안 위협의 주요 대상이 되고 있다. 따라서 빠르게 진화하는 악성 앱에 대응하기 위해 인공지능 구현기술 중 하나인 기계학습을 활용한 악성 앱 탐지 기법의 필요성이 대두되고 있다. 본 논문은 악성 앱의 탐지 성능을 향상할 수 있는 특성 선택 및 특성 추출을 이용한 특성 선별 방법을 제안하였다. 특성 선별 과정에서 특성 개수에 따라 탐지 성능이 향상되었으며, 권한보다 API가 상대적으로 좋은 탐지 성능을 보였고, 두 특성을 조합하 면 평균 93% 이상의 높은 탐지 정밀도를 보여 적절한 특성의 조합이 탐지 성능을 높일 수 있음을 확인하였다.

Mobile malicious apps are increasing rapidly, and Android, which accounts for most of the global mobile OS market, is becoming a major target of mobile cyber security threats. Therefore, in order to cope with rapidly evolving malicious apps, there is a need for detection techniques of malicious apps using machine learning, one of artificial intelligence implementation technologies. In this paper, we propose a selected feature method using feature selection and feature extraction that can improve the detection performance of malicious apps. In the feature selection process, the detection performance improved according to the number of features, and the API showed relatively better detection performance than the permission. Also combining the two characteristics showed high precision of over 93% on average, confirming that the appropriate combination of characteristics could improve the detection performance.

11

4,000원

Mental health problems leading to depression have become a critical concern due to the growing engagement of people on social media platforms. Several past approaches have been implemented by analyzing the pattern and behaviour of the posts by users on social networking sites. This research study proposed a system for predicting users who may be depressed, based on the characteristics of users who is already affected. A combination of both the tweet-level and the user-level architecture was used to generate a more robust and reliable system where semantic embeddings trained from advanced neural networks were adopted under the tweet-level. SVM with Word2Vec and TF-IDF has been used and yielded an accuracy of 98.14% and recall of 95.63%.

12

4,000원

본 논문에서는 생성 모델로 색칠하기 게임에서 사용 가능하도록 임의의 선화와 원하는 컬러링 스타일을 입력하 면 자동으로 컬러링 영상을 생성하는 신경망 모델인 FillingGAN을 제안한다. 제안된 모델은 스타일 영상의 특징 을 추출하는 오토 인코더 구조의 모듈과 추출된 스타일 영상의 특징을 선화에 적용해서 이미지를 생성하는 GAN 모델로 구성된다. GAN 모델은 선화에서 추출된 구조와 스타일 영상에서 추출된 색 정보를 이용해서 채색 영상을 생성하는 과정을 수행하며, 이를 위해서 선화의 구조와 스타일 영상의 색 정보를 유지하는 손실 함수를 설계한다. 우리의 모델은 선화의 고유한 특징을 보존하며 스타일이 적용된 이미지를 생성한다.

In this paper, we contribute to the field of game by presenting FillingGAN, an automatic coloring framework using a generative adversarial network (GAN). FillingGAN is devised to generate a coloring image from a line drawing by filling empty regions between the lines in the line drawing image from the coloring styles learned from sample coloring images. Our model consists of two style extracting modules and a GAN model. The style extracting modules are designed as an auto-encoder that extracts feature from input images. One module extracts color styles from coloring sample images and the other extracts structure from a line drawing. FillingGAN executes coloring process by applying the coloring styles from coloring samples to a line drawing. It determines the similarity between the generated coloring image and the input line drawing, and calculate the perceptual loss to preserve the structural similarity. FillingGAN generates coloring images with preserved details by adjusting the weights between the structural feature vectors and the style feature vectors. As a result, it generates a stylized image without distorting the unique features of a line drawing. Our framework can be improved to apply various styles including artistic media strokes to line drawings.

13

Local Prominent Directional Pattern을 이용한 얼굴 사진과 스케치 영상 성별인식 방법 KCI 등재

Farkhod Makhmudkhujaev, 채옥삼

한국융합보안학회 융합보안논문지 제19권 제2호 2019.06 pp.91-104

※ 기관로그인 시 무료 이용이 가능합니다.

4,600원

본 논문에서는 성별 인식을 위해 얼굴 영상을 효과적으로 기술하는 새로운 지역 패턴 방법 Local Prominent Directional Pattern (LPDP)를 제안한다. 제안된 LPDP 방법은 성별 인식에 중요한 얼굴 모양을 명확하게 구분하기 위해 주변 패턴이 누 적된 히스토그램을 통계적으로 분석하고 패턴 변화가 크게 발생하는 픽셀을 부호화 한다. 통계적인 정보를 사용하는 얼굴 모 양 구분에 중요한 뚜렷한 에지 방향 패턴 영역을 구분하는 중요한 정보를 제공 할 수 있다. 이는 뚜렷한 에지 방향 패턴이 나 타나는 영역의 주변도 유사한 에지 방향 패턴이 나타내기 때문에 통계적으로 특정 방향이 히스토그램에 많이 누적될 수 있기 때문이다. 또한 통계적인 방법은 주변 영역의 정보를 많이 수용하기 때문에 잡음으로 발생하는 에지 방향 변화 오류에 강력한 장점이 있다. 제안된 방법은 기존 방법들 보다 더 강력한 성별인식에 중요한 얼굴 모양 구분 능력을 보여주면서 국소적으로 발생하는 잡음에 견고함을 보여준다. 우리는 제안된 방법의 성능을 평가하기 위해 밝기, 표정, 연령, 머리 포즈가 변화하는 성 별 인식 데이터 셋에 다양한 실험을 실험 했고 기존 방법 보다 제안된 방법의 성능이 우수함을 입증했다.

In this paper, we present a novel local descriptor, Local Prominent Directional Pattern (LPDP), to represent the description of facial images for gender recognition purpose. To achieve a clearly discriminative representation of local shape, presented method encodes a target pixel with the prominent directional variations in local structure from an analysis of statistics encompassed in the histogram of such directional variations. Use of the statistical information comes from the observation that a local neighboring region, having an edge going through it, demonstrate similar gradient directions, and hence, the prominent accumulations, accumulated from such gradient directions provide a solid base to represent the shape of that local structure. Unlike the sole use of gradient direction of a target pixel in existing methods, our coding scheme selects prominent edge directions accumulated from more samples (e.g., surrounding neighboring pixels), which, in turn, minimizes the effect of noise by suppressing the noisy accumulations of single or fewer samples. In this way, the presented encoding strategy provides the more discriminative shape of local structures while ensuring robustness to subtle changes such as local noise. We conduct extensive experiments on gender recognition datasets containing a wide range of challenges such as illumination, expression, age, and pose variations as well as sketch images, and observe the better performance of LPDP descriptor against existing local descriptors

14

4,000원

정확한 인식률을 보이고 있는 상업적인 음성인식 시스템은 화자종속 고립데이터로부터 학습 모델을 사용한다. 그러 나 잡음 환경에서 데이터양에 따라 음성인식의 성능이 저하되는 문제점이 있다. 본 논문에서는 가우시안 분포에서 Maximum Log Likelihood를 이용한 벡터 양자화 기반 음성 인식 성능 향상을 제안한다. 제안하는 방법은 음성에 대한 특징 을 가지고 벡터 양자화와 Maximum Log Likelihood 음성 특징 추출 방법을 이용하여 유사 음성에 대한 음성 인식의 정확성 을 높이는 최적 학습 모델 구성 방법이다. 이를 위해 HMM을 기반으로 음성 특징을 추출하는 방법을 사용한다. 제안하는 방법을 사용하여 기존 시스템에서 생성되어 사용되는 음성 모델에 대한 부정확한 음성 모델에 대한 정확성을 향상시킬 수 있으므로 음성 인식에 강인한 모델을 구성할 수 있다. 제안하는 방법은 음성 인식 시스템에서 향상된 인식의 정확도를 보인다.

Commercialized speech recognition systems that have an accuracy recognition rates are used a learning model from a type of speaker dependent isolated data. However, it has a problem that shows a decrease in the speech recognition performance according to the quantity of data in noise environments. In this paper, we proposed the vector quantization based speech recognition performance improvement using maximum log likelihood in Gaussian distribution. The proposed method is the best learning model configuration method for increasing the accuracy of speech recognition for similar speech using the vector quantization and Maximum Log Likelihood with speech characteristic extraction method. It is used a method of extracting a speech feature based on the hidden markov model. It can improve the accuracy of inaccurate speech model for speech models been produced at the existing system with the use of the proposed system may constitute a robust model for speech recognition. The proposed method shows the improved recognition accuracy in a speech recognition system.

15

거리 기반의 특징 선택을 이용한 간질 분류 KCI 등재

이상홍

한국디지털정책학회 디지털융복합연구 제12권 제8호 2014.08 pp.321-327

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

특징 선택은 중복 또는 서로간의 관련이 없는 특징을 제거하여 분류 성능을 향상시키는 기술이다. 본 논문 에서는 가중 퍼지소속함수 기반 신경망 (Neural Network with Weighted Fuzzy Membership Functions; NEWFM)에서 제공하는 가중 퍼지소속함수의 경계합 (Bounded Sum of Weighted Fuzzy Membership functions, BSWFM)의 무게중 심간의 거리를 이용한 새로운 특징 선택을 제안하여 분류 성능을 향상시켰다. 이러한 거리 기반의 특징 선택을 이용 하여 초기 24개의 특징으로부터 무게중심간의 거리가 짧은 특징을 하나씩 제거되면서 분류 성능이 가능 높은 22개 의 최소 특징을 선택하였다. 이들 22개의 최소 특징을 NEWFM의 입력으로 사용하여 97.7%, 99.7%, 98.7%의 민감 도, 특이도, 정확도를 각각 구하였다.

Feature selection is the technique to improve the classification performance by using a minimal set by removing features that are not related with each other and characterized by redundancy. This study proposed new feature selection using the distance between the center of gravity of the bounded sum of weighted fuzzy membership functions (BSWFMs) provided by the neural network with weighted fuzzy membership functions (NEWFM) in order to improve the classification performance. The distance-based feature selection selects the minimum features by removing the worst features with the shortest distance between the center of gravity of BSWFMs from the 24 initial features one by one, and then 22 minimum features are selected with the highest performance result. The proposed methodology shows that sensitivity, specificity, and accuracy are 97.7%, 99.7%, and 98.7% with 22 minimum features, respectively.

16

서베일런스에서 피셔의 선형 판별 분석을 이용한 사람 검출의 성능 향상 KCI 등재

강성관, 이정현

한국디지털정책학회 디지털융복합연구 제11권 제12호 2013.12 pp.295-302

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

사람 검출은 정지된 영상 혹은 동영상으로부터 사람의 움직임이나 자세를 추정하고, 사람이 찾아질 경우 영 상 내 사람의 좌표, 동작 인식, 보안관련 인증 등을 알아내는 기술로 정의된다. 이러한 사람 검출은 다른 객체의 검 출이나 사람과 컴퓨터와의 상호작용, 동작 인식 등의 기초 기술로서 해당 시스템의 성능에 영향을 미치는 매우 중요 한 변수 중에 하나이다. 그러나 영상 내의 사람은 움직임, 자세, 크기, 빛의 방향 및 밝기, 다른 객체와의 중복 등의 환경적 변화로 인해 사람 모양이 다양해지므로 정확하고 빠른 검출이 어렵다. 따라서 본 논문에서는 피셔의 선형 판 별 분석을 이용하여 몇 가지 환경적 조건을 극복한 정확하고 빠른 사람 검출 방법을 제안한다. 제안된 방법은 사람 움직임 및 자세와 배경에 무관하게 빠른 시간 안에 사람을 검출하는 것이 가능하다. 이를 위해 계층적인 방법으로 사람 검출을 수행하며, 휴리스틱한 방법, 피셔의 판별 분석을 이용하여 사람 검출을 수행하고, 검색 영역의 축소와 선형 결정의 계산 시간의 단축으로 검출 응답 시간을 빠르게 하였다. 추출된 사람 영상에서 사람의 자세를 추정하고 사람의 영역을 검출함으로써 사람 정보의 사용에 있어 보다 많은 정보를 추출할 수 있도록 하였다.

Many reported methods assume that the people in an image or an image sequence have been identified and localization. People detection is one of very important variable to affect for the system's performance as the basis technology about the detection of other objects and interacting with people and computers, motion recognition. In this paper, we present an efficient linear discriminant for multi-view people detection. Our approaches are based on linear discriminant. We define training data with fisher Linear discriminant to efficient learning method. People detection is considerably difficult because it will be influenced by poses of people and changes in illumination. This idea can solve the multi-view scale and people detection problem quickly and efficiently, which fits for detecting people automatically. In this paper, we extract people using fisher linear discriminant that is hierarchical models invariant pose and background. We estimation the pose in detected people. The purpose of this paper is to classify people and non-people using fisher linear discriminant.

17

4,000원

Data mining and game sounds classification prerequisite to find a compact but effective set of features in the overall problem-solving process. As a preprocessing step of data mining, feature selection has tuned to be very efficient in reducing its dimensionality and removing irrelevant data at hand. In this paper we cast a feature selection problem on rough set theory and a conditional entropy in information theory and present an empirical study on feature analysis for classical instrument classification. An new definition of a significance of each feature using rough set theory based on rough entropy is proposed. Our results suggest that further feature analysis research is necessary in order to optimize feature selection and achieve better results for the musical instrument sound classification problem through Weka’s classifiers. The results show that the performance of the best 17 selected features among 37 features has 3.601 compared to 2.332 in standard deviation and 94.667 compared to 96.935 in average with four classifiers.

18

4,000원

인식 대상 학습 모델이 분류되어 있지 않거나 명확하게 분류되지 않은 경우 어휘 인식을 결정하지 못하여 인 식률이 저하되며 학습 모델 분류 형태가 변경되거나 새로운 학습 모델이 추가되면 인식 모델의 결정 트리 구조가 변 경되어야 하는 구조적 문제가 발생한다. 이러한 문제점을 해결하기 위하여 학습 모델 분류를 위한 결정 트리 학습 알 고리즘을 제안한다. 음운 현상이 충분히 반영된 음성 데이터베이스를 구성하고 학습 효과를 확보하기 위하여 학습 모 델 분류를 위한 결정 트리 방법을 사용하였다. 본 연구에서는 실내 환경에 대하여 어휘 종속 인식과 어휘 독립 인식 실험을 수행한 결과 실내 환경의 어휘 종속 실험에서는 98.3%의 인식 성능을 보였고, 어휘 독립 실험에서 98.4%의 인식 성능을 보였다.

Target learning model is not recognized in this category or not classified clearly failed to determine if the vocabulary recognition is reduced. Form of classification learning model is changed or a new learning model is added to the recognition decision tree structure of the model should be changed to a structural problem. In order to solve these problems, a decision tree learning model for classification learning algorithm is proposed. Phonological phenomenon reflected sound enough to configure the database to ensure learning a decision tree learning model for classifying method was used. In this study, the indoor environment-dependent recognition and vocabulary words for the experimental results independent recognition vocabulary of the indoor environment-dependent recognition performance of 98.3% in the experiment showed, vocabulary independent recognition performance of 98.4% in the experiment shown.

19

서베일런스에서 Adaptive Boosting을 이용한 실시간 헤드 트래킹 KCI 등재

강성관, 이정현

한국디지털정책학회 디지털융복합연구 제11권 제2호 2013.02 pp.243-248

※ 기관로그인 시 무료 이용이 가능합니다.

4,000원

본 논문에서는 복잡한 배경에서의 사람의 머리 추적에 있어서 효과적인 Adaptive Boosting에 의한 방법을 제 안한다. 하나의 특징 추출 방법은 사람의 머리를 모델링하기에는 부족하다. 따라서 본 연구에서는 여러 가지 특징 추 출 방법을 병행하여 정확한 머리 검출을 시도하였다. 머리 영상의 특징 추출은 sub-region과 Haar 웨이블릿 변환(Haar wavelet transform)을 이용하였다. Sub-region은 머리의 지역적인 특징을 나타내고, Haar 웨이블릿 변환은 얼굴의 주파 수 특성을 나타내기 때문에 이들을 이용하여 특징을 추출하면 효과적인 모델링이 가능해 진다. 실시간으로 입력되는 영상에서 사람의 머리를 추적하기 위하여 제안하는 방법에서는 3가지 형태의 Harr-wavelet 특징을 AdaBoosting 알고 리즘으로 학습한 후 결과를 이용하였다. 원래 AdaBoosting 알고리즘은 학습시간이 매우 길며 학습데이터가 변하면 다 시 학습을 수행해야 하는 단점이 존재한다. 이 단점을 극복하기 위하여 제안하는 방법에서는 캐스케이드를 이용한 AdaBoosting의 효율적인 학습방법을 제안한다. 이 방법은 머리 영상에 대한 학습시간은 감소시키며, 학습데이터의 변 화에도 효율적으로 대처할 수 있다. 이 방법은 학습과정을 레벨별로 분리한 후 중요도가 높은 학습데이터를 다음 단 계에 반복적으로 적용시킨다. 제안하는 방법이 적은 학습 시간과 학습 데이터를 사용해서 우수한 성능을 가지는 분류 기를 생성하였다. 또한, 이 방법은 다양한 머리데이터를 가진 실시간 영상데이터에 적용한 결과 다양한 머리를 정확 하게 검출 및 추적하였다.

This paper proposes an effective method using Adaptive Boosting to track a person's head in complex background. By only one way to feature extraction methods are not sufficient for modeling a person's head. Therefore, the method proposed in this paper, several feature extraction methods for the accuracy of the detection head running at the same time. Feature Extraction for the imaging of the head was extracted using sub-region and Haar wavelet transform. Sub-region represents the local characteristics of the head, Haar wavelet transform can indicate the frequency characteristics of face. Therefore, if we use them to extract the features of face, effective modeling is possible. In the proposed method to track down the man's head from the input video in real time, we ues the results after learning Harr-wavelet characteristics of the three types using AdaBoosting algorithm. Originally the AdaBoosting algorithm, there is a very long learning time, if learning data was changes, and then it is need to be performed learning again. In order to overcome this shortcoming, in this research propose efficient method using cascade AdaBoosting. This method reduces the learning time for the imaging of the head, and can respond effectively to changes in the learning data. The proposed method generated classifier with excellent performance using less learning time and learning data. In addition, this method accurately detect and track head of person from a variety of head data in real-time video images.

20

A High-Precision Feature Extraction Network of Fatigue Speech from Air Traffic Controller Radiotelephony Based on Improved Deep Learning

Zhiyuan Shen, Yitao Wei

[NRF 연계] 한국통신학회 ICT Express Vol.7 No.4 2021.12 pp.403-413

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Air traffic controller (ATC) fatigue is receiving considerable attention in recent studies because it represents a major cause of air traffic incidences. Research has revealed that the presence of fatigue can be detected by analysing speech utterances. However, constructing a complete labelled fatigue data set is very time-consuming. Moreover, a manually constructed speech collection will often contain only little key information to be used effectively in fatigue recognition, while multilevel deep models based on such speech materials often have overfitting problems due to an explosive increase of model parameters. To address these problems, a novel deep learning framework is proposed in this study to integrate active learning (AL) into complex speech features selected from a large set of unlabelled speech data in order to overcome the loss of information. A shallow feature set is first extracted using stacked sparse autoencoder networks, in which fatigue state challenge features from a manually selected speaker set of are exploited as the input vector. A densely connected convolutional autoencoder (DCAE) is then proposed to learn advanced features automatically from spectrograms of the selected data to supplement the fatigue features. The network can be effectively trained using a relatively small number of labelled samples with the help of AL sampling strategies, and the addition of a dense block to the convolutional automatic encoder can decrease the number of parameters and make the model easier to fit. Finally, the two above-mentioned features are combined using multiple kernel learning with a support-vector-machine classifier. A series of comparative experiments using the Civil Aviation Administration of China radiotelephony corpus demonstrates that the proposed method provides a significant improvement in the detection precision compared to current state-of-the-art approaches.

 
1 2 3 4 5
페이지 저장