Earticle

현재 위치 Home 검색결과

결과 내 검색

발행연도

-

학문분야

자료유형

간행물

검색결과

검색조건
검색결과 : 21
No
1

물류 산업에서 빅 데이터 분석을 위한 텍스트 시각화 도구 설계 및 개발 KCI 등재

이강수, 이수안, 강석, 박찬민, 김진호

한국EA학회 정보화연구 제13권 2호 2016.06 pp.355-365

※ 기관로그인 시 무료 이용이 가능합니다.

4,200원

최근에 일반적으로 처리가 불가능할 정도로 데이터양이 증가하기 때문에 많은 사람들이 대규모 데이터를 저장 및 분석하기 위한 빅 데이터 기술에 대한 관심을 가지고 있다. 또한 많은 사람들이 빅 데이터 처리를 통해 나온 결과를 시각화 하는 기술에 대해서도 관심을 가지고 있다. 본 논문에서는 물 류 산업의 고객 예측과 운영 효율 그리고 신규 사업 등의 의사 결정을 위해서 빅 데이터 분석을 수행 하였다. 물류와 관련된 뉴스, 보도자료, 소셜 미디어와 블로그 등의 텍스트 데이터를 수집하는 도구를 개발하였다. 또한 R과 Shiny를 사용하여 물류 데이터에 대한 대화형 웹 인터페이스의 시각화 도구를 설계 및 개발하였고, 이를 통해 효과적인 탄력적인 경영과 의사 결정에 도움을 줄 수 있다.

Recently many people have an interest in big data technologies to store and analysis of large-scale data because it increases the amount of data that can not be processed in traditional technology. Also, many people are interested in the techniques of data visualization through the results of big data processing. In this paper, we perform big data analysis in order to make decisions about customer prediction, operational efficiency, and new projects of the logistics industry. We developed a tool to collect text data such as news, press releases, social media and blogs related to logistics. In addition, we have designed and developed a visualization tool for interactive web interface for logistics data using R and Shiny. And it can help in effective management and flexible decision-making.

2

6,900원

이 연구에서는 텍스트가 나타내는 추상적인 내용을 구체적으로 시각화하여 표현하는 그래픽 조직자(Graphic Organizer)에 대해 살펴보고 이를 NIE(Newspaper in Education)에 적용할 수 있는 가능성에 대해 살펴보고자 하였다. NIE는 “신문을 활용한 교육”으로써 수업 시간에 신문 기사 및 정보를 활용하여 교육적인 효과를 높이는 것을 말한다. 신문은 다양한 형태의 텍스트와 여러 분야의 정보를 제공하기 때문에 학습에 대한 동기를 부여할 수 있다. 또한 신문은 실제로 일어나는 최신의 정보를 제공하기 때문에 현실적인 현상을 바탕으로 학습자의 지식을 넓히고 논리와 사고력을 향상시킬 수 있다는 점에서 긍정적인 교수・학습 자료이다. 이러한 점은 의사소통 능력의 배양과 해당 문화권에 대한 이해를 궁극적인 목표로 추구하는 한국어 교육에서도 가치 있게 논의할 수 있다. 그렇지만 신문이 제공하는 텍스트가 한국어 학습자들에게 추상적이고 어려울 수 있기 때문에 신문 기사 텍스트를 쉽게 이해할 수 있는 전략이 필요하다. 그래픽 조직자는 언어 텍스트의 추상적인 개념 및 정보 간의 관계를 선, 화살표, 도형, 공간 배열 등으로 구조화하여 표현하는 시각적인 도구이다. 그래픽 조직자는 읽기 텍스트의 내용을 구조적으로 파악할 수 있도록 시각적인 형태로 보여주기 때문에 NIE에서 제시하는 텍스트를 교수・학습하는 과정에서 효율적으로 활용할 수 있다. 이 연구는 한국어 학습자를 위한 NIE의 한 방법으로 추상적인 텍스트를 시각적으로 구조화하여 제시할 수 있는 그래픽 조직자에 대해서 논의하였다는 점에서 의의를 가질 것이다. 또한 이 연구에서 논의한 그래픽 조직자는 읽기 텍스트뿐만 아니라 향후 기능별 한국어 교재를 개발할 때 효율적으로 활용할 수 있을 것이다.

The purpose of this paper is to examine the graphic organizer which can visualize text in an abstract concept into a concrete concept and examine the possibility that applies to NIE. NIE stands for “Newpaper in Education” which enhances the educational effect using the newspaper article in the class. The newspaper can motivate the students to learn because it offers various types of texts and diverse fields of information. Besides, the newspaper is the positive teaching and learning material because it gives the latest information which happens in real time and widens the students’ knowledge and enhances logical thinking. That’s why NIE can be discussed worthily in the Korean Language Education that pursues the objectives about building up communicative competence and understanding about the Korean culture. However, the text in newspaper is abstract and difficult to the students who are learning Korean Language, so that the graphic organizer can be used as a strategy for easier learning. The graphic organizer can visualize the abstract concepts and the relationship among information of the text into the concrete concepts using lines, arrows, figures and space arrangement. The graphic organizer can be used effectively in NIE because it shows the reading text into the visualized structure. This study has significance of presenting the graphic organizer that make the abstract text into visualization as one of the teaching-learning methods in NIE for the studends who are learning Korean. Besides, it is hopeful that the graphic organizer may be applied to develop the functional Korean language textbooks

3

데이터 시각화를 통한 표준해사통신영어의 어휘 분석과 활용 방안 KCI 등재

권유민, 이진석, 김주성

한국해양경찰학회 한국해양경찰학회보 제10권 제1호 통권 제33호 2020.02 pp.1-22

※ 기관로그인 시 무료 이용이 가능합니다.

5,800원

오늘날 선상에서 사용 언어로써 영어 활용 능력은 선박의 안전한 항해와 원활한 선내 의사소통을 위한 해기사의 주요 능력 중 하나로 자리 잡았다. 특히 오늘날 국 제 해상 교역의 증가와 다국적 선원의 혼승 증가로 선상에서 영어 구술 능력은 매우 중요한 해기사의 능력 중 하나로 여겨지고 있다. 따라서 해기사 양성을 위한 교육기 관에서는 해사영어와 해사통신영어를 정규 교육 과정으로 운영하고 있으며, 승선 실 습 수행 전 이러한 교과과정을 이수하도록 하고 있다. 그러나 선박과 항해에 대한 실질적인 지식과 경험의 부재로 인하여 표준해사통신영어에 대한 이해 및 주요 어휘 의 사용에 대한 이해가 부족한 현실이다. 본 논문에서는 표준해사통신영어의 어휘 분석을 통하여 주요 어휘와 표현에 대하여 분석하고, 효율적인 교과 운영을 위한 시 청각 자료 구성의 기초자료로 사용하고자 하였다. 분석 도구로는 통계 및 데이터 마 이닝을 위한 언어인 R과 Textom을 이용하여 어휘 빈도 분석과 word cloud 구성, 어 휘의 활용도를 분석하였다. 분석 결과 주요 단어의 사용 빈도와 어휘 쌍을 구성하였 으며, 각 어휘의 연결 빈도를 도출하였다. 전체 SMCP에서 중복 제거 어휘는 1,020 개로 나타났으며, 노출빈도 상위 단어와 단어 쌍에 대하여 분석하면 Part A에서는 선박과 선박, 선박과 육상 사이의 조난·긴급·안전 통신과 원조 요청 등의 통신에서 위치관련 교신이 주를 이루며, Part B에서는 선박 운항, 선내 안전, 소화, 손상 제어, 좌초, 수색·구조, 화물 작업, 입·출항 관리 등 주로 선내 통신을 다루어 이와 관련된 단어가 상위 빈도를 차지한 것으로 분석되었다. 향후 본 논문의 결과를 바탕으로 해 사영어 전문 용어집 제작과 시청각 수업 자료의 구성 등에 활용할 수 있을 것으로 기대한다.

The ability to use English as a working language on board has become one of the major competencies of the seafarers for safe navigation and smooth onboard communication. Especially today, the ability of English oral skills is considered to be one of the crucial skills of marines with the increase of international maritime trade and multinational crew onboard. Therefore, educational institutes for training seafarers are operating regular English courses in Maritime English, and trainees are required to complete the curriculums prior to boarding a vessel. However, understanding of standard maritime English and the use of major vocabulary are abstruse to the trainees due to the lack of practical knowledge and experience of vessels and ship navigations. In this paper, vocabularies and expressions were analyzed through the lexical analysis of the IMO Standard Maritime Communication Phrases, and it is aimed to be used as the basic data for audiovisual data composition for efficient English course operation. The frequency of vocabulary, word cloud configuration, and vocabulary utilization were analyzed through the R and Textom as the analysis tools, which are languages for statistical and data mining. As a result of analysis, the frequency of use and the lexical pairs of the main words were constructed, and the connection frequency of each vocabulary was analyzed. The number of vocabulary was 1,020 in the SMCP except for duplicated words. As the Part A are composed regarding the communication such as distress, urgency, safety communication and request for assistance between ship to ship and ship to shore, words related on ship's position were mainly composed in the Part A. Meanwhile, words related on internal communications and words about reporting and reception were mainly composed in the Part B because it mainly deals with on-board communication such as ship navigation, onboard safety, fire extinguishing, damage control, search and rescue, cargo operations, and arrival and departure operations. Based on the results of this paper, it is expected that it would be useful for the construction of vocabulary book and audio-visual data composition for standard maritime English courses.

4

합성곱 신경망(CNN)은 다수의 2차원 필터가 계층적으로 중첩된 구조를 가지므로, 그 내부 구조와 데이터 흐름을 직관적으로 시각화하는 데 본질적인 어려움이 있다. 기존의 시각화 도구들은 각기 다른 표현 방식과 추상화 수준을 사용함으로써 서로 다른 모델 간 구조를 비교하거나 해석하는 데 한계를 가지며, 이러한 비표준화된 시각화 환경은 연구자가 모델을 이해하고 비교하는 과정에서 인지적 부담을 초래한다. 본 연구에서는 이러한 문제를 개선하기 위해 CNN 아키텍처 설명으로부터 구조 정보를 추출하고, 이를 구조화된 표현으로 변환한 뒤, 상호작용 가능한 웹 기반 3차원 시각화 프레임워크를 제안한다. 제안하는 시스템은 대형 언어 모델(LLM)을 활용하여 논문 텍스트를 JSON 형태의 중간 표현으로 변환하고, 사용자가 해당 프레임워크 상에서 CNN 아키텍처를 확대·축소 및 회전 등의 상호작용을 통해 탐색하고 비교할 수 있도록 한다. 이러한 접근은 모델 구현 파일에 의존하지 않고 논문에 기술된 아키텍처 구조를 직관적으로 분석할 수 있는 환경을 제공하며 복잡한 모델에 대한 이해를 지원하고 연구의 접근성을 제고하는 데 기여한다.

Convolutional Neural Networks (CNNs) consist of hierarchically stacked two-dimensional filters, which makes their internal structures and data flows inherently difficult to visualize intuitively. Existing visualization tools employ diverse representation styles and abstraction levels, limiting consistent comparison and interpretation across different models and increasing cognitive load for researchers. To address this issue, we propose an interactive web-based 3D visualization framework that extracts architectural information from CNN descriptions in academic papers and converts it into a structured representation. The proposed system utilizes a Large Language Model (LLM) to transform unstructured paper text into a JSON-based intermediate format, enabling users to explore and compare CNN architectures through interactions such as zooming and rotation within the framework. This approach supports intuitive analysis of architectures described in papers without relying on model implementation files and contributes to improving the accessibility of deep learning research.

5

Readability Visualization for Massive Text Data SCOPUS

Hyoyoung Kim, Jin Wan Park, Dongsu Seo

보안공학연구지원센터(IJMUE) International Journal of Multimedia and Ubiquitous Engineering Vol.9 No.9 2014.09 pp.241-248

※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.

In general, people read texts and decide by themselves to measure levels of understanda-bility and readability, which takes a lot of time and efforts. We believe visualizing readability gives intuitive impact on how difficult the texts will be before examining the texts further. Text visualization aims to provide structural characteristics of text contents in an efficient way. By using massive text data, such as books or documents, this study suggests readability meas-urement factors and formulas for the suggested methods that visualize texts by extracting a key factor ‘length’ for readability. In addition to the proposed methods, this study verifies effectiveness of visualization through the test of the case studies. The paper also includes case study findings that readers can have readability information not from independent texts, but from the comparison of previous texts, and therefore it becomes easier to accommodate diffi-cult level of new books.

6

문자정보의 시각화가 정보수용에 미치는 영향 KCI 등재

최은희, 이진호

한국브랜드디자인학회 브랜드디자인학연구 Vol.18 No.1 통권 제53호 2020.03 pp.185-194

※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.

효과적인 정보전달을 위해서는 정보의 시각화, 단순화, 간결화가 필수적이다. SNS 등에서 문자보다는 이모티콘으로 이미지화 되어 표현되고 있는 것처럼 대부분의 정보전달은 디지털시대의 뉴미디어 매체를 통하여 이미지로서 발전하고 있다. 본 연구의 목표는 일차적으로 연구자가 속한 대학의 '디자인과 창의적 발상' 수업을 통하여 정보의 시각화 유무에 따른 수용자의 태도에 대한 실험을 진행하고 상관관계를 알아보고자 하는데 있다. 연구 방법은 선행연구와 문헌 조사 이후 2019년도 D대학교 인문계열 1학년 대상의 수업에서 200여명의 학생을 대상으로 데이터를 수집하고 통계 분석하여 결과를 얻을 수 있었다. 그 이후 문자정보의 시각화와 비시각화에 따른 인지욕구와 감정강도의 상관관계를 살펴보았다. 수집된 자료는 SPSS 23.0 프로그램으로 분석하였고, 피험자의 성별분포와 정보의 시각화 여부 선호도를 확인하기 위하여 빈도분석을 실시하였다. 인지능력과 감정강도의 상관관계는, 신뢰도 부분에서 감정강도가 낮은 그룹은 비시각화 부분이 더 높은 지표를 나타내고, 감정강도가 높은 경우 시각화 유무에 따른 차이가 미미하게 나타나는 유의미한 결과를 얻을 수 있었다. 또한 시각화와 인지욕구 차이의 상호작용 효과는 유의미하지 않은 것으로 나타났다. 이에 본 연구에서는 문자정보의 시각화와 비시각화를 인지욕구와 감정강도의 상관관계를 가지고 비교한 결과 유용성, 선호도에서는 시각화가, 신뢰도에서는 비시각화가 우위에 있다는 것을 알 수 있었다. 연구 결과의 공유를 통하여 논의를 촉진시키고 정보의 시각화에 대한 객관적이고 정량적인 자료가 뒷받침 된다면 디자인 사고교육에도 많은 도움이 될 것으로 생각한다.

Most information delivery is done by images rather than text information, through new media in this digital era. Instead of texts, the information is made in images of emoticons in social media environment. The primary purpose of this study is to conduct an experiment to identify the correlation regarding attitude of the receivers toward visualization in 'Design and Creative Thinking' class for other departments than Design department to which the researcher belongs. After the previous research and literature review, the data were collected and statistically analyzed for 200 students in the course of the first year of D University in Humanities. The data collected in this study were analyzed using the SPSS 23.0 program, and frequency analysis was conducted to confirm the gender distribution of the study subjects and the preference of visualization. As a result of experiments that analyzed the number of cases of acceptance of information according to the presence or absence of visualization by measuring cognitive ability and emotional strength, the application of visualization in accepting information is less reliable than non-visualized information acceptance, and non-visualized information is likely to be significantly more reliable, which was contrary to the hypothesis in which visualized information is more reliable, useful and preferred over non-visualized information. This study is therefore to design classes based on reliability of non-visualization as well as usefulness and preference of utilizing visualized images to recognize the significance and value of visualization of information in design thinking education.

7

Text Mining and Visualization of Papers Reviews Using R Language

Li, Jiapei, Shin, Seong Yoon, Lee, Hyun Chang

[Kisti 연계] 한국정보통신학회 Journal of information and communication convergence engineering Vol.15 No.3 2017 pp.170-174

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Nowadays, people share and discuss scientific papers on social media such as the Web 2.0, big data, online forums, blogs, Twitter, Facebook and scholar community, etc. In addition to a variety of metrics such as numbers of citation, download, recommendation, etc., paper review text is also one of the effective resources for the study of scientific impact. The social media tools improve the research process: recording a series online scholarly behaviors. This paper aims to research the huge amount of paper reviews which have generated in the social media platforms to explore the implicit information about research papers. We implemented and shown the result of text mining on review texts using R language. And we found that Zika virus was the research hotspot and association research methods were widely used in 2016. We also mined the news review about one paper and derived the public opinion.

8

A Method of Mining Visualization Rules from Open Online Text for Situation Aware Business Chart Recommendation

권오병

[Kisti 연계] 한국전자거래학회 한국전자거래학회지 Vol.25 No.1 2020 pp.83-107

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

데이터의 성격과 시각화의 목적에 따라 비즈니스 차트를 선택하는 것은 비즈니스 분석에 유용한 지식이다. 그러나 현재 시각화 도구에는 상황에 맞는 비즈니스 차트를 선택할 수 있는 기능이 부족하다. 또한 매번마다 시각화 방법에 대해 전문가의 도움을 요청하는 것은 비용과 시간이 소요된다. 따라서 본 연구의 목적은 온라인으로 게시된 문서로부터 비즈니스 차트 선정 규칙에 대한 지식을 추출하여 비즈니스 차트 생산성을 향상시키는 방법을 제안하는 것이다. 이를 위해 인터넷에서 비즈니스 차트를 묘사하는 한국어, 영어 및 중국어 비정형 데이터를 수집하고 TF-IDF를 사용하여 컨텍스트와 비즈니스 차트 간의 관계를 계산했다. 또한 Galois 래티스를 사용하여 비즈니스 차트 선택 규칙을 생성했다. 제안된 방법으로 생성된 규칙의 품질을 평가하기 위해 실험군과 대조군에 대해 실험을 수행했다. 그 결과 제안된 방법으로 의미 있는 규칙이 추출되었음을 확인했다. 본 연구의 결과물로 시각화 전문가의 도움 없이도 사무직 직원들이 비즈니스 차트를 효율적으로 선택할 수 있을 것으로 기대된다. 또한 작업 중인 문서를 기반으로 비즈니스 차트를 추천함으로 직원 교육에 유용할 것이다.

Selecting business charts based on the nature of the data and the purpose of the visualization is useful in business analysis. However, current visualization tools lack the ability to help choose the right business chart for the context. Also, soliciting expert help about visualization methods for every analysis is inefficient. Therefore, the purpose of this study is to propose an accessible method to improve business chart productivity by creating rules for selecting business charts from online published documents. To this end, Korean, English, and Chinese unstructured data describing business charts were collected from the Internet, and the relationships between the contexts and the business charts were calculated using TF-IDF. We also used a Galois lattice to create rules for business chart selection. In order to evaluate the adequacy of the rules generated by the proposed method, experiments were conducted on experimental and control groups. The results confirmed that meaningful rules were extracted by the proposed method. To the best of our knowledge, this is the first study to recommend customizing business charts through open unstructured data analysis and to propose a method that enables efficient selection of business charts for office workers without expert assistance. This method should be useful for staff training by recommending business charts based on the document that he/she is working on.

9

텍스트 시각화 및 핵심어휘목록을 통한 통사론 교육: 중등영어 임용수험자를 중심으로

안영재, 이혜진

[NRF 연계] 미래영어영문학회 영어영문학 Vol.26 No.3 2021.08 pp.339-362

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

This corpus-driven research purports to generate the lists of the most frequent words and keywords in the discipline of syntax, one of the most difficult and abstruse subjects for secondary English teacher appointment exam (hereinafter, SETAE) preparers. Taking a purposive approach to the text analysis, the two most widely-used syntax textbooks for English education majors and the past 41 SETAE syntax test questions were compiled as a specialized syntax corpus. The Wordsmith 8.0, a software that investigates the lexical features and patterns within a corpus, was run and then parsed to locate the most frequent words along with keywords. The frequent words offered preliminary insights into discipline-specific features of syntax. Moreover, the keywords, calculated with respect to the BNC reference corpus, reflected the aboutness of the syntax corpus. The derived keywords, exhibiting the statistically outstanding distinction compared to the reference corpus, revealed a significant number of salient terms specific to syntax. The results were visualized in accordance with the most frequent and distribution of the terms. This study is expected to yield several implications that might support and better equip students especially in preparing their SETAE.

10

쓰기 수행 수준에 따른 중학생 논설문의 텍스트 시각화 분석

이슬기, 박영민

[NRF 연계] 학습자중심교과교육학회 학습자중심교과교육연구 Vol.17 No.15 2017.08 pp.401-422

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

본 연구의 목적은 학생 글에 사용된 어휘를 시각화함으로써 쓰기 수준에 따른 학생 글의 차이를 직관적이면서도 정량적인 방법으로 분석하는 데 있다. 이를 위해 2017년 4월부터 3개월 간 중학생 40명의 논설문을 임의로 수집한 후 텍스트 마이닝의 방법으로 분석하였으며, 질적 수준에 대한 평정 결과에 따라 집단별로 어떤 차이가 있는지를 텍스트 시각화의 방법으로 제시하였다. 본 연구에서 얻은 결과는 다음과 같다. 첫째, 중학생 논술문에서는 전반적으로 쉬운 어휘의 사용 빈도가 높았다. 이를 통해 판단하건대 중학생들이 논설문 쓰기에서 사용하는 어휘의 수준은 어휘 발달 연구에서 제안된 것보다 낮은 경향을 보였다. 둘째, 쓰기 수준이 높은 학생 글에서는 상대적으로 명사와 서술어가 다양하게 쓰였다. 어휘 다양성의 차이가 쓰기 수준의 차이와 관련이 있는 것으로 보인다. 셋째, 접속부사의 사용은 쓰기 수준별로 뚜렷한 차이를 보였다. 하위 집단의 글에서는 문장을 시간 순서로 나열하는 접속부사가 주로 쓰인 반면, 상위 집단의 글에서는 인과와 역접과 같은 논리적 순서를 드러내는 접속부사가 주로 쓰였다. 본 연구에서는 이러한 차이를 텍스트 시각화의 방법으로 직관적이면서도 정량적으로 파악할 수 있도록 하였는데, 이를 통해서 학생 글을 텍스트 시각화의 방법으로 분석하는 것이 학생 글 평가에 활용 가능한 방안임을 탐색할 수 있었다

The purpose of this study was to analyze differences in students' writing according to their writing proficiency intuitively and quantitatively by visualizing the vocabulary used in their writing. To this end, during three month since 2017 April, 40 essays of middle school students were randomly collected and analyzed using text mining technique, and differences according to the evaluation results were presented using text visualization. The following results were obtained in this study. First, high use frequency of generally easy vocabulary was found in the essays of middle school students. Judging from the finding, the level of vocabulary used by middle school students in their essay writing tended to be lower than the level suggested in studies of vocabulary development. Second, use of various nouns and predicates were relatively high in the essays of students with high level of writing proficiency. The difference in vocabulary diversity appears to be related to the difference in writing proficiency. Third, clear differences were seen in the use of conjunctive adverbs according to the writing proficiency. Conjunctive adverbs that list sentences in chronological order were mostly used in the essays of low writing proficiency group, while conjunctive adverbs that show logical order such as cause and effect and adversatives were mostly used in the essays of high writing proficiency group. This study made it possible to intuitively and quantitatively understand such differences by text visualization, and through the findings, it was confirmed that analyzing students' essays by using text visualization is a valid method that can be effectively used in evaluating students' essays.

11

단어 구름과 동적 그래픽스 기법을 이용한 영어성경 텍스트 시각화

장대흥

[Kisti 연계] 한국통계학회 The Korean journal of applied statistics Vol.27 No.3 2014 pp.373-386

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

단어 구름은 문자 텍스트 상의 복수개의 단어들을 대상으로 그 단어들의 출현 빈도에 비례하는 글자의 크기나 글자의 색깔로 중요도를 나타내는 텍스트 시각화 방법이다. 이 그림은 텍스트 상의 핵심단어를 재빨리 인지하고 단어들의 상대적 출현빈도수에 맞추어 배열하는 데 유용하다. 동적 그래픽스를 이용하여 텍스트 장들의 변화에 따른 핵심단어와 단어출현빈도의 패턴의 변하는 모습을 살필 수 있다. 행들이 텍스트 상의 장들이고 열들이 텍스트에 출현하는 단어들의 출현빈도수 순위들인 단어출현빈도행렬을 정의할 수 있고 이 행렬을 이용하여 단어출현빈도행렬그림을 그릴 수 있다. 동적 그래픽스를 이용하여 출현빈도수 순위의 변화에 따른 단어출현빈도행렬의 패턴의 변하는 모습을 살필 수 있다. 우리는 단어 구름과 동적 그래픽스 기법을 사용하여 영어성경 텍스트 시각화를 수행할 수 있다.

A word cloud is a visualization of word frequency in a given text. The importance of each word is shown in font size or color. This plot is useful for quickly perceiving the most prominent words and for locating a word alphabetically to determine its relative prominence. With dynamic graphics, we can find the changing pattern of prominent words and their frequencies according to the changing selection of chapters in a given text. We can define the word frequency matrix. In this matrix, rows are chapters in text and columns are ranks corresponding to word frequency about the words in the text. We can draw the word frequency matrix plot with this matrix. Dynamic graphic can indicate the changing pattern of the word frequency matrix according to the changing selection of the range of ranks of words. We execute an English Bible text visualization using word clouds and dynamic graphics technology.

12

캐릭터 넷을 통한 내러티브 텍스트 시각화 디자인 연구

전혜정, 박승보, 이오준, 유은순

[Kisti 연계] 한국콘텐츠학회 한국콘텐츠학회논문지 Vol.15 No.2 2015 pp.86-100

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

인터넷 발전과 스마트 혁명을 거치며 사용자가 생산하는 데이터양이 중가하고 그 유형도 다양해졌다. 이렇게 방대한 양의 데이터를 분석하고 새로운 가치로 활용한다는 개념의 빅데이터가 새로운 이슈로 부상하였다. 더욱이 빅데이터 속의 콘텐츠들을 검색하기 위해서는 동영상이 포함하고 있는 스토리에 대한 분석과 시각화에 대한 연구가 필요하다. 따라서 본 연구에서는 등장인물들 간의 대화를 분석하여 스토리를 모델링하는 캐릭터 넷(Character-net)이라는 인터페이스를 개발하였다. 캐릭터 넷은 스토리가 있는 동영상을 분석해서 인물들을 자동으로 추출할 수 있고, 등장인물들 간의 관계를 자동으로 모형화 할 수 있다. 이로써 기존 연구와는 다른 방법으로 스토리를 가시화하는 툴의 가능성을 발견할 수 있었다. 하지만 아직 활용하기 어렵고 한 눈에 스토리 특징을 파악하기 어렵다는 단점이 발견되었다. 이러한 캐릭터 넷을 개선하기 위해서는 정보 디자인을 접목하여 해결할 수 있을 것이라 가정하였다. 따라서 본고에서는 먼저 데이터 정보디자인 분야에서의 시각화 디자인들을 간략하게 소개하였다. 나아가 동영상 스토리를 시각화하는 연구 사례들을 살펴보았다. 그리고 캐릭터 넷의 핵심 아이디어와 기존 연구와의 기술적 차이점에 대해 소개한 뒤, 추가적으로 이를 디자인적 솔루션을 접목하여 개선할 수 있는 방법들을 모색하였다.

Through advances driven by the Internet and the Smart Revolution, the amount and types of data generated by users have increased and diversified respectively. There is now a new concept at the center of attention, which is Big Data for assessing enormous amount of data and enjoying new values therefrom. In particular, efforts are required to analyze narratives within video clips and to study how to visualize such narratives in order to search contents stored in the Big Data. As part of the research efforts, this paper analyzes dialogues exchanged among characters and offers an interface named "Character-net" developed for modelling narratives. The interface Character-net can extract characters by analyzing narrative videos and also model the relationships between characters, both in the automatic manner. This signifies a possibility of a tool that can visualize a narrative based on an approach different from those used in existing studies. However, its drawbacks have been observed in terms of limited applications and difficulty in grasping a narrative's features at a glace. It was assumed that Character-net could be improved with the introduction of information design. Against the backdrop, the paper first provides a brief explanation of visualization design found in the data information design area and investigates research cases focused on the visualization of narratives present in videos. Next, key ideas of Character-net and its technical differences from existing studies have been introduced, followed by methods suggested for its potential improvements with the help of design-side solutions.

13

텍스트 시각화 방법을 적용한 국어교사의 학생 글 평가 특성 분석

박영민

[NRF 연계] 청람어문교육학회 청람어문교육 Vol.71 2019.09 pp.331-356

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

이 글에서는 작문 교육의 국면에서 국어교사가 수행할 수 있는 자연 언어 처리 방법을 살펴보고, 자연 언어 처리 결과를 기반으로 한 텍스트 시각화 방법의 의의를 작문 지도의 국면과 작문 평가의 국면으로 구분하여 살펴보았다. 이를 통해서 알 수 있는 학생 텍스트에 대한 국어교사의 평가 특성에 대해서 논의하였다. 이를 위해 자연 언어 처리 과정을 ⑴ 텍스트 수집 단계, ⑵ 텍스트 전처리 단계, ⑶ 텍스트 분석 단계로 구분하고, 각 단계마다 국어교사가 수행해야 할 활동을 논의하였다. 작문 교육 국면에서 이루어지는 자연 언어 처리 과정은 일반적인 텍스트 분석과 과정과 동일함 면이 있지만 작문 교육의 상황에서 오는 특별한 사항이 추가적으로 존재한다. 이를 토대로 이어지는 장에서는 텍스트 시각화 방법이 주는 의의를 작문 지도 국면과 작문 평가 국면으로 구분하여, 국어교사가 지도할 단어의 선정과 제시의 방법, 텍스트 시각화 방법이 가지고 있는 장점, 이 방법을 적용할 때의 유의점 등에 대해서 논의하였다. 이후로도 이와 관련한 논의가 지속적으로 이루어져 작문 교육이 더욱 더 체계적으로 발전할 수 있기를 기대한다.

In this article, I look at the natural language processing(NLP) methods that Korean language teachers can perform in the context of writing education, and show the significance of text visualization methods based on the results of NLP. I saw it divided into two. I discussed the evaluation characteristics of Korean language teachers of students who can know. For this reason, NLP was divided into text collection, text preprocessing, and text analysis, and the activities that Korean language teachers should perform at each stage were discussed. NLP performed in the context of composition education has the same aspects as general text analysis and processes, but there are additional special items that come from the situation of composition education. In the chapter that follows this, the significance of the text visualization method is divided into a writing instruction phase and a writing evaluation phase, and there are methods for selecting and presenting words taught by Korean language teachers, and text visualization methods. There are advantages to be discussed, such as points to note when applying this method.

14

문서 요약 및 비교분석을 위한 주제어 네트워크 가시화

김경림, 이다영, 조환규

[Kisti 연계] 한국정보과학회 정보과학회논문지 Vol.44 No.2 2017 pp.139-147

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

문자 정보는 인터넷 공간에 통용되는 정보의 대다수를 차지하고 있다. 따라서 대용량의 문서의 의미를 빠르게 특히 자동적으로 파악하는 일은 빅 데이터 시대의 중요한 연구 주제중 하나이다. 이 분야의 대표적인 연구 중 하나는 문서의 의미를 요약해주는 주요 주제어의 자동 추출 및 분석이다. 그러나 단순히 추출된 개별 주제어들의 집합만으로 문서의 의미구조를 나타내기에는 부족함이 있다. 본 논문에서는 추출된 주제어들의 연관관계를 그래프로 표현하여 대상 문서의 의미구조를 보다 다양하게 표시하고 추상화할 수 있는 주제어 가시화 방법을 개발하였다. 먼저 각 주제어들 간의 연관관계를 추출하기 위해 주제어별 지배구간 모델과 단어거리 모델을 제안하였다. 이렇게 추출한 주제어 연결성과 그를 형상화한 그래프는 문서의 의미구조를 보다 함축적으로 담고 있으므로 문서의 빠른 내용파악과 요약이 가능하며 이 가시화 그래프를 비교함으로서 문서의 의미적 유사도 비교도 가능하다. 실험을 통하여 문서의 의미파악과 비교에 본 주제어 가시화 그래프는 일반적인 요약문이나 단순 주제어 리스트보다 더 유용함을 보였다.

Most of the information prevailing in the Internet space consists of textual information. So one of the main topics regarding the huge document analyses that are required in the "big data" era is the development of an automated understanding system for textual data; accordingly, the automation of the keyword extraction for text summarization and abstraction is a typical research problem. But the simple listing of a few keywords is insufficient to reveal the complex semantic structures of the general texts. In this paper, a text-visualization method that constructs a graph by computing the related degrees from the selected keywords of the target text is developed; therefore, two construction models that provide the edge relation are proposed for the computing of the relation degree among keywords, as follows: influence-interval model and word- distance model. The finally visualized graph from the keyword-derived edge relation is more flexible and useful for the display of the meaning structure of the target text; furthermore, this abstract graph enables a fast and easy understanding of the target text. The authors' experiment showed that the proposed abstract-graph model is superior to the keyword list for the attainment of a semantic and comparitive understanding of text.

15

Topographic non-negative matrix factorization에 기반한 텍스트 문서로부터의 토픽 가시화

장정호, 엄재홍, 장병탁

[Kisti 연계] 한국정보과학회 한국정보과학회 학술대회논문집 2006 pp.324-329

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

Non-negative matrix factorization(NMF) 기법은 음이 아닌 값으로 구성된 데이터를 두 종류의 양의 행렬의 곱의 형식으로 분할하는 데이터 분석기법으로서, 텍스트마이닝, 바이오인포매틱스, 멀티미디어 데이터 분석 등에 활용되었다. 본 연구에서는 기본 NMF 기법에 기반하여 텍스트 문서로부터 토픽을 추출하고 동시에 이를 가시적으로 도시하기 위한 Topographic NMF (TNMF) 기법을 제안한다. TNMF에 의한 토픽 가시화는 데이터를 전체적인 관점에서 보다 직관적으로 파악하는데 도움이 될 수 있다. TNMF는 생성모델 관점에서 볼 때, 2개의 은닉층을 갖는 계층적 모델로 표현할 수 있으며, 상위 은닉층에서 하위 은닉층으로의 연결은 토픽공간상에서 토픽간의 전이확률 또는 이웃함수를 정의한다. TNMF에서의 학습은 전이확률값의 연속적 스케줄링 과정 속에서 반복적 파리미터 갱신 과정을 통해 학습이 이루어지는데, 파라미터 갱신은 기본 NMF 기반 학습 과정으로부터 유사한 형태로 유도될 수 있음을 보인다. 추가적으로 Probabilistic LSA에 기초한 토픽 가시화 기법 및 희소(sparse)한 해(解) 도출을 목적으로 한 non-smooth NMF 기법과의 연관성을 분석, 제시한다. NIPS 학회 논문 데이터에 대한 실험을 통해 제안된 방법론이 문서 내에 내재된 토픽들을 효과적으로 가시화 할 수 있음을 제시한다.

16

시각화를 통한 문학 텍스트 구체화 교육의 가능성 고찰

박주형, 윤여탁

[NRF 연계] 국어국문학회 국어국문학 Vol.179 2017.06 pp.35-87

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

이 연구는 문학 텍스트를 매개로 이루어진 학습자의 시각화 경험이 문학 텍스트를 통해 환기된, 언표화되기 어렵거나 복잡 모호한 반응을 구체화하는 계기가 될 수 있다고 보고, 시각화를 통한 문학 텍스트 구체화 교육의 가능성을 문학‧미술 융합 교육의 관점에서 제시하였다. 사고의 측면에서 볼 때, 문학 텍스트를 매개로 이루어지는 시각적 표현은 텍스트가 환기시키는 이미지나 사상을 구조화하는 인지 작용의 결과이며, 문학 텍스트가 환기시키는 지각 경험에 부합할 때 유의미한 경험으로 축적될 수 있다. 또한 이 과정에서 소통되는 구상적이거나 추상적인 차원의 인식들은 시각적 표현을 조정하는 경험을 통해 보다 소통 가능한 것으로 발전할 수 있다. 한편, 시각화는 그것의 기호적・수사적 특성을 기반으로 문학 텍스트와 연결될 수 있는데, 다양한 시각화 양상 중 텍스트의 내용과 밀접한 관련성이 있으면서, 텍스트의 의미를 넘어서는 유의미한 정보를 포함하고 있는 것이 문학 텍스트의 구체화를 위한 계기를 마련해 줄 수 있다. 따라서 시각화를 통한 문학 텍스트의 구체화 가능성은 문학 텍스트의 이해 과정에서 학습자가 봉착할 수 있는 문제 상황과 조우할 수 있는 영역에서 보다 높아질 것으로 기대할 수 있다. 이러한 관점을 바탕으로 초등학생 학습자들의 시각화 양상을 살펴본 결과, 정서 체험과 연관된 상황을 시각적으로 표현하거나, 대상들 간의 비유적 관계를 형상으로 표현하는 활동, 언어적 반응을 시각적 표현과 결합시키는 활동에서 그러한 구체화의 가능성을 확인할 수 있었다. 그러나 이러한 가능성을 높이기 위해서는 시각적 표현의 의도를 상세화하거나, 예시적인 모델을 비계로 제공하거나, 조형 원리에 대한 이해를 제고하는 등의 교육적 처치가 필요하다는 점 역시 확인된다.

The learner's visualization experience through literary text can be the medium for concreting the ambiguous or complex response which is evoked through the text. From this perspective, this study presented the possibility of concretization education of literary text through visualization from the point of view of literature-art convergence education. In terms of thinking, visual expression through literary text is the result of the cognitive process of structuralizing images or ideas that the text evokes. And it can be accumulated as a meaningful experience for reader when it is suitable for the sensory experience evoked by the literary text. In addition, the concrete or abstract level of awareness communicated in this process can be developed to be more communicative through experience of coordinating visual expression. Visualization can be linked to the literary text based on its symbolic and rhetorical characteristics. Among various visualization aspects, the aspect that contains meaningful information transcending the meaning of the text, as well as closely related to the content of the text, becomes a medium for the concretization of the literary text. Therefore, it is expected that the possibility of the concretization of the literary text through visualization will be higher in the area where the learner can encounter the problem situation in the understanding process of the text. Based on these perspectives, this study analyzed the visualization aspects of elementary school students and found out the possibility of concretization of literary text in activities of visually expressing contexts of emotional experiences, expressing the metaphorical relationships in a plausible form, and combining visual form with linguistic expression. However, in order to increase this possibility, it is also confirmed that educational instruction such as refining the intention of visual expression, providing an illustrative model as a scaffold, or understanding the formative principle is needed.

17

빅데이터 분석 도구 R을 이용한 비정형 데이터 텍스트 마이닝과 시각화

남수태, 신성윤, 진찬용

[Kisti 연계] 한국정보통신학회 한국정보통신학회논문지 Vol.25 No.9 2021 pp.1199-1205

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

빅데이터 시대에는 단순히 데이터베이스에 잘 정리된 정형 데이터뿐만 아니라 인터넷, 소셜 네트워크 서비스, 모바일 환경에서 실시간 생성되는 웹 문서, 이메일, 소셜 데이터 등 비정형 빅데이터를 효과적으로 분석하는 것이 매우 중요하다. 빅데이터 분석은 데이터 저장소에 저장된 빅데이터 속에서 의미 있는 새로운 상관관계, 패턴, 추세를 발견하여 새로운 가치를 창출하는 과정이다. 빅데이터 분석 도구인 R 언어를 이용하여 비정형 논문 데이터를 빈도분석을 통해 분석결과를 요약과 시각화하고자 한다. 본 연구에서 사용된 데이터는 한국정보통신학회 학회지 논문 중에서 2021년 1월호-5월호 총 논문 104편을 대상으로 분석하였다. 최종 분석결과 가장 많이 언급된 키워드는 "데이터"가 1,538회로 1위를 차지하였다. 따라서 분석결과를 바탕으로 연구의 한계와 이론적 실무적 시사점을 제시하고자 한다.

In the era of big data, not only structured data well organized in databases, but also the Internet, social network services, it is very important to effectively analyze unstructured big data such as web documents, e-mails, and social data generated in real time in mobile environment. Big data analysis is the process of creating new value by discovering meaningful new correlations, patterns, and trends in big data stored in data storage. We intend to summarize and visualize the analysis results through frequency analysis of unstructured article data using R language, a big data analysis tool. The data used in this study was analyzed for total 104 papers in the Mon-May 2021 among the journals of the Korea Institute of Information and Communication Engineering. In the final analysis results, the most frequently mentioned keyword was "Data", which ranked first 1,538 times. Therefore, based on the results of the analysis, the limitations of the study and theoretical implications are suggested.

18

세계 영어의 지역적 차이에 대한 말뭉치 분석연구

김준기

[NRF 연계] 새한영어영문학회 새한영어영문학 Vol.49 No.2 2007.05 pp.143-155

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

A Corpus-based Study of the Dialectal Difference of the World EnglishKim, Joon-KiModality is used to represent a speaker’s subjective opinion or attitude such as volition, possibility, necessity, permission, inference, obligation, ability, etc. Modal auxiliaries (henceforth, modals) play the most important part in representing modality. They are used in epistemic and deontic usages. When modals are used in epistemic usage, they represent the meaning of possibility, necessity, probability, etc. When modals are used in deontic usage, they represent the meaning of ability, permission, volition, order, and obligation. The main purpose of this study is to examine the appropriateness of choosing English newspapers as authentic materials for teaching the concept of modality in English to students and to assess educational implications. The importance of this study is that teachers should make learners understand the basic meanings of modal auxiliaries regardless of a variety of Englishes just rather than provide various authentic materials which does not reflect on the appropriate uses of English modals.

19

내러티브 텍스트의 시적 변용과 영상화

최영승

[NRF 연계] 새한영어영문학회 새한영어영문학 Vol.49 No.1 2007.02 pp.165-184

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

The Poetic Transformation and Visualization of Narrative TextChueh, Young-SungCoppola reproduces Apocalypse Now visually drawing on Conrad’s Heart of Darkness and Eliot’s The Hollow Men. This study aims to investigate the process how the cinema is recreated from its source novel and the reason why it alludes to Eliot’s poetic text. The cinema is primarily visual while literature is generally verbal. The cinema produces meaning through other means than poetry and the novel. If the popular cultural text has the positive and specific features, another kind of reading is called for. The attention to the visual nature of the discourse should be paid. Coppola’s cinema represents Conrad’s narrative discourse in terms of Eliot’s rhythm. Therefore, the poetic transformation and visualization of narrative text form a good spectrum that shows the process of reproduction from high culture to popular culture. A close observation of the three texts clearly shows the process in which the conventional binary opposition between literary canons and a popular cultural text disappears. Eliot’s poem shows the textual values of interventions that accept and promote literary evolution in the process of transformation from Conrad’s verbal text to Coppola’s visual text with images and sounds.

20

텍스트 데이터를 활용한 체육교육과 교과과정 시각화 분석

최지상, 이병구

[NRF 연계] 한국교육과정평가원 교육과정평가연구 Vol.26 No.2 2023.05 pp.125-144

※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.

원문보기

본 연구의 목적은 워드 클라우드 기법을 활용하여 우리나라 체육교육학과 간의 교과과정을 시각적으로 비교ㆍ분석하는데 있다. 대입정보포털 ‘어디가(www.academyinfo.go.kr)’에 제공·소개하고 있는 학과정보를 이용하여 자료수집을 실시했다. 최종 연구대상으로 선별된 체육교육학과는 20개교 학과들로 총 828개의 개설 학과목을 워드 클라우드 기법을 기반으로 자료분석을 실시했다. 대체로 교원양성평가 기준에 따라 학과목을 운영하고 있었다. 그러나 실기보다는 학점 단위가 높은 체육이론 교과목의 단어 빈도가 높게 나타났다. 이어 동일 학과명에서도 ‘무용’, ‘선수’ 등 학과의 관심사와 연구분야에 따라 교과과정이 상이함을 워드 클라우드 결과로 확인할 수 있었다. 다만 계열에 따른 개설된 학과목의 차이는 클라우드 결과만으로 판명하기 어렵다. 따라서 체육교육이 다가올 시대에도 변함없이 교과로서 인정받기 위해서는 자율적인 변혁 의지와 함께 교과과정에 대해 자각하고 자강하여야 한다.

The purpose of this study is to visually compare and examine the curriculum between the departments of physical education in Korea using word cloud techniques. The department information provided and introduced on the university admission information portal ADIGA(www.academyinfo.go.kr) was used. There were 20 departments in the Department of Physical Education selected as the final research subject, and a total of 828 subjects were analyzed based on word cloud techniques. In general, subjects were operated according to the teacher training evaluation criteria. However, the frequency of words in physical education theory subjects, which have higher credit units than practical skills, was higher. And even if the name of the same department was the same, it could be confirmed through cloud results that the curriculum was different depending on the interests and research fields of departments such as "dance" and "athlete." But it is difficult to determine the difference in subjects according to the category through research results. Therefore, this study is meaningful in that it visually compared and analyzed the subjects opened between the departments of physical education by applying the word cloud technique.

 
1 2
페이지 저장