년 - 년
AI 기반 영상 편집 도구의 사용자 경험(UX) 비교 분석과 영상디자인 활용 방안 연구 KCI 등재
대한산업경영학회 산업융합연구(구 대한산업경영학회지) 제24권 제4호 2026.04 pp.23-31
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
본 연구는 2024~2025년 사이 실제 사용자 활용 빈도와 시장 확산성이 높은 AI 기반 영상 편집・생성 도구를 대상으 로, 사용자 경험(UX) 관점에서의 특성을 비교・분석하고 영상디자인 분야에서의 활용 방안을 도출하는 것을 목적으로 한다. 생성형 인공지능 기술의 발전은 영상 콘텐츠 제작의 자동화와 접근성을 향상시키며, 영상디자인 교육과 실무 환경에 새로운 제작 패러다임을 형성하고 있다. 이에 본 연구는 2025년 기준 글로벌 사용자 활용도, 기능 범위, 접근성, 연구 재현 가능성을 고려하여 Runway, CapCut, Synthesia, VEED.IO, Sora2의 5개 도구를 분석 대상으로 선정하였다. 연구 참여자들은 동일 한 영상 제작 과제를 수행한 후, 사용 편의성, 직관성, 학습 용이성, 창의성 지원, 작업 효율성 등 UX 요소에 대해 정량적 설문 과 정성적 인터뷰에 응답하였다. 분석 결과, 도구 간 UX 특성에는 유의미한 차이가 나타났으며, 각 도구는 사용자 숙련도와 영상디자인 목적에 따라 상이한 활용 적합성을 보였다. 본 연구는 AI 영상 편집 도구를 UX 중심으로 비교・평가함으로써, 영 상디자인 교육과 실무 현장에서의 합리적인 도구 선택 기준을 제시한다.
This study aims to compare and analyze the user experience (UX) characteristics of AI-based video editing and generation tools that showed high levels of user adoption and market diffusion between 2024 and 2025, and to derive implications for their application in the field of video design. Recent advances in generative artificial intelligence have enhanced the automation and accessibility of video content production, establishing a new production paradigm in both video design education and professional practice. Based on global usage prevalence, functional scope, accessibility, and research reproducibility as of 2025, five representative tools—Runway, CapCut, Synthesia, VEED.IO, and Sora2—were selected for analysis. Participants completed the same video production task using each tool and responded to both quantitative surveys and qualitative interviews focusing on key UX factors, including usability, intuitiveness, learnability, creative support, and task efficiency. The results revealed significant differences in UX characteristics among the tools, with variations depending on user proficiency and video design objectives. This study proposes UX-centered criteria for selecting AI video editing tools, contributing to more effective tool adoption in video design education and practice.
5,200원
본 연구는 영상편집의 패턴을 설정하기 위한 기초연구이다. 영상편집은 다양한 쇼트를 중심 으로 쇼트의 구분을 정확히 보고자 함이다. 리니어(Linear) 편집에서 넌리니어(Non-Linear) 편집으로의 전환은 편집에 있어서 새로운 편집보다는 재편집에 대한 기술적 발전을 이루어지게 하였다. 이는 편집기술의 발전은 초창기 생방송을 통해 전파를 보내는 것에서 녹화방송의 개념이 도입이 되었으며, 장비의 발전과 소형 화를 통해 개인 프러덕션이 탄생할 정도로 발전을 하고 있다. 방송제작 환경에서 디지털 환경의 변화는 영상/음향 신호 시스템의 리니어(Linear)시스템에 서 넌리니어(Non-Linear)시스템으로의 변화를 통해 제작인력의 창의성과 예술적 감성으로 인하 여 보다 다양한 형태의 제작양식을 창출해 내고 있다. 기술의 발전은 장르의 다양성도 나타나게 되었다. 드라마, 뉴스, 버라이어티 등 다양한 포맷 을 통해 수용자들의 선택의 폭이 넓어지기 시작 했다. 특히 버라이어티 장르는 나날이 발전을 하기 시작했으며, 다양한 콘텐츠를 통해 수용자들에게 전달되기 시작했다. 이에 본 연구에서는 다양한 쇼트의 구분을 하였다. 그 구분을 통해 색상의 정리를 통해 다양 한 구분을 하였다. 또한 이론적 배경을 통해 편집의 쇼트를 22가지로 구분하였으며, 22가지의 영상쇼트에 색상을 정리함으로서 편집의 패턴을 보고자 하였다. 하지만 정리의 한계와 영상편집의 쇼트의 구분을 통해 다음 연구에서는 패턴을 적용하여 어떤 분류와 형식이 존재하는지 연구할 것이다.
This study is the pattern for setting editing images based study. Video editing is the distinction between a variety of short shots around the box is exactly want. Linear (Linear) non-linear editing (Non-Linear) Edit to edit the transition to the new technical developments for editing, rather than re-editing was done become. Edit the early days of the development of this technology to send radio waves through a live recording from what has been the introduction of the concept of broadcasting equipment through the development and miniaturization of personal development and production is enough to be born. Broadcast production environment, changes in the digital environment, the video / audio signal of the linear system (Linear) systems, non-linear (Non- Linear) systems through changes in production due to labor than the creativity and artistic sensibility to create various types of production form off and be. The development of technology is the variety of genre appeared. Drama, news, variety of audiences through a variety of formats, including a wider range of choice began. In particular genre of variety from day to day and began to develop a variety of content to be delivered to the prisoners began. Chesterfield this study was the classification of the various short. Through its divisions through a variety of color breaks were clean. In addition, the theoretical background of the shot compilation was divided into 22 kinds, 22 kinds of colors to organize a short video compilation of patterns by patients. Limit theorem, but the distinction between short and video editing over the next study, which is applied to the pattern classification will study the existence and type.
원격 강의용 콘텐츠 제작 도구를 위한 동영상 생성 알고리즘 KCI 등재
한국정보교육학회 정보교육학회논문지 제22권 제5호 2018.10 pp.605-611
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
무크(Mook) 서비스 정착과 교육부의 원격 수업 확대 정책에 힘입어 대학을 중심으로 온라인 수업이 일반화하 고 있으며, 출석 정규 수업의 형태도 일부 원격 방식으로 전환되는 추세이다. 본 논문은 이러한 교육 환경에서 필요한 교육용 콘텐츠를 제작하는 도구에 관한 것으로 이에 필요한 오디오 동기화 방법과 편집 알고리즘을 제안 하였다. 제안한 알고리즘을 노트북 등 일반적인 PC 환경에서 그 성능을 측정하였다. 평가 결과, 안정적인 CPU 와 Ram 사용량을 보였다. CPU 사용은 평균 9.3% 점유율을 나타냈고 Ram은 평균 87Mega바이트의 안정적인 사용량을 기록하였다(CPU 2.60GHz, 820*600 영역 기준). 본 제안하는 방식을 사용하면 손쉽게 자기 PC로 원격 교육을 진행할 수 있을 것으로 판단된다.
On-Line Lectures are becoming more common due to the MOOK service and the expansion of national policy in Korea. Especially, It is being changed to new remote mixed style from traditional lecture in universities. We propose and implement a remote contents making tool with audio synchronization function based on more with less resources. To implement our proposed algorithm, we design an interactive interface to assign multiple cutting intervals and convert an input video to print a new result. In experimental, we can confirm our algorithm works properly with average performance value 9.3% cpu share ratio and 87mega byte ram usage(CPU 2.60GHz, 820*600 Area).
감성 전달을 위한 UCC 동영상 편집 방안에 관한 연구
[Kisti 연계] 한국디지털콘텐츠학회 디지털콘텐츠학회 논문지 Vol.12 No.4 2011 pp.449-456
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
UCC(User Created Contents) 동영상은 온라인상에서 사용자가 상업적인 의도없이 제작한 '감성 전달' 매체로서 최근 인터넷 대중문화의 핵심 콘텐츠로 급부상하였다. '감성 전달'은 UCC 동영상의 궁극적인 목적으로 '기 승 전 결'의 스토리텔링(Storytelling)이 뒷받침되어야 가능하다. 그러나 비상업적인 UCC 동영상은 불안정한 제작 환경, 다양한 변수들로 인하여 완성도 높은 작품을 기대하기 어렵다. 따라서 본 논문에서는 불안정한 UCC 동영상의 문제점을 최소화하기 위해 편집 단계에서 '세부적인 스토리텔링 편집' 과정을 제안하였다. 연구자의 실무 결과물과 UCC 동영상 공모전 수상작품을 토대로 연구한 이 제안은 공감대를 키우고, 창의적인 변화를 유도하며, 통일성있는 마무리 편집으로 작품의 완성도를 높임으로서 '감성 전달'의 기능을 확대할 수 있다.
UCC (User Created Contents) video rapidly rose as key contents of internet pop culture as a medium of 'sensibility delivery' that is created by a user without a commercial purpose. 'Sensibility delivery' is the ultimate purpose of UCC video and it can only be completed with 'well composed' storytelling. However, it is difficult to expect high degree of completion from noncommercial UCC videos due to an instable production environment and many variables. Thus, this study has been suggested 'detailed storytelling editing' process from editing phase in order to minimize problems of instable UCC videos. This suggestion based on actual results and award winning works of UCC videos of a team of research will raise a bond of sympathy, lead creative change, and expand the function of 'sensibility delivery' by increasing the degree of completion of work with unified final cut.
영상편집 소프트웨어의 시각 정체성과 서사 정체성 : 파이널컷프로X의 기호학적 분석 KCI 등재
한국영상문화학회 영상문화 제32호 2018.06 pp.137-159
※ 기관로그인 시 무료 이용이 가능합니다.
6,000원
본 논문은 대표적인 영상편집 소프트웨어인 애플의 파이널컷프로X를 기호로 분석하여 시각 정체성을 도출한다. 이를 위해 경쟁사 어도비의 제품인 프리미어프로CC와 이항대립하여 계열체의 체계와 통합체적 과정을 분석한다. 프리미어프로CC는 첫 출시 이래로 분리, 선형, 복잡, 수동, 전문의 특성을 지닌 기존 테이프기반 선형편집시스템을 재현한 인터페이스를 유지한다면 파이널컷프로X는 그러한 기존 인터페이스와 단절하고 통합, 비선형, 단순, 자동, 대중의 디지털 특성을 선택한다. 이로써 파이널컷프로X의 시각 정체성은 표현 면에서 하나의 스토리라인으로 통합된 편집과정과 단순하고 직관적인 작업공간이고 내용 면에서 자동화 기능으로 쉽게 스토리 창작을 즐길 수 있는 시민의 영상편집기라 할 수 있다. 파이널컷프로X는 단순히 마케팅 전략으로서 프리미어프로CC와 차별화한 것이 아니라 편집시스템 역사에서 불연속적 혁신을 확립하는 서사 정체성을 드러낸다. 이 서사 정체성은 기존에 불변하는 성격과 단절하고 변하는 자기성으로 스스로를 보존한 것으로 주체가 일생에 걸쳐 일관되게 추구한 스토리 안에서 이해될 수 있다. 1984년 최초의 개인용 컴퓨터 매킨토시 출시에 포함된 가상세계 속 개인의 자기실현이라는 스토리가 1999년 파이널컷프로로 이어져 누구나 전문가적 영상편집을 할 수 있는 가능성을 열었다. 2011년 파이널컷프로X는 기존의 테이프기반 선형편집시스템의 인터페이스와 단절하고 개인이 전문가의 도움 없이 스스로 스토리를 창작할 수 있는 환경을 제공함으로써 그 스토리를 이어간다. 이는 연극을 재현하였던 초창기 영화가 고전적 편집을 완성하면서 자기 정체성을 확립한 것처럼 테이프기반 영상편집시스템을 재현한 초창기 영상편집 소프트웨어가 디지털 인터페이스를 완성하면서 자기 정체성을 보존하는 것이다.
This paper is designed to produce visual identity by analyzing the Final Cut Pro X, a representative digital video editing software. For this purpose, it sets Final Cut Pro X and the rival product Premier Pro CC as a binary opposition and analyzes the paradigmatic system and the syntagmatic process of the Final Cut Pro X. If Premiere Pro CC maintains an interface that represents the existing tape-based linear editing system with separation, linearity, complexity, manual and professional characteristics since its first release, Final Cut Pro X disconnects that existing interface and selects the digital properties such as integration, non-linearity, simplicity, automatic and public characteristics. The visual identity of Final Cut Pro X is a simple and intuitive workspace and editing process with an integrated storyline in the plane of expression and a citizen's tool for story creation in the plane of content. Final Cut Pro X is not just differentiated from Premiere Pro CC by its marketing strategy, it reveals the narrative identity that establishes discontinuous innovation in the digital editing system history. This narrative identity can be understood in the story which the subject has consistently pursued throughout his / her life since it has disconnected the unchanging character, idem and preserved oneself with the changing character, ipseity. The story of individuals’ self-realization in the virtual world included in the 1984 release of the first personal computer Macintosh linked to Final Cut Pro in 1999, opening the possibility for anyone to do professional video editing. In 2011, Final Cut Pro X continues its story by providing an environment where individuals can create their own stories without the help of experts, disconnected from the difficult and complicated existing professional interfaces. Just as the early film that had represented the theater play established its own identity when completing classical editing, the earliest video editing software that had represented the tape-based video editing system preserved its identity when completing the digital interface.
영상편집 프로젝트 파일의 메타데이터 이상 탐지를 통한 구조 기반 변조 분석
[Kisti 연계] 한국정보보호학회 정보보호학회논문지 Vol.36 No.1 2026 pp.239-255
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
본 연구는 영상편집 프로젝트 파일의 메타데이터와 저장 구조를 정량적으로 분석하여 변조 여부를 탐지하는 실증적 연구이다. Adobe Premiere Pro와 CapCut을 대상으로 동일한 편집 시나리오를 적용해 정상·변조 파일 각 20개(총 40개)를 생성하고, 생성·수정일 불일치, 타임라인 구성 불일치, 클립 수 변동, 파일 구조 무결성 이상 등 네 가지 탐지 기준을 설정하였다. 분석 결과, 탐지 성공률은 '생성·수정일 불일치' 90%, '타임라인 구성 불일치' 85%, '클립 수 변동' 80%, '파일 구조 무결성 이상' 75%로 나타났다. 이를 통해 조작이 프로젝트 파일의 논리·물리적 구조에 미치는 영향을 규명하였으며, 단일 지표가 아닌 복수 기준을 병행하는 다층적 탐지 방식이 전체 효율을 높임을 확인하였다. 본 연구는 기존 영상 중심 포렌식 분석을 편집 단계의 메타데이터 기반 분석으로 확장하고, 프로젝트 파일의 증거 활용 가능성을 제시하였다.
This study conducts an empirical analysis to detect tampering in video editing project files by quantitatively examiningtheir metadata and storage structures. Using Adobe Premiere Pro and CapCut, 40 project files (20 original and 20 tampered) were created based on an identical editing scenario. Four detection criteria were established: mismatch between creation andmodification dates, timeline structure inconsistencies, changes in the number of clips, and file structure integrity issues. The detection success rates were 90% for date mismatch, 85% for timeline inconsistency, 80% for clip count variation, and 75%for structural anomalies. These results demonstrate that intentional manipulation affects both the logical and physical structures of project files. Moreover, a multi-layered detection approach using multiple indicators proved more effective thanrelying on a single metric. This study expands digital forensic analysis beyond finalized video outputs to the editing stage, providing empirical evidence for the use of project files as standalone digital evidence.
SNS 소통과 공유를 위한 영상 편집과 뷰티 크리에이터로서의 역량 강화 할동
국제보건미용학회 국제보건미용학회 학술컨퍼런스 메타버스타고 뷰티산업의 확장 2021.11 pp.95-97
※ 기관로그인 시 무료 이용이 가능합니다.
3,000원
360° VR 실사 영상과 3D Computer Graphic 영상 합성 편집에 관한 연구 KCI 등재
한국디지털정책학회 디지털융복합연구 제17권 제4호 2019.04 pp.255-260
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
본 연구는 360˚ 동영상과 3D 그래픽의 효율적인 합성 방법에 관한 연구이다. 먼저, 이안식 일체형 360˚ 카메라로 촬영한 영상을 스티칭하고, 영상에서 카메라 및 사물의 위치값을 추출하였다. 그리고 추출한 위치값의 데이터를 3D 프로그 램으로 불러와 3D 오브젝트를 생성하고, 자연스러운 합성을 위한 방법에 관하여 연구하였다. 그 결과 360˚ 동영상과 3D 그래 픽의 자연스러운 합성을 위한 방법으로 렌더링 요소와 렌더링 기법을 도출할 수 있었다. 첫째, 렌더링 요소로는 3D 오브젝트 의 위치와 재질, 조명과 그림자가 있었고, 둘째, 렌더링 기법으로는 실사 기반 렌더링 기법의 필요성을 찾을 수 있었다. 본 연구 과정 및 결과를 통해 360˚ 동영상과 3D 그래픽의 자연스러운 합성에 관한 방법을 제시함으로써, 360˚ 동영상 및 VR 영상 콘텐츠의 연구와 제작 분야에 도움이 될 것으로 기대한다.
This study is about an efficient synthesis of 360˚ video and 3D graphics. First, the video image filmed by a binocular integral type 360˚ camera was stitched, and location values of the camera and objects were extracted. And the data of extracted location values were moved to the 3D program to create 3D objects, and the methods for natural compositing was researched. As a result, as the method for natural compositing of 360˚ video image and 3D graphics, rendering factors and rendering method were derived. First, as for rendering factors, there were 3D objects’ location and quality of material, lighting and shadow. Second, as for rendering method, actual video based rendering method’s necessity was found. Providing the method for natural compositing of 360˚ video image and 3D graphics through this study process and results is expected to be helpful for research and production of 360˚ video image and VR video contents.
동영상 편집 기능이 영상디자인의 대중화에 미치는 현상에 관한 분석 연구 KCI 등재
한국디자인트렌드학회 한국디자인포럼 Vol. 16 2007.08 pp.271-280
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
유튜브 동영상 편집기법의 변화가 동영상 조회 수 및 수익에 미치는 영향 연구 - 방송사 뉴스채널 동영상 분석을 중심으로 - KCI 등재
한국브랜드디자인학회 브랜드디자인학연구 Vol.20 No.2 통권 제62호 2022.06 pp.117-130
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
본 연구는 국내 방송사들의 유튜브 채널 개설에 따른 전략 적 대응을 위하여, 유튜브 동영상 편집 기법 변화에 따른 조 회 수 및 수익에 미치는 인과성을 실증적으로 조사·분석하는 연구이다. 이를 위하여 4가지 연구가설에 따른 연구모형을 설 정하고, 뉴스 채널의 실험영상과 대조영상을 만들어 그 변화 의 요인들을 양적으로 측정하고자 A/B Testing을 실시하였 다. 연구결과는 첫째, 유튜브 동영상 섬네일의 시인성은 유튜 브 동영상 조회 수와 정(+)의 인과관계를 가지는 것으로 나타 동영상 섬네일의 시인성이 높은 영상이 그렇지 않는 영상보 다 조회 수가 2.8배 높은 결과를 나타냈다. 둘째, 중심인물의 이야기 기반 동영상 편집은 유튜브 동영상 조회 수에 정(+)의 인과관계를 가지는 것으로 나타났다. 이로 인해 중심인물이 이야기의 흐름을 잡고 까닭이나 맥락을 들어 자세히 설명하 는 동영상 편집기법이 그렇지 않은 동영상 편집기법 보다 조 회 수가 40배 이상 높다는 결과를 확인할 수 있었다. 셋째, 감 성적 동영상 편집은 유튜브 조회 수에 정(+)의 인과관계를 가 지는 것으로 나타났다. 이에 따라 동영상 분위기를 감정적으 로 선도하는 동영상 편집기법이 그렇지 않은 동영상 편집기 법 보다 조회 수가 30배 이상 높다는 결과를 확인할 수 있었 다. 넷째, 동영상 편집기법의 변화에 따른 조회 수 증가는 수 익에 정(+)의 인과관계를 가지는 것으로 나타났다. 이로 인해 동영상 편집기법의 변화를 적용한 콘텐츠가 그렇지 않은 콘 텐츠에 비해 수익이 12.9배 높은 것으로 나타나 그 인과성을 확인할 수 있었다. 이번 연구로 인해 그동안 암묵지 영역에 머물던 언론사들의 유튜브 동영상 편집기법의 제작방법 및 운영이 보다 객관적이고 실증적인 명시지 영역으로 옮겨갈 수 있을 것으로 판단된다.
This study empirically analyzes the causality of change s in YouTube video editing techniques on the number of views and profits in order to respond strategically to the opening of YouTube by domestic media companies. The results of the study were first, that the visibility of YouT ube video thumbnails had a positive causal relationship w ith the number of video views. Second, the story-based video editing of the central character was found to have a positive causal relationship on the number of YouTube video views. As a result, it was confirmed that the video editor method in which the central character captures the flow of the story and explains in detail the reason or cont ext was more than 40 times higher in number of views than the video editor method that does not. Third, emotio nal video editing was found to have a positive causal rela tionship on the number of views. Fourth, it was found th at the increase in the number of views according to the change in the video editing techniques had a positive cau sal relationship to the revenue. As a result, it was found that the content that applied the change in the video edito r method was 12.9 times higher in revenue than the conte nt that did not. With this study, Due to this study, it is believed that the methodology of the video editing techni ques of media companies, which have been in the tacit do main, will be able to move to a more objective and empiri cal domain.
비선형 편집 입문자를 위한 RPT 학습모형 절차 설계 및 평가 KCI 등재
국제인공지능학회(구 한국인터넷방송통신학회) 한국인터넷방송통신학회 논문지 제17권 제4호 2017.08 pp.163-172
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
최근 방송영상분야에서 영상편집 방법으로 비선형편집(Non-Linear Editing; NLE)을 주로 사용하고 있다. 기존의 선형편집에 비해 비선형편집은 컷(Cut)의 삽입과 삭제가 용이하고, 영상편집 시 원하는 위치의 영상에 바로 접근할 수 있다. 또한, 타이틀과 효과, 장면전환 효과를 적용할 수 있고 출력 전에 미리보기를 통해 적용한 타이틀과 효과를 확인하고 수정하는 것이 용이한 장점이 있다. 그러나, NLE편집을 처음 접하는 학생들이 그것을 배우는 것은 쉽지 않다. 본 논문에서는 비선형편집을 처음 접하는 학생들이 쉽게 배울 수 있는 기존의 상호동료교수법(Reciprocal Peer Tutoring; RPT)를 보완한 새로운 RPT 학습모형을 제시한다. 제안하는 교수학습모형을 적용한 실험집단과 적용하지 않는 비교집단으로 나누어 실험을 실시한다. 두 집단의 전체 평균, 성적 하위 집단의 학업 성취도, 표준편차, T검정과 함께 설문조사를 통한 만족도를 실시한다. 제안하는 학습모형을 적용한 실험집단이 통제집단에 비해 지표와 만족도에서 우월함을 보인다
In recent days, the Non-Linear Editing is mainly used in the field of broadcasting. In comparison to conventional editing, Non-Linear Editing can immediately access the image of the desired position and facilitate the insertion and deletion of video frame. Furthermore, it directly apply a title and transition effect to video frame. Moreover, it has an advantage of preview and easy modification in title effect, transition and editing prior to export. However, students who learn Non-Linear Editing first time are not easy to learn it. In this paper, we propose a new learning model based on Reciprocal Peer Teaching (RPT), which helps NLE beginners to understand Non-Linear editing more clearly. We divide the students into two groups i.e. control group and experimental group. The control group students do not apply proposed method while experimental group performs evaluation over our model. Furthermore, we carry out the experiments, which include the overall average of the two groups, academic achievement of students with low grades, standard deviation, T-test and satisfaction surveys. The experimental group shows the superiority in performed experiments and higher satisfaction ratings than the control group.
Audio-Based Video Editing with Two-Channel Microphone
보안공학연구지원센터(IJHIT) International Journal of Hybrid Information Technology Vol.1 No.3 2008.07 pp.71-80
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
Audio has a key index in digital videos that can provide useful information for video editing, such as capturing conversations only, clipping only talking people, and so on. In this paper, we are studying about video editing based on audio with a two-channel (stereo) microphone that is standard equipment on video cameras, where the video content is automatically recorded without a cameraman. In order to capture only a talking person on video, a novel voice/non-voice detection algorithm using AdaBoost, which can achieve extremely high detection rates in noisy environments, is used. In addition, the sound source direction is estimated by the CSP (Crosspower-Spectrum Phase) method in order to zoom in on the talking person by clipping frames from videos, where a two-channel (stereo) microphone is used to obtain information about time differences between the microphones.
A Study on the Application of AI-Based Video Editing in Broadcasting : Focusing on TV Programs
국제인공지능학회(구 한국인터넷방송통신학회) The International Journal of Advanced Smart Convergence Volume 14 Number 3 2025.09 pp.285-297
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
This study aims to examine how Artificial Intelligence (AI) can be practically applied to post-production in broadcasting and to evaluate its implications for future content creation. To this end, it focuses on the case of Earth Sweepers, a 2024 MBC reality variety program recognized as the first terrestrial TV show in Korea to adopt AI-based editing and de-identification under outdoor filming conditions. The research method involved analyzing the integration of two key technologies: automatic editing developed by the Electronics and Telecommunications Research Institute (ETRI) and de-identification technology developed by the Korea Electronics Technology Institute (KETI). Both tools were implemented as plug-ins for Adobe Premiere Pro to ensure seamless use within established editing workflows. The results revealed that AI significantly reduced production time—overall editing was shortened by approximately 35–40%, and de-identification achieved over 90% accuracy while cutting costs by more than 20 million KRW per episode. AI-assisted editing also improved multicamera alignment, scene segmentation, and highlight extraction, while de-identification enhanced both visual quality and compliance with regulatory requirements. The findings highlight the potential of human-AI collaboration in broadcasting and underscore the importance of establishing standardization, professional training, and institutional support to ensure sustainable adoption. This research provides valuable insights for guiding innovation and shaping the future ecosystem of AI broadcasting.
Promotional Video of Editing Techniques Utilizing Color and Brand Balance SCOPUS
보안공학연구지원센터(IJSEIA) International Journal of Software Engineering and Its Applications Vol.8 No.7 2014.07 pp.149-158
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
When we make public relation videos in this paper, we study this research to apply color that not only stimulating digital technology but also analogue sensibility for marketing. If there is a core color, we could establish brand identity and remain in our memory and send core messages. Color has different meaning with culture and country but color also has universal characteristic. Because color effects human's emotions and sensibility, we apply color to public relation video. Using it, we studied the methods of extract just one color. In other words, the purposes of this study are definite experiment and materialization about the method of expressing analogue sensibility. So, this study expression effects to use color. In future wealso use this to consider development possibility and expectation effectiveness.
틱톡(TikTok) 동영상 플랫폼 영상편집디자인 기능이 사용자 태도에 미치는 영향 KCI 등재
한국브랜드디자인학회 브랜드디자인학연구 Vol.19 No.4 통권 제60호 2021.12 pp.48-60
※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.
콘텐츠의 홍수 속에서 짧은 영상 콘텐츠를 선호하는 경향 이 점점 더 강해지고 있는 가운데 2016년 틱톡(TikTok) 동영 상 플랫폼은 이런 사용자들의 시청 욕구를 충족 시켜 주었다. 나아가 틱톡 동영상 플랫폼에서 제공하는 다양한 영상편집 디자인 기능은 틱톡 동영상 플랫폼 이용자에게 동영상 제작 의 즐거움 또한 제공해주었다. 본 연구는 틱톡 동영상 플랫폼 에서 제공하는 다양한 영상편집디자인 기능이 사용자의 만족 도, 몰입도 그리고 지속사용의도에 미치는 영향을 분석하였 다. 분석을 위해 틱톡 동영상 플랫폼 영상편집디자인 기능 사 용자를 총 8집단으로 나누고 1집단에 150명-153명씩 피실험 자를 배분하였고, 오류가 없는 총 1209명의 설문을 SPSS 21. 0 통계분석 프로그램을 사용하여 분석하였다. 분석 결과 틱톡 동영상 플랫폼 영상편집기능 중에 자동영상편집, 영상필터, 영상효과, 뷰티기능은 사용자 만족도, 몰입도 그리고 지속사 용의도에 높은 선호도를 보였으며, 반대로 자막, 사운드, 스티 커 기능은 가장 낮은 사용자 만족도, 몰입도 그리고 지속사용 의도를 보였다. 특히 틱톡 동영상 플랫폼에서 제공하는 영상 편집기능 중에서 사용자의 얼굴 윤곽 조정, 눈 크기 조정, 피 부 보정 등 인물 형태 변형이 가능한 뷰티 기능이 가장 높은 만족도, 몰입도, 지속사용의도 선호도를 보였다. 이러한 결과 는 향후 틱톡 동영상 플랫폼의 영상편집디자인 기능 추가에 있어 소비자의 특성을 이해하는데 기초 자료가 될 것이다.
While short image contents are more preferred in the fl ood of tight Life Style, TikTok, the 2016 satisfied such users’ desire of watching. Further, diverse image edition design functions provided by TikTok also provides TikT ok users with enjoyment of video production. This study analyzed the influence that diverse image edition design function provided by TikTok has on users’ satisfaction, immersion & intent of constant use. For analysis, users of TikTok image edition design function were divided in to 8 groups in all and 150-153 test subjects were distribu ted to 1 group. Error-free question papers of 1209 perso ns were analyzed by using SPSS 21.0 statistics analyzin g program to find the difference between groups. Accordi ng to analysis, out of TikTok automatic edition functions, automatic edition, image filter, image effect, beauty functi on were much preferred in users’ satisfaction, immersion & intent of constant use. On the contrary, caption, sound, sticker function showed the lowest users’ satisfaction, im mersion & intent of constant use. Out of image edition fu nction provided by TikTok , especially, beauty function which realizes transformation of personal shape such as adjustment of users’ face contour, adjustment of eye siz e, skin correction, etc., showed the highest satisfaction, fl ow & intent of constant use. This result will be the basic materials to understand consumers’ characteristics in add ing image edition design function of TikTok later.
비디오 편집 시스템을 활용한 원격 비디오 브라우징 서비스
[Kisti 연계] 한국멀티미디어학회 한국멀티미디어학회 학술대회논문집 2001 pp.303-307
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
최근 인터넷의 급속한 성장과 빠른 보급, 정보통신 분야의 기술퓨전 현상들은 인터넷을 이용한 다양한 콘텐츠의 개발을 가속화시키고 있다. 특히 멀티미디어 스트리밍 기술은 일반 사용자들에게 동영상은 물곤 풍부한 멀티미디어 데이터 전송을 통하여 능동적인 대화형 서비스를 제공할 수 있는 장점들을 가지고 있다. 비디오 편집시스템은 데이터의 종류에 따라서 제안된 알고리즘과 자동/수동 분류방식을 이용하여 장면전환 검출의 정확성과 효율성을 높이고자 하였으며, 사용자의 요구에 적합한 의미 정보들을 추출하고 편집하여 내용기반 데이터베이스 시스템을 구축하고자 하였다. 또한 원격에서 웹을 통한 브라우징 서비스를 통하여 데이터베이스로부터 사용자의 요구에 능동적으로 대처할 수 있는 비디오 브라우징 서비스를 제공하고자 하였다.
영상과 음성 정보를 이용한 비디오 편집 및 검색 시스템
[Kisti 연계] 한국정보과학회 한국정보과학회 학술대회논문집 2000 pp.228-230
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
동영상 데이터가 갖는 복잡하고 다양한 관계성 때문에 기존의 키워드 기반 정보 검색 방법에는 한계가 있으면 비디오 내용에 기반해 검색을 하는 내용기반 검색기법이 요구된다. 현재 MPEG-7에서도 비디오 내용 표현 방식에 관한 국제 표준화 작업이 시작되고 있다. 본 논문에서는 영상정보와 음성정보를 사용해 비디오의 원하는 부분을 내용에 기반해 검색할 수 있는 비디오 편집 및 검색 시스템을 개발하였다.
장르 특성 패턴을 활용한 매칭시스템 기반의 자동영상편집 기술
[Kisti 연계] 한국방송공학회 방송공학회논문지 Vol.25 No.6 2020 pp.861-869
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
본 논문은 영화 장르마다 나타나는 클라이맥스 패턴이 다름을 활용하여 사용자의 디바이스 내에 저장되어 있는 이미지들을 하나의 영상으로 자동생성해주는 애플리케이션 개발을 소개한다. 영화의 장르 특성 분류는 국내 영화 드라마, 액션, 공포와 국외 영화 드라마, 액션, 공포 장르를 분석하여 클라맥스 패턴 모델형을 만들었다. 클라이맥스 패턴은 영화의 특정 씬 부분에서 샷사이즈의 변화, 샷의 길이, 인서트샷 사용의 빈도를 특성 요소로 하였고, 결과를 시각화하였다. 장르별 시각화된 모델을 Firebase DB를 활용하는 템플릿으로 개발하였다. 사용자의 디바이스에 저장된 이미지를 선택하여 장르별 템플릿으로 개발된 클라이맥스 패턴 모델과 매칭하였다. 짧은 영상이지만 장르의 특성이 반영되어 감성스토리 영상을 자동생성할 수 있는 것이 본 애플리케이션의 특징이다. 최근 유튜브, 네이버와 같은 플랫폼 사업자들은 사용자가 스마트폰으로 직접 촬영한 사진이나 영상을 활용하여 자동으로 영상을 생성해주는 애플리케이션들을 매년 업그래이드하고 있으나, 영화와 같이 장르 특성을 갖는다거나, 스토리가 보이는 영상생성 기술을 포함한 애플리케이션은 아직 미흡하다. 제안한 자동영상편집은 감성전달이 가능한 영상편집 애플리케이션으로써의 발전 가능성이 있다고 예측한다.
We introduce the application that automatically makes several images stored in user's device into one video by using the different climax patterns appearing for each film genre. For the classification of the genre characteristics of movies, a climax pattern model style was created by analyzing the genre of domestic movie drama, action, horror and foreign movie drama, action, and horror. The climax pattern was characterized by the change in shot size, the length of the shot, and the frequency of insert use in a specific scene part of the movie, and the result was visualized. The model visualized by genre developed as a template using Firebase DB. Images stored in the user's device were selected and matched with the climax pattern model developed as a template for each genre. Although it is a short video, it is a feature of the proposed application that it can create an emotional story video that reflects the characteristics of the genre. Recently, platform operators such as YouTube and Naver are upgrading applications that automatically generate video using a picture or video taken by the user directly with a smartphone. However, applications that have genre characteristics like movies or include video-generation technology to show stories are still insufficient. It is predicted that the proposed automatic video editing has the potential to develop into a video editing application capable of transmitting emotions.
오픈 소스를 활용한 방송용 영상 편집 및 다기능 알람 시스템
[Kisti 연계] 한국컴퓨터정보학회 한국컴퓨터정보학회 학술대회논문집 2015 pp.39-40
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
본 논문에서는 대용량의 영상을 빠르고 정확하게 편집해야 하는 방송 환경에서 오픈 소스를 활용한 방송용 영상 편집 시스템과 편집 결과를 신속히 통보하는 다기능 알람 시스템에 대하여 다루었다. 디지털 방송에서 사용하는 영상은 매우 고품질이면서 동시에 매우 큰 데이터 용량을 가지고 있으며, 또 방송이라는 매체의 성격상 방송 시간에 맞춰 영상 편집을 완료해야 하는 시간적 부담이 있다. 본 논문에서 제안하는 방송 편집 시스템은 갈수록 그 기능이 다양하고 좋은 성능을 보여주는 오픈 소스 영상 처리 소프트웨어를 활용하는 동시에, 분산 처리 시스템을 도입하여 영상의 빠른 편집 기능을 저렴한 도입 비용으로 가능하도록 하였다. 오픈 소스 영상 처리 소프트웨어는 백엔드 모듈로 동작하며, 영상의 다양한 편집 기능들, 예를 들어 영상의 구간 자르기, 붙이기, 자막넣기, 로고 넣기, 비디오/오디오 포맷 변환 등을 수행한다. 사용자 편의성을 위한 프론트엔드 시스템은 자바 기반의 프레임워크를 사용하였으며, 비디오 편집 기능 결과 및 과정에 대한 다기능 알람 기능을 첨가하여 방송 종사자들의 업무 편리를 도모하였다. 본 시스템은 기존의 고가 방송 편집 장비와 비교하여 매우 저렴하면서도 안정된 성능을 보장하여 디지털 방송 시대에 활용 가능성이 매우 높다고 할 수 있다.
[Kisti 연계] 한국지능정보시스템학회 Journal of Intelligence and Information Systems Vol.29 No.4 2023 pp.15-30
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
최근 미디어 분야에도 인공지능(AI)을 적용한 다양한 서비스가 등장하고 있는 추세이다. 하지만 편집점을 찾아 영상을 이어 붙이는 영상 편집은, 대부분 수동적 방식으로 진행되어 시간과 인적 자원의 소요가 많이 발생하고 있다. 이에 본 연구에서는 Video Swin Transformer를 활용하여, 발화 여부에 따른 영상의 편집점을 탐지할 수 있는 방법론을 제안한다. 이를 위해, 제안 구조는 먼저 Face Alignment를 통해 얼굴 특징점을 검출한다. 이와 같은 과정을 통해 입력 영상 데이터로부터 발화 여부에 따른 얼굴의 시 공간적인 변화를 모델에 반영한다. 그리고, 본 연구에서 제안하는 Video Swin Transformer 기반 모델을 통해 영상 속 사람의 행동을 분류한다. 구체적으로 비디오 데이터로부터 Video Swin Transformer를 통해 생성되는 Feature Map과 Face Alignment를 통해 검출된 얼굴 특징점을 합친 후 Convolution을 거쳐 발화 여부를 탐지하게 된다. 실험 결과, 본 논문에서 제안한 얼굴 특징점을 활용한 영상 편집점 탐지 모델을 사용했을 경우 분류 성능을 89.17% 기록하여, 얼굴 특징점을 사용하지 않았을 때의 성능 87.46% 대비 성능을 향상시키는 것을 확인할 수 있었다.
Recently, various services using artificial intelligence(AI) are emerging in the media field as well However, most of the video editing, which involves finding an editing point and attaching the video, is carried out in a passive manner, requiring a lot of time and human resources. Therefore, this study proposes a methodology that can detect the edit points of video according to whether person in video are spoken by using Video Swin Transformer. First, facial keypoints are detected through face alignment. To this end, the proposed structure first detects facial keypoints through face alignment. Through this process, the temporal and spatial changes of the face are reflected from the input video data. And, through the Video Swin Transformer-based model proposed in this study, the behavior of the person in the video is classified. Specifically, after combining the feature map generated through Video Swin Transformer from video data and the facial keypoints detected through Face Alignment, utterance is classified through convolution layers. In conclusion, the performance of the image editing point detection model using facial keypoints proposed in this paper improved from 87.46% to 89.17% compared to the model without facial keypoints.
0개의 논문이 장바구니에 담겼습니다.
선택하신 파일을 압축중입니다.
잠시만 기다려 주십시오.