년 - 년
DEA-SBM 모델과 EM 군집 분석에 근거한 자동차 공급사슬 효율성 평가 및 벤치마킹 방안 KCI 등재
한국기업경영학회 기업경영연구 제23권 제1호 2016.02 pp.269-289
※ 기관로그인 시 무료 이용이 가능합니다.
5,700원
본 연구는 자동차 산업 1차 협력사를 대상으로 DEA-SBM 모델과 EM 군집분석을 활용하여 자동차 공급 사슬의 효율성 분석을 실시하였다. 협력사가 속한 EM 군집의 변동과 효율성의 변동에 대해 시계열 분석을 하였으며, EM 군집 및 업종을 고려한 효율성 분석을 통하여 개선이 필요한 집단을 선별하였다. 효율이 낮은 기업이 효과적으로 개선을 진행할 수 있도록 단계별 벤치마킹 경로를 제시하고 전략적 개선 방향과 함께 최종 개선 목표값을 제시하였다. 기존 선행 연구에서는 협력사가 속한 업종 및 규모에 대한 통합적 고려 없이 전체 협력사를 대상으로 산출된 효율성 점수만으로 개선 대상을 선정하였고, 전체 공급사슬의 경쟁력을 향상하기 위해 개선이 필요한 저 효율 기업에 대해 구체적인 개선 경로 및 방안을 제시하지 못 하였다. 자동차 샤시 부품 군에 속하는 64개 협력사의 3년간 재무 자료와 고객사 평가 점수를 투입변수와 산출변수로 사용하여 실증 연구를 실시한 결과, 규모와 업종에 따른 개선 대상 집단을 효과적으로 선별할 수 있었으며, 단계별 개선 경로를 도식화하여 개선의 방향성에 대해서 명확히 제시할 수 있었다. 협력사의 특성을 통합적으로 고려한 평가 방법과 단계별 구체적인 개선 방안을 제시한 점에서 기존 연구와 차별화되며, 자동차 산업 공급사슬 관리에 실효성 있는 모형으로 공헌할 것으로 판단된다.
The objective of this paper is to present the model for evaluating the supply chain efficiency and benchmarking process of the component suppliers in Korean automotive industry. The proposed method is theoretically based on the combination of the slacks based super efficiency model (DEA-SBM) and the expectation maximization (EM) algorithm clustering analysis. DEA-SBM is one of many data envelopment analysis (DEA) with input and output criteria to evaluate the relative efficiency in order to rank the decision making unit (DMU). EM algorithm makes cluster based on probability considering the maximization of log-likelihood. This method can be differentiated from previous approaches in the literature. This study uses time series data from 2012 to 2014 of 64 tier-one Korean domestic automotive chassis parts suppliers and adopts the DEA-SBM model in order to increase the analysis's accuracy and EM clustering to classify the suppliers by the actual input resources. It also suggests the practical stepwise benchmarking paths to the inefficient firms. In this study, three input factors are selected: number of employees, assets and cost of goods sold (COGS). The output factors include two financial and three non-financial data: sales, operating profit and three kinds of customer’s evaluation scores about suppliers, i.e quality, delivery and technology. Firstly, we conduct the EM clustering to classify the suppliers depending on three input factors, resulting in seven clusters. We analyze the cluster changes of each supplier over time so as to monitor the cause of changes. Then, we calculate the super-efficiency scores of the suppliers using DEA-Solver Pro 9.0 software. We also conduct the analysis about the changes of efficiency scores over time. From this analysis, we apparently recognize which suppliers get better or worse over three years. For the sake of the comprehensive evaluation, we use a hybrid approach, i.e. the efficiency analysis based on the EM cluster as well as on the industrial sectors. Since the value chain is totally different depending on the industrial sectors. The industrial sector is also a key perspective like EM cluster. The best efficient EM cluster was cluster 7 and the worst was cluster 4. On the other hand, the best efficient industrial sector was assembly sector and the worst was fastener sector. From the results of the analysis, we can find out the most inefficient cluster and industrial sectors which need to improve efficiency to increase the overall supply chain performance. In the case study, we selected the fastener sector as the improvement targets. We suggested the stepwise improvement paths. One stage is the efficiency improvement within the same EM cluster, which anticipated fast benchmarking and improvement due to the similarity of the firms' scales. Other stage is the improvement between different EM clusters. To obtain the final improvement targets, we adopt the DEA projection analysis within the selected industrial sector. We illustrate the benchmarking paths as the case of DMU #7 belonging to the fastener sector and EM cluster 5. As we have seen in the case study, this study provides the inefficient firms with a stepwise benchmark approach considering industrial sectors with final improvement target in terms of input and output factors to reach the highest level of efficiency frontier in the same industrial sector. With this proposed method, the buyer, specifically a car manufacturer, can precisely evaluate their suppliers, strategically select the improvement targets, practically determine improving direction and suggest effective benchmarking paths to suppliers.
창업보육센터의 중장기 발전 전략 : 창업기업 인터뷰와 선진국 창업보육센터 벤치마킹을 토대로
한국정보기술응용학회 JITAM Vol.30 No.4 2023.08 pp.1-9
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
The purpose of this study was to make a mid- and long-term development plan on the business incubator center after interviewing five startups that currently being occupied in or already left the center and reviewing benchmark on business incubator centers in developed countries such as USA, Sweden, and Israel. For the interview, the three startups currently being occupied in the center and the two companies already left the center were participated. The main strengths of the center from all of these five companies were easy accessibility to the equipment and space and at the same time trustworthy from the outside vendors and/or government, etc. USA is a leading country who has long history for the startups but mostly the private companies/organizations/individuals have supported the startups in terms of funding or consulting. Also, there are countless local governments nationwide who provide funding, education, and/or space for the small businesses. Mainly based on the interview and the benchmarking, the mid- and long-term development plan for the business incubator center was made. All six themes such as consortium for investment or a local network plan were derived for the development plan which was described in this study.
유럽연합(EU)과 한국의 배출권거래제 비교연구 : 탄소누출 대응을 중심으로 KCI 등재
아시아유럽미래학회 유라시아연구 제19권 제4호 통권 제67호 2022.12 pp.29-52
※ 기관로그인 시 무료 이용이 가능합니다.
6,100원
본 연구는 유럽연합 배출권거래제 4기와 제3차 계획기간 하 한국 배출권거래제를 비교하고, 이를 토대로 제3차 계획기간 이후 한국 배출권거래제에의 시사점을 도출하고자 하였다. 먼저 유럽연합 배출권거래제 4기에서는 탄소누출을 방지하기 위해 무상할당 방식과 탄소누출노출계수(Carbon Leakage Exposure Factor, CLEF), 탄소누출목록 (Carbon Leakage List) 등의 제도를 채택하고 있다. 무상할당의 경우 배출효율에 따른 할당방식인 벤치마크 방식을 사용함으로써 탄소누출의 위험성이 높 은 사업장 (Installation) 중 배출효율이 가장 높은 경우에 한정하는 형태를 보이고 있다. 이는 무상할당 을 원칙으로 하면서, 무상할당을 위해 배출량 기반 할당방식과 배출효율 기반 할당 방식을 함께 활용하 는 한국 배출권거래제(K-ETS)와 유럽연합 배출권거래제의 주요 차이점이다. 본고는 양 배출권거래제의 특징을 비교한 이후, 향후 한국 배출권거래제 추진방향을 다음과 같이 제 시한다. 첫째, 탄소누출에 대한 한국 배출권거래제의 문제의식을 실제 제도에 반영하기 위해 탄소누출 대응 을 위한 구체적 수단 마련이 필요하다. 본 연구에서는 탄소누출 대응을 위한 구체적 수단의 예시인 유럽연합의 탄소누출노출계수 산정 방식을 검토하고, 이를 한국 배출권거래제 하에서 활용하기 위한 선 행조건으로 배출효율 기반 할당방식 (Benchmark, BM)의 확대와 탄소누출목록과 유사한 방식의 탄소 누출 노출 업종 선별을 제시하였다. 둘째, 국내 배출권거래제 적용 업종 중 실제로 탄소누출에 고도로 노출된 업종을 선별하기 위한 새로 운 방법론의 개발이 필요하다. 본 연구에서는 유럽연합의 탄소누출목록 등재 여부 결정 방식이 한국에 적용될 수 있는지의 여부를 검토하고, 그 결과로 한국 배출권거래제 하에서 탄소누출 노출 업종을 선별 하기 위해서는 새로운 방법론을 제시하였다. 셋째, 한국 배출권거래제의 무상할당은 장기적으로 축소되어야 한다. 이는 탄소누출을 차단하기 위해 앞서 제시된 두 가지 방안의 안정적인 실행의 필수조건이자 향후 배출권거래제와 연관된 탄소국경조정 제도 등의 국제적 제도 시행 이후에도 한국의 대외 경쟁력을 유지하기 위해 반드시 달성되어야 하는 목 표이다. 본고에서는 이러한 목표 달성의 필요성을 검토하기 위해 유럽연합 배출권거래제와 연동되어 운영되는 탄소국경조정제도 하에서 한국이 부담하여야 할 잠재적 준수비용을 산출하였다. 산출 결과, 배출권거래가격이 유럽연합 수준과 현격한 차이를 보이는 현 상황에서는 수출 측면에서 한국의 대외 경쟁력 약화가 우려되며, 이와 함께 탄소누출 대응을 위한 구체적 수단의 이행에도 부담이 된다는 점을 확인하였다.
Based on the perception that carbon leakage, which is considered to be a significant matter in the European Union Emissions Trading System (EU ETS), may occur in the Korea Emissions Trading System (K-ETS) under the third compliance period, this paper compared the direction of Phase 4 of EU ETS with the 3rd compliance period of Korean Emissions Trading System (K-ETS) to suggest the future direction of the K-ETS after 3rd compliance period. Phase 4 of EU ETS uses the free allocation method, the carbon leakage exposure factor (CLEF), and carbon leakage list to prevent carbon leakage, and in the case of free allocation, the benchmark method based on the emission-efficiency tradeoff. At this point, there is a difference between the K-ETS, which uses the Grandfathering with the Benchmark for free allocation, and EU ETS. After comparing the characteristics of the two emission trading systems, this paper presents the future direction of the Korean emission trading system as follows. First, we prepare the means to respond to carbon leakage in order to reflect the recognition of the Korean emission trading system on carbon leakage. This study reviewed the European Union's carbon leakage exposure factor (CLEF) calculation method, which is an example of carbon leakage response, and suggested expanding the emission efficiency-based allocation method (BM) while at the same time selecting carbon leakage exposure industries as prerequisites for using similar policies with the CLEF under the Korean emission trading system. Second, it is necessary to develop a new methodology to select industries that are highly exposed to carbon leakage in the K-ETS. This study examined if the European Union's method of determining whether to register carbon leakage lists could be applied to Korea, and concluded that a new methodology should be proposed to select carbon leakage exposure industries under the Korean emission trading system. Third, the role of free allocation of the K-ETS should be reduced in the long run. This is a prerequisite for the stable implementation of the two measures mentioned earlier to block carbon leakage and is a goal that must be achieved to maintain Korea's external competitiveness with respect to the international systems such as the carbon border adjustment system. To examine the necessity of achieving these goals, this paper calculated the potential compliance costs that Korea must bear under the carbon border adjustment system operated in conjunction with the EU emission trading system. As a result of our calculation, Korea's external competitiveness is feared to weaken in terms of exports as the emission trading prices are significantly different from those in the European Union. It also appears rather burdensome to implement specific measures to cope with the carbon leakage.
JASMIN: Shielding Studies on High Energy Neutron Produced By 120 GeV Protons
대한방사선방어학회 대한방사선방어학회 학술발표회 논문요약집 2010년도 대한방사선방어학회 춘계 학술발표회 및 심포지움 2010.04 pp.94-95
OLTP서버 성능측정 및 규모산정을 위한 벤치마크 기준에 대한 고찰 KCI 등재후보
한국디지털정책학회 디지털융복합연구 제7권 제3호 2009.09 pp.25-33
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
Historically, performance prediction and sizing of server systems have been the key purchasing argument for customer. To accurate server's sizing and performance prediction, it is necessary to correctness guideline for sizing and performance prediction. But existing guidelines have many errors. So, we examine the benchmarks of performance organization such as SPEC and TPC. And then we consider to TPC-C and TPC-E benchmarks for OLTP server's sizing and performance prediction that is a basic concept of guidelines. Eventually, we propose improvement of errors in guidelines.
Benchmark of production yields of Bi(p,x) reactions with Monte Carlo codes
대한방사선방어학회 대한방사선방어학회 학술발표회 논문요약집 2018년도 대한방사선방어학회 추계 학술발표회 논문요약집 2018.11 pp.117-118
Court Ruling on the English Benchmark Requirement for Graduation in Taiwan SCOPUS KCI 등재
아시아영어교육학회 The Journal of AsiaTEFL Vol.16 No.1 2019.03 pp.345-348
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
The ILECA and IELTS Exams : A Benchmark Report SCOPUS KCI 등재
아시아영어교육학회 The Journal of AsiaTEFL Vol.17 No.2 2020.06 pp.742-749
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
Assessing Retirement WealthUsing Composite Measures of Adequacy Benchmark KCI 등재
한국응용경제학회 응용경제 제16권 제3호 2014.12 pp.125-163
※ 기관로그인 시 무료 이용이 가능합니다.
8,400원
본 연구는 예비은퇴자들의 은퇴자산의 적정성을 항상소득(permanent income)과 항상소비(permanent consumption)라는 두 벤치마크를 이용하여 평가하였다. 은퇴자산은 포함되는 자산의 종류 및 범위에 따라 확장적 정의, 중간적 정의, 제한적 정의의 세 가지를 사용하였고 각각에 따라 추정된 은퇴자산 적정성의 결정요인을 비교 분석하였다. Survey of Consumer Finances 자료를 이용한 분석결과, 예비은퇴자들의 은퇴자산을 어떻게 정의하는가와 어떠한 벤치마크를 이용하는지에 따라 은퇴자산의 적정성은 상당히 다르게 나타났다. 그러나 은퇴자산의 적정성을 결정하는 요인은 벤치마크의 종류와 은퇴자산의 정의와 관련 없이 일관성 있게 유사하게 나타났는데, 예상은퇴연령, 주관적 위험감수도, 확정급여형 퇴직연금의 소유여부, 비금융자산의 소유여부가 주요결정요인으로 밝혀졌다.
Based on two representative benchmarks, permanent income and permanent consumption, this study examines the adequacy of retirement wealth among pre-retirees who are currently employed. It compares the determinants of retirement wealth adequacy according to two benchmarks under various definitions of retirement wealth: broad, intermediate, and narrow wealth. Analysis using the Survey of Consumer Finances shows significant differences in the proportion of pre-retirees with adequate wealth according to benchmarks and wealth definitions. The important determinants of retirement wealth adequacy are fairly consistent regardless of the benchmarks and wealth definitions. The crucial determinants include planned retirement age, subjective risk tolerance, and ownership of defined-benefit plans and non-financial assets.
A Modified YoloV4 Network with Medium-Scale Challenging Benchmark for Efficient Animal Detection
한국차세대컴퓨팅학회 한국차세대컴퓨팅학회 학술대회 2023 한국차세대컴퓨팅학회 춘계학술대회 2023.06 pp.183-186
Animal detection and classification are crucial for effective wildlife management (WM) and reducing risks associated with animals related road accidents and attacks. Previous attempts trained the models using imbalanced data with fewer representative features and baseline models without improvement. This paper presents a new dataset of five animal classes captured in various poses, lighting conditions, and intraclass variations. The standard coupled detection head of the YoloV4 algorithm faces limitations when performing simultaneous classification and localization due to shared parameters and inputs. To address this issue, we propose a decoupled detection head (DDH) that handles these tasks separately, improving performance. We conducted extensive experiments using the proposed dataset. We found that the optimal backbone features marginally improve the performance of the modified network compared to state-of-the-art (SOTA) works in the subject domain. Our work contributes by addressing the limitations of the standard YoloV4 algorithm and proposing a new dataset for researchers to use in future studies.
4,500원
인공지능(AI)은 백오피스 자동화에서 공공 행정, 공공 서비스 제공, 정책 실행의 전략적 구성 요소로 자리잡았다. 그러나 국가마다 제도적 역량, 규제 전통, 데이터 성숙도, 디지털 인프라가 다르기 때문에AI 거버넌스에 대한 국가별 접근 방식은 크게 다르다. 이 논문에서는 싱가포르를 국제 벤치마크로 삼아 한국과 우즈베키스탄의AI 거버넌스를 비교 분석한다. 이 연구는 질적 비교 사례 연구 설계를 적용하고 국가AI 전략, 입법, 국제 지표, 다자 보고서, 동료 검토 공공 행정 문헌을 바탕으로 진행된다. 분석은 네 가지 차원으로 구성된다: 국가AI비전과 전략, 제도 및 규제 프레임워크, 정부에서의 부문별AI 적용, 도전 과제 및 가능성 조건. 결과에 따르면 싱가포르는 자발적 거버넌스, 강력한 투자, 실질적인AI 평가 메커니즘을 중심으로 한 유연한 도구 기반 벤치마크를 대표한다. 한국은AI 기본법, AI 안전 기관, 국가AI 인프라에 의해 뒷받침되는 법적으로 체계화되고 제도적으로 조정된 모델을 대표한다. 우즈베키스탄은 전자정부 서비스, 데이터 시스템, 법률 개정이 점진적으로 발전하고 있는 플랫폼 우선주의 및 역량 강화 모델을 대표한다. 여기서는AI 거버넌스가 순수한 기술적 과제라기보다는 근본적으로 거버넌스 과제라고 본다. 우즈베키스탄과 유사한 전환 경제에서 가장 시급한 우선순위는 데이터 거버넌스 강화, AI 안전 평가 역량 구축, 공무원 서비스에AI 문해력 내재, 국제 협력 심화, 현지 언어 및 평가 역량을 통한 디지털 주권 개발로 본다.
Artificial intelligence (AI) has moved from back-office automation to a strategic component of public administration, public service delivery, and policy implementation. Yet national approaches to AI governance differ substantially because countries vary in institutional capacity, regulatory tradition, data maturity, and digital infrastructure. This article develops a comparative analysis of AI governance in South Korea and Uzbekistan, using Singapore as an international benchmark. The study applies a qualitative comparative case study design and draws on national AI strategies, legislation, international indices, multilateral reports, and peer-reviewed public administration literature. The analysis is organized around four dimensions: national AI vision and strategy; institutional and regulatory framework; sectoral AI application in government; and challenges and enabling conditions. The findings show that Singapore represents a flexible, tool-based benchmark centered on voluntary governance, strong investment, and practical AI evaluation mechanisms. South Korea represents a legally codified and institutionally coordinated model, anchored by the AI Basic Act, presidential-level coordination, AI safety institutions, and national AI infrastructure. Uzbekistan represents a platform-first and capacity-building model, in which e-government services, data systems, and legal amendments are being developed incrementally. The article argues that AI governance is fundamentally a governance challenge rather than a purely technological one. For Uzbekistan and similar transitioning economies, the most urgent priorities are strengthening data governance, building AI safety evaluation capacity, embedding AI literacy in the civil service, deepening international cooperation, and developing digital sovereignty through local language and evaluation capabilities.
영국의 사회복지학 교과목 지침연구 - 학계와 실무와의 관계를 중심으로 -
[NRF 연계] 한국사회복지교육협의회 한국사회복지교육 Vol.29 2015.03 pp.19-40
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
영국에서 사회복지는 변화하는 사회 속에서 인간관계에서의 문제해결능력을 증진시키고, 개인의 복지를 이룩하기 위한 노력으로 뜻매김한다. 영국에서의 사회복지의 교과 과정은 교육부의 지침에서 비교적 자유롭고 자율적이며, 독립적으로 운영되고 있다. 영국에서 사회복지의 특징은 그 중심이 학교에서의 교육보다는 직업이나 직장으로 옮아가고 있다는 점이고, 인적서비스(human service)와의 관련성을 강조하고 있다는 점이다. 그 대표적인 사례가 전문직업능력기준이다. 전문직업능력기준은 사회복지 직무를 수행함에 있어서 사회복지사의 능력발전을 향한 기본계획을 제공함을 목적으로 한다. 여기에는 자격증 취득 전과 자격증 취득 후 과정으로 나누어지고, 이들 과정은 학생 단계에서부터 경력있는 사회복지사 단계를 거쳐, 고등단계(advanced level) 및 전략단계(strategic level)로 모두 9 단계로 되어 있다. 영국에서 사회복지는 자유주의의 전통에 기초하므로 교육에 관해서는 교수나 학교 또는 학계에 맡겨둔다는 점이다. 그러면서도 직업의 기준은 실무나 산업계에서 정하는 기준이 존재하므로 취업을 위해서는 그 기준을 충족시켜야 하기 때문에 자연스럽게 교육지침이 존재하는 것과 같은 효과를 누리고 있다.
Social welfare in the UK means that the social work profession promotes social change, problem solving in human relationships and the empowerment and liberation of people to enhance well-being. The curricula in the UK have been operated free from the UK government regulation. It means that the venue of welfare education has been shifted from university to workplace, which symbolizes the Professional Capabilities Framework(PCF). PCF has nine domains. The nine capabilities should be seen as interdependent, not separate. Throughout their careers, social work students and practitioners need to demonstrate integration of all aspects of learning, and provide a sufficiency of evidence across all nine domains: professionalism, values and ethics, diversity, rights, justice and economic well-being, knowledge, critical reflection and analysis, intervention and skills, contexts and organizations, professional leadership.
[NRF 연계] 한국경제학회 경제학연구 Vol.62 No.1 2014.03 pp.5-28
※ 협약을 통해 무료로 제공되는 자료로, 원문이용 방식은 연계기관의 정책을 따르고 있습니다.
한국의 2013년 출산율이 1.18명으로 10년째 OECD 회원국들 중 최하위를 기록했다. 노령화가 빠르게 진행되고 있는 상태에서 초저출산 현상이 겹쳐지고 있어 이대로 가다간 머지않아 한국경제는 성장을 멈추게 되고 더 나아가 국가의 생존까지 위태롭게 될 수 있다. 본고에서는 베커 교수의 가족경제론의 관점에서 한국의 과거 50년 출산율 추이와UN의 향후 50년 세계출산율 연구를 토대로 한국의 향후 50년 장기균형 출산율은2명 수준이 될 것으로 분석하였다. 향후 한국의 실제출산율은 여성임금의 가격효과와 소득효과, 그리고 정부의 출산장려정책에 달려있으나 결국은 2자녀 장기균형수준에 수렴해갈 것으로 전망하였다.
With a fertility rate of 1.18 in 2013, Korea ranks the lowest among themembers of OECD. Korea is now experiencing a trend of an aging population. If this trend continues for some time, Korea may not be able to survive on thisplanet. This paper attempts to demonstrate that Korea’s long-term benchmarkfertility rate is two children and argues using the results of the UN’s recentpopulation study that Korea’s actual fertility converges to its long-termbenchmark rate of 2.0 in the next 35 years towards 2050. Also this paperargues that the actual path of fertility movement depends on both thegovernment’s child-support policy and the relative strength of the income effectover the price effect of woman’s wage income.
대규모 언어 모델(LLM)의 포괄적 성능 비교 평가를 위한 평가 지표 및 데이터셋 개발 : 폐쇄형 LLM과 공개형 LLM의 비교를 중심으로 KCI 등재
한국경영정보학회 경영정보학연구 제26권 제3호 2024.08 pp.163-185
※ 기관로그인 시 무료 이용이 가능합니다.
6,000원
2020년 OpenAI가 1,750억 파라미터 규모의 GPT-3를 공개한 이후 간단한 작업부터 복잡한 작업에 이르기까지 다양한 다운스트림 작업에 대응하는 대규모 언어 모델(LLM)의 개발이 가속화되고 있다. LLM이 개발되고 고도화됨에 따라 LLM의 성능을 객관적으로 평가할 수 있는 평가 지표와 데이터셋이 개발되어 활용되고 있다. 이러한 데이터셋은 다양한 분야에 대해 LLM을 객관적으로 평가함에 있어 좋은 성과를 거두었으나, 규모 측면에서 개인이나 소규모 기관에서 활용하기 어렵고 실용적 측면에서 실제 사용자가 체감하는 바와 다소의 괴리를 가지고 있다. 이에 본 연구에서는 사용자의 활용 패턴을 반영하여 비교적 작은 양의 데이터를 활용해 LLM을 평가할 수 있는 평가 지표 및 데이터셋을 제시한다. 더 나아가, 가중치를 일반에 공개하는 공개형 대규모 언어 모델의 개발이 가속화되고 고성능의 공개형 LLM이 출시되고 있음에 따라 연구 수행 시점인 2024년 4월 기준 최신의 폐쇄형 LLM 4종과 공개형 LLM 6종에 대한 평가를 시행하고 폐쇄형 LLM과 공개형 LLM의 비교 평가 결과에 대해 논의한다. 연구 결과 새롭게 개발한 데이터셋이 작은 규모에도 불구하고 기존 데이터셋과 유사한 경향성을 보이는 것으로 나타났다. 상식 추론 및 글 스타일 변환과 같은 간단한 작업에서는 공개형 LLM이 폐쇄형 LLM과 대등하거나 우세한 성능을 보였으나 수학, 코딩, 이미지 질의응답 등의 복잡한 작업에서는 큰 성능 격차를 보임을 확인하였으며, 더 나아가 비교적 작은 규모의 LLM이 규모 대비 좋은 성능을 보임을 확인하였다.
The development of large language models (LLMs) has accelerated since OpenAI released GPT-3, which demonstrated generalizability and capability for various downstream tasks, thanks to its 175 billion parameters. Various metrics and datasets for LLM evaluation have been developed to objectively assess LLMs’ performance. Although existing evaluation metrics and datasets have widely been used across various fields, their large scale hinders their use in small organizations or by individuals. Furthermore, there is degree of discrepancy between evaluation results and actual user experiences. The study proposes evaluation metrics and datasets with relatively small amounts of data while reflecting real-world user experiences. In the process of testing the proposed metrics and datasets, the research evaluates and compares four closed-LLMs and six open-LLMs, which are latest as of April 2024. The results show that proposing datasets exhibited trends similar to existing datasets despite its smaller size, and furthermore, well reflected actual user experiences. Moreover, open-LLMs performed similar, or indeed, better than closed-LLMs in simple tasks while closed-LLMs performed significantly better in complex tasks such as mathematics, coding, and vision question-answering.
효율적인 웹 서버 관리를 위한 평가시스템 설계에 관한 연구 KCI 등재후보
한국융합보안학회 융합보안논문지 제7권 제3호 2007.09 pp.1-6
※ 기관로그인 시 무료 이용이 가능합니다.
4,000원
웹 서비스 제공에 중요한 비중을 차지하는 서버의 성능을 향상시키기 위해서는 현재 서버의 성능 평가 관리가 매우 중요하다. 이는 서버의 자원을 효율적을 활용하고 있는지에 대한 검증을 바탕으로 소프트웨어와 하드웨어적 성능 개선점을 보완하기 위한 방법이라 할 수 있다. 기존의 서버 성능 평가 방법으로는 서버와 클라이언트간 소프트웨어를 설치하여 클라이언트에서 서버로 접근하여 부하를 발생하고 이를 시뮬레이션 하여 패킷량 별로 발생하는 작업부하의 처리 효율을 평가하는 방식을 활용하였다. 이와 같은 방식은 서버 클라이언트가 발생하는 작업을 분석하기 위한 적절한 방법이지만, 단순한 작업 부하 처리 데이터만으로 평가함으로써 서버 성능을 개선하기 위한 명확한 데이터를 제공하지 못하는 한계성이 있다. 따라서 본 논문에서는 웹 서비스의 성격과 사용자 패턴을 구분하여 서버에 가해지는 작업부하 처리 평가를 함에 있어 보다 다양한 기준을 적용하여 측정하기 위한 방법을 제시하고자 한다.특히, 개인 관련 웹 서비스 환경과 비즈니스 형태의 웹 서비스에서의 평가 방법을 제시한다.
It is important to manage and correctly evaluate the performance of server in order to improve he performance of server. That is, with the aim of finding whether resources are properly utilized, it is a method to improve both software and hardware aspects. Conventional method in evaluating the server performance involved installing software between server and client and generating the load and by simulating the process, it evaluated the handing efficiency of work load generated per packet volume. Therefore, this paper aims to pro-vide precise measurement method by distinguishing the characteristics of web server and the users’ usage pattern and by evaluating the work load mana-gement through applying various standards. Specially, it presents the evaluation method from web service of personal and business.
데이터베이스 시스템을 위한 EBORD 성능 평가 방법론 KCI 등재
한국정보기술응용학회 JITAM Vol.12 No.2 2005.06 pp.29-43
※ 기관로그인 시 무료 이용이 가능합니다.
4,800원
The paper presents the EBORD (Extended Benchmark for Object-Relational Databases) benchmark, which is an extension of the BORD benchmark for object-relational databases. The EBORD benchmark is developed to evaluate the database common functions that should be supported in modern database systems. Besides the 36 test queries already defined in the BORD benchmark, totally 22 test queries in five categories are newly defined in order to measure the index-relevant performance issues and database import capabilities. The EBORD benchmark also features scalability, use of a synthesized database, and a query-oriented evaluation. In order to show the feasibility of the proposed benchmark, we implement it with two commercial database systems. The experimental results and analyses are also reported.
AI 모델 검증을 위한 기독교 벤치마크 개발을 위한 연구 KCI 등재
한국실천신학회 신학과 실천 제98호 2026.02 pp.571-595
※ 기관로그인 시 무료 이용이 가능합니다.
6,300원
인공지능 기술이 급속도로 발전함에 따라 인공모델에 대한 검증도 활발히 이루어지고 있다. 인공지능 모델 검증은 주로 벤치마크(Benchmark) 방법으로 이루어지는데, 표준 화된 테스트 세트와 평가기준을 적용한다. 대표적인 인공지능 벤치마크로는 MMLU, GPQA 등이 있는데, 새로운 인공지능 모델이 나올 때마다 벤치마크 검증 결과와 분 석이 함께 제시된다. 2025년 AI안전센터와 스케일AI는 기존의 벤치마크의 한계를 극 복하기 위해 전문가들로부터 질문을 수집한 인류 마지막 시험(Humanity’s Last Exam) 프로젝트를 실시하고, 연구한 결과를 보고하였다. 일반 분야와 달리 기독교 분 야에서의 벤치마크 연구는 제대로 이루어지지 않고 있다. 따라서 본 연구에서는 정통 기독교 성경과 교리에 기반한 벤치마크의 사례를 탐색하고 인공지능 모델의 정확성을 검증하기 위해 성경고사 문항을 토대로 기독교 벤치마크 기초 실험을 실시하였다. 실 험 결과로 세 가지 모델인 챗GPT, 제미나이, 클로드의 정확도가 94% 이상을 나타났 으며, 보정 에러는 7% 이하로 나타났다. 또한 거대언어모델별로 정확도와 보정 에러 의 근소한 차이가 나타났다. 전체 정확도는 제미나이, 클로드가 96%로 다소 높았고, 챗GPT는 94%로 상대적으로 낮았다. 주관식 문항은 세 모델 모두 100%의 정답률을 보였고 평균 확신도는 97% 이상으로 보고되었다. 보정 에러는 클로드, 제미나이, 챗GPT 순서로 나타났다. 세 모델 모두 보정 에러가 낮지만 높은 정답률과 확신도에 비 해 틀린 문항이 일부 나타난 것으로 보아 여전히 환각이 존재한다는 것을 알 수 있 다. 문항별 오답과 확신도를 분석한 결과 거대언어모델이 환각을 일으키는 현상이 일 부 분석되었다. 주요 사례는 제시한 성경구절에 명확하게 답이 나타나 있지 않아서 전체 맥락을 살펴보고 추론이 필요한 경우, 유사한 단어에서 정확한 단어를 선택해야 하는 경우, 성경구절을 명확히 참조하지 않고 문장의 앞부분만 계산해서 답을 선택한 경우, 한국어를 제대로 인식하지 못해 오답을 표기한 경우로 나타났다. 현재 개발된 인공지능 모델은 기독교 분야에서 잠재적 활용 가능성이 있으나 실제 활용에서는 주 의가 필요하다는 결론이 도출되었다. 또한 기독교 벤치마크를 개발할 때는 성경 구절 에 나타난 단순 지식을 묻는 문항 보다는 고차원적 추론이나 앞뒤 문맥의 맥락을 이 해해야 풀 수 있는 문항으로 개발할 필요가 있다는 점을 시사한다. 본 연구의 결과는 기독교 분야의 벤치마크 모델 개발의 기초자료가 된다.
As artificial intelligence technology advances rapidly, validation of AI models is also being actively conducted. AI model validation is primarily carried out through benchmarking methods, which apply standardized test sets and evaluation criteria. Representative AI benchmarks include MMLU and GPQA, and whenever a new AI model is released, benchmark validation results and analyses are presented together. In 2025, the Center for AI Safety and Scale AI conducted the Humanity’s Last Exam project, collecting questions from experts to overcome the limitations of existing benchmarks, and reported their findings. Unlike in general domains, benchmark research in the Christian domain has not been adequately conducted. Accordingly, this study explores benchmark cases grounded in orthodox Christian Bible and doctrine and conducted a foundational Christian benchmark experiment based on Bible examination items to verify the accuracy of AI models. The experimental results showed that the accuracies of the three models— ChatGPT, Gemini, and Claude—were at least 94%, and the calibration error was at most 7%. Small differences in accuracy and calibration error were observed across large language models. Overall accuracy was somewhat higher for Gemini and Claude (96%), whereas ChatGPT was relatively lower (94%). For the constructed-response items, all three models achieved a 100% correct answer rate, and the reported an average confidence was at least 97%. Calibration error was observed in the order of Claude, Gemini, and ChatGPT. Although all three models exhibited low calibration error, the presence of some incorrect items despite high accuracy and confidence indicates that hallucinations still exist. An analysis of item-level incorrect responses and confidence revealed some instances of hallucination in large language models. Major cases included situations in which the correct answer was not explicitly stated in the presented biblical passage and required inference from broader context, cases requiring selection of an exact term among similar words, cases in which the answer was chosen by considering only the beginning of a sentence without explicit reference to the biblical passage, and cases in which incorrect answers were produced due to inadequate recognition of Korean. It was concluded that, although currently developed AI models have potential applicability in the Christian domain, caution is required in their practical use. In addition, the findings suggest that, when developing a Christian benchmark, it is necessary to design items that require higher-order reasoning or understanding of contextual coherence rather than items that merely ask for simple knowledge stated in biblical passages. The results of this study provide foundational data for developing benchmark models in the Christian domain.
공중 인간 행동 인식을 위한 다양한 관점과 배경 벤치마크
한국차세대컴퓨팅학회 한국차세대컴퓨팅학회 학술대회 2023 한국차세대컴퓨팅학회 춘계학술대회 2023.06 pp.179-182
The aerial view diverse action recognition (AR) benchmark provides a valuable resource for researchers and developers in computer vision (CV) for human actions recognition (HAR) from an aerial perspective. With the increasing use of unmanned aerial vehicles (UAVs) for surveillance, delivery, search, and rescue, a robust understanding of human actions from an aerial view is crucial. Existing datasets lack representation of common outdoor actions and are unsuitable for intelligent UAVs. This article proposes a dataset that captured various actions from diverse viewpoints and in different environments. The dataset includes three viewpoints (Top, left, and right) allowing angle-invariant algorithm development. State-of-the-art algorithms (3D, and 2D convolutions with sequential learning) are evaluated on the dataset. The proposed model demonstrates exceptional performance with high accuracy (87.5%), precision (86.3%), and recall (87.2%) rates. The robustness of the model is showcased through real-time testing, indicating that the proposed dataset and model contribute to advancing research from drone view AR and have the potential to enhance surveillance and other UAV applications.
비상장 중소기업은 상장기업을 벤치마킹해야 하는가? - 전략적 성향에 대한 비교연구 -
한국기업경영학회 기업경영연구 제16권 제4호 2009.12 pp.183-203
※ 기관로그인 시 무료 이용이 가능합니다.
5,700원
최근 경기 침체기를 맞아 기업들은 원가절감을 통한 수익성제고를 우선시하는 양상을 보이고 있다. 그런데 전략적 프레임에 의해서 기업은 추구하고자 하는 경쟁우위와 여러 활동들을 결정하게 된다. 따라서 기업의 특성과 전략이 조화를 이루어야 양호한 기업 성과가 산출될 수 있다. 한편, 우리나라의 전체 기업 중 중소기업이 차지하는 비중은 2006년 말 현재 99.9%이며 지속적으로 증가하고 있는데 비해서 KOSPI 및 KOSDAQ 주식시장에의 상장률은 약 0.06%로서 대부분의 중소기업들이 비상장기업에 속하고 있다. 이러한 점을 감안할 때, 이들 비상장기업의 전략적 성향에 대한 연구는 학문적으로나 실무적인 측면에서 유용한 시사점들을 던져줄 수 있다. 특히 본 연구는 상장기업을 조사대상으로 하였던 기존연구와의 비교분석을 통하여 비상장 중소기업의 전략적 특징을 규명하고자 하였다. Porter의 본원적 전략에 기초한 군집분석 결과 비상장 중소기업도 상장기업과 유사하게 4개 유형의 전략집단(원가우위 전략집단, 차별화 전략집단, 듀얼 전략집단, 무전략집단)으로 분류되었다. 흥미로운 것은 우리나라의 비상장 중소기업들은 원가우위 전략을 가장 많이 채택하고 있지만, 목표하는 기업성과를 제대로 거두지 못하고 있다는 점이다. 반면, 차별화 전략을 채택한 비상장 중소기업들은 상대적으로 양호한 기업성과를 거두고 있었다. 이러한 조사결과는 경쟁우위 및 기업능력 측면에서 설명될 수 있다. 본 연구는 특히 비상장 중소기업의 경쟁우위의 원천을 규명함으로써, 보다 효과적인 전략수립 방향과 그에 따른 조직관리 방향을 제시하였으며 무조건적으로 상장기업을 모방하는 전략수립 행태는 지양되어야 함을 시사하고 있다.
Firms should decide the competitive advantages and behaviors depending on the strategic framework. Therefore the harmony of Firms’ characteristics with strategies can produce favorable performances. Meanwhile, The ratio of Korean SMEs (Small and Medium sized Enterprises) to total is 99.9% in 2006 and going up continuously. However, the ratio of IPO (Initial Public Offering) to KOSPI and KOSDAQ is only 0.06%, This means that most Korean SMEs are unlisted companies. Consideration of these aspects, the study on Korean unlisted companies can give lots of useful academic and managerial implications to us. To the best of my knowledge, this is the first attempt to generalize findings in the strategic tendencies of Korean unlisted companies through the comparative analysis with listed ones. Especially, this study tries to identify the strategic characteristics of Korean unlisted companies by comparative analysis to author’s prior study that investigated Korean listed companies. The results of this study revealed that the Korean unlisted companies were similar to listed ones in some aspects, but different in another aspects. First of all, the performances of unlisted companies were inferior to those of listed ones in every indices. This is due to the characteristics of SMEs mostly consisted with Korean unlisted companies. Second, the result of cluster analysis on the basis of Porter’s generic strategy showed that the Korean unlisted companies were classified into four strategic types (Cost leadership group, Differentiation group, Dual group, Stuck-in-the-middle group), similar to listed companies. However, the most unlisted companies were in Cost leadership strategic group. unlikely to listed ones. Meanwhile, the ratios of differentiation and dual strategic groups in unlisted companies were different with those in listed ones. That could be interpreted in the insufficiency of available resources, which was the characteristics of unlisted companies. The Existing studies have insisted that the strategic tendencies of firms influenced upon performances like effectiveness and adaptiveness and efficiency. And this opinion was supported in the study of Korean listed companies. However, it was not supported in Korean unlisted ones, and this is the interesting findings of this study. That is, the unlisted cost leadership strategic firms pursue to increase efficiency but they have different characteristics from listed ones in aspect of the source of competitive advantages. The result of it, they do not achieve favorable goals. On the contrary, the unlisted differentiation strategic firms achieved favorable goals very well. But the performances of them were inferior to those of listed differentiation strategic firms, though I expected better performances owing to the relative advantages. The capabilities can be taken as the alternative explanation for this result, but further researches are needed for exact analysis. This study has some fruitful implications. First of all, This study identified the relationships between firms’ performances and strategies. Especially, as identifying the sources of competitive advantages which should be important considerations for employment of strategy for unlisted companies, indicates the effective strategic direction which reflect the characteristics of unlisted companies, rather than strategic planning behavior of undifferentiated imitating listed ones. Also, this study implies the effective organization management ways to strategic planning by identifying the characteristics of unlisted companies’ performances based on Poter’s generic strategies. For examples, in the case of employment of differentiation strategy, it will be effective to manage organization focusing on market sensing. Meanwhile in the case of employment of cost leadership strategy, it will be effective to manage organization focusing on interaction within and between organizations. This study has some limitations also. First, The points of data gathering time for Korean listed companies and unlisted ones were not accorded, so the effects of extraneous variables could be involved. Second, because this study was a cross sectional research, the inference of causality relationships between variables was limited. Third, subjective judgement could be involved because data collecting was depending on single informant. Strategy can be forecasted on th basis of internal resources and capabilities of firm, so the researches on what strategy will be effective in what conditions are necessary. This is the matter of the capabilities of firm. Deep analysis on the requirements for IPO can give more effective comparative analysis frame by making the relative characteristics of Korean listed companies also. And considerations of control variables like the turbulence of environment in designing research models can increase the possibility of generalization of the researches.
0개의 논문이 장바구니에 담겼습니다.
선택하신 파일을 압축중입니다.
잠시만 기다려 주십시오.