• Title/Summary/Keyword: 텍스트 수집

Search Result 691, Processing Time 0.029 seconds

A Study on the Factors of Well-aging through Big Data Analysis : Focusing on Newspaper Articles (빅데이터 분석을 활용한 웰에이징 요인에 관한 연구 : 신문기사를 중심으로)

  • Lee, Chong Hyung;Kang, Kyung Hee;Kim, Yong Ha;Lim, Hyo Nam;Ku, Jin Hee;Kim, Kwang Hwan
    • Journal of the Korea Academia-Industrial cooperation Society
    • /
    • v.22 no.5
    • /
    • pp.354-360
    • /
    • 2021
  • People hope to live a healthy and happy life achieving satisfaction by striking a good work-life balance. Therefore, there is a growing interest in well-aging which means living happily to a healthy old age without worry. This study identified important factors related to well-aging by analyzing news articles published in Korea. Using Python-based web crawling, 1,199 articles were collected on the news service of portal site Daum till November 2020, and 374 articles were selected which matched the subject of the study. The frequency analysis results of text mining showed keywords such as 'elderly', 'health', 'skin', 'well-aging', 'product', 'person', 'aging', 'female', 'domestic' and 'retirement' as important keywords. Besides, a social network analysis with 45 important keywords revealed strong connections in the order of 'skin-wrinkle', 'skin-aging' and 'old-health'. The result of the CONCOR analysis showed that 45 main keywords were composed of eight clusters of 'life and happiness', 'disease and death', 'nutrition and exercise', 'healing', 'health', and 'elderly services'.

Three Newspapers Research from The Perspective of Disability : Focusing on The Types of Disabilities on The Disabled Person Welfare Law (3개 신문사 기사에 나타난 장애관 연구 : 장애인복지법상 장애 종류를 중심으로)

  • Lim, Ok-Hee;Cho, Won-Il
    • Journal of Korea Entertainment Industry Association
    • /
    • v.14 no.7
    • /
    • pp.487-500
    • /
    • 2020
  • This research analyzed articles about the disability under the 「The Disabled Person Welfare Law」 in a major daily newspaper. A total of 7,684 articles on disability were collected from homepages of the three newspapers , , and . Through network text analysis and content analysis, we considered about "The perspective of Disability" based on "Multiple Disability Model". As a result of this research, when comparing individual models versus social models, individual models have a higher rate 64.31% than social models 35.69%. According to the newspapers, the major perception of Disability is a traditional individual model, which means disability must be solved by individuals. In addition, due to low social and institutional supports, the public's attention and consideration required for the disabled, socially weak people. This research implied that despite the changing times of looking at disability, three newspapers are still staying in the traditional paradigm. Therefore, It is required that viewing a disability from the perspective on disabled people, and a mature awareness that recognizes the diversity of individual needs. The significance of this study can be found in the fact that no attempt has been made to treat the disability perspectivec in newspaper articles as quantitative and qualitative data.

Trend Analysis of Sports for All-Related Issues in Early Stage of COVID-19 Using Topic Modeling (토픽 모델링을 활용한 코로나19 초기 생활체육 이슈 분석)

  • Chung, Yunkil;Seo, Sumin;Kang, Hyunmin
    • Journal of Intelligence and Information Systems
    • /
    • v.28 no.3
    • /
    • pp.57-79
    • /
    • 2022
  • COVID-19, which started in December 2019, has had a great impact on our lives in general, including politics, economy, society, and culture, and activities in sports and arts have also been significantly reduced. In the case of sports, sports for all fields in which ordinary citizens participate were particularly affected, and cases of infection in places closely related to people's lives, such as gyms, table tennis, and badminton clubs, also amplified the social fear of the spread of COVID-19. Therefore, in this study, we analyzed news articles related to sports for all at the time when COVID-19 was first spread, and investigated what issues were emerging and being discussed in the sports for all field under the COVID-19 situation. Specifically, we collected news articles dealt with sports for all issues under the COVID-19 situation from Korea's leading portal news sites and identified key sports for all issues by performing topic modeling on these articles. Through the analysis, we found meaningful issues such as COVID-19 outbreak in sports facilities and support for sports activities. In addition, through wordcloud analysis of these major issues, we visually understood the issues and identified the changes in these issues over time.

Civic Participation in Smart City : A Role and Direction (스마트도시 구현을 위한 시민참여의 역할과 방향에 관한 연구)

  • Nam, Woo-Min;Park, Keon Chul
    • Journal of Internet Computing and Services
    • /
    • v.23 no.6
    • /
    • pp.79-86
    • /
    • 2022
  • This study aims to analyze the research trends on the civic participation in a smart city and to present implications to policy makers, industry professionals and researchers. As rapid urbanization is defining development trend of modern city, urban problems such as transportation, environment, and energy are spreading and intensifying around the city. Countries around the world are introducing smart cities to solve these urban problems and to achieve sustainable development. Recently, many countries are modifying urban planning from top-down to down-up by actively engaging citizens to participate in the urban construction process directly and indirectly. Although the construction of smart cities is being promoted in Korea to solve urban problems, awareness of smart cities and civic participation are low. In order to overcome this situation, discussions on ideas and methods that can increase civic participation in smart cities are continuously being conducted. Therefore, in this study, by collecting publication containing both 'Smart Cities' and 'Participation (Engagement)' in Scopus DB, the topics of related studies were categorized and research trends were analyzed using topic modeling. Through this study, it is expected that it can be used as evidence to understand the direction of civic participation research in smart cities and to present the direction of related research in the future.

A Multi-speaker Speech Synthesis System Using X-vector (x-vector를 이용한 다화자 음성합성 시스템)

  • Jo, Min Su;Kwon, Chul Hong
    • The Journal of the Convergence on Culture Technology
    • /
    • v.7 no.4
    • /
    • pp.675-681
    • /
    • 2021
  • With the recent growth of the AI speaker market, the demand for speech synthesis technology that enables natural conversation with users is increasing. Therefore, there is a need for a multi-speaker speech synthesis system that can generate voices of various tones. In order to synthesize natural speech, it is required to train with a large-capacity. high-quality speech DB. However, it is very difficult in terms of recording time and cost to collect a high-quality, large-capacity speech database uttered by many speakers. Therefore, it is necessary to train the speech synthesis system using the speech DB of a very large number of speakers with a small amount of training data for each speaker, and a technique for naturally expressing the tone and rhyme of multiple speakers is required. In this paper, we propose a technology for constructing a speaker encoder by applying the deep learning-based x-vector technique used in speaker recognition technology, and synthesizing a new speaker's tone with a small amount of data through the speaker encoder. In the multi-speaker speech synthesis system, the module for synthesizing mel-spectrogram from input text is composed of Tacotron2, and the vocoder generating synthesized speech consists of WaveNet with mixture of logistic distributions applied. The x-vector extracted from the trained speaker embedding neural networks is added to Tacotron2 as an input to express the desired speaker's tone.

Analysis of domestic and foreign future automobile research trends based on topic modeling (토픽모델링 기반의 국내외 미래 자동차 연구동향 비교 분석: CASE 키워드 중심으로)

  • Jeong, Ho Jeong;Kim, Keun-Wook;Kim, Na-Gyeong;Chang, Won-Jun;Jeong, Won-Oong;Park, Dae-Yeong
    • Journal of Digital Convergence
    • /
    • v.20 no.5
    • /
    • pp.463-476
    • /
    • 2022
  • After industrialization in the past, the automobile industry has continued to grow centered on internal combustion engines, but is facing a major change with the recent 4th industrial revolution. Most companies are preparing for the transition to electric vehicles and autonomous driving. Therefore, in this study, topic modeling was performed based on LDA algorithm by collecting 4,002 domestic papers and 68,372 overseas papers that contain keywords related to CASE (Connectivity, Autonomous, Sharing, Electrification), which represent future automobile trends. As a result of the analysis, it was found that domestic research mainly focuses on macroscopic aspects such as traffic infrastructure, urban traffic efficiency, and traffic policy. Through this, the government's technical support for MaaS (Mobility-as-a-Service) is required in the domestic shared car sector, and the need for data opening by means of transportation was presented. It is judged that these analysis results can be used as basic data for the future automobile industry.

A Study on Construction of Digital Museum Archiving Regarding Dance Costume (무용공연작품 의상을 위한 디지털 뮤지엄 아카이빙 구축)

  • Jeong, Yu-Jin;Yoo, Ji-Young;Baek, Hyun-Soon
    • Journal of Korea Entertainment Industry Association
    • /
    • v.13 no.1
    • /
    • pp.81-88
    • /
    • 2019
  • This article aims to identify the characters and theme shown in dance costume and utilize them from an educational perspective by constructing digital museum archiving, which can be systematically collected, classified and stored from dance costume. It deals with definition of digital museum archiving as theoretical background and examples of how to create digital museum archiving as research content. The role that archiving plays in digital museum and effectiveness have been demonstrated. Archive is a term used to indicate extensive material and its storage and referred to as an integrative model of display in the computer-generated space. When it comes to producing dance costume as a form of digital museum, the museum is to be made in the computer-generated area of dance costume. The museum shows each division of major, medium and minor classification. The major classification divides genre of dance performance into Korean dance, modern dance and ballet. The middle involves choreographers, costume designers. The minor categorization includes newspaper, interviews, performance pictures, and programs. Digital museum has the value of space utilization, creation, culture, utilization of multiple educational programs, offering of digital museum content, two-way communication, and program development of the new display form.

Analysis of articles on water quality accidents in the water distribution networks using big data topic modelling and sentiment analysis (빅데이터 토픽모델링과 감성분석을 활용한 물공급과정에서의 수질사고 기사 분석)

  • Hong, Sung-Jin;Yoo, Do-Guen
    • Journal of Korea Water Resources Association
    • /
    • v.55 no.spc1
    • /
    • pp.1235-1249
    • /
    • 2022
  • This study applied the web crawling technique for extracting big data news on water quality accidents in the water supply system and presented the algorithm in a procedural way to obtain accurate water quality accident news. In addition, in the case of a large-scale water quality accident, development patterns such as accident recognition, accident spread, accident response, and accident resolution appear according to the occurrence of an accident. That is, the analysis of the development of water quality accidents through key keywords and sentiment analysis for each stage was carried out in detail based on case studies, and the meanings were analyzed and derived. The proposed methodology was applied to the larval accident period of Incheon Metropolitan City in 2020 and analyzed. As a result, in a situation where the disclosure of information that directly affects consumers, such as water quality accidents, is restricted, the tone of news articles and media reports about water quality accidents with long-term damage in the event of an accident and the degree of consumer pride clearly change over time. could check This suggests the need to prepare consumer-centered policies to increase consumer positivity, although rapid restoration of facilities is very important for the development of water quality accidents from the supplier's point of view.

Detecting Weak Signals for Carbon Neutrality Technology using Text Mining of Web News (탄소중립 기술의 미래신호 탐색연구: 국내 뉴스 기사 텍스트데이터를 중심으로)

  • Jisong Jeong;Seungkook Roh
    • Journal of Industrial Convergence
    • /
    • v.21 no.5
    • /
    • pp.1-13
    • /
    • 2023
  • Carbon neutrality is the concept of reducing greenhouse gases emitted by human activities and making actual emissions zero through removal of remaining gases. It is also called "Net-Zero" and "carbon zero". Korea has declared a "2050 Carbon Neutrality policy" to cope with the climate change crisis. Various carbon reduction legislative processes are underway. Since carbon neutrality requires changes in industrial technology, it is important to prepare a system for carbon zero. This paper aims to understand the status and trends of global carbon neutrality technology. Therefore, ROK's web platform "www.naver.com." was selected as the data collection scope. Korean online articles related to carbon neutrality were collected. Carbon neutrality technology trends were analyzed by future signal methodology and Word2Vec algorithm which is a neural network deep learning technology. As a result, technology advancement in the steel and petrochemical sectors, which are carbon over-release industries, was required. Investment feasibility in the electric vehicle sector and technology advancement were on the rise. It seems that the government's support for carbon neutrality and the creation of global technology infrastructure should be supported. In addition, it is urgent to cultivate human resources, and possible to confirm the need to prepare support policies for carbon neutrality.

A Generation and Matching Method of Normal-Transient Dictionary for Realtime Topic Detection (실시간 이슈 탐지를 위한 일반-급상승 단어사전 생성 및 매칭 기법)

  • Choi, Bongjun;Lee, Hanjoo;Yong, Wooseok;Lee, Wonsuk
    • The Journal of Korean Institute of Next Generation Computing
    • /
    • v.13 no.5
    • /
    • pp.7-18
    • /
    • 2017
  • Recently, the number of SNS user has rapidly increased due to smart device industry development and also the amount of generated data is exponentially increasing. In the twitter, Text data generated by user is a key issue to research because it involves events, accidents, reputations of products, and brand images. Twitter has become a channel for users to receive and exchange information. An important characteristic of Twitter is its realtime. Earthquakes, floods and suicides event among the various events should be analyzed rapidly for immediately applying to events. It is necessary to collect tweets related to the event in order to analyze the events. But it is difficult to find all tweets related to the event using normal keywords. In order to solve such a mentioned above, this paper proposes A Generation and Matching Method of Normal-Transient Dictionary for realtime topic detection. Normal dictionaries consist of general keywords(event: suicide-death-loop, death, die, hang oneself, etc) related to events. Whereas transient dictionaries consist of transient keywords(event: suicide-names and information of celebrities, information of social issues) related to events. Experimental results show that matching method using two dictionary finds more tweets related to the event than a simple keyword search.