Search | Korea Science

Term Frequency-Inverse Document Frequency (TF-IDF) Technique Using Principal Component Analysis (PCA) with Naive Bayes Classification

J.Uma;K.Prabha
- International Journal of Computer Science & Network Security
- /
- v.24 no.4
- /
- pp.113-118
- /
- 2024
Pursuance Sentiment Analysis on Twitter is difficult then performance it's used for great review. The present be for the reason to the tweet is extremely small with mostly contain slang, emoticon, and hash tag with other tweet words. A feature extraction stands every technique concerning structure and aspect point beginning particular tweets. The subdivision in a aspect vector is an integer that has a commitment on ascribing a supposition class to a tweet. The cycle of feature extraction is to eradicate the exact quality to get better the accurateness of the classifications models. In this manuscript we proposed Term Frequency-Inverse Document Frequency (TF-IDF) method is to secure Principal Component Analysis (PCA) with Naïve Bayes Classifiers. As the classifications process, the work proposed can produce different aspects from wildly valued feature commencing a Twitter dataset.
https://doi.org/10.22937/IJCSNS.2024.24.4.12 인용 PDF

A Study of Correlation Analysis between Increase / Decrease Rate of Tweets Before and After Opening and a Box Office Gross (개봉 전후 트윗 개수의 증감률과 영화 매출간의 상관관계)

Park, Ji-Yun;Yoo, In-Hyeok;Kang, Sung-Woo
- Journal of the Korea Safety Management & Science
- /
- v.19 no.4
- /
- pp.169-182
- /
- 2017
Predicting a box office gross in the film industry is an important goal. Many works have analyzed the elements of a film making. Previous studies have suggested several methods for predicting box office such as a model for distinguishing people's reactions by using a sentiment analysis, a study on the period of influence of word-of-mouth effect through SNS. These works discover that a word of mouth (WOM) effect through SNS influences customers' choice of movies. Therefore, this study analyzes correlations between a box office gross and a ratio of people reaction to a certain movie by extracting their feedback on the film from before and after of the film opening. In this work, people's reactions to the movie are categorized into positive, neutral, and negative opinions by employing sentiment analysis. In order to proceed the research analyses in this work, North American tweets are collected between March 2011 and August 2012. There is no correlation for each analysis that has been conducted in this work, hereby rate of tweets before and after opening of movies does not have relationship between a box office gross.
https://doi.org/10.12812/ksms.2017.19.4.169 인용 PDF KSCI

Method for Spatial Sentiment Lexicon Construction using Korean Place Reviews (한국어 장소 리뷰를 이용한 공간 감성어 사전 구축 방법)

Lee, Young Min;Kwon, Pil;Yu, Ki Yun;Kim, Ji Young
- Journal of Korean Society for Geospatial Information Science
- /
- v.25 no.2
- /
- pp.3-12
- /
- 2017
Leaving positive or negative comments of places where he or she visits on location-based services is being common in daily life. The sentiment analysis of place reviews written by actual visitors can provide valuable information to potential consumers, as well as business owners. To conduct sentiment analysis of a place, a spatial sentiment lexicon that can be used as a criterion is required; yet, lexicon of spatial sentiment words has not been constructed. Therefore, this study suggested a method to construct a spatial sentiment lexicon by analyzing the place review data written by Korean internet users. Among several location categories, theme parks were chosen for this study. For this purpose, natural language processing technique and statistical techniques are used. Spatial sentiment words included the lexicon have information about sentiment polarity and probability score. The spatial sentiment lexicon constructed in this study consists of 3 tables(SSLex_SS, SSLex_single, SSLex_combi) that include 219 spatial sentiment words. Throughout this study, the sentiment analysis has conducted based on the texts written about the theme parks created on Twitter. As the accuracy of the sentiment classification was calculated as 0.714, the validity of the lexicon was verified.
https://doi.org/10.7319/kogsis.2017.25.2.003 인용 KSCI

Measuring Similarity Between Movies Based on Sentiment of Tweets (트위터를 활용한 감성 기반의 영화 유사도 측정)

Kim, Kyoungmin;Kim, Dong-Yun;Lee, Jee-Hyong
- Journal of the Korean Institute of Intelligent Systems
- /
- v.24 no.3
- /
- pp.292-297
- /
- 2014
As a Social Network Service (SNS) has become an integral part of our everyday lives, millions of users can express their opinion and share information regardless of time and place. Hence sentiment analysis using micro-blogs has been studied in various field to know people's opinion on particular topics. Most of previous researches on movie reviews consider only positive and negative sentiment and use it to predict movie rating. As people feel not only positive and negative but also various emotion, the sentiment that people feel while watching a movie need to be classified in more detail to extract more information than personal preference. We measure sentiment distributions of each movie from tweets according to the Thayer's model. Then, we find similar movies by calculating similarity between each sentiment distributions. Through the experiments, we verify that our method using micro-blogs performs better than using only genre information of movies.
https://doi.org/10.5391/JKIIS.2014.24.3.292 인용 PDF KSCI

Exploring Opinions on COVID-19 Vaccines through Analyzing Twitter Posts (트위터 게시물 분석을 통한 코로나바이러스감염증-19 백신에 대한 의견 탐색)

Jung, Woojin;Kim, Kyuli;Yoo, Seunghee;Zhu, Yongjun
- Journal of the Korean Society for information Management
- /
- v.38 no.4
- /
- pp.113-128
- /
- 2021
In this study, we aimed to understand the public opinion on COVID-19 vaccine. To achieve the goal, we analyzed COVID-19 vaccine-related Twitter posts. 45,413 tweets posted from March 16, 2020 to March 15, 2021 including COVID-19 vaccine names as keywords were collected. The 12 vaccine names used for data collection included 'Pfizer', 'AstraZeneca', 'Modena', 'Jansen', 'NovaVax', 'Sinopharm', 'SinoVac', 'Sputnik V', 'Bharat', 'KhanSino', 'Chumakov', and 'VECTOR' in the order of the number of collected posts. The collected posts were analyzed manually and automatedly through keyword analysis, sentiment analysis, and topic modeling to understand the opinions for the investigated vaccines. According to the results, there were generally more negative posts about vaccines than positive posts. Anxiety about the aftereffects of vaccination and distrust in the efficacy of vaccines were identified as major negative factors for vaccines. On the contrary, the anticipation for the suppression of the spread of coronavirus following vaccination was identified as a positive social factor for vaccines. Different from previous studies that investigated opinions about COVID-19 vaccines through mass media data such as news articles, this study explores opinions of social media users using keyword analysis, sentiment analysis, and topic modeling. In addition, the results of this study can be used by governmental institutions for making policies to promote vaccination reflecting the social atmosphere.
https://doi.org/10.3743/KOSIM.2021.38.4.113 인용 PDF KSCI

Public Opinion on Lockdown (PSBB) Policy in Overcoming COVID-19 Pandemic in Indonesia: Analysis Based on Big Data Twitter

Suratnoaji, Catur;Nurhadi, Nurhadi;Arianto, Irwan Dwi
- Asian Journal for Public Opinion Research
- /
- v.8 no.3
- /
- pp.393-406
- /
- 2020
The discourse on the lockdown in Indonesia is getting stronger due to the increasing number of positive cases of the coronavirus and the death rate. As of August 12, 2020, the confirmed number of COVID-19 cases in Indonesia reached 130,718. There were 85,798 victims who have recovered and 5,903 who have died. Data show a significant increase in cases of COVID-19 every day. For this reason, there needs to be an evaluation of the government policy of the Republic of Indonesia in dealing with the COVID-19 pandemic in Indonesia. An evaluation of policies for handling the pandemic must include public opinion to determine any weaknesses of this policy. The development of public opinion about the lockdown policy can be understood through social media. During the COVID-19 pandemic, measuring public opinion through traditional methods (surveys) was difficult. For this reason, we utilized big data on social media as research data. The main purpose of this study is to understand public opinion on the lockdown policy in overcoming the COVID-19 pandemic in Indonesia. The things observed included: volume of Twitter users, top influencers, top tweets, and communication networks between Twitter users. For the methodological development of future public opinion research, the researchers outline the obstacles faced in researching public opinion based on big data from Twitter. The research results show that the lockdown policy is an interesting issue, as evidenced by the number of active users (79,502) forming 133,209 networks. Posts about the lockdown on Twitter continued to increase after the implementation of the lockdown policy on April 10, 2020. The lockdown policy has caused various reactions, seen from the word analysis showing 14.8% positive sentiment, 17.5% negative, and 67.67% non-categorized words. Sources of information who have played the roles of top influencers regarding the lockdown policy include: Jokowi (the president of the Republic of Indonesia), online media, television media, government departments, and governors. Based on the analysis of the network structure, it shows that Jokowi has a central role in controlling the lockdown policy. Several challenges were found in this study: 1) choosing keywords for downloading data, 2) categorizing words containing public opinion sentiment, and 3) determining the sample size.
https://doi.org/10.15206/ajpor.2020.8.3.393 인용 PDF KSCI

Relations between Reputation and Social Media Marketing Communication in Cryptocurrency Markets: Visual Analytics using Tableau

Park, Sejung;Park, Han Woo
- International Journal of Contents
- /
- v.17 no.1
- /
- pp.1-10
- /
- 2021
Visual analytics is an emerging research field that combines the strength of electronic data processing and human intuition-based social background knowledge. This study demonstrates useful visual analytics with Tableau in conjunction with semantic network analysis using examples of sentiment flow and strategic communication strategies via Twitter in a blockchain domain. We comparatively investigated the sentiment flow over time and language usage patterns between companies with a good reputation and firms with a poor reputation. In addition, this study explored the relations between reputation and marketing communication strategies. We found that cryptocurrency firms more actively produced information when there was an increased public demand and increased transactions and when the coins' prices were high. Emotional language strategies on social media did not affect cryptocurrencies' reputations. The pattern in semantic representations of keywords was similar between companies with a good reputation and firms with a poor reputation. However, the reputable firms communicated on a wide range of topics and used more culturally focused strategies, and took more advantages of social media marketing by expanding their outreach to other social media networks. The visual big data analytics provides insights into business intelligence that helps informed policies.
https://doi.org/10.5392/IJoC.2021.17.1.001 인용 PDF KSCI HTML

Analysis of YouTube's role as a new platform between media and consumers

Hur, Tai-Sung;Im, Jung-ju;Song, Da-hye
- Journal of the Korea Society of Computer and Information
- /
- v.27 no.2
- /
- pp.53-60
- /
- 2022
YouTube realistically shows fake news and biased content based on facts that have not been verified due to low entry barriers and ambiguity in video regulation standards. Therefore, this study aims to analyze the influence of the media and YouTube on individual behavior and their relationship. Data from YouTube and Twitter are randomly imported with selenium, beautiful soup, and Twitter APIs to classify the 31 most frequently mentioned keywords. Based on 31 keywords classified, data were collected from YouTube, Twitter, and Naver News, and positive, negative, and neutral emotions were classified and quantified with NLTK's Natural Language Toolkit (NLTK) Vader model and used as analysis data. As a result of analyzing the correlation of data, it was confirmed that the higher the negative value of news, the more positive content on YouTube, and the positive index of YouTube content is proportional to the positive and negative values on Twitter. As a result of this study, YouTube is not consistent with the emotion index shown in the news due to its secondary processing and affected characteristics. In other words, processed YouTube content intuitively affects Twitter's positive and negative figures, which are channels of communication. The results of this study analyzed that YouTube plays a role in assisting individual discrimination in the current situation where accurate judgment of information has become difficult due to the emergence of yellow media that stimulates people's interests and instincts.
https://doi.org/10.9708/jksci.2022.27.02.053 인용 PDF KSCI HTML

Word Clustering Scheme for Twitter Sentiment Analysis Based on POS (트위터 감정 분석을 위한 POS 기반의 단어 군집화 기법)

Kim, Se-Jun;Lim, Hwan-Hee;Lee, Byung-Jun;Kim, Kyung-Tae;Youn, Hee-Yong
- Proceedings of the Korean Society of Computer Information Conference
- /
- 2019.01a
- /
- pp.31-32
- /
- 2019
본 논문에서는 최근 빅데이터 활용 분야의 큰 이슈인 트위터 메시지의 효율적인 감정 분석을 위한 POS 기반의 단어 군집화 기법을 제안하였다. 기존에 군집화를 통한 다양한 텍스트 감정 분석 기법이 제시되어 왔으나, 군집화 된 기능과 분류 결과 간의 관련성에 대한 연구는 미흡하였다. 또한 모든 단어에 대한 감정 분석은 노이즈로 작용될 수 있는 단어로 인해 정확도가 감소할 수 있다. 본 논문에서는 이를 해결하기 위하여 Chi Square 기법을 통하여 분석 결과에 영향을 미치는 단어에 가중치를 부여함으로써 정확도를 향상시킨다.
PDF

The Brand Personality Effect: Communicating Brand Personality on Twitter and its Influence on Online Community Engagement (브랜드 개성 효과: 트위터 상의 브랜드 개성 전달이 온라인 커뮤니티 참여에 미치는 영향)

Cruz, Ruth Angelie B.;Lee, Hong Joo
- Journal of Intelligence and Information Systems
- /
- v.20 no.1
- /
- pp.67-101
- /
- 2014
The use of new technology greatly shapes the marketing strategies used by companies to engage their consumers. Among these new technologies, social media is used to reach out to the organization's audience online. One of the most popular social media channels to date is the microblogging platform Twitter. With 500 million tweets sent on average daily, the microblogging platform is definitely a rich source of data for researchers, and a lucrative marketing medium for companies. Nonetheless, one of the challenges for companies in developing an effective Twitter campaign is the limited theoretical and empirical evidence on the proper organizational usage of Twitter despite its potential advantages for a firm's external communications. The current study aims to provide empirical evidence on how firms can utilize Twitter effectively in their marketing communications using the association between brand personality and brand engagement that several branding researchers propose. The study extends Aaker's previous empirical work on brand personality by applying the Brand Personality Scale to explore whether Twitter brand communities convey distinctive brand personalities online and its influence on the communities' level or intensity of consumer engagement and sentiment quality. Moreover, the moderating effect of the product involvement construct in consumer engagement is also measured. By collecting data for a period of eight weeks using the publicly available Twitter application programming interface (API) from 23 accounts of Twitter-verified business-to-consumer (B2C) brands, we analyze the validity of the paper's hypothesis by using computerized content analysis and opinion mining. The study is the first to compare Twitter marketing across organizations using the brand personality concept. It demonstrates a potential basis for Twitter strategies and discusses the benefits of these strategies, thus providing a framework of analysis for Twitter practice and strategic direction for companies developing their use of Twitter to communicate with their followers on this social media platform. This study has four specific research objectives. The first objective is to examine the applicability of brand personality dimensions used in marketing research to online brand communities on Twitter. The second is to establish a connection between the congruence of offline and online brand personalities in building a successful social media brand community. Third, we test the moderating effect of product involvement in the effect of brand personality on brand community engagement. Lastly, we investigate the sentiment quality of consumer messages to the firms that succeed in communicating their brands' personalities on Twitter.
https://doi.org/10.13088/jiis.2014.20.1.067 인용 PDF KSCI

Search Result 93, Processing Time 0.02 seconds

이메일무단수집거부

이용약관

제 1 장 총칙

제 2 장 이용계약의 체결

제 3 장 계약 당사자의 의무

제 4 장 서비스의 이용

제 5 장 계약 해지 및 이용 제한

제 6 장 손해배상 및 기타사항

Detail Search

Image Search (β)