Youngseok Choi scite author profile

Previous studies on predicting the box-office performance of a movie using machine learning techniques have shown practical levels of predictive accuracy. Their works are technically-and methodologically-oriented, investigating what algorithms are better at predicting the movie performance. However, the accuracy of prediction model can also be elevated by taking other perspectives. For example, it is possible to increase the model accuracy by introducing unexplored features that might be related to the prediction of the outcomes. In this paper, we examine multiple approaches to improve the performance of the prediction model. First, we add a new feature derived from the theory of transmedia storytelling. Such theory-driven feature selection not only increases the forecast accuracy, but also enhances the interpretability of a prediction model. Second, we use an ensemble approach, which has rarely been adopted in the research on predicting box-office performance. As a result, the proposed model, Cinema Ensemble Model (CEM), outperforms the prediction models from the past studies using machine learning algorithms. We suggest that CEM can be extensively used for industrial experts as a powerful tool for improving decision-making process.2

show abstract

The role of e-participation and open data in evidence-based policy decision making in local government

Sivarajah

Weerakkody

Waller

et al. 2015

Journal of Organizational Computing and Electronic Commerce

View full text Add to dashboard Cite

Gas-generating polymeric microspheres for long-term and continuous in vivo ultrasound imaging

Min¹,

Kang²,

Koo³

et al. 2012

Biomaterials

View full text Add to dashboard Cite

Data properties and the performance of sentiment classification for electronic commerce applications

Choi

Lee

2017

Inf Syst Front

View full text Add to dashboard Cite

Sentiment classification has played an important role in various research area including e-commerce applications and a number of advanced Computational Intelligence techniques including machine learning and computational linguistics have been proposed in the literature for improved sentiment classification results. While such studies focus on improving performance with new techniques or extending existing algorithms based on previously used dataset, few studies provide practitioners with insight on what techniques are better for their datasets that have different properties. This paper applies four different sentiment classification techniques from machine learning (Naïve Bayes, SVM and Decision Tree) and sentiment orientation approaches to datasets obtained from various sources (IMDB, Twitter, Hotel review, and Amazon review datasets) to learn how different data properties including dataset size, length of target documents, and subjectivity of data affect the performance of those techniques. The results of computational experiments confirm the sensitivity of the techniques on data properties including training data size, the document length and subjectivity of training /test data in the improvement of performances of techniques. The theoretical and practical implications of the findings are discussed.

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Youngseok Choi

Helpfulness of Online Consumer Reviews: Readers' Objectives and Review Cues

A decision support system for vessel speed decision in maritime logistics using weather archive big data

Genomics-based screening of differentially expressed genes in the brains of mice exposed to silver nanoparticles via inhalation

Activation of AMPK by berberine induces hepatic lipid accumulation by upregulation of fatty acid translocase CD36 in mice

Predicting movie success with machine learning techniques: ways to improve accuracy

The role of e-participation and open data in evidence-based policy decision making in local government

Gas-generating polymeric microspheres for long-term and continuous in vivo ultrasound imaging

Data properties and the performance of sentiment classification for electronic commerce applications

Contact Info

Product

Resources

About