1Department of Computer Science, Universitas Amikom Yogyakarta, Yogyakarta, Indonesia
2Department of Computer Science, Universidad de Murcia, Spain
BibTex Citation Data :
@article{JMASIF82051, author = {Heri Santosa and Kusrini Kusrini and Rodrígo Martínez-Béjar}, title = {Comparative Evaluation of FinBERT and IndoBERT for LSTM-Based Stock Price Prediction in Indonesia}, journal = {Jurnal Masyarakat Informatika}, volume = {17}, number = {2}, year = {2026}, keywords = {Stock price prediction, sentiment analysis, LSTM, FinBERT, IndoBERT}, abstract = { Stock price prediction remains challenging due to the nonlinear and dynamic behavior of financial markets, where price movements are influenced not only by historical market data but also by investor sentiment reflected in financial news. This study investigates the effectiveness of sentiment integration for stock price prediction by comparing FinBERT, a financial domain-specific language model, and IndoBERT, a language-specific Indonesian language model, within an LSTM-based forecasting framework. Experiments were conducted using daily stock price data and financial news collected for three Indonesian stocks (BBCA, BBRI, and BSDE) over the period 2019–2025. Sentiment scores extracted from financial news were aggregated on a daily basis and incorporated as additional features alongside historical price variables. Model performance was evaluated using MAE, RMSE, MAPE, R², and Directional Accuracy. The results indicate that sentiment-enhanced models do not consistently improve numerical forecasting accuracy compared with the baseline LSTM model. However, sentiment integration does not consistently improve directional prediction, although limited stock-specific differences are observed under certain stock-specific conditions. Comparative analysis further shows that neither FinBERT nor IndoBERT consistently outperforms the other across all datasets and metrics, suggesting that sentiment effectiveness is highly dependent on linguistic context and market characteristics. These findings highlight that sentiment information should be incorporated selectively rather than assumed to universally improve stock forecasting performance in emerging markets. }, issn = {2777-0648}, pages = {198--217} doi = {10.14710/jmasif.17.2.82051}, url = {https://ejournal.undip.ac.id/index.php/jmasif/article/view/82051} }
Refworks Citation Data :
Stock price prediction remains challenging due to the nonlinear and dynamic behavior of financial markets, where price movements are influenced not only by historical market data but also by investor sentiment reflected in financial news. This study investigates the effectiveness of sentiment integration for stock price prediction by comparing FinBERT, a financial domain-specific language model, and IndoBERT, a language-specific Indonesian language model, within an LSTM-based forecasting framework. Experiments were conducted using daily stock price data and financial news collected for three Indonesian stocks (BBCA, BBRI, and BSDE) over the period 2019–2025. Sentiment scores extracted from financial news were aggregated on a daily basis and incorporated as additional features alongside historical price variables. Model performance was evaluated using MAE, RMSE, MAPE, R², and Directional Accuracy. The results indicate that sentiment-enhanced models do not consistently improve numerical forecasting accuracy compared with the baseline LSTM model. However, sentiment integration does not consistently improve directional prediction, although limited stock-specific differences are observed under certain stock-specific conditions. Comparative analysis further shows that neither FinBERT nor IndoBERT consistently outperforms the other across all datasets and metrics, suggesting that sentiment effectiveness is highly dependent on linguistic context and market characteristics. These findings highlight that sentiment information should be incorporated selectively rather than assumed to universally improve stock forecasting performance in emerging markets.
Article Metrics:
Last update:
Last update: 2026-08-12 03:16:46
The authors who submit the manuscript must understand that the article's copyright belongs to the author(s) if accepted for publication. However, the author(s) grant the journal right of first publication with the work simultaneously licensed under a Creative Commons Attribution-ShareAlike 4.0 International License. Authors should also understand that their article (and any additional files, including data sets, and analysis/computation data) will become publicly available once published under that license. By submitting the manuscript to Jmasif, the author(s) agree with this policy. No special document approval is required.
The author(s) guarantee that:
The author(s) retain all rights to the published work, such as (but not limited to) the following rights:
Suppose the article was prepared jointly by more than one author. Each author submitting the manuscript warrants that all co-authors have given their permission to agree to copyright and license notices (agreements) on their behalf and notify co-authors of the terms of this policy. Jmasif will not be held responsible for anything arising because of the writer's internal dispute. Jmasif will only communicate with correspondence authors.
Authors should also understand that their articles (and any additional files, including data sets and analysis/computation data) will become publicly available once published. The license of published articles (and additional data) will be governed by a Creative Commons Attribution-ShareAlike 4.0 International License. Jmasif allows users to copy, distribute, display and perform work under license. Users need to attribute the author(s) and Jmasif to distribute works in journals and other publication media. Unless otherwise stated, the author(s) is a public entity as soon as the article is published.