{"doi":"10.1109/aide57180.2022.10060044","title":"An N-gram-Based BERT model for Sentiment Classification Using Movie Reviews","abstract":null,"journal":"2022 International Conference on Artificial Intelligence and Data Engineering (AIDE)","year":2022,"id":597697,"datarank":0.3453877639491069,"base_score":2.302585092994046,"endowment":2.302585092994046,"self_citation_contribution":0.3453877639491069,"citation_network_contribution":0.0,"self_endowment_contribution":0.3453877639491069,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":9,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":1531275,"name":"Ashok Kumar Jayaraman","orcid":null,"position":1,"is_corresponding":false},{"id":1029781,"name":"Erik Cambria","orcid":"0000-0002-3030-1280","position":2,"is_corresponding":false},{"id":1531276,"name":"Gayathri Ananthakrishnan","orcid":null,"position":3,"is_corresponding":false},{"id":1531277,"name":"Satanik Mitra","orcid":null,"position":4,"is_corresponding":false},{"id":1531274,"name":"Tina Esther Trueman","orcid":null,"position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"An N-gram-Based BERT model for Sentiment Classification Using Movie Reviews","abstract":"An abundance of product reviews and opinions is being produced every day across the internet and other media. Sentiment analysis analyzes those data and classifies them as positive or negative. In this paper, a classification model is proposed for n-gram sentiment analysis using BERT. Specifically, the large IMDB movie review dataset is used that contains 50K instances. This dataset is tokenized and encoded into unigrams, bigrams, and trigrams and their combinations such as unigram and bigram, bigram and trigram, and unigram, bigram, and trigram. The proposed BERT model employs on these extracted features. Then, this model is evaluated using the F1 score and its micro, macro, and weighted-average scores. The model shows comparable results to state-of-the-art methods for all n-gram features. In particular, the model achieves 94.64% highest accuracy for the combination of bigram and trigram features, and 94.68% unigram, bigram, and trigram features than other n-gram features.","is_dataset_classified":null,"base_score":2.302585092994046,"endowment":2.302585092994046,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"23304386","pmcid":null,"openalex_id":"https://openalex.org/W4327499987","authors":[],"funders":[],"total_grants":0,"fwci":null,"citation_percentile":null,"influential_citations":0,"citation_trend":[{"year":2023,"count":1},{"year":2024,"count":3},{"year":2025,"count":4},{"year":2026,"count":1}],"oa_status":"closed","license":"https://doi.org/10.15223/policy-029","oa_locations":[{"url":"http://xplorestaging.ieee.org/ielx7/10059460/10059641/10060044.pdf?arnumber=10060044","host_type":"publisher"},{"url":"https://doi.org/10.1109/aide57180.2022.10060044","host_type":""}],"fields_of_study":["Sentiment Analysis and Opinion Mining","Topic Modeling","Stock Market Forecasting Methods"],"mesh_terms":[],"keywords":["Bigram","Trigram","n-gram","Computer science","Artificial intelligence","Speech recognition","Language model","Natural language processing","Pattern recognition (psychology)"],"sdg_mappings":[],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-07-28T13:58:03.575493Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}