{"doi":"10.1504/ijiids.2019.10026240","title":"Improving named entity recognition and disambiguation in news headlines","abstract":null,"journal":"International Journal of Intelligent Information and Database Systems","year":2019,"id":631304,"datarank":0.0,"base_score":0.0,"endowment":0.0,"self_citation_contribution":0.0,"citation_network_contribution":0.0,"self_endowment_contribution":0.0,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":0,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":null,"is_data_producer":false,"deposit_databanks":null,"is_oa":false,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":null,"fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":1635982,"name":"Rajdeep Niyogi","orcid":null,"position":1,"is_corresponding":false},{"id":1635981,"name":"Jayendra Barua","orcid":null,"position":0,"is_corresponding":false}],"reference_count":0,"raw_metadata":{"has_enrichment":true,"resolved":true,"title":"Improving named entity recognition and disambiguation in news headlines","abstract":"In this paper, we present a framework for extraction and disambiguation of hyphenated and partially named entities in news headlines. The direct application of state-of-the-art named entity detection and disambiguation approaches on news headlines results in significantly degraded performance due to different headline formatting in comparison with regular text; hyphenated mentions; and partial entity mentions. In this paper, we introduce a novel framework that assists existing named entity recognition and disambiguation systems to deal with introduced challenges. In particular, we deal with hyphenated entity mentions and partial entity mentions present in news headlines. We modify the hyphenated and partial entity in a way that increases the probability of disambiguation to correct entity in knowledge base. Our framework leverages headlines of recent past to improve the entity mentions in headlines. The experimental results showed that presented framework improves the F1-score of mention detection by 12% and 9% in state-of-the-art Stanford and Illinois NER systems, whereas F1-score of disambiguation is improved by 9%, 12%, 7% and 5% in AIDA, Wikifier, TagMe, and YODIE state-of-the-art NED systems respectively.","is_dataset_classified":null,"base_score":0.0,"endowment":0.0,"datacite_reuse_total":0,"file_count":0,"downloads":0,"views":0,"has_version_chain":false,"is_dataset":false,"is_oa":false,"pmid":"19910364","pmcid":null,"openalex_id":"https://openalex.org/W3000685722","authors":[],"funders":[],"total_grants":0,"fwci":0.0,"citation_percentile":0.18765771,"influential_citations":0,"citation_trend":[],"oa_status":"gold","license":null,"oa_locations":[{"url":"https://doi.org/10.1504/ijiids.2019.10026240","host_type":"journal"},{"url":"https://doi.org/10.1504/ijiids.2019.104530","host_type":"GOLD"},{"url":"https://doi.org/10.1504/ijiids.2019.10026240","host_type":"publisher"},{"url":"http://www.inderscienceonline.com/doi/full/10.1504/IJIIDS.2019.10026240","host_type":"publisher"}],"fields_of_study":["Topic Modeling","Natural Language Processing Techniques","Web Data Mining and Analysis","Computer Science"],"mesh_terms":[],"keywords":["Named-entity recognition","Computer science","Entity linking","Headline","Named entity","Information retrieval","Disk formatting","Natural language processing","Knowledge base","Artificial intelligence","Information extraction","Linguistics","Task (project management)"],"sdg_mappings":[{"sdg_number":0,"sdg_label":"Quality Education"}],"linked_datasets":[],"clinical_trials":[],"software_tools":[],"database_accessions":[],"source":"live","citation_network_status":"fetched"},"created_at":"2026-08-05T23:17:18.107628Z","pmid":null,"pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}