{"doi":"10.1109/tsp.2024.3517323","title":"Large-Scale Independent Vector Analysis (IVA-G) via Coresets","abstract":"Joint blind source separation (JBSS) involves the factorization of multiple matrices, i.e. “datasets”, into “sources” that are statistically dependent across datasets and independent within datasets. Despite this usefulness for analyzing multiple datasets, JBSS methods suffer from considerable computational costs and are typically intractable for hundreds or thousands of datasets. To address this issue, we present a methodology for how a subset of the datasets can be used to perform efficient JBSS over the full set. We motivate two such methods: a numerical extension of independent vector analysis (IVA) with the multivariate Gaussian model (IVA-G), and a recently proposed analytic method resembling generalized joint diagonalization (GJD). We derive nonidentifiability conditions for both methods, and then demonstrate how one can significantly improve these methods’ generalizability by an efficient representative subset selection method. This involves selecting a <italic xmlns:mml=\"http://www.w3.org/1998/Math/MathML\" xmlns:xlink=\"http://www.w3.org/1999/xlink\">coreset</i> (a weighted subset) that minimizes a measure of discrepancy between the statistics of the coreset and the full set. Using simulated and real functional magnetic resonance imaging (fMRI) data, we demonstrate significant scalability and source separation advantages of our “coreIVA-G” method vs. other JBSS methods.","journal":"IEEE Transactions on Signal Processing","year":2024,"id":466874,"datarank":0.24141568686511508,"base_score":1.6094379124341003,"endowment":1.6094379124341003,"self_citation_contribution":0.24141568686511508,"citation_network_contribution":0.0,"self_endowment_contribution":0.24141568686511508,"citer_contribution":0.0,"corpus_percentile":null,"corpus_rank":null,"citation_count":4,"citer_count":0,"citers_with_citation_signal":0,"citers_with_endowment":0,"datacite_reuse_total":0,"is_dataset":false,"is_dataset_confidence":0.9243,"is_data_producer":false,"deposit_databanks":null,"is_oa":true,"file_count":0,"downloads":0,"has_version_chain":false,"published_date":"2024-01-01","fair_score":null,"fair_percentile":null,"algorithm_id":"datarank_citation_only_1hop_v6","ranking_scope":"data_only","authors":[{"id":1106825,"name":"Hanlu Yang","orcid":"0000-0001-7903-6257","position":1,"is_corresponding":false},{"id":366126,"name":"Trung Vu","orcid":"0000-0003-2180-5994","position":2,"is_corresponding":false},{"id":227761,"name":"Vince D. Calhoun","orcid":"0000-0001-9058-0747","position":3,"is_corresponding":false},{"id":428411,"name":"Tülay Adalı","orcid":"0000-0003-0594-2796","position":4,"is_corresponding":false},{"id":731461,"name":"Ben Gabrielson","orcid":"0000-0001-9217-6641","position":0,"is_corresponding":true}],"reference_count":36,"raw_metadata":{"citation_network_status":"fetched"},"created_at":"2026-07-19T02:05:06.743102Z","pmid":"40994814","pmcid":null,"fwci":null,"citation_percentile":null,"influential_citations":0,"oa_status":null,"license":null,"views":0,"total_file_size_bytes":0,"version_count":0,"fair_f":null,"fair_a":null,"fair_i":null,"fair_r":null,"fair_zscore":null,"fair_rationale":null,"fair_model":null,"fair_agent_version":null,"fair_fulltext_source":null,"fair_has_llm":null,"fair_computed_at":null,"clinical_trials":[],"software_tools":[],"db_accessions":[],"linked_datasets":[],"topics":[]}