Advancing FAIR data towards comparable, organized, predictive AI-ready data for community validation.

Wood-Charlson EM, Åkerström WN, Anderson L, Borton MA, Burley SK, Byers N, Chandramouliswaran I, Costes SV, Dehal PS, Doktycz M, Eloe-Fadrosh E, Henry C, Fagnan K, Fiehn O, Halappanavar M, Hurwitz B, Joachimiak MP, Jungbluth S, Koblitz J, Mahmud G, McCue LA, Metz TO, Mouncey N, Mungall CJ, Nelson TM, Skye V, Saravia-Butler AM, Saripalli VR, Tringe SG, Van Den Bossche T, Arkin AP

Commun Biol 9 (1) - [2026-07-18; online 2026-07-18]

The interrogation of data across biological and environmental systems has become increasingly complex. Fortunately, communities are adopting the FAIR (Findable, Accessible, Interoperable, Reusable) data principles for individual datasets, and continue to develop domain-specific, machine-actionable standards. However, integrating FAIR data for meta-analysis across data resources is still challenging. Understanding how disparate datasets are organized remains a manual, time-consuming process. Updating FAIR databases to reflect changes in knowledge is slow, allowing stale annotations and incorrect relationships to propagate, amplified by Artificial Intelligence (AI) systems that harvest data. Building on FAIR, we argue that data should be iteratively updated and improved. FAIR + COPE (Comparable, Organized, Predictive, Engaged) takes FAIR data and makes it Comparable, rapidly Organized (applying / updating standards) for Predictive models, which can be validated and improved by an Engaged community. We provide examples of FAIR + COPE resources and science use cases that highlight the importance of FAIR + COPE in scientific research.

Bioinformatics (NBIS) [Collaborative]

Bioinformatics Support and Infrastructure [Collaborative]

Bioinformatics Support, Infrastructure and Training [Collaborative]

PubMed 42471458

DOI 10.1038/s42003-026-10694-y

Crossref 10.1038/s42003-026-10694-y

pmc: PMC13380610
pii: 10.1038/s42003-026-10694-y


Publications 9.5.1