CORE (COnnecting REpositories) is an open scholarly infrastructure that indexes millions of open access research papers and metadata from repositories and journals worldwide. Its goal is to improve the discoverability and reuse of research outputs and support machine access to scholarly content in line with open access and open science principles. This talk will provide an overview of CORE and its services for repositories, including compliance monitoring, metadata validation, and tools to improve interoperability and discoverability. It will also present research by the Big Scientific Data and Text Analytics Group (BSDTAG), showcasing recent innovations such as CORE-GPT, a system for trustworthy question answering over scholarly literature; SDG: Classify, which maps research papers to UN Sustainable Development Goals; and SoFAIR, which addresses reproducibility and research software management. Finally, the talk will discuss how CORE enables external research and innovation in areas such as training large language models, plagiarism detection, library discovery, and the construction of scholarly graphs, fostering a globally connected and machine-readable open research ecosystem
The session will also briefly cover:
- Persistent identifiers: how CORE’s OAI‑based resolver can turn existing repository identifiers into cost‑free, globally resolvable PIDs that complement DOIs and keep repositories central in dissemination and assessment workflows.
- Bots and machine access: how CORE’s own harvesters interact responsibly with repositories, how they distinguish “bots for good” from abusive scrapers, and practical advice for repository managers on configuring access so beneficial services can work effectively without exposing content to misuse.
The webinar is designed as an accessible introduction for OAA members, with plenty of time for questions and discussion.