Skip to content

Latest commit

 

History

History
308 lines (207 loc) · 35.5 KB

File metadata and controls

308 lines (207 loc) · 35.5 KB

Ricardo Hartley Belmar

MSc, MSc, Dr.

Open science & research information governance / PID traceability / Responsible research assessment

Open Science Expert at the Interdisciplinary Transformation University Austria (IT:U) in Linz, where I design a university-wide Open Science strategy and build the systems that carry it. The strategy rests on three pillars, concept and operation, engagement, and training and support, with a transversal axis on the use of AI in research, run over three horizons on the SCOPE framework (INORMS Research Evaluation Group, 2021) and with indicators that follow the responsible-use principles of DORA, the Leiden Manifesto and CoARA. Underneath it I build the working parts: a research information system whose provenance ordering is enforced at every write path rather than merely documented, a research management system covering the grant lifecycle, machine-actionable data management planning, reproducible computational environments, and the container platform they all run on.

Previously I completed a Short-Term Scientific Mission at the Open Innovation in Science Center, Ludwig Boltzmann Gesellschaft (Vienna), building a baseline of institutional identifiers and reproducible PID-based traceability workflows aligned with Austria's ERA-NAP 2026–2028 priorities.

My work centers on research information governance, metadata quality, and responsible assessment, bridging technological, practical, and regulatory interfaces so institutions can operationalize ambitious open science commitments.

ricdho@duck.com (alias) · ORCID 0000-0001-5058-9309 · Substack

Experience

Open Science Expert · Interdisciplinary Transformation University Austria (IT:U)

May 2026 – Present · Linz, Austria

Design, implement and maintain a university-wide Open Science strategy: three pillars (concept and operation, engagement, training and support) plus a transversal axis on AI in research, run over three horizons on the SCOPE framework (INORMS Research Evaluation Group, 2021), with more than fifteen indicators carrying operational definitions, baselines, targets, deadlines and named owners in a single source table, governed by the responsible-use principles of DORA, the Leiden Manifesto and CoARA. Build and operate the systems underneath it: a current research information system whose provenance ordering is enforced at every write path rather than merely documented, a research management system covering the grant lifecycle from call to post-award budget, a daily aggregator of open and forthcoming funding calls, a Renku 2.0 deployment for reproducible environments, machine-actionable data management planning validated against the RDA DMP Common Standard, and bilingual documentation generators for the team's APIs and automation pipelines. Support researchers through workshops, data management plan clinics and a champions network. Represent IT:U in national and international Open Science initiatives, and contribute to grant writing, ideation and mentorship across academic projects.

Short-Term Scientific Mission (STSM) Researcher · Open Innovation in Science Center, Ludwig Boltzmann Gesellschaft

August 4 – October 28, 2025 · Vienna, Austria

Developed a comprehensive baseline of institutional identifiers, audited coverage of scholarly records across systems (ORCID, ROR, Crossref, DataCite, OpenAlex), and designed reproducible workflows for PID-based traceability of research outputs. Prepared integration guidelines to expand ORCID adoption among LBG researchers and compared information consistency between open infrastructures and institutional systems. The mission supported LBG's open science strategy and alignment with Austria's National Action Plan for the European Research Area (ERA-NAP) 2026–2028.

Strengthened my fluency with Microsoft-based ecosystems - including SharePoint, Power BI, and Azure Active Directory - to interface PID data with internal reporting suites, and piloted automation orchestrations using n8n and complementary integration platforms to streamline identifier governance tasks across teams.

Independent Researcher · Remolino Consulting

May – October 2025

Conducting research and consultancy on research infrastructures, metadata systems, and responsible research assessment.

Researcher · Universidad Central de Chile

2016–2025

Conducted research and institutional projects focused on research governance, scholarly communication, and Open Science policy development.

Researcher and Director · InES Open Science Project, Universidad Central de Chile

2022–2024

Led institutional policy development in Open Science, designed data governance models, and implemented research output traceability systems. The project, funded by Chile's national agency ANID, received the highest score in its 2021 cohort and resulted in the university's first Open Science policy.

Associated Researcher · Fundación Data Observatory, Chile

2024

Contributed to national FAIR data strategies and supported institutional adoption of open science policies.

Director, Institute for Research and Postgraduate Studies · Universidad Central de Chile

2021–2022

Oversaw research development and postgraduate programs in the Faculty of Health Sciences, promoting open science and institutional research capacity-building.

Coordinator, Institute for Research and Postgraduate Studies · Universidad Central de Chile

2016–2021

Coordinated research management efforts, supported data-driven accreditation processes, and promoted open infrastructure adoption.

Education

  • Dr. in Applied Cellular and Molecular Biology, Universidad de La Frontera, Chile (2017)
  • Master in Biological Sciences, Universidad de Tarapacá, Chile (2013)
  • Master in Sciences, Universidad de La Frontera, Chile (2013)
  • Diploma in Scientific Research and Open Knowledge Generation, Universidad del Desarrollo, Chile (2022)
  • Postgraduate Diploma in Data Science and Engineering, Universidad de Chile (2018)
  • Licentiate in Medical Technology, Universidad de Chile (2006)
  • Professional Title of Medical Technologist, Universidad de Chile (2007)

Skills

  • Research Infrastructure & Metadata Systems: ORCID, ROR, Crossref, DataCite, OpenAlex, Metadata standards (Dublin Core, DataCite Metadata Schema, Schema.org)
  • Data Analysis & Visualization: R, Python, JavaScript, D3.js, Observable, API integration and data pipelines
  • Data Management & FAIR Principles: Data Management Plans (DMPs), FAIR data principles implementation, Repository management, Research data governance
  • Languages: Spanish: Native, English: Technical

Publications

Peer-reviewed

Policy papers

Guides

Datasets

Working papers

Talks & presentations

  • Entities and/or Content: Metadata Quality in Open Science Infrastructure · First Digital Humanities Seminar - Chile 2026: Technologies, Memories and Possible Futures, Panel E: Open Science and Reproducibility, Pontificia Universidad Católica de Chile, Campus Lo Contador (April 9, 2026)
  • Tracing Research Outputs: FAIRness, Responsibility, and Reuse Across Infrastructure Layers · Open Science Festival, University of Vienna (September 8, 2025)
  • Research Traceability for Responsible Assessment · FoReA-Treffen Q3 2025 (September 17, 2025)
  • Open Science on Closed Infrastructures? Traceability, Reuse, and Research Assessment · IT:U Interdisciplinary Transformation University, Linz (October 15, 2025)
  • Making Research Count: Challenges in the Evaluation and Valuation of Scientific Work · Open Innovation in Science Center (OIS) (October 16, 2025)
  • Las métricas de datos: Avanzando la gestión responsable de los datos abiertos · Make Data Count & OpenLab Ecuador (March 13, 2025)
  • Más allá de la indexación: ciencia abierta, fuentes abiertas y el valor de los metadatos en la visibilidad académica · Semana Internacional de Acceso Abierto, Pontificia Universidad Católica Argentina (October 24, 2024)
  • Ciencia abierta y el impacto de las publicaciones: Análisis comparativo del sentido y uso de citas en universidades chilenas acreditadas en investigación · VII Seminario Internacional FLACSO Argentina (October 23-24, 2024)
  • Evaluación del impacto de publicaciones científicas y el fomento de prácticas de Ciencia Abierta en universidades para la acreditación en investigación: un análisis más allá de la indexación · Seminario Internacional CNA Chile (October 23-24, 2024)
  • Empowering Global Knowledge through Open Access · IUCN Open Access Week (October 22, 2024)
  • How would having more access to scientific data and research affect scientific credibility and societal impact? · International Conference on Societal Impact of Social Sciences, AESIS Network, Cape Town (October 16-18, 2024)
  • Significando la Ciencia Abierta desde el reuso: ejemplos de desafíos en la calidad de metadatos en productos de investigación · Foro Latinoamericano de Ciencia Abierta, Quito (September 13, 2024)
  • El desafío de la trazabilidad de los productos de investigación: Una aproximación desde la experiencia del proyecto InES Ciencia Abierta de la Universidad Central de Chile · Foro Latinoamericano de Ciencia Abierta, Quito (September 12, 2024)
  • Laboratorio de Datos: Meta ciencia con datos abiertos · Foro Latinoamericano de Ciencia Abierta, Quito (September 11-12, 2024)
  • A Decade of Data Citations: From Principles to Action · FORCE11 Annual Conference, UCLA, Los Angeles (August 1-3, 2024)
  • Governance for AI in Scientific Publications · FSCI Campus Course, UCLA, Los Angeles (July 29-31, 2024)
  • Characterizing Research Outputs and Their Use with OpenAlex and Scite · OpenAlex Virtual User Conference (May 30, 2024)
  • Equity Challenges for National PID Strategies · RDA 22nd Plenary Meeting (May 22, 2024)
  • Characterizing the Use of Research Products from Chilean Universities · 10th Congress of University and Specialized Libraries (April 10, 2024)
  • FAIR Principles and their Implementation in Chile · Data Observatory and UC Chile (January 23, 2024)
  • Importance of Persistent Identifiers · InES Open Science Workshop Series, Universidad Autónoma de Chile (January 12, 2024)
  • Presentation of the Survey Results on Perception and Practices of Open Science in Ecuador · Days Discovering Open Science, OpenLab Ecuador (September 29, 2023)
  • The Value of Digital Repositories Beyond Theses · Days Discovering Open Science, OpenLab Ecuador (August 31, 2023)
  • It is not just counting and graphing: Enriching the Analysis Methodologies of Types of Article Licenses from Universities belonging to the CUP · 1st Open Government Congress, Academic Network Open Government Chile (April 26-28, 2023)
  • Proposal for an Institutional Structure for Open Science at Universidad Central de Chile · 1st Open Government Congress, Academic Network Open Government Chile (April 26-28, 2023)

Writing

Tools & experiments

Active · 2026 · SvelteKit, Hono, Cloudflare Workers, D1 + Vectorize

Web-based tool that verifies academic bibliographies against six scholarly APIs (Crossref, OpenAlex, Open Library, OpenAIRE, Internet Archive, ISBNdb). Parses references in APA, MLA, Chicago, and Vancouver formats, classifies them as verified, partial, not found, or likely fabricated, detects duplicates, and generates downloadable reports. Includes a Microsoft Word add-in for in-document verification and an OAI-PMH endpoint for library harvesting. Built with SvelteKit and Hono, running entirely on Cloudflare (Workers, Pages, and D1 + Vectorize as a semantic cache). Available in English, Spanish and German.

Active · 2026 · Crossref, Metadata Quality, FAIR, React

Dashboard that audits how complete Crossref metadata deposits are for SciELO Chile journals. Evaluates 37 fields per article across 4 FAIR-inspired categories, with interactive visualizations including rain cloud plots, radar charts, heatmaps, and year-based filtering. Built as a hub to add more publisher/registry audits.

Active · 2026 · FAIR, OAI-PMH, DataCite, Custom API

Evaluates metadata quality of digital repositories against the FAIR Principles. Connects to OAI-PMH endpoints, DataCite records, or any custom REST API (CSW, STAC, CKAN, and more), harvests metadata, and runs a structured assessment across 14 sub-principles. Classifies persistent identifiers, detects license types, validates controlled vocabularies (DCMI, ISO 639), and checks community standards compliance. Includes prioritized recommendations, exportable reports (CSV, JSON, TXT), and interactive repository structure visualizations. The FAIR assessment engine runs server-side on Cloudflare Workers.

Active · 2026 · DataCite, Metadata Quality, React, Mapbox

Scans DataCite for DOIs pointing to problematic files: OS artifacts (.DS_Store, Thumbs.db), security risks (.env, credentials), and dev debris (node_modules, pycache) in research repositories. Enriched with geographic data from ROR and cross-referenced against re3data repository metadata. Interactive dashboard with Mapbox world map, category breakdowns, and repository drill-downs.

Active · 2026 · Crossref, OpenAlex, Metadata Quality, FAIR

Do publishers care about metadata? Timeline evidence comparing Crossref deposits vs OpenAlex enrichment across 15 major publishers (2015–2025). Evaluates 47 metadata fields in 5 categories (FAIR + Evaluation/CoARA), with gap analysis radar charts, enrichment matrices, and per-publisher drill-downs showing mandate compliance trends.

Active · 2025 · React, Recharts, Retraction Watch, OpenAlex

Interactive dashboard analyzing citation persistence of 60,921 retracted articles. Reveals that 56% of citations occur post-retraction with no decay over 15 years, geographic disparities in detection speed, and flat supporting/contradicting ratios over time. Built with data from Retraction Watch, OpenAlex, and Scite.ai. Also in https://observablehq.com/d/cecaefb7d49c7727

Active · 2026 · RDA maDMP, Horizon Europe, ANID, FAIR

Analyzes the gap between Horizon Europe data management plans (DMPs) and the RDA maDMP standard. Evaluates 6,119 DMPs against 48 RDA properties, measuring what proportion is mentioned in text (77.5%) versus machine-actionable (22.7%). Includes comparison with Chile's ANID template (1/48 fields machine-actionable). Evidence that mandates generate narrative documents, not connected entities.

Active · 2026 · ORCID, PID Adoption, Metadata Quality, React

Do ORCID consortia drive meaningful adoption, or just empty accounts? Compares profile completeness across 88 institutions and 200,595 profiles from 6 continents. Key finding: consortium institutions average 41.2% completeness with 49% empty profiles, versus 40.4% and 51% for non-consortium, a gap of just +0.7pp. Pre-2018 profiles reach ~50% (intrinsic motivation); post-2022 drop below 30% (mandate-driven bulk creation).

Active · 2026 · MCP, Docker, Node.js, ORCID

MCP Server for real-time research data. Provides live access to authoritative academic sources (ORCID, OpenAlex, Crossref, DataCite, ROR, and PubMed) via a Dockerized Node.js server compatible with Claude Desktop.

Active · 2026 · MCP, Docker, Node.js, OpenAIRE

MCP server exposing the OpenAIRE Graph API (100M+ open access research products) to AI assistants. Provides 14 tools for searching publications, datasets, software, organizations, projects, and persons. Dockerized Node.js server with in-memory cache, rate limiting, and retry logic. Built as a companion to MCP CRIS Live focused on open science infrastructure.

OpenAIRE Research Assistant: Demo

Active · 2026 · Cloudflare Workers AI, LLM, Agentic AI, OpenAIRE

Chat interface where Cloudflare Workers AI (Llama 3.1 8B) uses the OpenAIRE MCP tools in an agentic loop to answer research questions in natural language. The LLM decides which tools to call, executes the OpenAIRE searches, and synthesizes the results into a readable response. Deployed entirely on free Cloudflare infrastructure (Pages + Workers AI).

Active · 2026 · Sonification, Data Viz, R, Accessibility

Sonification project exploring data inequality patterns through audio. Transforms statistical disparities into sound compositions to make data accessible through auditory perception.

PID Traceability Workflows

Active · 2025 · PIDs, Research Infrastructure, Python, API Integration

Development of reproducible workflows for tracking research outputs across persistent identifier systems (ORCID, ROR, Crossref, DataCite, OpenAlex). Built for institutional research information governance.

CRIS for IT:U

Active · 2026 · SvelteKit, PostgreSQL, pgvector, Crossref

Institutional current research information system, built for IT:U and deployable by any institution from a ROR identifier alone. Every field is enriched from the source closest to the primary deposit: the DOI registration agency is authoritative, aggregators may only add what it does not carry, and that ordering is enforced at every write path rather than merely documented, with each output displaying the chain it was built from. A discovered layer and a curated layer coexist without overwriting each other. Publishes OAI-PMH, accepts COAR Notify, reports under DORA, and runs in English, German and Spanish.

RMS for IT:U

Active · 2026 · Hono, SvelteKit, PostgreSQL, Grant lifecycle

Research management system companion to the CRIS: where the CRIS records what was published and by whom, this records how research is funded and administered. Its spine is one grant path with two named gates, institutional eligibility and the funder's decision, so an application rejected in-house is never collapsed with one the funder turned down. Access answers two separate questions, what a role may do and which projects it may see, resolved in one place. The budget is a ledger where planned and committed lines are reconciled by period and category.

OpenCalls (IT:U)

Active · 2026 · GitHub Actions, Node.js, EU Funding & Tenders, FFG

Open and forthcoming research funding calls from five portals, collected once a day, normalised to one schema and published as a single filterable page for the Grant Office. The past-calls registry is a committed file, so the history of every call the page has listed is the git history; a call that leaves its portal is archived only once its deadline has passed, and one that reappears keeps its original first-seen date. A second view compares the richness of what is collected against the RIS Synergy funding schema.

Open Science Strategy (IT:U)

Active · 2026 · Open Science policy, SCOPE, INORMS, DORA

University-wide Open Science strategy: three pillars, concept and operation, engagement, and training and support, plus a transversal axis on AI in research, run over three horizons on the SCOPE framework (INORMS Research Evaluation Group, 2021). More than fifteen indicators carry an operational definition, a baseline, a target, a deadline, a data source and a named owner in one source table that every other document links back to instead of restating. Indicators are transparent, contextualised, read alongside their context and reviewed yearly.

maDMP Template (IT:U)

Active · 2026 · RDA DMP Common Standard, maDMP, Horizon Europe, FAIR

Machine-actionable data management plan template with tiered fields and a Horizon Europe structure: one structured source and a renderer, rather than a prose form that has to be re-read to be reused. The generated skeleton validates against the RDA DMP Common Standard v1.2 schema, and two worked examples are validated with it. It is the instrument behind the strategy's machine-actionable DMP indicator.

Renku 2.0 Deployment (IT:U)

Active · 2026 · Renku, Kubernetes, k3s, Podman

Deployment of Renku 2.0 for reproducible computational environments, in two routes rather than one: k3s with its bundled containerd where the institutional VM permits it, and a kind cluster whose nodes are Podman containers where it does not. Renku ships as a Helm chart and not as an image, so the choice of what provides the single-node cluster is the decision that has to be taken before the machine is provisioned, and it is documented as such alongside the bootstrap scripts.

DART Documentation Generators

Active · 2026 · Static site generator, Python, OpenAPI, n8n

Two generators that turn one source folder per item into a single self-contained bilingual reference page: one for the team's API integrations, built from sanitised OpenAPI specs with an interactive try-it panel per endpoint, and one for its automation pipelines, where each pipeline keeps its workflow files, its own changelog and its catalogue entry in the same folder so the two cannot drift apart. New entries arrive as a filled-in template through an issue form, which is screened automatically and opens a pull request.

Active · 2026 · FAIR, OAI-PMH, DataCite, Open source

Open, client-side twin of Repo MetAudits: FAIR metadata scoring for DataCite and OAI-PMH repositories, running entirely in the browser. The published, open-source counterpart of the hosted evaluator.

Active · 2026 · SvelteKit, Citation Verification, Open source

Open-source companion of BiblioHelp: the reference-verification app and Word add-in, as code and a static build. The published twin of the hosted tool.

Active · 2026 · CoARA, Research Assessment, Open source

Self-assess your institution against the ten CoARA commitments and generate a prioritised, editable action plan, entirely in your browser. The open twin of the Reform Assessment toolkit.

Reform Assessment

Draft · 2026 · Barcelona Declaration, Open Research Information, Taxonomy, Research policy

Contribution to Task Force 2 of Working Group 7 of the Barcelona Declaration on Open Research Information: eight benefits of opening information about research, arranged in three axes (quality and trust, collaboration and innovation, impact and relevance) and published as an interactive explorer rather than a PDF. One machine-readable taxonomy is the single source for the explorer, the written brief and the flow map, so the three cannot drift apart. Readers propose changes and examples through a pre-filled issue form that lands in a versioned register. Draft at v0.1, not yet reviewed by the task force.

Engagements

  • Member, Open Science Monitoring Initiative (OSMI), WG1 & WG4 · WG1: Defining open science monitoring needs and WG4: Building shared infrastructures and tools
  • Member, Barcelona Declaration on Open Research Information, WG7 · Promoting interoperability, openness, and inclusivity in research information ecosystems
  • Advisory Board Member, Cross-Domain Interoperability Framework (CDIF) - CODATA · Supporting interdisciplinary metadata interoperability frameworks

Training

  • Skills4EOSC First Cohort Participant (2025): Participant in the Train of Trainers Learning Path: Open Science and Research Data Management in the Social Sciences and Humanities, part of the European Skills4EOSC program. Focus on FAIR-enabling services, interoperability infrastructures, and applied data governance for research institutions.
  • Essentials 4 Data Support (2023): Completed official RDNL training for research data support professionals. Focused on FAIR principles, metadata quality, data sharing, and the implementation of robust Data Management Plans (DMPs). Only active Latin American participant in the cohort.
  • CWTS Scientometrics Summer School (2020): Intensive 75-hour course on scientometrics, bibliometric indicators, knowledge flows, and research evaluation methodologies. Covered theory and practical application for policy and infrastructure development.
  • FORCE11 Scholarly Communication Institute – FSCI (2017–2025): Completed multiple tracks including: Governance for AI in Scientific Publications, Forensic Scientometrics, FAIR Data in the Scholarly Communication Lifecycle, Metadata Governance, Research Reproducibility, Open Science in the Global South, Data visualization in R and D3.js, Using APIs (ORCID, Sherpa Romeo, Unpaywall) for institutional analysis
  • DataCite Certified Core Training (2024): Certification focused on the use of persistent identifiers (PIDs), metadata best practices, and repository integration with DataCite infrastructure.

Generated from data/cv.json · 2026-08-24