<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0" xmlns:media="http://search.yahoo.com/mrss/">
<channel>
<title><![CDATA[ Invest in Open Infrastructure ]]></title>
<description><![CDATA[ Helping you invest in the open technology that research relies on. ]]></description>
<link>https://investinopen.org</link>
<image>
    <url>https://investinopen.org/favicon.png</url>
    <title>Invest in Open Infrastructure</title>
    <link>https://investinopen.org</link>
</image>
<lastBuildDate>Thu, 10 Sep 2026 12:26:23 +0000</lastBuildDate>
<atom:link href="https://investinopen.org/rss/" rel="self" type="application/rss+xml"/>
<ttl>60</ttl>

    <item>
        <title><![CDATA[ Fondo de IOI para la Adopción en Redes: actualización sobre los primeros seis meses de avance de LA Referencia ]]></title>
        <description><![CDATA[ Adopción en red con una sólida base técnica. ]]></description>
        <link>https://investinopen.org/fondo-de-ioi-para-la-adopcion-en-redes-la-referencia-6meses-es/</link>
        <guid isPermaLink="false">6a91cc6aa644ab00015706dd</guid>
        <category><![CDATA[  ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Mon, 31 Aug 2026 12:30:33 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p><em>Este artículo también está disponible en </em><a href="https://investinopen.org/blog/ioi-fund-for-network-adoption-la-referencia-6mo-en" rel="noreferrer"><em>inglés</em></a>.</p><p>A finales de 2025, <a href="https://investinopen.org/ioi-fund/"><u>Invest in Open Infrastructure lanzó el Fondo de IOI para la Adopción en Redes</u></a>, un nuevo tipo de inversión diseñado para acelerar la adopción global de infraestructura abierta para la investigación y el intercambio de datos. En lugar de financiar herramientas individuales o instituciones aisladas, el Fondo invierte en <em>redes</em>: consorcios de instituciones e investigadores con relaciones ya establecidas, una gobernanza activa de sus miembros y un compromiso demostrado con la ciencia abierta. La lógica es que fortalecer la infraestructura compartida que una red ya opera puede multiplicar el impacto en todas sus instituciones miembro al mismo tiempo.</p><p>El Fondo surgió del diálogo directo con las comunidades de investigación para identificar los proyectos de infraestructura abierta que más necesitan apoyo sostenido, y reúne contribuciones de un grupo de financiadores interesados en fortalecer el ecosistema global de investigación abierta. Cada red seleccionada recibe hasta 1,5 millones de dólares durante un período de dos a tres años, junto con apoyo continuo del equipo de IOI para la implementación en ámbitos como gobernanza, participación comunitaria y modelos de negocio. El objetivo de IOI con el Fondo no es solamente permitir que las organizaciones beneficiarias desarrollen proyectos capaces de generar un gran impacto en sus redes, sino también contribuir a la resiliencia de esas redes mucho después de finalizada la subvención.</p><p><a href="https://investinopen.org/our-work/la-referencia/"><u>LA Referencia</u></a> fue seleccionada como una de las dos primeras beneficiarias del Fondo, <a href="https://investinopen.org/blog/empowering-networks-advancing-openness-invest-in-open-infrastructure-announces-inaugural-grantees-of-the-ioi-fund-for-network-adoption/"><u>entre más de 100 postulaciones provenientes de 22 países</u></a>. En conjunto, las dos primeras organizaciones beneficiarias del Fondo están en condiciones de beneficiar a más de 1.300 instituciones de investigación en 36 países de América Latina y África. Este artículo ofrece una visión general de lo que LA Referencia ha realizado durante los primeros seis meses de esta inversión. En los próximos meses publicaremos también una actualización sobre la otra organización beneficiaria inaugural, UbuntuNet Alliance.</p><figure class="kg-card kg-image-card"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/08/lareferencia.png" class="kg-image" alt="Logo de LA Referencia" loading="lazy" width="1038" height="240" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/08/lareferencia.png 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/08/lareferencia.png 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/08/lareferencia.png 1038w" sizes="(min-width: 720px) 720px"></figure><h1 id="introducci%C3%B3n-al-proyecto-de-la-referencia"><strong>Introducción al proyecto de LA Referencia</strong></h1><p>LA Referencia es una red regional de cooperación que agrega y proporciona infraestructura de&nbsp; acceso abierto a la producción académica y científica de repositorios nacionales de América Latina y España, basada en un modelo de gobernanza federada y en una infraestructura de propiedad pública y no comercial. Con el apoyo del Fondo de IOI, LA Referencia está desarrollando un proyecto para el período 2026–2028 destinado a ampliar el alcance regional y la profundidad técnica de esta infraestructura mediante cuatro grandes objetivos: incorporar más países y repositorios a la red, modernizar la forma en que los investigadores descubren contenidos en múltiples idiomas, construir un sistema soberano y verificable de identificadores persistentes y poner en marcha un repositorio regional compartido junto con un programa de capacitación para datos de investigación. El avance completo del proyecto está documentado en detalle en el <a href="https://la-referencia-ioi.github.io/reports/?ref=investinopen.org"><u>sitio de informes del proyecto de LA Referencia</u></a>, y el relato que sigue se basa directamente en el <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/executive-mid-year-report-2026-07-es.md?ref=investinopen.org"><u>informe ejecutivo del equipo correspondiente a enero–julio de 2026</u></a>.</p><p>El proyecto está organizado en tres componentes complementarios que conectan contenidos, tecnologías de descubrimiento, infraestructura de identificadores y capacidades humanas:</p><ul><li><strong>Componente A — Expansión de la infraestructura de agregación de LA Referencia:</strong> ampliar la cobertura regional, mejorar y normalizar los metadatos e introducir búsqueda semántica multilingüe en la capa de descubrimiento de LA Referencia.</li><li><strong>Componente B — Identificadores ARK descentralizados y preservación de metadatos (dARK):</strong> desarrollar dARK 2.0, una infraestructura modular basada en blockchain e ipfs para asignar, publicar y resolver identificadores persistentes.</li><li><strong>Componente C — Repositorio regional de datos&nbsp; datos para instituiones sin infraestructura local y soporte para Dataverse:</strong> poner en marcha un repositorio regional para datos de investigación basado en Dataverse, acompañado de un programa federado de capacitación y curaduría.</li></ul><p>En el resto de este artículo profundizamos en los avances realizados en cada uno de estos componentes y compartimos productos y resultados del trabajo de LA Referencia que ya pueden ser explorados.</p><h1 id="componente-a-expansi%C3%B3n-de-la-infraestructura-de-agregaci%C3%B3n-de-la-referencia"><strong>Componente A: Expansión de la infraestructura de agregación de LA Referencia</strong></h1><p>El descubrimiento depende de contar con buenos metadatos para las colecciones, por lo que el Componente A combina la ampliación de la cobertura regional de LA Referencia con el trabajo sobre calidad de metadatos y una nueva capacidad de búsqueda multilingüe que permite descubrir contenidos independientemente del idioma utilizado en la búsqueda o del idioma del contenido.</p><h3 id="avances-hasta-la-fecha"><strong>Avances hasta la fecha</strong></h3><p>Durante el período del informe, LA Referencia implementó una metodología regional para identificar instituciones y repositorios candidatos a ser cosechados con el objetivo de ampliar su colección. Este trabajo incluyó la verificación de servicios OAI-PMH, la revisión de metadatos, la actualización de la Interoperability and Registry Database (IRD) de COAR y la verificación cruzada de información institucional con ROR. El trabajo de curaduría se completó para los países miembros Ecuador, Uruguay y Costa Rica, y para los países no miembros Colombia, Paraguay, Venezuela, Bolivia y Cuba; los registros de otros países se encuentran en proceso.</p><p>En total, el equipo actualizó los registros de 362 repositorios: 187 fueron editados, 77 incorporados como nuevos y 98 archivados. Venezuela fue el país con mayores avances, con dos repositorios cosechados exitosamente como experiencia piloto para la incorporación de contenidos provenientes de países que todavía no operan como nodos completos de LA Referencia.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/a1-expansion-regional/es/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explorar el informe completo de avances sobre la expansión de la cobertura regional</a></div><p>En relación con esta expansión geográfica, LA Referencia está desarrollando una solución para un problema común identificado en su comunidad: muchos investigadores realizan búsquedas en un idioma mientras que el trabajo que necesitan está descrito en otro. Un registro de un repositorio en español podría no aparecer nunca para una persona que realiza una búsqueda en inglés, o viceversa, porque la búsqueda tradicional por palabras clave solo puede hacer coincidir las palabras exactas presentes en la página. El Componente A busca cerrar esa brecha mediante búsqueda semántica multilingüe, que encuentra resultados haciendo coincidir significados en lugar de palabras exactas. Para determinar la mejor manera de hacerlo, el equipo realizó una evaluación rigurosa y reproducible, comparando aproximadamente treinta modelos diferentes de IA en dos colecciones de prueba de entre 30.000 y 50.000 documentos, con miles de búsquedas de prueba en hasta 10 idiomas. Este trabajo produjo resultados prometedores: en el piloto, la búsqueda semántica y la búsqueda combinada recuperaron aproximadamente el doble de resultados relevantes entre los primeros 10 resultados —pasando de un promedio de 4,5 a 9,1— en comparación con la búsqueda únicamente por palabras clave. Las mayores mejoras se observaron en idiomas que suelen presentar mayores dificultades para la búsqueda por palabras clave, como chino, japonés y árabe. El equipo <a href="https://la-referencia-ioi.github.io/reports/productos/a4-seleccion-modelos/es/index.html?ref=investinopen.org"><u>evaluó distintos modelos</u></a> y seleccionó uno que ofrecía un equilibrio entre resultados de alta calidad, eficiencia y velocidad para operar sobre millones de registros.</p><p>Esta evaluación ya fue incorporada a un sistema beta funcional dentro de la plataforma de búsqueda de LA Referencia. Cuando una persona realiza una búsqueda, puede elegir entre la coincidencia tradicional por palabras clave, una coincidencia “basada en significado” que funciona entre distintos idiomas o un modo combinado que integra ambas aproximaciones. La reindexación se vuelve considerablemente más rápida cuando se reutilizan embeddings generados previamente. Reconstruir el índice completo para un conjunto de prueba de aproximadamente 28.000 registros requirió alrededor de 5 horas la primera vez, pero solamente unos 10 minutos cuando los datos basados en significado ya estaban almacenados, lo que hace posible realizar actualizaciones regulares a medida que la colección continúa creciendo. Paralelamente, el equipo incorporó <a href="https://la-referencia-ioi.github.io/lareferencia-anubis-demo/?lang=es&ref=investinopen.org"><u>un WAF abierto que filtra bots automatizados y scrapers</u></a>, junto con una forma de aplicar esa misma protección simultáneamente en múltiples sitios de LA Referencia.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://semantic.lareferencia.info/?ref=investinopen.org" class="kg-btn kg-btn-accent">Probar la demostración de búsqueda semántica e híbrida</a></div><p>También es posible profundizar en los detalles técnicos consultando el <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-a-mid-year-report-2026-07-es.md?ref=investinopen.org"><u>informe completo del Componente A de LA Referencia</u></a>.</p><h3 id="pr%C3%B3ximos-pasos"><strong>Próximos pasos</strong></h3><p>LA Referencia continúa revisando e incorporando repositorios de otros países miembros y de nuevos países fuera de su red. El equipo planea ampliar la evaluación de la búsqueda semántica multilingüe a toda la colección regional y consolidar la versión beta de búsqueda semántica antes de avanzar hacia un despliegue más amplio.</p><h1 id="componente-b-identificadores-ark-descentralizados-dark-y-preservaci%C3%B3n-de-metadatos"><strong>Componente B: Identificadores ARK descentralizados (dARK) y preservación de metadatos</strong></h1><p>El Componente B comprende el desarrollo de dARK 2.0, la próxima generación de la infraestructura de LA Referencia para asignar, publicar y resolver identificadores persistentes ARK. Este componente contempla una <a href="https://la-referencia-ioi.github.io/reports/productos/b1-dark-2/es/index.html?ref=investinopen.org"><u>infraestructura regional, soberana y sostenible</u></a>, construida específicamente para la región a la que sirve LA Referencia.</p><h3 id="avances-hasta-la-fecha-1"><strong>Avances hasta la fecha</strong></h3><p>Los identificadores persistentes son los que permiten que una producción de investigación continúe siendo localizable y citable a largo plazo mediante un enlace permanente que sigue conduciendo al artículo, conjunto de datos o informe correcto años después, incluso si el archivo cambia de ubicación o si cambia el sistema que lo aloja. Actualmente, gran parte de esta infraestructura es gestionada por un pequeño grupo de proveedores externos. dARK 2.0 representa la visión de LA Referencia de una infraestructura que la propia región pueda operar, gobernar y considerar confiable, sin depender de una única organización como centro de control.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/b1-dark-2/es/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explorar dARK: identificadores persistentes como bien público regional</a></div><p>Para lograrlo, el equipo dividió dARK 2.0 en componentes separados que pueden ser verificados, actualizados o escalados de manera independiente. Un registro compartido y resistente a manipulaciones —un tipo de blockchain— mantiene el registro autoritativo de qué identificadores existen y hacia dónde apuntan, mientras que los metadatos descriptivos de cada registro se almacenan por separado, permitiendo que cualquiera pueda verificar que no han sido alterados. Esta separación mantiene la capa de “prueba” simple y confiable, mientras que los datos descriptivos pueden continuar creciendo sin necesidad de modificar esa capa de “prueba”. El equipo también construyó todo el flujo de procesamiento que requiere el sistema: reservar nuevos identificadores, preparar y almacenar sus metadatos, publicar la referencia en el registro y permitir que cualquier persona consulte un identificador y lo resuelva hacia el registro correcto.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/b2-infraestructura-multisitio/es/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explorar la arquitectura técnica de dARK 2.0</a></div><p>El equipo sometió este flujo de trabajo a pruebas de estrés con millones de registros y ejecutó cientos de verificaciones automatizadas para confirmar su robustez, además de construir herramientas administrativas para que el personal pueda monitorear identificadores, actividad y consultas. El flujo de dARK ya está listo para ser desplegado y conectado con el sistema regional de cosecha de LA Referencia, por lo que el proceso de cosecha se está integrando con el flujo de dARK para permitir la asignación de identificadores a medida que los registros son procesados, incluso en colecciones grandes y en constante cambio.</p><p><a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-b-mid-year-report-2026-07-es.md?ref=investinopen.org"><u>Profundizar en los detalles en el informe completo del Componente B</u></a>.</p><h3 id="pr%C3%B3ximos-pasos-1"><strong>Próximos pasos</strong></h3><p>El siguiente paso consiste en distribuir la infraestructura dARK entre distintas instituciones en lugar de ejecutarla desde un único lugar. Se está preparando una instalación equivalente en RNP, en Brasil, y la institución socia IBICT está preparando la migración de sus identificadores desde la versión original 1.0 de dARK. Con múltiples instituciones utilizando el sistema, LA Referencia planea validar su funcionamiento entre distintos sitios y formalizar el modelo de seguridad y gobernanza de esta infraestructura regional federada.</p><h1 id="componente-c-repositorio-regional-para-datos-sin-infraestructura-institucional-y-soporte-para-dataverse"><strong>Componente C: Repositorio regional para datos sin infraestructura institucional y soporte para Dataverse</strong></h1><p>El Componente C combina un nuevo repositorio regional basado en Dataverse con un programa federado de capacitación y curaduría, cuyo objetivo es ofrecer un espacio para aquellos datos de investigación que carecen de infraestructura institucional, al mismo tiempo que se difunden las capacidades necesarias para operar servicios de datos de manera sostenible en toda la región.</p><h3 id="avances-hasta-la-fecha-2"><strong>Avances hasta la fecha</strong></h3><p>El equipo comenzó con un diagnóstico regional que identificó aproximadamente 96 repositorios de datos registrados y alrededor de 30 instalaciones de Dataverse en América Latina, concentradas principalmente en Brasil, Argentina, México y Colombia. El diagnóstico se basó en cuestionarios y entrevistas que abordaron gobernanza, infraestructura, curaduría, preservación, soporte e interoperabilidad. Este análisis confirmó que la región ya cuenta con una experiencia y conocimientos significativos, lo que dio forma a una estrategia de colaboración entre pares que concibe al repositorio como un servicio sociotécnico: software e infraestructura combinados con políticas, equipos multidisciplinarios, apoyo a investigadores y capacitación continua.</p><p>El principal resultado técnico es una versión de prueba funcional y en línea del repositorio —un “Dataverse Alpha”— que funciona sobre la propia infraestructura de LA Referencia. El repositorio está disponible en inglés, español y portugués, incluye campos para indicar el país de origen de un conjunto de datos y con cuáles de los 17 Objetivos de Desarrollo Sostenible de las Naciones Unidas se relaciona, y utiliza licencias abiertas (CC0 y CC BY) para que los datos puedan ser reutilizados libremente. Ya se han cargado conjuntos de datos de ejemplo para que el equipo pueda probar la carga, descripción, búsqueda y acceso a archivos antes de que comiencen a recibirse depósitos institucionales reales.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://data.lareferencia.info/?ref=investinopen.org" class="kg-btn kg-btn-accent"> Explorar el Dataverse Alpha de LA Referencia</a></div><p>Junto con el desarrollo técnico, el equipo está elaborando un MOOC asincrónico de 16 horas, <em>Introducción a la Gestión y Curaduría de Datos en Dataverse</em>, que abarca fundamentos y publicación, metadatos y documentación, verificación técnica y formatos, y ética, licenciamiento y aprobación. Este es uno de los tres productos de capacitación planificados; los otros incluyen capacitación práctica en curaduría y apoyo directo a las instituciones piloto. El curso será construido e instalado en un entorno Moodle de producción, con materiales en español y portugués.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/c2-curso-curaduria-datos/es/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Leer más sobre el MOOC y sus módulos</a></div><p>Invitamos a los lectores a explorar todos los detalles técnicos en el <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-c-mid-year-report-2026-07-es.md?ref=investinopen.org"><u>informe completo del Componente C</u></a>.</p><h3 id="pr%C3%B3ximos-pasos-2"><strong>Próximos pasos</strong></h3><p>El equipo finalizará los contenidos del curso, validará de extremo a extremo los flujos de trabajo del repositorio, seleccionará las primeras instituciones piloto y transformará las lecciones aprendidas durante la fase Alpha en un modelo estable para el servicio regional.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explorar el sitio completo de informes del proyecto de LA Referencia</a></div><h1 id="el-proyecto-de-la-referencia-apoyando-los-objetivos-del-fondo-de-ioi-para-la-adopci%C3%B3n-en-redes"><strong>El proyecto de LA Referencia: apoyando los objetivos del Fondo de IOI para la Adopción en Redes</strong></h1><p>A seis meses de iniciado el proyecto, LA Referencia ha pasado del diseño y el trabajo de diagnóstico a sistemas funcionales: una experiencia de descubrimiento multilingüe con mejoras medibles, una arquitectura funcional para identificadores persistentes soberanos y verificables, y un repositorio regional de datos en funcionamiento acompañado de un proceso de capacitación diseñado para perdurar más allá de una única subvención. Nada de esto fue construido para una sola institución; cada mejora contribuye a una infraestructura compartida y federada de la que ya dependen los repositorios miembros de LA Referencia en toda América Latina.</p><p>Esta es la premisa del Fondo de IOI para la Adopción en Redes: que la forma más rápida y duradera de hacer crecer la adopción de infraestructura abierta es invertir en redes que ya cuentan con la gobernanza, la confianza y el alcance necesarios para poner nuevas capacidades al servicio de muchas instituciones simultáneamente. Los primeros seis meses de LA Referencia muestran esta tesis en acción. La base técnica que están construyendo permitirá avanzar hacia la próxima fase del proyecto: expandirse hacia nuevos países e idiomas, validar dARK entre distintas instituciones e incorporar las primeras instituciones piloto al Dataverse regional.</p><hr><p><em>Leé más sobre el proyecto en el</em> <a href="https://la-referencia-ioi.github.io/reports/index.html?ref=investinopen.org" rel="noreferrer"><em><u>sitio de informes del proyecto de LA Referencia</u></em></a><em>, incluido el</em> <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/executive-mid-year-report-2026-07-es.md?ref=investinopen.org"><em><u>informe ejecutivo completo</u></em></a><em>en el que se basa este artículo, o conocé más sobre</em> <a href="https://investinopen.org/our-work/la-referencia/"><em><u>el trabajo de IOI con LA Referencia</u></em></a> <em>y el</em> <a href="https://investinopen.org/ioi-fund/"><em><u>Fondo de IOI para la Adopción en Redes</u></em></a><em>.</em></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ IOI Fund for Network Adoption: LA Referencia’s six-month progress update ]]></title>
        <description><![CDATA[ Network adoption with a strong technical foundation. ]]></description>
        <link>https://investinopen.org/blog/ioi-fund-for-network-adoption-la-referencia-6mo-en/</link>
        <guid isPermaLink="false">6a91c03ea644ab0001570692</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Mon, 31 Aug 2026 12:30:19 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p><em>This article is also available in </em><a href="https://investinopen.org/fondo-de-ioi-para-la-adopcion-en-redes-la-referencia-6meses-es/" rel="noreferrer"><em>Spanish</em></a><em>. </em></p><p>In late 2025, Invest in Open Infrastructure launched the <a href="https://investinopen.org/ioi-fund/"><u>IOI Fund for Network Adoption</u></a>, a new kind of investment designed to accelerate the global adoption of open infrastructure for research and data sharing. Rather than funding individual tools or single institutions, the Fund invests in <em>networks</em>: consortia of institutions and researchers with established relationships, active member governance, and a demonstrated commitment to open science. The logic is that strengthening the shared infrastructure that a network already runs can multiply impact across every member institution at once.</p><p>The Fund grew out of direct engagement with research communities to identify the open infrastructure projects most in need of sustained support, and it pools contributions from a group of funders all interested in strengthening the global open research ecosystem. Each selected network receives up to $1.5 million over two to three years, paired with ongoing implementation support from IOI staff on governance, community engagement, and business modeling. IOI’s goal with the Fund is not only to enable grantees to do projects that can have a big impact on their networks, but also to support the resilience of these networks long after the grant ends.</p><p><a href="https://investinopen.org/our-work/la-referencia/"><u>LA Referencia</u></a> was named one of the Fund's two inaugural grantees,<a href="https://investinopen.org/blog/empowering-networks-advancing-openness-invest-in-open-infrastructure-announces-inaugural-grantees-of-the-ioi-fund-for-network-adoption/"> <u>selected from more than 100 applications spanning 22 countries</u></a>. Together, the Fund's first two grantees are positioned to benefit more than 1,300 research institutions across 36 countries in Latin America and Africa. This post provides an overview of what LA Referencia has done with the first six months of that investment. Look for an update post about our other inaugural grantee, UbuntuNet Alliance, in the coming months.&nbsp;</p><figure class="kg-card kg-image-card"><a href="https://new.lareferencia.info/?ref=investinopen.org"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/08/lareferencia-1.png" class="kg-image" alt="Logo of LA Referencia" loading="lazy" width="1038" height="240" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/08/lareferencia-1.png 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/08/lareferencia-1.png 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/08/lareferencia-1.png 1038w" sizes="(min-width: 720px) 720px"></a></figure><h2 id="an-introduction-to-la-referencia%E2%80%99s-project">An introduction to LA Referencia’s project</h2><p>LA Referencia is a regional cooperation network that aggregates and provides open access to the scholarly and scientific output of national repositories across Latin America and Spain, built on a model of federated governance and publicly owned, non-commercial infrastructure. With IOI Fund support, LA Referencia is running a 2026–2028 project to expand that infrastructure's regional reach and technical depth with four major goals: bringing more countries and repositories into the network, modernizing how researchers discover content in multiple languages, building a sovereign and verifiable system of persistent identifiers, and standing up a shared regional repository and training program for research data. The project's full progress is documented in detail on<a href="https://la-referencia-ioi.github.io/reports/en/?ref=investinopen.org"> <u>LA Referencia's project reporting site</u></a>, and the account below draws directly on the team's own<a href="https://github.com/LA-Referencia-IOI/reports/blob/main/executive-mid-year-report-2026-07-en.md?ref=investinopen.org"> <u>executive report for January–July 2026</u></a>.&nbsp;</p><p>The project is organized into three complementary components that connect content, discovery technology, identifier infrastructure, and human capacity:</p><ul><li><strong>Component A — Expanding LA Referencia’s Aggregation Infrastructure:</strong> growing regional coverage, improving and normalizing metadata, and introducing multilingual semantic search to LA Referencia's discovery layer.</li><li><strong>Component B — Decentralized ARK Identifiers and Metadata Preservation (dARK):</strong> building dARK 2.0, a modular, blockchain-backed infrastructure for assigning, publishing, and resolving persistent identifiers.</li><li><strong>Component C — Regional Orphan Data Repository &amp; Dataverse Support:</strong> standing up a Dataverse-based regional repository for research data alongside a federated training and curation program.</li></ul><p>In the rest of this post, we dive deeper into the progress made on each of these components and share explorable outputs of LA Referencia’s work to date.&nbsp;</p><h2 id="component-a-expanding-la-referencia%E2%80%99s-aggregation-infrastructure">Component A: Expanding LA Referencia’s Aggregation Infrastructure</h2><p>Discovery depends on good metadata for the collections, so Component A pairs the expansion of LA Referencia's regional coverage with metadata quality work and a new multilingual search capability that enables discovery regardless of the language of the search or the content. </p><h3 id="progress-to-date">Progress to date</h3><p>Over the reporting period, LA Referencia enacted a regional methodology for identifying candidate institutions and repositories for harvesting to bolster their collection. This work involved verifying OAI-PMH services, reviewing metadata, updating COAR's Interoperability and Registry Database (IRD), and cross-checking institutional information against ROR. Curation work was completed for member countries Ecuador, Uruguay, and Costa Rica, and non-member countries Colombia, Paraguay, Venezuela, Bolivia, and Cuba; records for another countries are in progress.&nbsp;</p><p>In total, the team updated records for 362 repositories: 187 edited, 77 newly added, and 98 archived. Venezuela was the country with the greatest progress, with two repositories successfully harvested as a pilot for incorporating content from countries that do not yet operate as full nodes in LA Referencia.&nbsp;</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/a1-expansion-regional/en/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explore the full progress report on expanding regional coverage</a></div><p>Related to this geographic expansion, LA Referencia is building a solution for a common problem they encountered in their community: many researchers search in one language while the work they need is described in another. A record from a Spanish-language repository might never surface for someone searching in English, or vice versa, because traditional keyword search can only match the exact words on the page. Component A set out to close that gap with multilingual semantic search, which finds results by matching meaning instead of exact wording. To figure out the best way to do this, the team ran a rigorous, reproducible evaluation, comparing roughly thirty different AI models across two test collections of 30,000 to 50,000 documents, with thousands of test searches in as many as 10 languages. This work resulted in promising outcomes: in the pilot, semantic and blended search surfaced roughly twice as many relevant results in the top 10 (up from an average of 4.5 to 9.1) compared with keyword search alone, with the biggest gains for languages that keyword matching tends to handle poorly, like Chinese, Japanese, and Arabic. The team <a href="https://la-referencia-ioi.github.io/reports/productos/a4-seleccion-modelos/en/index.html?ref=investinopen.org"><u>evaluated different models</u></a> and chose one that balanced high-quality results with efficiency and speed to run on millions of records.&nbsp;</p><p>That evaluation has now been built into a beta working system on LA Referencia's search platform. When someone runs a search, they can choose the traditional keyword match, a "meaning-based" match that works across languages, or a blended mode that combines both. Reindexing becomes substantially faster when previously generated embeddings are reused. Rebuilding the full index for a test set of about 28,000 records took roughly 5 hours the first time, but only about 10 minutes once the meaning-based data was already stored, making regular updates practical as the collection keeps growing. Alongside this, the team added <a href="https://la-referencia-ioi.github.io/lareferencia-anubis-demo/?lang=en&ref=investinopen.org"><u>a open WAF that screens out automated bots and scrapers</u></a>, plus a way to apply that same protection across multiple LA Referencia sites at once.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://semantic.lareferencia.info/?ref=investinopen.org" class="kg-btn kg-btn-accent">Try a demo of semantic and hybrid search</a></div><p>You can also dive into the technical details by reading LA Referencia’s <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-a-mid-year-report-2026-07-en.md?ref=investinopen.org"><u>full Component A report</u></a>.</p><h3 id="whats-next">What's next</h3><p>LA Referencia continues to review and add repositories from additional member countries and new countries beyond their network. The team plans to expand its multilingual semantic search evaluation across the full regional collection and consolidate the semantic search beta ahead of broader rollout.</p><h2 id="component-b-decentralized-ark-identifiers-dark-and-metadata-preservation">Component B: Decentralized ARK Identifiers (dARK)&nbsp; and Metadata Preservation</h2><p>Component B involves building dARK 2.0, the next generation of LA Referencia's infrastructure for assigning, publishing, and resolving persistent ARK identifiers.&nbsp; This component accounts for <a href="https://la-referencia-ioi.github.io/reports/productos/b1-dark-2/en/index.html?ref=investinopen.org"><u>regional, sovereign, and sustainable infrastructure</u></a> built specifically for the region served by LA Referencia.</p><h3 id="progress-to-date-1">Progress to date</h3><p>Persistent identifiers are what keep a piece of research findable and citable for the long term by using a permanent link that still leads to the right paper, dataset, or report years later, even if the file moves or the system hosting it changes. Today, most of that infrastructure is run by a handful of external providers. dARK 2.0 is LA Referencia's vision for an infrastructure the region can run, govern, and trust for itself without relying on one single organization as a locus of control.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/b1-dark-2/en/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explore dARK: Persistent identifiers as a regional public good</a></div><p>To do that, the team split dARK 2.0 into separate components that can each be checked, updated, or scaled independently. A shared, tamper-evident ledger (a type of blockchain) maintains the authoritative record of which identifiers exist and what they point to, while the descriptive metadata for each record is stored separately, allowing anyone to verify it hasn't been altered. That split keeps the "proof" layer simple and trustworthy, while the descriptive data can keep growing without needing to touch the "proof" layer. The team also built the full pipeline this system requires: reserving new identifiers, preparing and storing their metadata, publishing the reference to the ledger, and allowing anyone to look up an identifier and resolve it to the correct record.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/b2-infraestructura-multisitio/en/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Explore the dARK 2.0 technical architecture</a></div><p>The team stress-tested this workflow against millions of records and ran hundreds of automated checks to confirm it holds up, and built administrative tools so staff can monitor identifiers, activity, and lookups. The dARK pipeline is now ready to be deployed and connected to LA Referencia's regional harvesting system, so the harvesting workflow is being integrated with the dARK pipeline to support identifier assignment as records are processed, even across large, constantly changing collections.&nbsp;</p><p><a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-b-mid-year-report-2026-07-en.md?ref=investinopen.org"><u>Dive into the details in the full Component B report</u></a>.</p><h3 id="what%E2%80%99s-next">What’s next</h3><p>The next step is spreading the dARK infrastructure across institutions rather than running it from one place. A matching setup is being prepared at RNP in Brazil, and partner institution IBICT is preparing to migrate its identifiers over from the original 1.0 version of dARK. With multiple institutions using the system, LA Referencia plans to validate cross-site operation and formalize the security and governance model for this federated regional infrastructure. </p><h1 id="component-c-regional-orphan-data-repository-dataverse-support">Component C: Regional Orphan Data Repository &amp; Dataverse Support</h1><p>Component C pairs a new Dataverse-based regional repository with a federated training and curation program, aimed at giving research data lacking institutional infrastructure a home while spreading the skills needed to run data services sustainably across the region.</p><h3 id="progress-to-date-2">Progress to date</h3><p>The team began with a regional diagnostic that identified roughly 96 registered data repositories and around 30 Dataverse installations across Latin America, concentrated in Brazil, Argentina, Mexico, and Colombia, drawing on questionnaires and interviews covering governance, infrastructure, curation, preservation, support, and interoperability. That diagnostic confirmed the region already has significant experience and expertise, which has shaped a peer-collaboration strategy treating the repository as a socio-technical service: software and infrastructure paired with policy, multidisciplinary teams, researcher support, and ongoing training.</p><p>The headline technical result is a live, working test version of the repository (a "Dataverse Alpha") running on LA Referencia's own infrastructure. The repository works in English, Spanish, and Portuguese, includes fields to tag a dataset's country of origin and which of the UN's 17 Sustainable Development Goals it relates to, and uses open licenses (CC0 and CC BY) so the data can be freely reused. Sample datasets are already loaded so the team can test uploading, describing, searching, and accessing files before real institutional deposits start coming in.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://data.lareferencia.info/?ref=investinopen.org" class="kg-btn kg-btn-accent">Explore Dataverse Alpha at LA Referencia</a></div><p>Alongside the technical build, the team is developing a 16-hour asynchronous MOOC, <em>Introduction to Data Management and Curation in Dataverse</em>, covering foundations and publication, metadata and documentation, technical verification and formats, and ethics, licensing, and approval. This is one of three planned training products; the others include hands-on curation training and direct support for pilot institutions. The course will be built and installed on a production Moodle environment with materials in Spanish and Portuguese.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/productos/c2-curso-curaduria-datos/en/index.html?ref=investinopen.org" class="kg-btn kg-btn-accent">Read more about the MOOC and its modules</a></div><p>We encourage readers to explore the full technical details in the <a href="https://github.com/LA-Referencia-IOI/reports/blob/main/component-c-mid-year-report-2026-07-en.md?ref=investinopen.org"><u>full Component C report</u></a>.</p><h3 id="what%E2%80%99s-next-1">What’s next</h3><p>The team will finish course content, validate repository workflows end-to-end, select the first pilot institutions, and turn the lessons from the Alpha into a stable model for the regional service.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://la-referencia-ioi.github.io/reports/en/?ref=investinopen.org" class="kg-btn kg-btn-accent">Explore the full project reporting website</a></div><h1 id="la-referencia%E2%80%99s-project-supporting-the-goals-of-the-ioi-fund-for-network-adoption">LA Referencia’s Project: Supporting the Goals of the IOI Fund for Network Adoption</h1><p>Six months in, LA Referencia has moved from design and diagnostic work to functioning systems: a measurably better multilingual discovery experience, a working architecture for verifiable, sovereign persistent identifiers, and a live regional data repository paired with a training pipeline built to outlast any single grant. None of this was built for one institution; every improvement contributes to shared, federated infrastructure that LA Referencia's member repositories across Latin America already rely on.&nbsp;</p><p>That is the premise of the IOI Fund for Network Adoption: that the fastest, most durable way to grow open infrastructure adoption is to invest in networks that already have the governance, trust, and reach to put new capacity to work across many institutions at once. LA Referencia's first six months show that thesis in motion. The technical foundation that they are building will enable the project's next phase, expanding to new countries and languages, validating dARK across institutions, and bringing its first pilot institutions onto the regional Dataverse.</p><hr><p><em>Read more about the project on</em><a href="https://la-referencia-ioi.github.io/reports/en/?ref=investinopen.org"><em> <u>LA Referencia's project reporting site</u></em></a><em>, including the full</em><a href="https://github.com/LA-Referencia-IOI/reports/blob/main/executive-mid-year-report-2026-07-en.md?ref=investinopen.org"><em> <u>executive report</u></em></a><em> this post is based on, or learn more about</em><a href="https://investinopen.org/our-work/la-referencia/"><em> <u>IOI's work with LA Referencia</u></em></a><em> and the</em><a href="https://investinopen.org/ioi-fund/"><em> <u>IOI Fund for Network Adoption</u></em></a><em>.&nbsp; </em></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Charting a sustainable future for TDCC-NES ]]></title>
        <description><![CDATA[ IOI is partnering with TDCC-NES, a collaborative network across the Dutch natural and engineering sciences domain that addresses FAIR data and other digital-related research challenges, to co-develop a post-NWO sustainability and co-investment model with its community. ]]></description>
        <link>https://investinopen.org/our-work/charting-a-sustainable-future-for-tdcc-nes/</link>
        <guid isPermaLink="false">6a5f3e0ddca8a0000125669a</guid>
        <category><![CDATA[ Our Work ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Tue, 21 Jul 2026 10:11:37 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <!-- START case study metadata -->
<p><strong>Duration:</strong> April 2026 to December 2026<br>
<strong>Team:</strong> Emmy Tsang, Emma Green, Sarah Lippincott<br>
<strong>Funder:</strong> TDCC-NES<br>
<strong>Skillset:</strong> Strategic Planning, Convening / Facilitation, Fiscal Assessment &amp; Planning</p>
<!-- END case study metadata →
--><h2 id="overview"><strong>Overview</strong></h2><p>TDCC-NES is the Thematic Digital Competence Centre for the Natural &amp; Engineering Sciences in the Netherlands, a network organisation that helps Dutch researchers and digital support experts explore digital competency challenges and find solutions through connection and collaboration, reducing fragmentation and duplication across the domain. Funded by NWO as part of its investment in digital research infrastructure, TDCC-NES engaged IOI to develop a credible long-term sustainability model ahead of the end of the NWO funding window.</p><h2 id="what-we-will-do"><strong>What we will do</strong></h2><p>Through a phased engagement, IOI is building a shared, evidence-based understanding of TDCC-NES's value proposition, stakeholder landscape, and sustainability options, through stakeholder interviews, document review, and a discovery findings presentation to the governing board. In late 2026, we will convene key stakeholders in a sustainability workshop to co-develop the foundations of a co-investment model and build co-ownership of the path to 2030.</p><h2 id="outcomes"><strong>Outcomes</strong></h2><p>Phase 1 delivered a discovery findings report and board presentation that established a shared evidence base for the board's sustainability conversation. Further outcomes will be added as the engagement progresses.</p><h3 id="facing-a-funding-transition-lets-design-your-next-model-together">Facing a funding transition? Let's design your next model together.</h3><p>Many publicly funded infrastructure and network organisations face the end of an initial grant or funding window. IOI helps organisations like yours understand your value to stakeholders and design sustainable co-investment models.<a href="https://investinopen.org/free-consult/"> <u>Book a free consult</u></a>.</p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Wellcome partners with IOI to map the data resources powering open research ]]></title>
        <description><![CDATA[ IOI&#39;s new initiative that maps the data resource ecosystem ]]></description>
        <link>https://investinopen.org/blog/wellcome-partners-with-ioi-to-map-the-data-resources-powering-open-research/</link>
        <guid isPermaLink="false">6a59cd496d07bb0001928f10</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Fri, 17 Jul 2026 06:56:40 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>We're excited to announce a new partnership with Wellcome, the global charitable foundation supporting science to solve urgent challenges. Our work will help Wellcome to gain deeper insight into the data resource ecosystem, providing a snapshot of the resources their researchers rely on.</p><p>Data resources that researchers rely on are critical to advancing science and discovery. That research often starts with data that research funders, such as Wellcome, did not create and does not maintain, including biodata repositories, population cohort databases, and environmental and climate data resources, built and sustained by organizations that often operate out of public view. At a moment when the research community is paying closer attention than ever to the resilience of the data infrastructure it depends on, we are supporting Wellcome to gain&nbsp;a clearer, deeper picture of that landscape. In turn, this will enable collective conversations within the research community about how we sustain these resources.</p><p>IOI is leading that effort, drawing on our track record of working alongside infrastructure providers, not just evaluating them from the outside. We'll be engaging directly with the teams running these resources to understand their governance, funding models, technical dependencies, open policies, sustainability outlooks and the broader ecosystem they operate within, the kind of insight that only comes from sustained, trust based engagement.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/07/markus-winkler-kA7zREkzrBw-unsplash-2.jpg" class="kg-image" alt="scrabble pieces spelling the word data" loading="lazy" width="1920" height="1280" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/07/markus-winkler-kA7zREkzrBw-unsplash-2.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/07/markus-winkler-kA7zREkzrBw-unsplash-2.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/07/markus-winkler-kA7zREkzrBw-unsplash-2.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/07/markus-winkler-kA7zREkzrBw-unsplash-2.jpg 1920w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Photo by </span><a href="https://unsplash.com/@markuswinkler?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Markus Winkler</span></a><span style="white-space: pre-wrap;"> on </span><a href="https://unsplash.com/photos/a-wooden-block-spelling-data-on-a-table-kA7zREkzrBw?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Unsplas</span></a><span style="white-space: pre-wrap;">h</span></figcaption></figure><p>The foundation for this work is<a href="https://infrafinder.investinopen.org/?ref=investinopen.org"> <u>Infra Finder</u></a>, IOI's open discovery platform for research infrastructure. We'll build verified, publicly accessible profiles for each resource in scope, using our established research and diligence process.This is also a significant expansion for Infra Finder, which has grown to include nearly 150 profiles of open infrastructure since its 2024 launch. Infra Finder is more than a directory. It's the foundation of our annual<a href="https://investinopen.org/data-room/state-of-oi/"> <u>State of Open Infrastructure</u></a> report, which tracks how the field is evolving year over year. Adding this cohort brings these data resources into that same yearly monitoring and analysis, giving Wellcome and the broader community longitudinal visibility into how they change and how they're funded over time.</p><p>We're looking forward to the work ahead. Stay tuned for more in the coming months! </p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ The Data Resilience Funding Landscape: A Preliminary Analysis ]]></title>
        <description><![CDATA[ Mapping where the money is (and isn&#39;t) going in the data rescue movement. ]]></description>
        <link>https://investinopen.org/blog/data-resilience-funding-landscape/</link>
        <guid isPermaLink="false">6a4c03fa34b98500012af011</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Katherine E Skinner ]]></dc:creator>
        <pubDate>Tue, 07 Jul 2026 14:13:01 +0000</pubDate>
        <media:content url="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/07/christian-j-w8u0d3UPovw-unsplash.jpg" medium="image"/>
        <content:encoded><![CDATA[ <p>Between January 2025 and the time of writing (July 2026), hundreds of organizations, from grassroots volunteer networks to major research institutions, have mobilized to rescue, preserve, and provide access to at-risk federal data in a rapid response to a sustained and accelerating assault on these knowledge assets. While this wave of work began with many ad hoc and siloed efforts, it has developed into a movement, and invested stakeholders have stepped forward to take on “hub” roles, helping to coordinate and guide the rescue efforts. A sense of movement generosity has also been building among the players, perhaps most evident in hard-hit fields such as environmental and climate science. Those with resources and know-how are sharing wherever they can, particularly via their time, energy, and networks.&nbsp;</p><p>Funders have responded to the crisis, with the Robert Wood Johnson Foundation, Hewlett Foundation, David and Lucile Packard Foundation, John D. and Catherine T. MacArthur Foundation, Alfred P. Sloan Foundation, Mellon Foundation, and others directing significant resources toward data rescue and resilience efforts. Some funders are working together in formal coordination groups (e.g., Funders for the Future of Public Data (3FPD) and the Portfolio to Protect Science) in hopes of moving resources in ways that reduce duplication, “noise,” while increasing investment alignment.</p><p>But a close look at where that money is flowing reveals concerning patterns that could subvert or undercut this work, limiting its ability to outlast the crisis that prompted it. Concretely, this includes dangerously low investments made in open infrastructures – including the technical platforms, standards, and software required to produce, share, discover, access, and preserve knowledge.</p><h2 id="crisis-reveals-what-was-always-true"><strong>Crisis reveals what was always true</strong></h2><p>Funding is moving in familiar, system-driven ways that we have witnessed many times before. As my colleagues and I have noted elsewhere (e.g., <a href="https://educopia.org/blog/why-are-so-many-scholarly-communication-infrastructure-providers-running-a-red-queens-race/?ref=investinopen.org"><u>Red Queen’s Race</u></a>, <a href="https://investinopen.org/blog/the-emperors-new-clothes-common-myths-hindering-open-infrastructure/"><u>Emperor’s New Clothes</u></a>), stakeholders in the knowledge and information ecosystem have for years been rewarded for ways of working that provide short-term gains but that tend to fail to add up to long-term success.</p><p><strong>Incentive structures, including funding and career advancement, encourage stakeholders to pursue innovation over maintenance, to build competing siloes rather than grow interoperability and coordination, and to misinterpret bare survival of open infrastructure communities and services as though that survival somehow signifies success. </strong></p><p>The resulting fragility of open infrastructure components has been persistent, and the current political crisis of 2025-2026 hasn’t created precarity so much as it has exposed it. Data is disappearing because the infrastructure in which it is nested was already brittle. This moment is an opportunity, not just to rescue data, but to correct a long-standing dependency on temporary and inadequate funding for infrastructure-level work.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/07/christian-j-w8u0d3UPovw-unsplash-1.jpg" class="kg-image" alt="A close up of petrified wood, showing intricate marble-like details of the rock-like texture in reds, blues, yellows, and whites. " loading="lazy" width="2000" height="1333" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/07/christian-j-w8u0d3UPovw-unsplash-1.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/07/christian-j-w8u0d3UPovw-unsplash-1.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/07/christian-j-w8u0d3UPovw-unsplash-1.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w2400/2026/07/christian-j-w8u0d3UPovw-unsplash-1.jpg 2400w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Photo by </span><a href="https://unsplash.com/@chrisk91?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Christian J.</span></a><span style="white-space: pre-wrap;"> on </span><a href="https://unsplash.com/photos/a-close-up-of-a-red-and-white-marble-w8u0d3UPovw?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Unsplash</span></a></figcaption></figure><h2 id="tracking-investment-flows"><strong>Tracking investment flows</strong></h2><p>IOI has been mapping the emerging data resilience ecosystem, tracking more than 85 “data rescue” and “data resilience” projects, initiatives, forums, and funding efforts that have launched or significantly evolved since 2025. The landscape clusters into recognizable functional layers: active rescue and collection efforts; tools, repositories, and discovery systems; monitoring and tracking; coordination and strategy; and advocacy and fundraising. Across these clusters, coordination-style work — convenings, roadmaps, governance documentation, advocacy campaigns — dominates the funded project list. Rescue and collection efforts, including some remarkable grassroots mobilizations, are also well represented in the list.</p><p>What is almost entirely absent in the projects and initiatives we are tracking is investment in the technical infrastructure layer:&nbsp; the tools, pipelines, standards, systems, and people that make any of the other work durable. Of the 87 projects and initiatives we are tracking, 34 involve coordination activities, 28 involve advocacy, and 33 involve data collection, while only 16 are actively building and maintaining technical tools, and more than half of those are building for their own project's needs rather than building shared, reusable infrastructure.&nbsp;</p><p>Those who are conducting “data rescue” or aiming for “data resilience,” depend on tools, pipelines, and standards to do their work, but those elements often are not being resourced through these funding streams. Rescued data has to be collected, and then it has to go somewhere; it has to be validated, described, made findable, and kept accessible. The systems that do that work, from web archiving tools to identifier infrastructure, from metadata standards to format normalization, and from discovery layers to storage and preservation activities, are either assumed to already exist and be adequately supported, or they are simply not visible in the investment picture.&nbsp;</p><p>The danger is hidden: much of the content that is being rescued is flowing onto known and trusted infrastructures, e.g., Internet Archive, ICPSR and Data LUMOS, and Texas Advanced Computing Center (TACC). The longevity of these environments is unknown and they are not being funded to take on a permanent or future-facing role for this content. What does that mean for the maintenance and reliability of these resources over time?</p><p>The infrastructure layer is chronically underfunded relative to the work it enables, with costs hidden through volunteerism and institutional subsidy. The data resilience funding landscape is replicating that failure under crisis conditions.</p><h2 id="learning-from-previous-experiences"><strong>Learning from previous experiences</strong></h2><p>This is, of course, far from the first time activist communities have mobilized to rescue endangered digital collections at scale. For example, the 2016 data rescue movement, responding to the first round of federal data threats, activated DataRefuge, Preserving Electronic Government Information (PEGI), and the Environmental Data and Governance Initiative (EDGI), among other efforts. Libraries, cultural heritage professionals, and researchers have also mobilized around a myriad of events over time, from the Estonian cyberattacks of 2007 to the conflict zones in Sudan, Syria, Ukraine, and Afghanistan; and from responses to the British Library ransomware attack of 2023 to rescue efforts in the wake of wildfires, floods, and other major natural disasters. Each of these efforts generated hard-won knowledge about what works — distributed custody models, community triage processes, the limits of dark archiving sans deliberate attention to access, and the differences between rescue and preservation.&nbsp;</p><p>That knowledge is not being systematically drawn on in many of today’s projects, and the current US focus is producing efforts that are, in some cases, re-learning lessons that have already been learned elsewhere. We need to actively encourage and fund work that connects today's actors to this history.&nbsp;</p><p>For example, do we know what parts of the 2016-17 data rescue tooling and environment still exist, and what parts do not? Do we know what characteristics differentiate still-active groups like EDGI, ICPSR, and Internet Archive from other groups and toolsets that have sunset? Knowing more about the past can help us to invest more wisely and sunset more adeptly, establishing stronger scaffolding for the field. That requires directing some portion of the funding towards practical analysis of these past efforts.&nbsp;</p><h2 id="duplication-of-backbone-human-infrastructure"><strong>Duplication of backbone (human) infrastructure</strong></h2><p>Our analysis shows a third gap that is perhaps the most costly today: instead of reinforcing the successful networks that already exist and distilling models from those that we can activate in other fields and disciplines, much of the funding available today is flowing to external contractors and consultants to build landscape analyses in order to better understand where the connectors and touchpoints between current efforts are.&nbsp;</p><p>Unintentionally, these external groups are both duplicating and undermining long-standing, often mature or maturing networks on the ground. While the contractors and consultants entering this space for the first time gain remarkable resourcing, existing network players, including strong hubs of activation, struggle for survival, with staff living from contract to contract, and no organizational runway available to them to enable the strategic planning or expansion work they may be uniquely suited to do. In other words, rather than directing funds to strong cornerstone players (e.g., EDGI, PEDP) in support of the (volunteer-driven) bridge building they already have underway, resources are going to external and/or new groups.</p><p>This problem is structural, not intentional, but it still deletes resources from the grassroots actors and specialists in the field in at least three ways. It happens first as dollars flow out of the information management ecosystem to fuel the work of often expensive, deliberately temporary players that promise to provide a system-level view and unbiased recommendations. It happens again, as those consultants build their own maps and recommendations only by drawing heavily on the time and knowledge of the very specialists they are shadowing — asking them to explain what they do, who the key voices are, who is and isn’t connected, and where funding is most needed. And then it happens again as the specialists don’t spare attention for funding calls and proposal deadlines because they are spending their time doing the work.&nbsp;</p><h2 id="an-unresolved-strategic-question"><strong>An unresolved strategic question</strong></h2><p>Underneath the funding gaps lies a more fundamental question that the field has not yet answered: what is today’s data rescue and resilience effort actually building toward? There is a significant difference between building <strong><u>survival infrastructure</u></strong> to keep specific repositories and datasets alive through the current period of threat, versus building <strong><u>perpetuity infrastructure</u></strong> – systems and services that repositories and datasets can plug into for long-term, stable access regardless of political conditions or other threats. Both are needed. Neither is being resourced with that distinction in mind.</p><p>Key, instructive models and seasoned players, ones that have actively solved this problem for adjacent challenges, seem to be missing from the current networks that are assembling: for example, CLOCKSS and Portico for journal content, DataCite for research data identifiers, Crossref for publication metadata, OpenAlex as an open catalog of scholarly works, and Digital Preservation Coalition and National Digital Stewardship Alliance for community standards. Each of these longstanding and well-embedded infrastructure elements has succeeded by defining a function, building infrastructure to serve it collectively, and creating a governance and funding model that distributes the cost across many stakeholders. None of them emerged from crisis response; all of them were built in periods of relative stability. Many of them could provide strong, grounded perspective and leadership to help tie today’s activity to years of preservation practice.&nbsp;</p><p>The current moment is not stable, but it is not too early to ask what long-term preservation and access infrastructure for federal data looks like, and to begin directing some portion of crisis-response investment toward building it.&nbsp;</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/07/tim-mossholder-2RUZ3S32hi8-unsplash.jpg" class="kg-image" alt="Wood texture on a tree. The bark diverts in a curve to reinforce a weak spot. " loading="lazy" width="2000" height="1333" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/07/tim-mossholder-2RUZ3S32hi8-unsplash.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/07/tim-mossholder-2RUZ3S32hi8-unsplash.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/07/tim-mossholder-2RUZ3S32hi8-unsplash.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w2400/2026/07/tim-mossholder-2RUZ3S32hi8-unsplash.jpg 2400w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Photo by </span><a href="https://unsplash.com/@timmossholder?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Tim Mossholder</span></a><span style="white-space: pre-wrap;"> on </span><a href="https://unsplash.com/photos/brown-tree-trunk-2RUZ3S32hi8?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"><span style="white-space: pre-wrap;">Unsplash</span></a></figcaption></figure><h2 id="what-this-means"><strong>What this means</strong></h2><p>The data resilience field is doing to data rescue what scholarly communications has done to open infrastructure for thirty years: funding the visible and easy-to-narrate activities while underinvesting in the operational and technical layer that makes those activities durable. Coordination without tools is a map without roads. Rescue without perpetuity infrastructure is a temporary holding action, not a long-term solution.</p><p>Three specific redirections could materially improve the current investment picture.&nbsp;</p><ul><li>First, dedicated funding for the tools layer, including web archiving software, ingest pipelines, metadata infrastructure, and identifier systems needs to appear in funding portfolios alongside rescue and coordination projects. Both types of funding are needed, and neither can be substituted for the other.&nbsp;</li><li>Second, structured knowledge exchange with communities that have done this work before, e.g., the digital preservation community, the 2016-2017 data rescue efforts, SUCHO and other more recent engagement, should itself be an active, funded transfer-of-knowledge activity. Where are there tools, models, and examples that we might build on, and who can quickly help us identify their value? Funding these groups to do this research could vastly improve our readiness to act in the near future.</li><li>Third, instead of funding external groups to tell us what our experts know, let’s fund the lynchpins that are already in place and doing the work, and let’s free up some of their time to help surface and advance the models they are already successfully using. Too many of the organizations best positioned to lead this work, including the established data stewards and the coordination hubs, are operating on project-to-project funding with no runway for strategic planning. Shoring up those existing human and organizational network agents so that they can shift from crisis response to longer-term strategizing is likely to be the best investment we can make.&nbsp;</li></ul><p>The urgency of this moment is real, and so is the opportunity it offers. The infrastructure we build in response to this crisis will shape what is possible for data access and preservation for a generation. Building it on the same underfunded, fragmented, misaligned model that has characterized open infrastructure for decades would be a preventable mistake.</p><hr><p><em>IOI is actively mapping the data resilience landscape and tracking investment trends across the ecosystem, and as we do so, we are grateful for the help and perspective of many other groups, including American Geophysical Union, Center for Open Science, Data Rescue Network, Internet Archive, Public Environmental Data Partners, and The Data Foundation. Please contact </em><a href="mailto:katherine@investinopen.org" rel="noreferrer"><em>katherine@investinopen.org</em></a><em> to contribute to or build on this analysis.</em></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ LA Referencia ]]></title>
        <description><![CDATA[ Scaling federated open science infrastructure across Latin America ]]></description>
        <link>https://investinopen.org/our-work/la-referencia/</link>
        <guid isPermaLink="false">6a3cbdd3a978af0001e1f28b</guid>
        <category><![CDATA[ Collective Funding ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Tue, 30 Jun 2026 19:02:36 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <h2 id="overview">Overview</h2><!-- START case study metadata -->
<p><strong>Duration:</strong> Ongoing<br>
<strong>Team:</strong> Katherine Skinner, Kaitlin Thaney, Lauren Collister, Madelyn Waterbury<br>
<strong>Funders:</strong> <em>Wellcome, Digital Science, Arcadia, Karger Publishers Foundation, Kahle Austin Foundation, EBSCO, Lyrasis</em><br>
<strong>Skillset:</strong> <em>Business Development Advisory, Capacity Building, Collective Funding</em></p>
<!-- END case study metadata --><p>Latin America has built one of the most successful federated open science networks in the world. <a href="https://www.lareferencia.info/?ref=investinopen.org" rel="noreferrer">LA Referencia</a>'s aggregation and discovery platform connects national repositories across Latin America and Spain, making research from more than 500 institutions easier to share, preserve, and discover. It's a model of what regional coordination can accomplish, and the region is ready to take it further.</p><p>LA Referencia's proposal to the IOI Fund for Network Adoption stood out from more than 100 applications across 22 countries for the scale of its ambition, its track record of effective delivery, and its clear vision for the future. </p><h2 id="what-the-project-will-do">What the project will do</h2><p>LA Referencia will extend its platform to an additional 10 Latin American countries, creating a more inclusive and equitable regional network and bringing national science agencies and their repositories into a shared infrastructure.</p><p>The platform itself is getting a major upgrade. An AI-powered multilingual semantic search system will help researchers discover and connect work across languages, supported by metadata enrichment that opens the door to new metrics and more transparent approaches to research assessment throughout the region. The project will publish close to 6 million enriched metadata records through this system.</p><p>Underneath it all, LA Referencia is building a decentralized persistent identifier system and long-term metadata preservation based on blockchain technology and already running at IBICT in Brazil. A regional Dataverse repository for orphan datasets, along with expanded training and translated documentation, rounds out the work.</p><h2 id="where-ioi-comes-in"><strong>Where IOI comes in</strong></h2><p>Infrastructure at this scale thrives on more than good engineering. It needs a story the world can hear, a business model built to last, and connections to the people who can help it grow. That's where our support is focused.</p><p>We're helping LA Referencia organize and amplify its communications so the project's progress reaches funders, institutions, and peer networks across Latin America and beyond. We're also providing project monitoring to support a shared picture of what's working and where momentum is building.</p><p>Alongside that day-to-day work, we're researching the strategic questions that will shape the platform's future, including licensing models and business model options for the long term, and connecting LA Referencia with other projects and partners across our network.</p><h2 id="what%E2%80%99s-next">What’s next?</h2><p>By the end of the project, LA Referencia's platform will reach further than ever before, with a multilingual discovery layer that helps researchers find work unbound by linguistic barriers and a decentralized preservation network that keeps research outputs permanently accessible and verifiable. Technical capacity grows right along with it.</p><p>The larger outcome is strengthened digital sovereignty for Latin America: research infrastructure that the region's own institutions build, govern, and sustain, with scholarship in Spanish and Portuguese as discoverable as work published anywhere else in the world.</p><h3 id="have-a-vision-for-your-region">Have a vision for your region?</h3><p>Some of the most exciting open infrastructure work happens when institutions pool their strengths across borders. If your network has an ambitious idea and needs the funding, business planning, or connections to bring it to life, we want to hear about it. Get in touch to learn how IOI's Coordinated Funding efforts pair investment with hands-on support. </p><div class="kg-card kg-button-card kg-align-left"><a href="https://investinopen.org/contact-us/" class="kg-btn kg-btn-accent">Contact Us</a></div><hr><h2 id="outputs">Outputs: </h2><ul><li><a href="https://investinopen.org/blog/empowering-networks-advancing-openness-invest-in-open-infrastructure-announces-inaugural-grantees-of-the-ioi-fund-for-network-adoption/" rel="noreferrer">Empowering networks, advancing openness: Invest in Open Infrastructure announces inaugural grantees of the IOI Fund for Network Adoption</a></li></ul> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Infra Finder ]]></title>
        <description><![CDATA[ Building a public tool to help funders and adopters find the open infrastructure they need ]]></description>
        <link>https://investinopen.org/our-work/infra-finder/</link>
        <guid isPermaLink="false">6a3ec205a978af0001e1f43a</guid>
        <category><![CDATA[ Ecosystem Intelligence ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Tue, 30 Jun 2026 02:45:01 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <!-- START case study metadata -->
<p><strong>Duration:</strong> 2024 – present<br>
<strong>Team:</strong> Chrys Wu, Lauren Collister, Emmy Tsang, Sarah Lippincott, Katherine Skinner, Gail Steinhart<br>
<strong>Funders:</strong> Mellon Foundation, Arcadia, Digital Infrastructure Insights Fund, National Science Foundation, Sustaining Circle members<br>
<strong>Skillset:</strong> <em>Ecosystem Intelligence, Landscape Analysis</em></p>
<!-- END case study metadata --><h2 id="overview">Overview</h2><p>The open infrastructure landscape is large and growing but the information needed to navigate it is scattered, inconsistent and held by no neutral party with a view across the whole. There are platforms, standards, repositories, and persistent identifier services doing important work, but they can be difficult to discover and even harder to evaluate side by side. For funders, that made it difficult to direct resources well. For adopters, it made decision-making slower and riskier than it needed to be. And for the infrastructures themselves, it made visibility a persistent challenge.</p><p>IOI, funded by the Mellon Foundation, set out to build something practical: a public tool that gave both funders and adopters the clear, verified, comparable information they needed.</p><h2 id="what-we-did">What we did</h2><p>Before building anything, we needed to understand what funders and adopters actually needed. IOI’s position at the intersection of both stakeholder groups meant that we knew who to talk to and what the real questions were. We spoke with more than 40 institutional budget holders and 26 infrastructure service providers in one-on-one meetings, focus groups, and user testing sessions.&nbsp;</p><p>From there, we built <a href="https://infrafinder.investinopen.org/solutions?ref=investinopen.org"><u>Infra Finder</u></a> as a public-facing product: a searchable, filterable database of open infrastructures that could grow with input from the community. Each entry covers the areas that matter most to evaluators: technical attributes, governance, community engagement, key policies, and funding models — each area a different facet of what it means to be "open".</p><p>Getting the information right required more than a one-time data collection effort. We also built the workflows to source, validate, and keep that information current, working directly with infrastructure providers to make sure what users find in the tool is accurate and up to date. A comparison view lets users look at up to four services side by side, with links to deeper detail where they need it.</p><h2 id="outcomes"><strong>Outcomes</strong></h2><p>Infra Finder launched with coverage of 57 infrastructure services and has since grown to more than 130 thanks to community input.&nbsp;</p><p>For funders, it offers a clearer view of the landscape than existed before, with enough detail to inform real investment decisions. For adopters, it shortens the research process and reduces the risk of choosing infrastructure that turns out to be a poor fit.</p><p>For the infrastructures themselves, it creates a new kind of visibility. As Niels Stern, Managing Director of the OAPEN Foundation, put it: "Participating in this tool provides us with the opportunity to disseminate a lot of information in a very concise way and be more transparent about what we do. This is useful because we can share our entries with institutions, funders, and anyone who wants to work with us or is interested in learning more about our work."</p><p>That transparency matters. Open infrastructure often does essential work quietly, without the visibility that helps funders find it or institutions trust it. Infra Finder was built on the premise that better information, in one place, for everyone, changes that dynamic.</p><h3 id="impact-by-the-numbers">Impact by the numbers</h3><p>In the two years since its inception, Infra Finder has become a go-to resource for open infrastructure diligence in the scholarly community. The site has received over 30,000+ visits, 40,000+ unique pageviews from 130 countries globally since launch. Of these numbers, close to 7,000 returning visits (23% of total visits), highlighting its value as a trusted resource.</p><h2 id="finding-your-way-through-the-landscape">Finding your way through the landscape&nbsp;</h2><p>The open infrastructure ecosystem does not stand still. Platforms emerge, mature, merge, and sometimes close. Funding models shift. The expectations placed on infrastructure by institutions, by policy, and by researchers keep growing. Keeping track of what exists, what it does, and whether it is still the right fit for your needs is genuinely hard work.</p><p>If you're trying to make better decisions about where to invest or what to adopt, <a href="https://infrafinder.investinopen.org/solutions?ref=investinopen.org"><u>Infra Finder</u></a> is a great place to start. And if you need support thinking through what those decisions actually require, we would love to chat!</p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ UbuntuNet Alliance ]]></title>
        <description><![CDATA[ Expanding open access to African scholarship through a regional repository network ]]></description>
        <link>https://investinopen.org/our-work/ubuntunet-alliance/</link>
        <guid isPermaLink="false">6a3cbfd6a978af0001e1f2b4</guid>
        <category><![CDATA[ Capacity Building ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Tue, 30 Jun 2026 02:44:29 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <h2 id=""></h2><!-- START case study metadata -->
<p><strong>Duration:</strong> Ongoing<br>
<strong>Team:</strong> Katherine Skinner, Kaitlin Thaney, Jerry Sellanga, Lauren Collister<br>
<strong>Funders:</strong> <em>Wellcome, Digital Science, Arcadia, Karger Publishers Foundation, Kahle Austin Foundation, EBSCO, Lyrasis</em><br>
<strong>Skillset:</strong> <em>Business Development Advisory, Capacity Building, Collective Funding</em></p>
<!-- END case study metadata --><h2 id="overview">Overview</h2><p>Research from across Africa is growing in volume, ambition, and relevance, and the world benefits every time it becomes easier to find. Open institutional repositories make that happen, keeping scholarship discoverable, citable, and preserved for the long term.</p><p><a href="https://ubuntunet.net/?ref=investinopen.org" rel="noreferrer">UbuntuNet Alliance</a> is well positioned to build that repository layer at regional scale. Its 16 National Research and Education Networks (NRENs) serve roughly 846 institutions across Eastern and Southern Africa, and they hold something no other provider can replicate: deep trust within their research communities. The Alliance's proposal to the IOI Fund for Network Adoption stood out from more than 100 applications across 22 countries for its ambition and its practicality in equal measure.</p><h2 id="what-the-project-will-do">What the project will do</h2><p>Its member NRENs, with coordination by UbuntuNet Alliance, will roll out DSpace as a managed cloud service, giving institutions repository services on infrastructure they already know and trust. More than 560 librarians will train as repository managers and metadata librarians, creating or strengthening over 400 open institutional repositories across the region.</p><p>The project will also build out AfricArXiv as a pan-African aggregator for open research metadata, and train more than 450 data management champions to support a target of over 960 curated open datasets that can be cited, discovered, and reused worldwide.</p><h2 id="where-ioi-comes-in"><strong>Where IOI comes in</strong></h2><p>The strongest open infrastructure pairs good funding with a solid business model, sound governance, and a community that owns it. That's why the Fund was designed to deliver both. Every grant comes with hands-on strategic support, and that's where our work with UbuntuNet Alliance is focused.</p><ul><li><strong>Business planning that fits the region: </strong>We're running business planning workshops and training with NRENs across Eastern and Southern Africa, drawing on IOI's experience in nonprofit revenue models, collective funding approaches, and funder diversification. The goal is for each network to develop a sustainable model that serves its library and researcher communities well beyond the grant period.</li><li><strong>Listening before designing:</strong> We're conducting interviews across the region to build a deeper understanding of local needs and trends in repository development, so that business models and services reflect what institutions actually want rather than assumptions about what they need.</li><li><strong>Connecting to the global community:</strong> DSpace is a mature open-source platform with an active worldwide community and established governance. We're connecting UbuntuNet Alliance directly into that community, so the region's voice shapes the platform's direction and UbuntuNet Alliance's teams can draw on global expertise.</li></ul><h2 id="what%E2%80%99s-next"><strong>What’s next?</strong></h2><p>By the end of the project, the region will have a trusted metadata aggregator of its own, hundreds of new and strengthened repositories, and NRENs equipped with the business models to sustain these services for the long haul. The work aligns with the African Union's Agenda 2063 and the UNESCO Recommendation on Open Science, and it strengthens South-South research collaboration in the process.</p><p>The larger outcome is the one that matters most: African scholarship that is visible, valued, and verifiable, available on infrastructure that the region's own institutions run and govern.</p><h3 id="building-something-at-network-scale">Building something at network scale? </h3><p>Networks and consortia hold a kind of trust that makes open infrastructure adoption move faster and stick. If your network is thinking about how to scale repository services, strengthen sustainability, or connect with funders who back this work, we would love to talk. Reach out to learn more about the IOI Fund for Network Adoption and what integrated funding and strategic support can look like. </p><div class="kg-card kg-button-card kg-align-left"><a href="https://investinopen.org/contact-us/" class="kg-btn kg-btn-accent">Contact Us</a></div><hr><h2 id="outputs">Outputs: </h2><ul><li><a href="https://investinopen.org/blog/empowering-networks-advancing-openness-invest-in-open-infrastructure-announces-inaugural-grantees-of-the-ioi-fund-for-network-adoption/"><u>Empowering networks, advancing openness: Invest in Open Infrastructure announces inaugural grantees of the IOI Fund for Network Adoption</u></a></li></ul><p><a href="https://investinopen.org/strategic-support" rel="noreferrer">&lt;&lt; Back to Our Work</a></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ We have updated our privacy policy ]]></title>
        <description><![CDATA[ Please read the current privacy policy and note your rights under it. ]]></description>
        <link>https://investinopen.org/blog/we-have-updated-our-privacy-policy-2026/</link>
        <guid isPermaLink="false">6a3a8663a1c1c600016376ed</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Invest In Open Infrastructure ]]></dc:creator>
        <pubDate>Tue, 23 Jun 2026 13:15:49 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>Today we published an updated version of our <a href="https://investinopen.org/ioi-privacy-policy/">privacy policy</a>. This affects everyone who uses our websites, engages with our services, or interacts with us online and offline.</p><p><strong>Why this change?</strong></p><p>People, and their privacy, safety and rights, remain at the heart of IOI's work. As our organization and operations have grown, we've revisited the data we collect, how we store it, and how we work with it, to make sure our practices stay robust and continue to serve our community and everyone we interact with. This update reflects that review.</p><p><strong>What should I do?</strong></p><p>Please read the current <a href="https://investinopen.org/ioi-privacy-policy/">privacy policy</a> and note your rights under it. You can find <a href="https://web.archive.org/web/20260623101854/https://investinopen.org/privacy-policy/" rel="noreferrer">the previous version via the Wayback Machine</a>.</p><p>If you have any questions, email us at data-request [at] investinopen [dot] org.</p><p>We review this policy on a regular basis.</p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Building Open Infrastructure That Lasts: A Spotlight on Digital Scholar ]]></title>
        <description><![CDATA[ A case study with Sharon Leon, Co-CEO of the Corporation for Digital Scholarship (Digital Scholar) ]]></description>
        <link>https://investinopen.org/blog/digital-scholar-building-infrastructure/</link>
        <guid isPermaLink="false">6a15e2d2ee259b0001ccfbd2</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Lauren Collister ]]></dc:creator>
        <pubDate>Thu, 28 May 2026 12:39:01 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>In a landscape where open source infrastructure routinely outlasts the funding that created it, the Corporation for Digital Scholarship (also known as Digital Scholar) has spent more than 15 years exploring a different path. Built around flagship tools used by millions of researchers worldwide, and now extending that experience to help other projects find their footing, Digital Scholar offers a model worth understanding. We spoke with Co-CEO Sharon Leon about the organization's origins, its distinctive approach to sustainability, and what they have learned from years of keeping infrastructure — and the humans behind it — going.</p>
<!--kg-card-begin: html-->
<div style = "text-align: center;">
  <img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/05/Leon-Headshot-2019.jpg" width="300" alt="Photograph of Sharon Leon, a white person with gray hair and black square glasses, smiling, wearing a black button up shirt. In the background, a greenscape and a brick academic building.">
  </div>
<!--kg-card-end: html-->
<p><em>Sharon Leon is Co-CEO of the Corporation for Digital Scholarship. She has worked across Digital Scholar's software projects since their inception, originally as faculty at the Roy Rosenzweig Center for History and New Media at George Mason University and then at Michigan State University.</em></p><h2 id="a-humanistic-approach-to-open-infrastructure-sustainability-beyond-grant-funded-beginnings">A humanistic approach to open infrastructure sustainability beyond grant-funded beginnings</h2><p><a href="https://digitalscholar.org/?ref=investinopen.org"><u>Digital Scholar</u></a> was founded in 2009, which surprises people when Leon mentions it. The original purpose was straightforward: find a way to sustain core open source software, specifically <a href="https://omeka.org/?ref=investinopen.org"><u>Omeka</u></a> and <a href="https://zotero.org/?ref=investinopen.org"><u>Zotero</u></a>, so that they could continue to exist without depending indefinitely on grants. The founders' approach was shaped by who they were. Leon and her colleagues were trained as historians, and that perspective shaped how they thought about the people who would use their tools, even as that user base eventually extended far beyond digital humanities or the academy.</p><p>"Zotero has 17 million users around the world," Leon notes. "They're certainly not all humanities scholars. They're in the legal field, the sciences, the corporate world, all over. Our perspective shows us that the users of our software are not necessarily folks who can build their own tools. They're not necessarily super technically adept, though many are." Digital Scholar's insistence on interfaces and documentation that make research tools genuinely accessible — resisting the temptation to assume users can just write a script when they encounter a problem — flows directly from that founding orientation. "What we are keeping in mind are the humans doing the digital work, in addition to the humanistic subject matter," Leon says.</p><p>Their work has expanded beyond those two initial projects. Digital Scholar now stewards five core projects: the original two (Omeka and Zotero), with <a href="https://pressforward.org/?ref=investinopen.org"><u>PressForward</u></a>, <a href="https://tropy.org/?ref=investinopen.org"><u>Tropy</u></a>, and <a href="https://sourceryapp.org/?ref=investinopen.org"><u>Sourcery</u></a> joining through the years. The staff has grown as well, expanding from a very lean operation to a team of roughly 35 people.</p>
<!--kg-card-begin: html-->
<div style = "text-align: center;">
  <a href="https://digitalscholar.org/?ref=investinopen.org"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/05/digital-scholar-logo.png" width="400" alt="Digital Scholar Logo"></a>
  <p align=center><i><a href="https://infrafinder.investinopen.org/comparisons/share/490dc36f-405e-4a2f-b0a7-47a7a21da91a?m=f&ref=investinopen.org">Explore Digital Scholar’s entries on IOI’s Infra Finder.</a> </i> </p>
  </div>
<!--kg-card-end: html-->
<p></p><h1 id="building-sustainability-through-services">Building sustainability through services</h1><p>The sustainability solution the founders landed on was to offer services alongside the software in response to user needs. "People are willing to pay a little bit for storage, or for hosting, or some other helpful service," Leon explains. "And that, in turn, allows us to pay the project teams to keep the software going."&nbsp;</p><p>Central to this approach is a core question: what do people find valuable enough that they're willing to pay for it? The signals come from users themselves, sourced from the Digital Scholar team monitoring closely requests that surface in forums, discussion posts, and GitHub issues; these requests reveal where genuine needs aren't yet being met. "In the case of Omeka, it was a clear sense that not all of our users had the time or interest in setting up a server themselves. We thought, wouldn't it be easier if they could just press a button and start?" That signal led to a popular hosted Omeka service. Similarly, Digital Scholar offers paid storage upgrades for Zotero, making it possible for individuals and groups to sync and share materials.</p><p>The organization's origin story includes an unusual institutional detail that is fundamental to how it operates in order to sell these services. Digital Scholar is not a 501(c)(3) like many nonprofit organizations in the open infrastructure space in the United States. Rather, it is a nonprofit, nonstock corporation in the Commonwealth of Virginia, meaning the state confers nonprofit status, but the organization carries no IRS tax-exempt designation. "We don't have a membership model," Leon says. "We actually sell services. And selling services is a different kind of model; we pay income tax, we collect sales and remit tax, we do all of those things that a regular business would. But we have no owners, no shareholders, and no way to do anything with the income except support our core software projects and support the field."</p><p>This structure also lets Digital Scholar play a direct role in strengthening the broader research ecosystem. Because its business purpose includes support for open source software and open access work in digital scholarship and cultural heritage, Digital Scholar can make direct charitable donations to other nonprofits or it can directly underwrite projects at other institutions (such as its <a href="https://digitalscholar.org/blog/dhnow-and-community/?ref=investinopen.org"><u>support</u></a> of the <a href="https://digitalhumanitiesnow.org/?ref=investinopen.org"><u>Digital Humanities Now</u></a> news outlet at <a href="http://cds.library.northeastern.edu/?ref=investinopen.org"><u>Centers for Digital Scholarship</u></a> at the<a href="https://library.northeastern.edu/?ref=investinopen.org"> <u>Northeastern University Library</u></a>). This model also lets Digital Scholar explore ways to contribute in-kind resources and staff time, or offer access to its business systems and financial infrastructure. This flexibility has opened the door to a new kind of work.</p><p></p><h1 id="extending-the-model-fiscal-support-and-organizational-mentorship">Extending the model: fiscal support and organizational mentorship</h1><p>In recent years, Digital Scholar has begun offering a different kind of support to other projects in the open infrastructure ecosystem. Known within the organization as Slipstream, it is not yet a formal, open program, but two current relationships illustrate the shape of what Digital Scholar is developing.</p><p>The first is with <a href="https://mukurtu.org/?ref=investinopen.org"><u>Mukurtu CMS</u></a>, a cultural heritage platform that has been around nearly as long as Omeka. Mukurtu began offering managed <a href="https://mukurtuhosting.org/?ref=investinopen.org"><u>hosting</u></a> to generate revenue outside of grant funding, and turned to Digital Scholar to handle the operational side: purchasing, server procurement, invoicing, and contracting. "We have that business infrastructure that we can share with them," Leon says. "Net income from running all of those systems get transmitted back to the project to support it." For a mature project like Mukurtu that doesn't need intensive guidance, the relationship is largely practical; Leon describes it as a shared business layer that eliminates duplication of administrative effort.</p><p>The relationship with <a href="http://rightsstatements.org/?ref=investinopen.org"><u>RightsStatements.org</u></a> is more intensive. RightsStatements.org is a critical piece of rights infrastructure for open access to digital cultural heritage, originally born from collaboration among major aggregators including DPLA and Europeana. But over time, through COVID and shifting institutional priorities, its governance structures had become fragile. The project's Interim Steering Committee was dedicated but lacked an independent legal entity and the organizational capacity to sustain the work.</p><p>When RightsStatements.org put out a call for potential new institutional homes, Digital Scholar responded. "We realized how important they are to the global working of understanding access to digital cultural heritage materials," Leon says. The result is an initial three-year relationship in which Digital Scholar is working with the Interim Steering Committee to rebuild governance, creating structures that allow the community to participate in decisions about the statements, their updates, and their translations. The goal is, ultimately, to set the project up for success as its own independent organization.</p><p>The two relationships represent distinct points on a spectrum: with Mukurtu, they are offering shared services building on existing expertise, while with <a href="http://rightsstatements.org/?ref=investinopen.org"><u>RightsStatements.org</u></a>, they are participating in an intense hands-on project of building a governance framework from the ground up. Through these two projects, Digital Scholar is discovering how much capacity it actually has for this kind of support; that knowledge will shape its offerings going forward.</p><p></p><h1 id="planning-infrastructure-for-the-future">Planning infrastructure for the future</h1><p>Open source infrastructure, Leon observes, faces a particular challenge: the software can persist indefinitely, but to do so, the sustainability of the people who work on it must be considered. "Infrastructure may not be coming to the end of its life," she says. "But the humans whose careers have been centered around supporting it — at some point, those people have to be allowed to move on, whether to retirement or to something else. But infrastructure is infrastructure, and people rely on it."</p><p>Her advice: think about sustainability from the beginning. Not as an afterthought, not as a future grant application, but as a foundational design question. There are still too many pieces of the infrastructure the research community relies on that haven't accounted for the long-term.&nbsp;</p><p>For their part, Digital Scholar is entering a new strategic planning cycle, and the agenda ahead reflects the same user-driven logic that has guided the organization since its founding. Leon shared a few developments on the horizon that users can look forward to.&nbsp;</p><p>PressForward, a WordPress plugin for aggregating and curating scholarly gray literature, is in the early stages of being extended into a standalone web service; fewer people run WordPress blogs, and content now moves through Substacks, newsletters, and static site generators. "The world of scholarly communications has changed, and the tool needs to meet users where they are," she says. The team is also thinking carefully about preservation and portability: connectors for ArchivesSpace and Archivematica extend Omeka’s reach into the archival ecosystem, and static site exporters now allow Omeka sites to migrate to Hugo for retirement and/or preservation. On artificial intelligence, Digital Scholar is watching carefully and beginning to support some applications through opt-in plugins — but nothing by default. For instance, Tropy is developing a computer-generated transcription service for research materials. "The users will guide us on the uses they feel comfortable with," Leon says.</p>
<!--kg-card-begin: html-->
<div style = "text-align: center;">
  <img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/05/mcf-doc-03-detail.png" width="600" alt="A split page of text from the software program Tropy. On the left, a type written and hand-annotated page of French text. A handwritten note on the side is highlighted in blue. On the right, computer text transcription in a serif font. A highlighted portion corresponds to the light blue highlighted handwritten text. "></a>
  <p align=center><i>An example of the new handwriting transcription feature in Tropy.</i></p>
  </div>
<!--kg-card-end: html-->
<p></p><p>At the same time that they are cultivating their own tools and communities, Digital Scholar is also actively growing its role as a support and mentorship organization for the broader open infrastructure ecosystem. What Digital Scholar offers (business infrastructure, governance support, hard-won institutional knowledge) exists in service of a field that still has too many projects struggling to survive the gap between initial grant funding and genuine sustainability.&nbsp;</p><p>Building on 15 years of foundational support for critical open infrastructure tools, Digital Scholar is responding to the needs of researchers, creators, and stewards to strengthen the research ecosystem from the ground up.&nbsp;</p><hr><p><em>The Corporation for Digital Scholarship (Digital Scholar) stewards Omeka, Zotero, PressForward, Tropy, and Sourcery, and provides organizational support to RightsStatements.org and Mukurtu. Learn more at</em><a href="https://digitalscholar.org/?ref=investinopen.org"><em> </em></a><a href="http://digitalscholar.org/?ref=investinopen.org"><em><u>digitalscholar.org</u></em></a><em> and visit their entries in </em><a href="https://infrafinder.investinopen.org/comparisons/share/490dc36f-405e-4a2f-b0a7-47a7a21da91a?m=f&ref=investinopen.org"><em><u>Infra Finder</u></em></a><em>. </em></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Surfacing Shared Incentives for Curated Collections and AI Tools ]]></title>
        <description><![CDATA[ We describe the incentives that could lead to the success of a commons-based approach for engagement between AI companies and open collections stewards. ]]></description>
        <link>https://investinopen.org/blog/surfacing-shared-incentives/</link>
        <guid isPermaLink="false">6a05ee968b83cb0001f78169</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Sarah Lippincott ]]></dc:creator>
        <pubDate>Tue, 19 May 2026 12:05:48 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>As part of our research for the <a href="https://investinopen.org/strategic-support/building-resilient-infrastructure-through-dialogue-growth-and-exchange-bridge/"><u>BRIDGE project</u></a>, we recently published a landscape review of the current state of interactions between AI companies and curated collections.&nbsp;</p><p>One thread that we found was that the extant literature nearly uniformly advocates commons-based approaches as alternatives or complements to market-based strategies for open curated collections. The commons approach emphasizes collective coordination, social norm cultivation, and institutional trust-building rather than monetization of individual collections. Unlike licensing deals struck between AI firms and individual publishers, commons-based approaches seek to preserve the open, shared character of knowledge resources while ensuring that the entities profiting most from them contribute meaningfully in return. This perspective positions small collections not as individual vendors but as stewards of interconnected knowledge resources requiring<a href="https://doi.org/10.2139/ssrn.4836354?ref=investinopen.org"> <u>collective protection</u></a>.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/05/pexels-andreas-staver-341167683-28495533-1-.jpg" class="kg-image" alt="Dual bridges spanning the scenic Salt River Canyon in Arizona with striated rock faces and desert grasses." loading="lazy" width="2000" height="1279" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/05/pexels-andreas-staver-341167683-28495533-1-.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/05/pexels-andreas-staver-341167683-28495533-1-.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/05/pexels-andreas-staver-341167683-28495533-1-.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w2400/2026/05/pexels-andreas-staver-341167683-28495533-1-.jpg 2400w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Photo by </span><a href="https://www.pexels.com/photo/historic-salt-river-canyon-bridge-in-arizona-28495533/?ref=investinopen.org"><u><span class="underline" style="white-space: pre-wrap;">Andreas Staver</span></u></a><span style="white-space: pre-wrap;">.</span></figcaption></figure><p>The success of a commons-based approach assumes that AI companies do have incentives to engage with open collections stewards—incentives not predicated solely on legal risk. These incentives include:&nbsp;</p><p><strong>Ensuring high-quality data sources remain online.</strong> Excessive bot traffic and scraping threatens to push some small collections offline, as many lack the resources to "continue adding more servers, deploying more sophisticated firewalls, and hiring more operations engineers in perpetuity" (<a href="https://glamelab.org/products/are-ai-bots-knocking-cultural-heritage-offline/?ref=investinopen.org"><u>Weinberg, 2025</u></a>;<a href="https://www.eff.org/deeplinks/2025/06/keeping-web-under-weight-ai-crawlers?ref=investinopen.org"> <u>Grant, 2025</u></a>). If key data sources go dark, AI companies lose access to the very content that makes their models useful.</p><p><strong>Encouraging competition and discouraging consolidation.</strong> Investment in the commons enables everyone, not only the wealthiest corporations, to build and refine models that lead to innovation. Even larger companies share a broad interest in encouraging innovation in the sector that they can benefit from in the future. &nbsp;In its announcement of its support for Harvard’s Institutional Data Initiative, Microsoft cited the motivation to grow “a vibrant, competitive AI economy” by expanding access to the data resources needed to build LLMs (<a href="https://blogs.microsoft.com/on-the-issues/2024/12/12/supporting-new-open-data-initiatives-institutional-data-initiative-and-core/?ref=investinopen.org"><u>Davis, 2024</u></a>). Openly licensed datasets can encourage competition and offer smaller players a way in. Adoption of the Model Concept Protocol (or MCP, initially developed by Anthropic and later donated to the Linux Foundation) by the major AI market players is an example of industry-wide cooperation in this vein.</p><p><strong>Retain scraping access.</strong> The relationship between AI companies and the broader web is already showing signs of strain, and the consequences of ignoring this dynamic are already visible. A closing off of the web in response to AI crawlers, especially through blunt approaches that do not distinguish them from other machines, is affecting crawling for legitimate and widely accepted purposes, such as archiving and research. As of December 2025, around 5.6M websites had blocked OpenAI's GPTBot, a nearly 70% increase over the previous six months (<a href="https://www.theregister.com/2025/12/08/publishers_say_no_ai_scrapers/?ref=investinopen.org"><u>Claburn, 2025</u></a>). As scraping restrictions increase and more organizations adopt brute force approaches to thwarting bots, AI companies risk losing access to important sources of high-quality, novel training data. Investing in the health of the commons, and in norms that distinguish reasonable use from indiscriminate extraction, can help arrest this trend.</p><p><strong>Sustain high-quality, diverse training data.</strong> Well-stewarded commons provide not just volume but also the metadata, documentation, and quality control that make training data more valuable. Companies may see value in building and sustaining resources that provide them with access to high-value, unique, or novel datasets. Some analyses have speculated that the open web will become polluted by low-quality, machine-generated content, making curated collections increasingly valuable data sources.<a href="https://doi.org/10.48550/arXiv.2508.06470?ref=investinopen.org"> <u>Noroozian et al. (2025)</u></a> write that AI model developers should have a vested interest in making curated collections data "identifiable, visible, and discoverable" in order to avoid 'model collapse' or increasingly more repetitive, biased, and less capable AI" caused by the growing presence of synthetic data across the web. Looking further ahead,<a href="https://scholarlykitchen.sspnet.org/2026/01/21/guest-post-ai-isnt-going-to-pay-for-content-at-least-not-how-youre-hoping-it-will/?ref=investinopen.org"> <u>Woahn (2026)</u></a> predicts that "The next improvements in model capability will come from: highly specialized domain corpora; well-structured technical datasets; targeted refreshes rather than massive new ingestions; data with deep internal organization, not broad volume."</p><p><strong>Foster ongoing human contributions to the open web.</strong> The commons is sustained by the continuous labour and ingenuity of human creators. Without new approaches for providing permission, credit, and compensation, these creators have diminishing incentives to openly share their work, and AI models lose access to original content (<a href="https://arxiv.org/abs/2303.09001v2?ref=investinopen.org"><u>Chan et al., 2023</u></a>,<a href="https://www.amazon.science/publications/fairness-and-welfare-quantification-for-regulating-large-language-models?ref=investinopen.org"> </a><a href="https://arxiv.org/abs/2303.11074?ref=investinopen.org"><u>Huang &amp; Siddarth, 2023</u></a>).<a href="https://doi.org/10.1162/99608f92.c17c3adb?ref=investinopen.org"> <u>Borgman and Groth (2025)</u></a> argue that scholars participate in a gifting economy in which they volunteer labour (such as sharing data) "with the expectation that these gifts create indebtedness, encourage reciprocity, and enhance reputations." To build trust among scholars and collections stewards, AI companies may need to more visibly and concretely adopt the norms of a gifting economy, for example, by ensuring proper attribution.</p><p><strong>Create a positive public image and consumer trust.</strong> Beyond practical considerations, AI model developers may have a reputational incentive to demonstrate a commitment to "ethical" or "responsible" AI, including appropriate data-harvesting practices. The non-profit Fairly Trained, for example, was launched to certify AI model developers and products that adhere to standards for their training data (<a href="https://www.wired.com/story/ai-executive-ed-newton-rex-turns-crusader-stand-up-for-artists/?ref=investinopen.org"><u>Knibbs, 2024</u></a>). They also have an interest in providing consumers with reliable information from robust sources to increase adoption and engagement with their platforms.</p><p><strong>Mitigate regulatory and legal risks.</strong> Finally, the legal landscape surrounding AI and data use remains in flux. Depending on the outcomes of several lawsuits and pending legislation, AI model developers may need to fundamentally alter how they harvest data. If they cannot rely on fair use justifications for scraping copyrighted data, for example, they will be increasingly reliant on openly licensed and public domain data. The strongest incentive for change could be future government policy that regulates the use of openly available data, for example, by strengthening creator opt-outs.</p><p>Despite these shared incentives, the challenge lies in bridging the gap between curated collections stewards and AI companies and developing the sociotechnical infrastructure needed to facilitate cross-sector engagement. Trust between commons communities and AI companies is severely eroded. Open source developer communities have expressed "deep frustration with what they view as AI companies' predatory behaviour toward open source infrastructure," undermining the relationship-building these approaches require (<a href="https://arstechnica.com/ai/2025/03/devs-say-ai-crawlers-dominate-traffic-forcing-blocks-on-entire-countries/?ref=investinopen.org"><u>Edwards, 2025</u></a>). Philosophically, there's tension between ideals of openness and the need for protection. The "open with thoughtfulness" paradigm (<a href="https://rosalynmetz.substack.com/p/openness-has-limits"><u>Metz, 2025</u></a>) requires continuous judgment calls that may fragment the commons into incompatible governance zones.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://doi.org/10.5281/zenodo.19458128?ref=investinopen.org" class="kg-btn kg-btn-accent">Read the full landscape review</a></div><p>Addressing these challenges will require deliberate effort on multiple fronts. The commons needs norms, governance frameworks, and contribution models developed with input from a range of stakeholders, including AI companies and technology platforms, as well as researchers, creators, and open curated collections stewards. The BRIDGE project is ongoing and represents IOI’s current efforts to identify and foster relationships between these currently disparate groups, and to pilot new approaches to encourage reciprocity that benefits all stakeholders in the AI economy.&nbsp;</p><hr><p><em>This post is an excerpt from our longer work, "</em><a href="https://doi.org/10.5281/zenodo.19458128?ref=investinopen.org"><em><u>Sustaining the Commons in the AI Economy: A Landscape Scan of Challenges and Strategies for Bridging AI Companies and Open Curated Collections</u></em></a><em>."&nbsp;</em></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Highlights from Sustaining the Commons in the AI Economy ]]></title>
        <description><![CDATA[ A conversation with IOI’s Research team about the latest report ]]></description>
        <link>https://investinopen.org/blog/highlights-from-sustaining-the-commons-in-the-ai-economy/</link>
        <guid isPermaLink="false">69ef7e01b701030001323bd9</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Lauren Collister ]]></dc:creator>
        <pubDate>Tue, 28 Apr 2026 09:22:49 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>Last week, Invest in Open Infrastructure (IOI) released our first publication from our<a href="https://investinopen.org/strategic-support/building-resilient-infrastructure-through-dialogue-growth-and-exchange-bridge/"> <u>Building Resilient Infrastructure through Dialogue, Growth, and Exchange (BRIDGE) project</u></a>, which aims to bring open research infrastructures, cultural collections, and commercial tech organizations together to build mutually beneficial partnerships in the AI era. <a href="https://zenodo.org/records/19458128?ref=investinopen.org" rel="noreferrer"><em>Sustaining the Commons in the AI Economy</em> </a>is an overview of our research on the key themes and current approaches to resolving the current tensions in this ecosystem.&nbsp;&nbsp;</p><p>As part of our project release, IOI’s Chrys Wu interviewed the report’s authors (Sarah Lippincott, Lauren Collister, and Katherine Skinner) about their perspectives on the main takeaways from this report. This interview is now available to watch or listen to.&nbsp;</p><figure class="kg-card kg-embed-card"><iframe width="200" height="113" src="https://www.youtube.com/embed/bOqR2V_zBfM?feature=oembed" frameborder="0" allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture; web-share" referrerpolicy="strict-origin-when-cross-origin" allowfullscreen="" title="Sustaining the commons in the AI economy - Roundtable discussion"></iframe></figure><p>We hope you enjoy this interview with the authors, and we would like to hear your take-aways as well. What parts of the report most resonate with you? Send us a message at research @ investinopen dot org, or respond to <a href="https://www.linkedin.com/posts/invest-in-open_highlights-from-sustaining-the-commons-in-activity-7454929939272724480-XKsn?ref=investinopen.org" rel="noreferrer">our LinkedIn post</a> to share your view.&nbsp;</p><p>Read about the full report at <a href="https://investinopen.org/blog/sustaining-the-commons-in-the-ai-economy/"><em><u>Sustaining the Commons in the AI Economy: A Landscape Scan of Challenges and Strategies for Bridging AI Companies and Open Curated Collections.</u></em></a></p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Sustaining the Commons in the AI Economy: A Landscape Scan of Challenges and Strategies for Bridging AI Companies and Open Curated Collections ]]></title>
        <description><![CDATA[ Read initial findings from our BRIDGE project. ]]></description>
        <link>https://investinopen.org/blog/sustaining-the-commons-in-the-ai-economy/</link>
        <guid isPermaLink="false">69e62e6006b2fb0001bf78b9</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Lauren Collister ]]></dc:creator>
        <pubDate>Tue, 21 Apr 2026 13:30:46 +0000</pubDate>
        <media:content url="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash-1.jpg" medium="image"/>
        <content:encoded><![CDATA[ <p>As part of our <a href="https://investinopen.org/strategic-support/building-resilient-infrastructure-through-dialogue-growth-and-exchange-bridge/"><u>Building Resilient Infrastructure through Dialogue, Growth, and Exchange (BRIDGE) project</u></a>, Invest in Open Infrastructure (IOI) is pleased to release the first outcome of our research: a landscape scan of the challenges and strategies for bridging AI companies and open curated collections.&nbsp;</p><div class="kg-card kg-button-card kg-align-center"><a href="https://doi.org/10.5281/zenodo.19458128?ref=investinopen.org" class="kg-btn kg-btn-accent">Read the full report</a></div><hr><p></p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash.jpg" class="kg-image" alt="An image of a suspension bridge over water, fading into mist." loading="lazy" width="2000" height="1334" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w2400/2026/04/modestas-urbonas-vj_9l20fzj0-unsplash.jpg 2400w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Photo by</span><a href="https://unsplash.com/@modestasu?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"> <u><span class="underline" style="white-space: pre-wrap;">Modestas Urbonas</span></u></a><span style="white-space: pre-wrap;"> on</span><a href="https://unsplash.com/photos/golden-gate-bridge-san-francisco-california-vj_9l20fzj0?utm_source=unsplash&utm_medium=referral&utm_content=creditCopyText"> <u><span class="underline" style="white-space: pre-wrap;">Unsplash</span></u></a></figcaption></figure><p>Curated collections, such as digital archives, open access journals, scientific data repositories, preprint servers, and knowledge graphs, form part of the digital commons: the shared pool of open resources made freely accessible online. Used regularly by researchers, the public, and commercial entities, this digital commons benefits every sector of society. Maintained largely by academic institutions, nonprofits, governments, and volunteer communities, these curated collections represent a public good built on decades of labour, public and private funding, and an ethos of open knowledge sharing. Their sustainability is inseparable from the sustainability of open science and democratic access to information.</p><p>That infrastructure is now under strain. The rapid expansion of AI development has turned open curated collections into an increasingly valuable source of training data. Automated bots now generate traffic that, in some cases, exceeds human visits, overwhelming servers, inflating bandwidth costs, and triggering outages. Meanwhile, a surge of AI-generated content submissions threatens to overwhelm editorial and curation workflows at repositories that rely on community contributions.&nbsp;</p><p>To understand this landscape, key themes explored in the report include:</p><ul><li><strong>The strain on infrastructure</strong>: How AI bot traffic is overwhelming servers, inflating costs, and triggering service disruptions at open curated collections.</li><li><strong>The limitations of current approaches</strong>: Why technical, legal, and market-based mechanisms face significant challenges in protecting the commons at scale.</li><li><strong>The risks of defensive restrictions</strong>: How access controls intended to protect collections may paradoxically accelerate data consolidation among well-resourced corporations.</li><li><strong>The tension of voluntary compliance</strong>: Many current approaches rely on voluntary compliance, with legal repercussions as enforcement mechanisms; what else could be done to encourage cooperative behaviour? Wrestling with whether enlightened self-interest can succeed where legal and technical frameworks have fallen short.</li><li><strong>The possibilities of the commons</strong>: AI companies and collection stewards both have much to gain from exploring the commons as a shared space for investment and cooperation.&nbsp;</li></ul><p>The report points toward a promising, if demanding, path: commons-based governance grounded in reciprocal norms and shared interests across sectors. AI companies have concrete reasons to want the commons to survive. The loss of reliable data sources reduces the quality and diversity of training data; an increasingly walled-off web raises legal and regulatory risks; growing public frustration with extractive practices creates reputational pressure. Investment in the commons, properly framed, is investment in the quality of AI itself.</p><p>Yet a central tension remains unresolved. Whether enlightened self-interest will prove more effective than the legal and technical mechanisms that have already fallen short is an open question. To answer it requires engaging the curators and consumers of open collections as stewards of the digital commons, co-creating partnership models that could align open knowledge strategies with commercial demand.</p><div class="kg-card kg-button-card kg-align-center"><a href="https://doi.org/10.5281/zenodo.19458128?ref=investinopen.org" class="kg-btn kg-btn-accent">Read the full report</a></div><p>We're eager to hear from interested parties and continue the conversation. If you'd like to be part of further discussions on this topic, please contact us at research [at] investinopen [dot] org.</p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Open infrastructure for better health: Proyecto ARPHAI and digital healthcare transformation in Argentina ]]></title>
        <description><![CDATA[ Proyecto ARPHAI shows how open, locally hosted infrastructure can enable secure, data-driven public health in resource-constrained settings ]]></description>
        <link>https://investinopen.org/blog/open-infrastructure-for-better-health-proyecto-arphai-and-digital-healthcare-transformation-in-argentina/</link>
        <guid isPermaLink="false">69e617b306b2fb0001bf7893</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Jerry Sellanga ]]></dc:creator>
        <pubDate>Mon, 20 Apr 2026 13:12:23 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>Digitized health records enable monitoring of health trends, support more targeted resource allocation, and enable rapid responses to disease outbreaks. In Argentina, use of technology in healthcare research has been hampered by reductions in federal science funding and runaway inflation in the last few years, which have made acquiring and maintaining digital health infrastructure increasingly prohibitive.</p><p>In response,<a href="https://www.ciecti.org.ar/arphai/?ref=investinopen.org"> <u>Proyecto ARPHAI</u></a> sought to develop a project to accelerate and promote the ethical use of digitized health records nationwide. Led by CIECTI (the Interdisciplinary Centre for Science, Technology, and Innovation Studies), ARPHAI researches tools based on artificial intelligence and data science that can be applied to electronic medical records. By strengthening both the analytical value of health records and the technical and human capacity to work with health data, ARPHAI contributes to earlier detection of epidemic outbreaks and supports more timely, preventive public health decision-making.</p><h2 id="collaboration-with-ioi"><strong>Collaboration with IOI</strong></h2><p>In 2023, IOI opened the call for proposals for the<a href="https://investinopen.org/funding-pilots/oi-fund/"> <u>Open Infrastructure Fund</u></a>, which aimed to support the development of open research infrastructure services to strengthen sustainability, resilience, and adoption of open infrastructure. Proyecto ARPHAI was selected as one of the eight grantees of the fund, receiving funding of around US$18,000 to develop and sustain the infrastructure for processing and storing sensitive data from electronic health records in Argentina.</p><h2 id="implementation-and-impact"><strong>Implementation and Impact</strong></h2><p>The one-and-a-half-year project, which commenced in March 2024, was implemented in collaboration with the<a href="https://supercomputo.unc.edu.ar/?ref=investinopen.org"> <u>Centro de Computación de Alto Desempeño (CCAD)</u></a>, a high-performance computing facility at the Universidad Nacional de Córdoba (UNC), Argentina. A key deliverable was the acquisition of a high-capacity 'fatnode' server — affectionately nicknamed <em>Gordito</em> (Spanish for 'chubby') — equipped with 1 terabyte of RAM and support for 8 next-generation GPUs. The hardware provides the core computing infrastructure required to host and process electronic health records in a secure environment. The servers will support data storage, analysis, and the development of artificial intelligence models, enabling researchers to work with sensitive health data while maintaining responsible data governance. The server runs within CCAD's Kubernetes Cluster, hosting both Project Jupyter tools for interactive computing and Ollama for running large language models. The additional computing capacity has also benefited the broader teaching and research community at CCAD, extending its value well beyond health data applications.</p><p>“<em>This server is a key piece for the JupyterHub at CCAD, which was recently inaugurated, because it is the one that makes it possible to run notebooks with one terabyte of RAM and up to four GPUs, providing computing power for the largest projects. The facility has also been a big boost for our wider research community at CCAD as we now have more computing power available even for researchers in other areas beyond healthcare</em>,” commented Nicolás Wolovick, Director of CCAD.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/Screenshot-2026-04-20-at-15.19.21.png" class="kg-image" alt="Photo off the implementation team at the University of Cordoba" loading="lazy" width="932" height="675" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/04/Screenshot-2026-04-20-at-15.19.21.png 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/Screenshot-2026-04-20-at-15.19.21.png 932w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">From left to right, Nicolás Wolovick (Director of CCAD), Verónica Xhardez and Laura Alonso Alemany (researchers from ARPHAI) at the Data Center located on the grounds of the National University of Córdoba.</span></figcaption></figure><p>To ensure that sensitive data stored on national servers yields actionable policy insights, the ARPHAI team organized two training sessions for government health officials in May and July 2025. These interactive sessions gave officials a solid foundation for working with sensitive electronic health data and for leveraging insights derived for the public good. Extensive documentation and training materials in Spanish were developed and made freely available to the Argentine research community under CC-BY licences, hosted on both the<a href="https://github.com/ARPH-AI?ref=investinopen.org"> <u>ARPHAI repository</u></a> and<a href="https://zenodo.org/communities/arphai/?ref=investinopen.org"> <u>Zenodo</u></a>.</p><p>The importance of using local servers when dealing with a population’s health data is key to ensuring digital sovereignty and the protection of personal data, and strengthens states’ ability to manage data autonomously and in line with public interest objectives.</p><p><em>“It was a great pleasure for us to share our experience and help pave the way for those who face the daily challenge of working with sensitive data. It was doubly rewarding, not only because of the interest and feedback generated by the training, but also because it was an achievement in itself to get people so busy with the demands of daily management to take a moment to reflect on and analyze the impact of their work and their approach to it. Today, in a context where science funding faces significant challenges (particularly in countries like Argentina), these kinds of initiatives are especially valuable, as they contribute to sustaining and strengthening established research groups with demonstrated expertise and ongoing activity,” remarked Sabrina Lopez from Proyecto ARPHAI.</em></p><h2 id="lessons-learnt"><strong>Lessons Learnt</strong></h2><p><strong>Technical investment must be matched by human investment</strong><em>.</em> The project design accounted not only for hardware but also for the capacity building needed to ensure long-term sustainability. Projects that focus solely on the technical side often face sustainability challenges. Community buy-in and well-trained personnel are equally critical to success.</p><p><strong>The importance of partnerships in funding and sustaining open infrastructure.</strong> In recent years, funding for the science and education sector in Argentina has been cut. This, in turn, has created a massive financial deficit. From the project, we can see exemplary collaboration between different stakeholders (universities, research organizations, infrastructure services, and non-profits) to work together for the common good by pooling resources, which needs to be emulated.</p><p>ARPHAI's vision was realized through its partnership with CCAD, not in isolation. CCAD hosts the project’s infrastructure and data and supports the development of an active community of practice. CCAD had supported ARPHAI's research for two years before the IOI funding was secured — a testament to the importance of cultivating strong, mutually beneficial relationships in advancing societal impact.</p><p><strong>Embed, don't impose.</strong> Infrastructure investments are more likely to see organic adoption when they flow through organizations and networks already embedded in their communities. These organizations bring the trust and contextual understanding that funders can't manufacture — they know what to build, for whom, and why it will actually get used.</p><h2 id="conclusion"><strong>Conclusion</strong></h2><p>Proyecto ARPHAI demonstrates what becomes possible when technical ambition is matched by the right partnerships, development of human capacity, and a commitment to openness. The challenge now is to scale this model to ensure that other sectors beyond healthcare can leverage open infrastructure for broader, sustained impact.</p> ]]></content:encoded>
    </item>
    <item>
        <title><![CDATA[ Infrastructure Showcase: 2i2c ]]></title>
        <description><![CDATA[ A conversation with 2i2c’s Jim Colliander on the recent strategic consulting partnership with IOI aimed at strengthening their business development capacity and accelerate progress toward product-market fit. ]]></description>
        <link>https://investinopen.org/blog/infrastructure-showcase-2i2c/</link>
        <guid isPermaLink="false">69e22f5206b2fb0001bf7867</guid>
        <category><![CDATA[ Blog ]]></category>
        <dc:creator><![CDATA[ Jerry Sellanga ]]></dc:creator>
        <pubDate>Fri, 17 Apr 2026 13:43:00 +0000</pubDate>
        <media:content url="" medium="image"/>
        <content:encoded><![CDATA[ <p>Last year, Invest in Open infrastructure embarked on a four-month engagement (June-October 2025) with <a href="https://2i2c.org/?ref=investinopen.org"><u>2i2c</u></a> aimed at providing executive coaching and strategic advisory services to strengthen its business development capacity and accelerate progress toward product-market fit. 2i2c is a non-profit organization that designs, develops, and operates interactive computing environments that facilitate workflows for open science and education in the cloud.</p><p>IOI's Director of Development, Emma Green led the engagement, working with 2i2c's Business Development Manager, Jim Colliander, and supported by Chrys Wu, Solutions Strategist at IOI. Jim reflected on what the partnership looked like in practice, what it made possible, and the key takeaways from the engagement. Below are some highlights.</p><div class="kg-card kg-callout-card kg-callout-card-yellow"><div class="kg-callout-text">For infrastructures looking to improve health and operational sustainability and/or scale services, IOI can provide critical capacity, business model guidance, and strategic consulting. Be it strengthening your governance, financial sustainability, organizational structure, and/or stakeholder engagement, we have over 20 years of experience in our team and are here to help — <a href="https://form.asana.com/?k=N0ZA9_gjiVKCTZh7PZyf2Q&d=1204039279428915&ref=investinopen.org"><u>get in touch today</u></a> if you’d like to find out more.</div></div><p><strong>When IOI first began working with 2i2c, what were the most urgent business or growth challenges you were trying to solve, and how were those affecting your ability to operate sustainably?</strong></p><p><strong>Jim:</strong> The core challenge was straightforward but not easy: we needed a credible path to organizational sustainability. Since 2i2c’s formation in 2020, we have built an initial base of recurring revenue. However, much of that recurring revenue had been secured through founders' networks and warm referrals; a pattern the team recognized as unsustainable.&nbsp;</p><p>At the same time, we were navigating a significant internal transition. In early 2025, we moved from offering a managed JupyterHub service to a membership model, partly to make the full range of value we provide more explicit. That shift raised harder questions about what kind of organization we actually wanted to be. Some of the team saw our path to scale as becoming more like a SaaS product. Others felt it required more consultancy-style, bespoke engagement. IOI helped us work through that tension and arrive at a shared definition of what scaling actually means for 2i2c — which turned out to be important for our next steps.</p><p><strong>What key assumptions about your role and approach to market, pricing, or go-to-market approach did IOI help you test or rethink, and which of those turned out to be the most important?</strong></p><p><strong>Jim: </strong>IOI guided 2i2c to take a more hypothesis-driven approach to go-to-market strategy, starting with identifying a key assumption that 2i2c had made: that 2i2c could successfully sell directly to Premier-tier customers through cold outreach. IOI helped us think about go-to-market with improved terminology and a scientific approach that centers user needs through user discovery processes.</p><p>What we discovered surprised us. Customers were not only seeking access to hubs. They were looking for genuine engagement with the 2i2c team, connections to other 2i2c communities, and a sense of participation in something larger.</p><p>On a personal level, the engagement also shifted my perception of my role. I had been conflating three distinct functions: the Farmer (focused on retention and expansion), the Hunter (focused on net new sales), and the Business Development Lead (focused on strategy and building growth systems). Recognizing which hat I was wearing at any given moment changed how I thought about scaling sales.</p><p><strong>How did IOI’s engagement influence conversations/enhance alignment between business development, product, and engineering within 2i2c?</strong></p><p><strong>Jim: </strong>We discovered that we had real internal misalignment — not just on strategic direction, but on how commitments from the sales side were being translated into scoped work for product and engineering, and on how customer feedback was feeding back into what we built.</p><p>IOI recommended a recurring customer insights meeting that brings together business development, product, and delivery teams to create a structured cross-functional feedback loop. This has been implemented and we are already seeing improvements in the business-to-product feedback loop within 2i2c. IOI introduced and documented sales ceremonies. We now have a daily Business Development standup, a bi-weekly business strategy meeting, and a weekly engagement meeting. IOI’s Emma joined some of 2i2c’s BD meetings near the end of the engagement to provide real-time coaching and follow-up guidance.</p><figure class="kg-card kg-image-card kg-card-hascaption"><img src="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/jacky-watt-1pkqYTQxwyI-unsplash.jpg" class="kg-image" alt="Photo by Jacky Watt on Unsplash" loading="lazy" width="1920" height="1277" srcset="https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w600/2026/04/jacky-watt-1pkqYTQxwyI-unsplash.jpg 600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1000/2026/04/jacky-watt-1pkqYTQxwyI-unsplash.jpg 1000w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/size/w1600/2026/04/jacky-watt-1pkqYTQxwyI-unsplash.jpg 1600w, https://storage.ghost.io/c/e3/0f/e30f355d-2d35-423e-a82d-26c811a07bf5/content/images/2026/04/jacky-watt-1pkqYTQxwyI-unsplash.jpg 1920w" sizes="(min-width: 720px) 720px"><figcaption><span style="white-space: pre-wrap;">Vingage VW Beetle</span></figcaption></figure><p><strong>Can you describe a concrete change in how 2i2c now approaches the BD role, sales, partnerships, or product design that grew directly out of this coaching and advisory work?</strong></p><p><strong>Jim: </strong>IOI helped us develop clear written role definitions distinguishing the Account Manager function from the Business Development Lead function. That clarity reduced some friction.</p><p>We also adopted a SAM/TAM horizon framework — mapping our Serviceable and Total Addressable Markets across a three-horizon planning timeline — which shifted our thinking from quarterly wins to a two-to-three year view of market expansion. That longer lens has been useful.</p><p>We built a practice of customer discovery. After IOI coached us on survey design, bias, and validation thresholds, I ran a structured customer survey experiment with eleven respondents. The results shaped our product direction and improved our internal consensus.</p><p><strong>Looking ahead, how has this work reshaped your confidence to contribute to and drive and what ‘scaling’ realistically means for 2i2c as an open infrastructure organization?</strong></p><p><strong>Jim:&nbsp; </strong>The most clarifying moment in the entire engagement was probably the reframing of what "scale" actually means for an organization like ours. Our first target isn't AWS-level growth — it's building repeatable, legible processes for five to ten customers of the same type. That sounds modest, but getting there requires exactly the kind of disciplined system-building we needed..</p><p>Emma's coaching shifted something more fundamental too. My focus used to be on getting things done; now it's on building evidence-based systems and knowledge that the whole team can work from. That's a different job in some important ways. The 2i2c team is more focused now on learning and building strong internal habits than on task completion — and I think that's the right foundation for where we want to go.</p><p><strong><em>This post is part of our “Infrastructure Showcase” series. To stay updated on posts from this series and more from Invest in Open Infrastructure, please </em></strong><a href="https://share.investinopen.org/newsletter?ref=investinopen.org"><strong><em><u>sign up for our newsletter</u></em></strong></a><strong><em>. Interested in IOI’s strategic consulting services to further research infrastructure health, sustainability, and growth on these topics or related areas of work? </em></strong><a href="https://form.asana.com/?k=N0ZA9_gjiVKCTZh7PZyf2Q&d=1204039279428915&ref=investinopen.org"><strong><em><u>Get in touch!</u></em></strong></a><strong><em>&nbsp;&nbsp;</em></strong></p> ]]></content:encoded>
    </item>

</channel>
</rss>
