Video summary
The 13th CONNECT Community Call focused primarily on the new statistical features introduced to Opener Connect for organizations with premium plans, alongside recent improvements in page efficiency and bot defense mechanisms. Since March 2026, four service releases have been implemented, with a specific highlight given to statistics available as of June that were further refined recently. These tools allow institutions to monitor key indicators regarding openness, findability, and fairness by calculating percentages of publications with persistent identifiers or open licenses. Additionally, the system tracks how research objects are interconnected, such as linking publications to datasets or software, providing a visual representation of scientific collaboration through maps based on co-author countries and project grants.
Users can interact deeply with these data visualizations by viewing charts in full screen, printing them, downloading images in various formats, or exporting the underlying CSV and Excel files for further analysis. The interface also offers flexible visualization options, allowing users to switch between bar charts and box plots depending on preference. Beyond general statistics, new indicators detail where researchers publish their work across different open access models like Diamond, Gold, and Transformative journals, as well as which repositories host their datasets and software. For instance, the system can show that Harvard Dataverse is currently hosting more data for a specific institution compared to others, while also listing software hosted on platforms like Zenodo or Kodiosan.
The second major segment of the call addressed the three-year history and future roadmap of the Netherlands Research Portal, which aims to unify Dutch research outputs in one place. Originally launched two decades ago with the goal of reducing subscription costs through green open access, the portal has evolved under OpenAire Connect to serve as a federated infrastructure that enriches local repositories while maintaining institutional autonomy. The project is now supported by a mandate from university boards and funded via the Dutch Open Science Fund to ensure its long-term sustainability. This initiative integrates with the European Open Science Cloud node, enabling seamless data transfer between storage facilities like SURF Research Drive and compute resources, thereby creating a robust ecosystem where metadata remains in the central graph while full texts are preserved locally or nationally.
During the Q&A session, participants clarified how organizations can be included in search results for portals that previously operated on a repository-only basis rather than an organization-specific list. The presenters explained that administrators can manually add specific repositories to the configuration of a portal like the Dutch Research Portal, which will subsequently make affiliated organizations searchable and include their products even if not directly deposited in institutional archives. Furthermore, users were informed about granular control over harvesting settings, allowing repository managers to select exactly which datasets they wish OpenAire to index rather than submitting everything automatically. The session concluded with an invitation for continued collaboration on expanding the federation's capabilities and utilizing the extensive graph data, which currently contains millions of records across European partners including Italy, Poland, Portugal, and Greece.
Read the full video transcript
Okay. So, I'm Allesi. Uh, I'm a
researcher at the Italian National
Research Council and I'm the service
manager of Opener Connect. So, most of
you already know me. Uh, the agenda was
already presented. So, I will go
straight to the first item which are the
novelties of the service since March
2026,
which is when we had our last community
call.
So we performed uh four service releases
and today we will focus on something
that was released in June and improved
uh yesterday basically which is about
the availability of uh statistics and
charts for organizations that have uh a
premium plan of open air connect.
Other news are um improvements on the
efficiency of the pages especially for
the the loading of the plug-in for
accessibility
and also we had to defend against the
bots. Uh this is a problem that many
websites and repositories and uh in
general scholar communication services
are having uh in these uh these months.
So we tried our best to to defend
against the AI bots and more um more
solution more technology to to defend
from that will be is being implemented
and
so let's go to the main functionality I
would like to present you today.
So
if you go now on the landing page of an
organization in a premium gateway,
uh the organizations
features a new tab statistics
and if you click on it, you go to the
statistics section where you can find
different types of indicators
uh indicators about um openness,
findability and fairness.
These are uh somehow a percentage of how
many uh publications data are open with
respect to the total. How many
publications data sets have a persistent
identifier so they are findable and um
also how fair the data sets are.
And we also provide indicators that
tells you how the research objects of
your um of the of the organizations are
linked to each other. So we track the
number of publications that are linked
to to research data, publications that
are linked to software and data sets
that are linked to publications
because you know very well that science
is connected.
Okay. And this is the these are
examples. I do not remember uh exactly
if that's the organization we saw in the
previous slide or not. But in any case,
uh you can see that we have the
openness, findability and fairness score
and the number of
uh of linked research outputs for a
specific institution.
And um in case a chart has a note, a
descriptive note, you can see this icon
and if you click it, the card indicators
will uh change, let's say. and the
description of the indicators will show
up. Something that I would like to ask
to you that if you find notes that I'm
that are not clear enough
uh let us know because we can change it.
So we try to do our best to make things
as clear as possible but sometimes it's
um uh
it's better to to get the feedback from
uh from you because we may give for
granted uh concepts that are that are
not. So your feedback on this would be
very helpful.
Um you will also see indicators about
collaborations.
So uh it's easy if I show you this slide
with a screenshot. So on the left you
can see a map that highlights
uh the countries that collaborate with
the current institution based on the uh
authors of the publications. So we start
from the publications of an institution
and we look at the at the organizations
of the co-authors. So based on the
number we plot the information in this
map. So where do you see um a darker
color there is a stronger collaboration
on the right. Instead we focus on
collaborations based on project grants.
So basically here the colors depend on
the number of projects this institution
is um
participating together with other
institutions in other countries.
For all the charts you have the
possibility to um to view it in full
screen. You can print the chart. You can
download the image in different formats
and you can also download the data that
is used to create this uh this chart. In
particular, you can download as a CSV,
as an Excel file or you can view uh the
data just below the the chart and you
have these three lines on the top with
uh with all these options.
H
then um I wanted to have some indicators
about
um
about where researchers uh publish.
So uh I included three indicators about
the journals
and I have one indicator for diamond
open access journals, one for gold and
one for transform transformative
journals
which you can see here.
And um basically if you hover on one of
the colored bar you can see the specific
number of publications of the
institution in that journal as you can
see here
and
I was not really sure on the
on the size on the format of the charts.
So uh they may be too big but we can
shrink them a little bit. So if you are
in your gateway and you think really
that they are too big, we can decide to
make it smaller.
Um
we can also provide a different types of
view. So instead of the bars we can have
the boxes and I will show you an example
in a in few slides. So if you think that
the other visualization with the boxes
is preferable, we can also change that.
And coming back to the researchers
habits. So one thing is you know the the
journals and the journal business models
that the researchers choose and another
one is uh where researchers decide to
deposit their data sets and their
software. So we also have charts that
show
that this aspect. So on the left you can
see um a box a block chart about the
repositories where the data sets of this
institution
um are hosted.
In this case you can see Harvard the
data versse is winning over the others
and on the right you can see the
repositories of the of the software. So
in this case there were only two Zenodo
and Kodosian.
Um as before if you over with a mouse on
a on a block on a box you can see the
specific number of research products
and as in any other charts you can click
on the embed functionality. This will
open a new window with
uh basically the JavaScript uh code that
you can include in your website. And
this is useful if you want to embed any
of these charts in your institutional
website for example.
So this was the presentation of the new
functionality. I don't know if there are
any questions or suggestions that you
would like to share.
>> Uh, hi Allesia. I see that there is a
question in the chat. Uh, Maurice is
asking if you could show where we can
find these charts.
>> Oh yes. Let me share the screen again
and I will dive you live. I'll
Okay.
So, let me pick our product.
So, under the search um button, you have
the possibility to search for the
organizations.
And here you have the list of the uh
organizations that are related to the
Aurora University Alliance and the um
subscribed organizations have the
dedicated statistics tab here. So if you
click it, you can find the statistics.
I already see that we have to fix this
because I see plus one twice.
So, so if you see Okay. And this doesn't
work, but it was working yesterday. Let
me pick another one just to double
check.
Okay, it works for it.
So if you see anything weird like plus
one uh written in slightly different
ways or any other problems with the
charts, let us know and uh and we will
uh fix it as soon as possible.
I can see that there are no any other
questions for now. Thank you Alysia for
your presentation and we can move
forward.
>> I have another I have another question.
Sorry. So [clears throat] I I I looked
at the the Netherlands portal but I
cannot find it in the list uh of uh of
objects to find. Should I enable
something in in the in a management
interface or
>> Okay. So the the opportunity to search
for the organization is available when
um we have something in the inclusion
criteria by organization. So usually
what happens is that for example
university alliances they want all the
research products of
uh of a specific university to be
included in the gateway and they specify
the list of organization and those are
searchable. So in the case of the
Netherlands, this is not the case
because we worked with the repository
point of view. So um what we could do in
fact is to include all the organizations
based on the Netherlands that we have in
the open air graph to the configuration
of the Netherlands portal. So
the consequence will be the
organizations will be available in the
search options
and all the products that have an
affiliation to those organizations will
also be included even if they are not
included in the repositories.
Sometimes this is something that happens
>> because researchers are lazy and forget
to
uh deposit something in the institution
or repository. So maybe we can find it
in cross ref for example.
>> No let's let's talk about further uh
yeah on this uh solution. Yeah.
>> Yeah.
>> Great. Thanks.
>> If there are no
there is another question in the chat.
Uh it says that they cannot see uh cano
if I pronounce it correctly. Royal
Academy of Netherlands in the Dutch
research portal.
uh this as an organization or as a
source
because as a source we have to
one institutional Chris and the
institutional repository.
>> Uh if Zor would like to explain further
question.
>> Hi uh thanks for your answer. Actually,
I don't see uh Canvas. Uh I'm not sure
whether it's tracked or not. And uh as
Morris mentioned, I don't see uh search
based on I mean uh the uh organizational
um uh button on the Canary research on
the uh open research portal or Dutch
research portal. And uh when I see when
I go to publications I see the name of
the universities but not can away is
there maybe there are underlining data
but uh yeah I don't know how does it
work and uh what should be done to be
part of it.
>> Uh for sure we start thank you for sure
we start a conversation to include the
organizations in this menu.
Uh
and basically what you can do if you
want to search for scholarly works of a
specific institution for now what you
can do is do a search select the
advanced search
and then
you can select organization
is
no
and then we have several options. I
think it's this one Royal Academy. Yes.
Uh, sorry Allesia if you you thought
that you're sharing your screen. You're
not.
>> I'm not. Oh, yes. I'm sorry.
I I'm sorry. I thought I was sharing.
Um
so
if I search for the data sources for
example I can find the two data sources
for uh
and for example let's pick uh the cris
so this is what we collect from the cris
system of
here you can see related organization so
if we go there you have the list of
publications research data research
software linked to the organization
and you can of course view them all.
So the the same list you can get it by
basically using the the advanced search.
What we will we can work in the future
is the possibility to add uh the search
by organizations directly in the menu.
And this is uh this will let's say make
it simple.
>> Uh okay thank you. So I see that this is
a kind of uh I should look into that
deeply but I think this is based on
aggregated level right because for
example at there are 12 research
institutes
uh that they have their own pure. So uh
I need to look into that. But is it also
possible to the aggregate that level as
well to have that option?
>> Uh I don't know what kind of
presentation Maurice had in mind. Maybe
you're going to address these.
>> I will ask Morris about it.
>> No, it's fine. It's fine. Sorry. Yeah,
let's let's if if we need to have a
discussion on that bilaterally. It's
fine. That's uh let's let's do that
discussion. Uh it's it's it's supposed
to be part of the of the DERF project to
add as many uh uh repositories as
possible uh from the Netherlands and
also from uh the the Royal Dutch Society
of uh uh science and and and arts,
right? Arts and Sciences. Um and uh yes,
our our list is perhaps not not complete
uh as as yet as of yet, but uh um we we
have to work with you uh on that to uh
to make it as complete as possible.
>> Okay. Thank you.
>> So yes, please.
>> We'll keep in touch. Thanks.
>> Yeah.
>> And there is also one more question in
the chat from Hanek.
It says our UVA there and HVA research
database is at sources.
>> So I I guess it means that they can find
uh their repositories uh in the gateway.
So more than a question is a uh is a
comment I would say.
>> Okay.
Okay. Perfect. I cannot see any other
questions in the chat. So we can proceed
with uh Maurice's presentation.
>> Yes. Great. Thanks. Um and and uh for
reference uh Allesia I I see that in the
Netherlands portal there are 150
affiliations available. So um where we
have to look into why that isn't showed
up in at the source at the s uh what is
it search u sub menu. But anyway, let's
um continue my presentation. So, thank
you for for having me uh everyone. Um um
I I'm asked to talk something about the
third anniversary of the Netherlands
research portal. And frankly, uh uh I I
I forgot a bit about it that it's a
third anniversary anniversary already.
Um,
so, uh, so what what happened in the
meantime? Uh, I'm sharing my screen,
right? Yeah.
>> Uh, no, we cannot, uh, see.
>> Oh, I Yeah, sorry. I need to press the
share button and then it activates
sharing, I saw.
>> Yeah. Okay.
>> Now, now it's sharing, I guess. Good.
Um, and uh uh I was like, oh yeah, what
what's happened in that uh in the in
that meantime? Uh but we'll get to get
uh get to that um in a in a few clicks.
Uh first uh I want to show you what it
is and I think I am um talking to an
audience that I can perhaps go quickly
over the whole this this whole part. uh
then I want to show you something about
the history then about some context and
then about the projects where uh this uh
Netherlands research portal uh lives in.
Um so first the what what is it? Well,
it's supposed to be all Dutch research
in one place. Uh where we present all
the all uh all the projects, all the
publications, all research data, all
software, all researchers maybe in the
future. And uh uh because it's it's in
this beautiful graph from open air uh we
want to have relationships between all
these uh things altogether. Um and uh
later on uh I want to tell you something
about what we want to do with this these
enrichments putting them back into the
sources where they came from. Um so in
the portal um yeah you obviously know
because I'm preaching to the choir um uh
we want to stimulate or we want to
present uh the repositories that we
harvest and where you can also um link
back into uh where you is a place where
you can deposit um your research your
data your software etc. um search it but
also link it. That's the links between
all these um uh projects, persons, uh uh
outputs, etc. Um
uh well this this is a this is the
manual approach where you can offer um a
researcher or research administrators to
uh to do so. Uh but obviously the graph
is doing a lot of automated work uh by
itself. So what I said before um uh
there's a s a list of suggested uh
repositories where you can deposit and
yes or if we have your cal repositories
in there and you deliver these nice
descriptions
um uh where we can present uh your
repository in this way it will appear
here as well. um of course search uh and
uh and a neat thing that open air did is
uh they put also organizations here in
this in this list. So you can even see
where um which organizations are
involved in this uh publication. So that
was a pretty neat thing we did with
Aurora and they implemented it
throughout all the all the portals. Um
and uh now also um
this was the linking part where you can
find sources and link them to other
entities that you find. Uh you can do
this by hand one by one or by bulk. It's
pretty neat uh feature to uh to make uh
more meaningful connections to the
graph. uh and this information will be
brought back into your system uh um uh
into uh into your repository in the end
uh if you would like. Um so for portal
management if you log in uh here is the
part where I can add sources uh for
example the other repositories
um yeah don't show it here up here but
there is a add button and I can find the
repository and add them to the
Netherlands research portal and then in
the next round of indexing will be taken
up into the uh into the sub graph. uh
but also uh here are the affiliations
where we can select the Dutch
affiliations
um in this portal etc etc. Um also uh
affiliation management. So uh there is
this open orcs database where open air
is collecting all kinds of name
variations of all your universities or
your university where you can um try to
uh uh merge or split or um think all
kinds of conflicts that happen with
records. uh to add uh records to your uh
name to your um affiliation for example.
So when you press on the affiliation
button the statistics that Allesia just
showed you uh are being uh even more up
to date and uh with with the newest uh
changes that you made here.
Um also for repository managers there is
interface where you can login and see uh
which when your repository has uh
aggregated and uh which date uh of the
aggregation is being indexed in the
graph. Um so you can see here that uh
for example the repository is now uh
from May uh is in the index and not from
the from July.
Um
so you can uh keep a little bit track on
that. Um but also uh what I said before
the graph is enriching your records
built on uh information from other
records that other people and parties uh
deliver. um where you can for example um
uh enrich uh your your projects for
example that can be associated to your
publications um and that would be very
interesting information to put back into
your repository or Chris system
um
okay so how how did this come three
years already I said so um uh let let me
just show you a bit of history Um the
Netherlands research portal didn't came
there overnight. It was uh it started 20
years ago. Um in 2006 and I don't know
how much time you have or half an hour.
Um it started with Daret there was the
Dutch academic repositories network. Um
and uh um there is um uh built on on the
um open archives model uh for um where
repositories built and uh H university
uh built one and the the mission was to
build a pressure mechanism for
commercial publishers to reduce the
prices on uh on um on the um
subscriptions.
Uh the idea was that well uh with the
mandate of our repositories green open
access we can publish ourselves if you
want. Um so um was a was a mechanism to
reduce costs. Um and that helped uh in
in the beginning days. Um then uh Donet
uh joined the driver project. It was a
predecessor of open air uh where we
drawn up uh the landscape repositories
in in uh in Europe. Um and was the start
actually of uh of of open air.
Um then uh uh the project thereet ended
uh and nurses from the from the the
Royal Dutch uh Academy of uh Arts and
Sciences took over uh this website um or
this this part called Nars.
Um and uh 13 years later um they
decommissioned uh the portal uh where it
was it was felt not not part of the uh
of the overall strategy uh because it
was part of Dons um and Dons is a data
archive uh and they felt to uh go back
to the core of their business uh and to
decommission uh the portal. So we were
left uh stuck. Uh but in the meantime,
open air developed uh so enormously that
they uh had this uh new tool open air
connect where we and we had this strong
um repository culture already uh and we
use the we use the same standards that
we easily could migrate into uh open air
connect. Um the strategy was before was
that uh Nars was the the proxy where it
collected all the repositories from the
Netherlands um and offered one place for
open air to harvest. Now we in this in
this mini project that we did uh with
not only me but Dong from Leen
University and KSB bars from Dons to ask
repositories to be harvested um by open
air directly uh and offering them this
interface where they could manage their
repository but also get their
enrichments back.
So three years later then um uh it was
it was like okay now now we built this
uh we put this uh online
um uh and uh we need to have uh finances
for the next uh for the next round. So
uh we we
uh it was like in between in limbo what
h would happen with this um because it
was a just a a nice project but there
was no governance behind it. There was
no financial support. So with this
initial uh idea, we had to go to the
library directors
um and and ask them uh our library
directors in the Netherlands and ask
them how to continue do they want to
continue still with this or uh should we
uh should we abandon uh the project
altogether? Uh but they said this
decided no this is this is a very core
um uh thing to keep uh keep alive. Um,
it was also part of the I didn't put the
picture on that of the the Dutch open
science strategy to have uh from the
universities of the Netherlands to have
a strong federated um repository
infrastructure or publishing platform
infrastructure they called it. Um and
that's uh where but also a central place
where everything can come together and a
and a place where all these uh research
is being presented
um and it very much look like uh the the
the portal. So uh we we have written uh
a proposal uh for the uh Dutch uh open
science fund uh from 1.5 million uh to
work on several aspects uh of this uh
infrastructure that we that we've that
we don't want to lose.
Um and I will come back to this. This
was formulated in the derf project and
derf and there is a Dutch translation of
there. So there is a uh link in history
of the name. Um
and also I want to tell you something
about uh the EOS node uh the Dutch EOS
node that is using data from this
repository network.
So
uh why why why do we want to go all
through this this hassle? So we want uh
basically uh not only green open access
but we also want to have transform the
closed research information
infrastructures to u open research
information commons and for the
governance what we need to do is
collective action uh is clarity on who
does what what are the roles and
commitments that everybody uh has to
take in this ecosystem. So there is a
role for for the libraries, there is a
role for the universities, for the
repository managers, for the people
working in these repositories, but also
um
uh the u the metadata uh uh specialists
at our universities. Um but also for for
open air there is a role. Um uh so also
we need agreement on who maintains what.
So uh open air maintains for example the
aggregation infrastructure
universities they need to maintain the
repositories but also people need to
maintain all the um uh what is it the
affiliation uh fixes that I that I
showed you before and um to have real
commitment to this collective value uh
over this local optimization. So put
effort in in a bigger in the bigger
picture than only um uh fixing uh the
things in in a in a local in your local
situation.
So I will go very broadly over this. Uh
it's well the the title says it it's a
very rich and fragmented landscape and
we need uh to have something that that
brings us together and and the research
portal Netherlands research portal is
that uh that very thing that that helps
us bring uh the collective uh um idea in
this uh because like uh Zor already was
an very good examp I use you as example
Uh but uh apparently you want to be part
of the Netherlands support because you
said I I'm missing out there is missing
out stuff that it means that that it
means that that is that it is important
for you uh to do so. Um and uh yeah
I I explained uh before so it's it was
part of of a whole um builtup of of uh
of policy lines and principles that uh
for started for example back in 2006
but also uh the national open science
agenda uh the uh the digital comments
that been uh presented by the Dutch
government that we want to be more
independent of of big tech to provide uh
research information but also in um
uh getting access to our own research
that we that we have produced.
Um
uh what I want to show you a little bit
and and remember uh maybe you can
research this a little bit later because
um at first this looks very stupid but
uh the hawk the global open research
commons is a model that shows you what
is needed to make um this infrastructure
or this ecosystem sustain on the long
run. Um and what I said before it's not
only about uh the metadata and and the
PDFs etc. And it's not only about the
the portal that is uh being kept alive
and the repositories uh and and the and
the compute infrastructure but it's also
about the standards uh making sure how
the information flow flows be uh behind
uh all goes through all these uh steps.
Um
uh but it's most importantly it's the
white part here. It's about um how can
you make this uh infrastructure
sustainable, financially sustainable?
How um can you make sure it's being
used? So you have you need to have
engagement uh with your users and your
your user group to make sure that what
you are doing is still relevant for that
uh for that group. Uh, of course the
governance uh structures um it needs not
only be a play toy of people uh at the
workforce at the work floor but also uh
it needs to be some uh of some
importance at the board level. Um the
rules of participation and access who
who when when are you allowed to be in
the for example the Netherlands re
research portal? what kind of what are
the the rules of participation in order
to to uh to be allowed in there? Um what
kind of um uh yeah what's that called?
Um
uh what what do you need to do for the
for the community for the for the
commons to be sustainable? Um and also
what I said it's it's also the big part
is human capacity. um there are people
behind all these systems that that uh
maintain it uh that ask questions to
researchers etc. So these are um this is
a um um framework where you can see
where you can plot you can you can see
the the weak spots where haven't you
thought of when you when you are
implementing your um ecosystem or
infrastructure
um so um what at at surf I mean I also
work at surf for for a bit uh they they
are um thinking on on this u digital
autonomy. Uh they want to promote uh
more the uh thinking of the commons and
the research commons. What are the
components in there the building blocks
but also have this workflow perspective
because that reflects uh the user and
how the user is using all these kind of
services in this ecosystem.
uh I won't go over this in detail but uh
for example here is the the workflow
from it goes from funding to eventually
impact and reporting uh and you need uh
common building blocks uh over there but
also federative access over all these
systems
um okay now I will go because that was
that was the three years of the research
portal um and now I will talk about the
three years onward work because we got a
great mandate uh from all the um
university boards in the Netherlands to
continue uh building this uh um Dutch
repository federation and Chris systems
um but [clears throat] it wouldn't fit
in the name. So um
uh we want to work together on a
federated open uh ecosystem
um putting all data software and
publications together in one place but
uh have autonom autonomy uh each
university and each uh research facility
needs to have their own autonomy how to
uh
safeguard this information. So this
project will run for three four years.
Um and it u like I said it was a it
should be uh the robust backbone of open
science uh uh in in the Netherlands. Um
and uh what it will do is to um
integrate all these systems under one
shared governance. Uh we want to enable
enriched metadata feedback via the open
air graph. We want to increase the full
text collections um uh coming in uh in
these repositories. We want to preserve
all that information metadata and the
full text uh via the national library
eo. Um and we want to uh maximize
discoverability through multiple
channels like also open Alex and Google
uh and the EOS nodes etc. Um and also we
want to have a uh uh develop the portal
a little bit further with the help of uh
open air. Uh for example, one thing that
was missing is an expert finder uh that
uh that was currently in in the in the
Narsis uh portal and we want to bring
that uh back.
So all these things are puzzles pieces
that uh need to fit together in order to
um uh get to get successfully to to the
end. Um
and this is a little bit of the picture
how that how that flows. We start here
at the the left part with repositories.
uh we go follow this arrow is being
harvested uh and validated by open air
provide and being put into into the open
air graph and reached uh uh within the
open air graph with all the other
sources. put arrows in there, but they
are uh and that information being put
can be put back into the repositories if
they want to um and if the information
is uh good enough to ingest back into uh
these uh repositories and uh Chris
systems then from this graph we can put
it into the research portal but also we
use that information to uh preserve that
in the national library uh archive
and uh at the same time uh here we want
to um fetch the full texts uh because we
have a special law in the Netherlands,
the Tarvena Amendment for author rights
that allows us to um get the uh not only
the author manuscript but also the
published version uh of uh of of a
publication from a Dutch author and put
them into a repository for as a copy as
a backup to safeguard uh that
information.
Um and uh at the same time here at five,
we want to have this information that's
in in these systems uh being um uh
aggregated and indexed in all other um
indexes and cataloges etc that have high
priority. Uh so it's not uh it has to be
done independently by all these systems
uh because of this federated uh nature.
So not only not not via only one single
point of failure the graph to push that
into these cred logs but have these
systems operate uh autonomously.
So who are uh working on this? Uh uh the
team leads for each of these uh packages
come from uh these universities.
Lighting University, Delha, National
Library, University Amsterdam, V and
Surf
uh but also um all other universities uh
are involved uh and of course uh open
air
um uh what we want to do uh for each of
the teams we want to make uh make sure
that we can monitor these things uh via
datadriven dashboards. Uh and we're
working on that uh as well. Um
and then I will go over to EOS. So there
is one one project to to make advance of
that. Uh EOS is um is the European open
science cloud node. Uh one of the nodes
is uh the Dutch node. So what we can see
here in so um Sturf is doing a pilot uh
trying to uh build uh very the the very
core
uh of that node uh with their own um
infrastructures uh like research cloud
research drive uh store the data um uh
but also then have the research objects
uh and that's where open air comes in
here to see to have a catalog uh with
all data sets that we can be used and
transferred and we'll go to another
picture to show that into these
infrastructures um and have services and
tools um presented in a in a catalog
what's what's being made available uh
for surf in the next that's phase one in
the next phase they want to uh uh and uh
roll that out to the other um um
uh research teams uh the thematic um
digital curation centers where they um
also have these kind of um research
objects etc. but also the human capacity
and engagement. Um to to follow up on
that
um
this yeah what I
uh so it connects uh all all these kind
of um um services and tools uh that that
are already existing u and uh in for
within the Netherlands. And the
beautiful thing is that this is one node
and uh it should also connect to the
European node so that uh the European
researchers can find uh services
available uh also in the Netherlands for
them to to use compute services, data
services uh etc.
uh and have have this uh unified access
uh I think via these uh um access
management uh systems um to use the
services that that are available in the
Netherlands. So where does open air fit
in? So what we saw the Netherlands
research portal is a one of the services
and tools um that connects all the
repositories to to bring all that uh
metadata together from publications but
also from data sets and in this this
case we focus on data sets. Uh this
being put into the uh into the subgraph
or in the at least also in the bigger
graph open air. Um and this is can you
can see this as a catalog for
uh data sets in the from the
Netherlands. Um these data sets uh can
then be uh presented to the other
services for example the uh storage and
and comput facilities uh to make sure
that the data can be reused
uh transferred and reused. So what what
it basically is in the workflow
perspective is that you
uh find fair data at least has to be
interoperable and reusable.
Um and you find this ideally via this uh
this path that path is being presented
in these uh in these in the in the Dutch
node where you can select the data that
you want. Then the data is being
transferred into a selected storage. Uh
in this case the surf research drive but
it can be also if your institution
offers also uh a storage facility can be
uh added to that as well and then uh you
can connect that uh storage facility to
your compute facility where you can uh
work with the data to analyze. So that's
the that's the idea the flow uh of data
of information that uh that we want to
build uh where where this uh fits in.
>> Maurice that's it.
>> Thank you Maurice. Can you go back one
slide?
>> Back
>> because Yes. Perfect. Because in your
No, no, no. That's
>> Yeah.
>> Because in your workflow perspective I
think you missed one last part.
Because after you analyze the data,
>> maybe you process the data, you
transformed it.
>> Yeah. Yeah.
>> So you want to publish it again.
>> Yeah. Sure.
>> As a new version, etc. And you use the
repositories of your federation to store
the data
>> which gets you know in the graph again.
>> Yeah.
>> Absolutely. Yeah.
>> It's a complete
>> Yeah. It's a full circle. Yeah. Yeah.
Yeah. No, absolutely. Absolutely. I just
I just wanted to show that. Yeah, but I
could I could add at add that to have it
have it full circle. Yeah. But uh now
now it's a little bit bigger. [laughter]
Yeah. Yeah. That's that's that's uh
that's that's uh pretty nice. Yeah. Let
me uh in the next version of this uh
presentation I uh I will make it a
circle.
Yes. So that the federation is the the
starting point and also the final point.
So that's
>> yes
>> and in the end there is there is no
start and there is no end
>> because research is always evolving.
>> Any other questions?
Should I stop sharing?
I see I see Hanukkah there
and Manuel.
Hi.
Any questions?
>> No. Okay.
Are you going to use uh the graph uh
information to make your analysis uh
Manuel?
>> Uh yeah, I think we are thinking about
it, but we are only now starting to
think about open air. So, we're slowly
getting there.
>> Yeah.
Yeah. Well, that's uh it
>> Yes. If you if you have questions on how
to use the the graph, feel free to to
contact Maurice and me and we can
provide support if needed.
We have plenty of documentation
available online but uh I know that
different use cases may require you know
uh specific uh tips let's say because
the graph is huge that and also if you
look only at the the Netherlands only at
the Netherlands subgraph still we're
talking about more than three millions
metadata records so it's not something
um mold. [laughter]
>> Yeah.
>> I think I have one question about the
last uh slide you showed.
>> Yeah.
>> If we um if we decide to import data, is
it possible to import a specific set via
the API or something like that? Do you
know if that's possible the filter and
then
>> um let me see
uh this this this slide you mean?
>> Yeah. And then the last step analyze and
then maybe import via the surf resource
cloud the data which is um relevant for
us and then
we don't want to import all the data
maybe but that's import
[laughter]
>> ah
and then filtering but
>> yeah sorry yeah maybe maybe I uh no I
understand yeah this this uh this
picture might might no what I what what
I meant was uh find a fair data and that
means one or more data sets that you
want to find and it's not a full graph
that you that you are
>> putting into uh No, it's no I understand
the confusion. Yeah. Yeah. No, it's it's
about here you find a data set uh in in
in the search interface and then you
click on on the data sets that you need
for your analysis. That's that these
data sets are being transferred into
research drive and that's being an
analyzed not the metadata from the
>> Okay. Yeah.
>> The graph. No, it's Yeah.
>> No, that's very good question to make
that clear. Yeah. Yeah. Yeah. So, the
Yeah. the metadata stays here in the
open air graph uh and the catalog also.
But it's it's it's a way to select um
data set you're interested in uh to
reuse for your analysis. Uh for example,
you need kind of a data that's uh
historical data on on weather uh
patterns or so. Um and you find them
here and you transfer them into your own
storage here and then you can analyze
them.
in in a comput facility.
>> Okay.
>> Yeah.
Um yeah, it was I was if uh and if you
if you want to uh
if I if I go back into the presentation
a little bit further where I showed you
um about your own Chris system or
repository and maybe that's maybe it
could also be a question for you that
you that you might ask can can I offer
only one stat here. Uh there is um in
this harvest management system where you
can uh I think yeah if you click on
update and then you see the details of
your repository
then you can I think also select a
specific set that you want to uh get
harvested by open air instead of
everything. So that might be another
question that you
might have.
>> Okay.
>> Yes. Each source is each repository can
decide uh if open air has to harvest
everything or only one set or multiple
sets.
So we have these three these three
options.
>> Yeah. So maybe I can show you here.
Um you update here. I think it's
>> it's in update interfaces if I remember.
>> Oh here. Yeah. So you can select the set
that you want to um get harvested and
you can add multiple right. So
you can add several sets
um and several interfaces etc to be
harvested. So
it's easy peasy. Yeah, open air makes it
really uh easy for you. So it's uh
[laughter]
just uh just ask.
>> Oh, we try to do
to do our best. Sometimes we're not very
good at telling people uh how many
functionality
and how many options are available.
Yeah. Because you have to consider that
most of the services that open offers
are the results of research. So we
started from research projects. So and
all these tools services that you see
are really the results of research
activities from um from partners that
are spread across all Europe. So Athens,
Italy, Poland, Portugal. Uh so
with open air we are pushing them to
become products. So actually usable by
uh by others in production level uh
settings let's say and that is great
because
>> everything you saw
>> is really the results of research in
scholarly communication and open
science.
>> Yeah. So open air is actually the
perfect example for commons in working.
So it's uh
>> and in fact we leazed with operas for
the for the use nodes on schol on
scholarly comments.
>> Yeah.
>> So
>> I think I think it's time one minute
past.
>> Yes. I do not see any more merily but I
do not even see any other questions in
the chat.
>> No we don't have any other questions
and we can close the today's session.
Thank you both for your presentation.
Thank you all for your questions and
feel free to contact us either Lesie or
Maurice and also in open air help desk
and see you at our next session on
November.
Have a nice afternoon.
>> Thank you very much.
>> Thank you. Bye.
>> Bye bye. Bye-bye.