WS2: Difference between revisions

From EERAdata Wiki
Jump to navigation Jump to search
Access restrictions were established for this page. If you see this message, you have no access to this page.
No edit summary
No edit summary
 
(93 intermediate revisions by 5 users not shown)
Line 1: Line 1:
The second workshop of EERAdata is organized as an online webinar and hackathon during November, 30, to December, 4th, 2020.  
The second workshop of EERAdata is organized as an online webinar and hackathon during November and December 2020. It is dedicated to metadata.


== Objective ==
== Objective ==
* Discuss and develop metadata standards for FAIR and open data in the low carbon energy research community.
* Discuss and develop metadata standards for FAIR and open data in the low carbon energy research community. '''The focus is on the link between metadata and how humans explore data.'''
* Identify gaps and needs that hinder their standardized realization and implementation.  
* Identify gaps and needs that hinder their standardized realization and implementation.  
* Jointly work on a community paper and/or join the discussions in use cases. We start with a commented and revised draft. [https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit?usp=sharing Link to draft of the community paper]
* Jointly work on a community paper and/or join the discussions in use cases. We start with a commented and revised draft. [https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit?usp=sharing Link to draft of the community paper]
Line 8: Line 8:
[[File:EERAdata community paper.png|thumb|Workshop concept of EERAdata]]
[[File:EERAdata community paper.png|thumb|Workshop concept of EERAdata]]


== Read aheads ==
== Read & watch aheads ==


'''Obligatory'''
To prepare for the paper writing hackathon, please participate in the preparatory 'Briefing Workshop' or watch the recorded videos. The workshop took place Thursday, 5th of November, 10-12 CET. The workshop introduced recent work on metadata and outlined the tasks and procedures on how to join the writing team for the community.


Participation in preparatory workshop (or watching the recorded videos). This workshop will take place Thursday, 5th of November, 10-12 CET. The workshop will introduce recent work on metadata, and it will outline the tasks and procedures on how to join the writing team for the community. A draft is open for commenting by 23.11.2020, we highly welcome any feedback! Link to [https://www.eeradata.eu/event/1645:preparatory-workshop-fair-and-open-metadata-for-low-carbon-energy-research.html register for the preparatory workshop].
{| class="wikitable"
|-
! Agenda & Material of Briefing Workshop !! Full video recording: [https://kaltura.hvl.no/media/EERAdata+Workshop+to+prepare+for+the+hackathon/0_xxdhcjl7 link]
|-
|Slides: [[File:EERAdata_WS2_wierling.pdf|80px|thumb|none]] || Introduction & Talk ''Advancing metadata standards for low carbon energy research'', August Wierling, HVL.
|-
| Slides: [[File:EERAdata KIT Suess.pdf|80px|thumb|none]] || ''Overview of the Helmholtz Metadata Collaboration Platform'', Wolfgang Süß, KIT.
|-
| Slides: [[File:Talk_Celino_ENEA.pdf|80px|thumb|none]] || ''Metadata in materials: The new path forward'', Massimo Celino, ENEA.
|-
| Chat Notes: [[File:EERAdata chat notes.pdf|80px|thumb|none]] || ''Panel discussion with Q&A''
|-
| Slides[[File:WS2 prep schwanitz.pdf|80px|thumb|none]]  || ''Briefing for EERAdata paper writing workshop (Who? How? When? What?)'', Valeria Jana Schwanitz, EERAdata PI.  
|}


'''Suggested read on the history of metadata'''
* Invitation to [https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit# comment the draft] before the writing workshop
* Suggested read on the history of metadata
Metadata - [https://www.springer.com/gp/book/9783319408910 Shaping Knowledge from Antiquity to the Semantic Web by Richard Gartner]


Metadata - [https://www.springer.com/gp/book/9783319408910 Shaping Knowledge from Antiquity to the Semantic Web by Richard Gartner]
There is also a [[References on metadata|list of interesting references]].


== Join and collaborate interactively! ==
== Join us and collaborate interactively! ==
=== Becoming a co-author  ===
=== Becoming a co-author  ===
Workshop participants are invited to become co-authors of the planned community paper
Workshop participants are invited to become co-authors of the planned community paper.
 
'''How to join the team writing the community paper?'''                                                         
'''How to join the team writing the community paper?'''                                                         
Briefing is provided during the preparatory workshop (also recorded). Includes recommendation for read and watch aheads.
Briefing is provided during the preparatory workshop (also recorded). Includes recommendation for read and watch aheads.
Line 34: Line 50:
* Recommendations on how to proceed.
* Recommendations on how to proceed.


Invitation to [https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit# comment the draft] before the writing workshop, e.g.:
'''Invitation to [https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit# comment the draft] before the writing workshop''', e.g.:
* Review of key questions,
* Review of key questions,
* Suggestions of literature, illustrative cases, metadata perspectives & issues,
* Suggestions of literature, illustrative cases, metadata perspectives & issues,
Line 44: Line 60:
Using above prior comments & contributions provided by 23.11.2020, the EERAdata team will revise the draft before the workshop starts. During the workshop, the discussion & writing of the paper continues in small writing teams. A prior registration to teams is therefore recommended. However, teams can change anytime. The EERAdata team is strongly interested in inviting a large group of contributing authors. It is not only in the interest of the community project as such, but it increases the credibility of the main output - the draft (or road towards) a low carbon energy ontology in support of metadata”.  
Using above prior comments & contributions provided by 23.11.2020, the EERAdata team will revise the draft before the workshop starts. During the workshop, the discussion & writing of the paper continues in small writing teams. A prior registration to teams is therefore recommended. However, teams can change anytime. The EERAdata team is strongly interested in inviting a large group of contributing authors. It is not only in the interest of the community project as such, but it increases the credibility of the main output - the draft (or road towards) a low carbon energy ontology in support of metadata”.  


'''Co-authorship is granted to active participants of the EERAdata workshop as well as to other contributors to the community paper.'''  
'''Co-authorship is granted to active participants of the EERAdata workshop as well as to other contributors to the community paper.'''


=== EERAdata wiki - How to? ===
=== EERAdata wiki - How to? ===
Line 66: Line 82:


=== Just have fun! ===
=== Just have fun! ===
==== Metadata memory game ====
'''Read before playing:''' This is a little memory game where players need to '''identify triples''', that is three tiles that belong to one data set. All of them relate to low carbon energy. Some tiles have the form of a picture, others show metadata descriptions (in various formats) and some are even sound files. So, the game is to have fun, while learning about metadata. Try out if you can find all triples. Note if you have opened 3 tiles and they do not match, all will automatically close. Just as you know it from a real-life memory game.
'''Read before playing:''' This is a little memory game where players need to '''identify triples''', that is three tiles that belong to one data set. All of them relate to low carbon energy. Some tiles have the form of a picture, others show metadata descriptions (in various formats) and some are even sound files. So, the game is to have fun, while learning about metadata. Try out if you can find all triples. Note if you have opened 3 tiles and they do not match, all will automatically close. Just as you know it from a real-life memory game.
[http://eeradata.webfactional.com/memory/memory.html '''Link''']
[https://eeradata-platform.eu/memory/memory.html '''Link''']


== Agenda and notes ==
== Agenda and notes ==
Line 75: Line 90:
{| class="wikitable"
{| class="wikitable"
|-
|-
! Time slot !! Topic
! 30th November !! Framing workshop (Part 1): Online hackathon to write a community paper
'''Link to the recording of the first part of the session: [https://kaltura.hvl.no/media/Session+30-11_firstpart.mp4/0_6m7tx64w Link1]
Link to the recording of the second part of the session [https://kaltura.hvl.no/media/Session+30-11_finalpart.mp4/0_bhas9l73 Link2]
|-
| 10.00-10.15 || Welcome and introduction to the first draft of the planned paper: “[https://docs.google.com/document/d/1oS_t7yb9_xDM-RYDPTd48B7O9Sq9dXiz3ibUO34G7oA/edit# Advancing metadata for low carbon energy research - current state and call for action]”, Valeria Jana Schwanitz & August Wierling, HVL.                                                                                                         
|-
| 10.15-10.30 ||Discussion of the structure, collection of comments and notes
|-
| 10.30-10.45 || Creation of writing teams in break out groups. Preferences are collected beforehand if possible.
|-
| 10.45-15.00  || Work in '''writing teams''' (break out groups in zoom). The work-flow, coffee and lunch breaks etc. are organized by the groups themselves.
|-
| 15.00-16.00 || Wrap up day 1 of the framing workshop: Reports from the work done in break out groups, collection of feedback and comments from all. Suggestions of questions to be investigated in use case workshops.
|}
 
{| class="wikitable"
|-
! 1st December !! Use case “EU policy and energy research taxonomies” '''
Link to the recording of the session: [https://kaltura.hvl.no/media/Session+01-12.mp4/0_c80q6nes Link]
|-
| 10:00am - 2:00pm || The workshop focuses on metadata, exploring how to link ontologies used for policy-making and those used in research domains. The workshop will use the software "GitMind" to work on the ontology, incorporating suggestions from workshop participants. The participants will form one of the writing teams for the community paper. [[WS2UC4|Link to workshop page with agenda, notes and results]]   
|-
| 10:00 - 11:00 || '''Intro to the session'''
* Welcome and introduction to the topic
* Present gitmind, create logins
* Present the 4 ontology islands
* Share the ontology
* Present the tasks
|-
| 11:00-13:00 || '''Time for individual work'''
(incl. break/Lunch time). Space for discussion as requested.
|-
| 12:45 -  ||  '''Collecting of contributions'''
(e.g., gitmind maps elaborating on the 4 ontology islands)
|-
| 13:00-14:00 || ''' Discuss results with everybody '''
|-
|}
 
 
{| class="wikitable"
|-
! 2nd December !! Use case “Buildings efficiency”
Link to the recording of the session: upcoming
|-
| 10:00-12:00 (GMT+1) || The aim of this session is to build up on discussions that took place in 1st EERADATA workshop together with the participation of the invited experts, utilizing the experience related to the FAIR assessment in order to: a) Highlight main issues with FAIRness, b) Jointly come up with possible remedies and solutions, and c) Identify pointers to metadata standardization, if and how it can be achieved in Buildings Efficiency domain. [[WS2UC1|Link to workshop page with agenda, notes and results]]                                                                                                         
|-
|10.00-10.10 || Introduction: What does the “Buildings Efficiency” use case of EERADATA aim? Mehmet Efe Biresselioğlu, IUE                                                                                                         
|-
| 10.10-10.25 || Standardized Flexibility: FAIR data principles in H2020 projects - ECHOES and ENCHANT - databases, Jens Olgard Dalseth Røyrvik, NTNU Samfunnsforskning
|-
| 10.25-10.40 ||  Achieving Fair Principles: An Overview of the Hotmaps Project and Hotmaps Building Stock EU28 Dataset, Simon Pezzutto, EURAC
|-
| 10.40-10.55  || openENTRANCE CS1 residential energy demand response: The importance of nomenclature and data transparency for open source energy system model platform, Ryan O’Reilly, Energieinstitut an der JKU
|-
| 10.55-11.10 || Exceed database: Lessons learned and continuous improvement, Daniele Antonucci, EURAC
|-
|-
| 10.00-10.20 || '''Welcome and introduction''' with workshop goals and procedures: “'''EERAdata - Towards Utopia for low carbon energy research'''”, Valeria Jana Schwanitz, HVL & PI EERAdata. Link: [http://eeradata.webfactional.com/WS1/EERAdata_WS1.pdf]. Main points:
| 11.10-11.15 ||Short Break
* In the energy system the data revolution offers prospects, but we are fare away from harvesting them. The time researchers spent with data governance (finding, cleaning, revising of formats) and bureaucracy exceeds time spent on exciting stuff (thinking, creating new insights, collaborating, and discussing with others) by far. 
* There is a lack of joint standards and common metadata formats to support researchers in finding and reusing heterogeneous data. Machine-actionability is also an issue.
* The vision of developing a one-stop-entry point for energy research is clear before our eyes: being able to search for data, access rich metadata, choose & select datasets, crunch data real-time online, being able to link to "My-researcher-space", a platform that offers a personalized workspace for data analysis, paper writing, and research collaboration.
|-
|-
| 10.20-12.15 || '''Online lectures''':
| 11.15-12.00 ||Panel Discussion with Q&A
“'''The EOSC Nordic: machine-actionable FAIR maturity evaluations & the FAIRification of data repositories'''” - Andreas Jaunsen, Nordforsk & EOSC-Nordic. Link: [http://eeradata.webfactional.com/WS1/EOSC-Nordic_FAIR_data_presentation_2_June_2020.pdf]. Main points:
|}
* Goal and vision of EOSC: Enable researchers to access data across domains and disciplines as easy as possible, support locating the relevant data. All European data should be available to researchers, but this does not mean that all data is stored centrally. Instead, databases should be interconnected.
* FAIR maturity evaluations: Close to 100 repositories have been tested using an automated tool. Most of them only score with 0.10 out of 1.
* ''FAIR Evaluation Tool:'' [https://fairsharing.github.io/FAIR-Evaluator-FrontEnd/#!/collections/new/evaluate]
* ''EOSC Website:'' [https://www.eosc-portal.eu/]


'''OpenAIRE: Open Access Infrastructure for Research in Europe'''”, Ilaria Fava, OpenAIRE. Link: [http://eeradata.webfactional.com/WS1/20200602_OpenAIRE-EERAdata.pdf]. Main points:
{| class="wikitable"
* Goal and vision of OpenAIRE: To "Bridge the worlds where Science is performed and where Science is published", by monitoring, accelerating and supporting Open access research and publishing.
|-
* OpenAIRE consists of 50 European partners, including 34 National Open Access Desks (NOADs, that provide support on issues related to Open Science Policies, Open Science Infrastucture, Open research Data and Open Access to publications
! 3rd December !! Use case “Metadata user stories” '''
* Lessons learned: "Research is global, support is local". Regional differences in culture and maturity of open access infrastructure require support strategies specifically tailored to each region.
Link to the recording of the session: [https://kaltura.hvl.no/media/Session+03-12.mp4/0_mc5hoazq Link]
* ''OpenAIRE Website:'' [https://www.openaire.eu]
|-
* ''OpenAIRE Connect:'' Platform that allows to connect with the research community of a specific research field [https://connect.openaire.eu/]
| 10:00-12:00 (GMT+1) || The goal of this workshop is to understand how users find, request, use and also provide data. The workshop connects to functional specifications needed for developing the EERAdata platform and other infrastructure in support of data FAIRification work flows. Participants will have the possibility to comment on first mock-ups of the platform and try out some functionalities. For example, using the ontologies produced during UC “EU policy and energy research taxonomies” and participating in virtual discussions on the EERAdata Community Forum. [[WS2UC Metadata User Stories|Link to workshop page with agenda, notes and results]] Link to zoom meeting: upcoming.
* ''OpenAIRE Provide:'' platform that allows open access publishing of Data [https://provide.openaire.eu]
|-
* Further Links copied from Conference chat:
|10:00-10:15 || '''Welcome & Introduction'''
** ''Working Group on Rewards'': [https://ec.europa.eu/research/openscience/index.cfm?pg=rewards_wg]
Aims and intended role of the platform in the community (Manfred Paier, AIT)                                                                                                   
** ''Open Science Policy Platform'': [https://ec.europa.eu/research/openscience/index.cfm?pg=open-science-policy-platform]
|-
** ''Clarivate Data Citation index'': [https://clarivate.com/webofsciencegroup/solutions/webofscience-data-citation-index/]
| 10:15-10:30 || '''Platform elements'''
Status quo (initial mock-ups, Community Forum) (Astrid Unger, AIT)
|-
| 10:30-11:30 ||  '''User stories'''
Experiences from UC4, UC1, other Use Cases and Gap Analysis


A short break of 15 min -
1. Searching for data


“'''Community-driven metadata and ontologies for Materials Science and their key role in artificial-intelligence tools'''”, Luca Ghiringhelli, FHI Berlin. Link: [http://eeradata.webfactional.com/WS1/Ghiringhelli_FAIRmetadata_EERAdata.pdf]. Main points:
2. Linking existing datasets
* The attributes of a data object can be '''data''' or '''metadata''', depending on the context
* "An Ontology is a '''formal''' (= machine readable) '''representation''' (= concepts, properties, relations, functions, constraints, axioms are explicitly defined) of the '''knowledge''' (= domain specific) of a '''community''' (= consensual) for a '''purpose''' (= question driven)."
* Definition of FAIR data:
** Findable: unique names, human-readable descriptions
** Accessible: URL, accessible via API
** Interoperable: typed, extensible schema -> ontologies
** reusable: hierarchical schema -> data-analztics
* ''NOMAD Meta Info:'' [https://metainfo.nomad-coe.eu]


“'''Metadata practices from IRP Wind'''”, Anna Maria Sempreviva, DTU. Link: [http://eeradata.webfactional.com/WS1/EERAdata_Sempreviva.pdf]. Main points:
General discussion
* Alternative interpretation of FAIR data: Reusability of Data is the final goal with Findability, Accessibility and Interoperability being prerequisites for Reusability. Reusability of Data for multiple purposes multiplies the value of the Data
|-
* Open data = available data <-> FAIR data = findable data
| 11:30-11:45  || '''Implications for the data model'''
* '''Issue:''' How to make data findable but safe (in regards to data protection, competitive advantages, etc)?
Top-down metadata, bottom-up metadata (Michael Barber, AIT)
** '''Solution:''' Create a searchable data catalogue of '''distributed''' data
* How to create a taxonomy?
** Expert elicitation: Group of experts creates a taxonomy which is then reviewed by wider research community
*** "top-down" approach
*** + clearly defined, controlled vocabulary
*** - static, unable to adapt to new trends
** Taxonomy based of author keywords: Map keywords used by authors along similarities in meaning, frequency of usage
*** "bottom-up" approach
*** + adaptable, able to track new trends
*** - Mix of disciplines, models, etc; Many errors and ambiguities; single generic words with a broad range of possible interpretations.
* ''IRP Wind Website:'' [https://www.irpwind.eu/]
|-
|-
| 12.15-13.00 || ''Lunch Break''. Play the EERAdata game “Utopia and metadata”. Or any time.
| 11:45-12:00 || '''Way forward'''  
for designing and implementing the platform (Manfred Paier, AIT)
|-
|-
| 13.00-14.00 || '''Online lectures''':
|}


“'''Humanities and data: for a community-driven path towards FAIRness'''”, Elena Giglia, UNITO. Re-using presentation held at the Open Science Conference 2020 in Berlin. Presentation stored at zenodo. Link: [https://zenodo.org/record/3776849#.XthYBnYza90]. Main points:
{| class="wikitable"
* The Data Management Lifecycle: '''Identify''' research data -> '''Plan''' data management -> '''collect/produce & Structure & Store''' -> '''Deposit for Preservation, Cite & Share''' -> '''Dissimination'''
|-
** AT which phase to apply the FAIR principles?
! 4th December !! Art workshop “Collaborative mosaic about FAIR data and Sustainable Development Goals”
* "There is value and risk at being a first mover (regarding implementation of FAIR principles), but there is a higher risk at being a follower"
|-
* What is data in the humanities:
| 9-15 || The aim is to create a mosaic which reflects the views of different researchers and data stakeholders on the relevance of FAIR data and the achieving of the Sustainable Development Goals. Participants are ask to contribute their retrospective from the year 2030. The artist Barbara Bellier will connect the individual pieces creating the collaborative mosaic.
** Never "raw" data
|}
** Data is always an expression of the method
** there is always a choice ( methodological, epistemological, political,...)
** There is always an interpretation, subjectivity (Data are not generated by a machine)
** there is always a discussion
* preliminary issues of FAIRness:
** what language?
** Lack of skills among researchers
** registry of existing tools
** need to preserve specificity of how we do research in the humanities
** services and tools need to be sustainable
** Time consuming and no incentive or reward to apply FAIR principles


“'''RISIS - An e-Infrastructure for the STI-Policy research community'''”, Thomas Scherngell, AIT. Link: [http://eeradata.webfactional.com/WS1/RISIS2_eera_data.pdf]. Main points:
{| class="wikitable"
* What is RISIS: First pan-European research infrastructure to study research and innovation dynamics and policies
** Set of interlinked databases on: Firm Innovation capabilities, R&D output, Public research and Higher Education, Policy Learning.
** Not all data accessible, but interlinking mechanisms are fully public
* While Metadata descriptions in, for instance, PDF format are not machine readable, it is also important to have a qualitative format such as PDF that is easily understandable by humans, for instance to present your data/ work to the outside.
* ''RISIS Website:'' [https://risis.eu/]
* ''RISIS Knowmak tool:'' Provides Indicators in Key Enabling Technologies and Societal Grand Challanges [https://www.knowmak.eu/]
* ''SIPER:'' Science and Innovation Policy Repository [http://datasets.risis.eu/metadata/siper]
|-
|-
| 14.00-14.30 || ''Break'' - Game “Utopia and metadata”. Or any time.  
! 4th December !! EERAdata participates in [https://conference.codata.org/FAIRconvergence2020/programme/ International FAIR Convergence 2020 Symposium]
|-
|-
| 14.30-16.00 || '''Discussion''' to compile a to-do list for work in use cases on the second day. Serves as a guiding and aligning process. Lead by WP2, August Wierling/Valeria, HVL.  What is the take-to-day-2 message for your use cases?
| 15-17 || '''M4M Workshops: Making domain-relevant machine-actionable metadata at scale'''. There is no FAIR data without machine-actionable metadata. Semantically-rich, domain-relevant, machine-actionable metadata is fundamental for FAIR interoperation of data, algorithms and compute resources. However, FAIR-enabling metadata schema often do not yet exist for many communities and need to be created de novo by the relevant domain experts. In other cases there may be already existing metadata schema but they are difficult to find, assess and reuse. In an effort to support convergence of metadata components, the GO FAIR initiative has launched a systematic and scalable approach to the creation, registration and reuse of machine-actionable metadata called Metadata for Machines (M4M) Workshops. M4M workshops bring together representatives from given scientific communities and FAIR metadata experts to jointly define standards for the metadata that they wish to use in practice. We have found scientists to be highly motivated to develop and use metadata standards through our M4M workshops, and tooling such as the CEDAR Workbench to be effective in the collaborative development of new metadata templates. Thirteen M4M Workshops that have been delivered and in this Session, the most recent workshops (for VODAN Africa, DeiC and ZonMW) will be reviewed. This will be followed by a focused discussion on a proposed M4M workshop for EERA Data Consortium to be run early in 2021.  
* ''Use case 1:'' The main issue for us is re-usability. We need to assess the databases that were chosen previously. Learn from best practices.
 
* ''Use Case 2:'' Our main issue is privacy/ sensitivity of data. Security should come first. There is a tradeoff between universal metadata language and domain-specific language.
The '''agenda''', roughly, is:
* ''Use case 3:'' Linguistics is a problem. As soon as we change the application of the material, we also change related metadata. Find similarities of already existing metadata.
* to use the first hour to recap recent M4M accomplishments (VODAN Africa, DeiC, ZonMW COVID Program), and
* ''Use case 4:'' We face different languages and terminologies. The range of interpretation of the same terms is broad. We should address low hanging fruits but also aim at cracking hard nuts to improve FAIR/O principles.
* to use the second hour to scope plans for an M4M workshop with the EERA Data Consortium
* ''General:'' EERAData is probably more about asking the right questions to the energy research community than providing the right answers. There are already a lot of answers out there, we need to link them to our data issues. We envision being able to suggest low carbon energy metadata standards for and with the research community. Colleagues working on the EERAdata platform will join in on all use case discussions on day 2.  
* Motto: '''"'''Ontology is a formal representation of the matured knowledge of a community on a specific purpose'''"'''.
|}
|}
{| class="wikitable"
|-
! 7th December !! Framing workshop (Part 2): Online hackathon to write a community paper
'''Link to the recording of the first part of the session: [https://kaltura.hvl.no/media/Session+07-12_firstpart.mp4/0_o8ef6c8v Link1]
Link to the recording of the second part of the session: upcoming
|-
| 10.00-10.45 || Welcome and input from use case workshops presented by use case leaders (see above).                                                                                                         
|-
| 10.45-11.00 || Guidance for writing teams
|-
| 11.00-11.15 ||  Creation of writing teams in break out groups (continuation of day 1).
|-
| 11.15-15.00  || Work in break out groups. The work-flow, coffee and lunch breaks etc. are organized by the groups themselves.
|-
| 15.00-16.00 ||Wrap up day 2 of the framing workshop: Reports from the work done in break out groups. Collection of feedback and comments from all. Guidance on how to finalize the paper collaboratively after the workshop.
|}
== Metadata knowledge base ==
[[Domain_specific_classification_schemes]]

Latest revision as of 13:42, 8 September 2021

The second workshop of EERAdata is organized as an online webinar and hackathon during November and December 2020. It is dedicated to metadata.

Objective[edit]

  • Discuss and develop metadata standards for FAIR and open data in the low carbon energy research community. The focus is on the link between metadata and how humans explore data.
  • Identify gaps and needs that hinder their standardized realization and implementation.
  • Jointly work on a community paper and/or join the discussions in use cases. We start with a commented and revised draft. Link to draft of the community paper
Workshop concept of EERAdata

Read & watch aheads[edit]

To prepare for the paper writing hackathon, please participate in the preparatory 'Briefing Workshop' or watch the recorded videos. The workshop took place Thursday, 5th of November, 10-12 CET. The workshop introduced recent work on metadata and outlined the tasks and procedures on how to join the writing team for the community.

Agenda & Material of Briefing Workshop Full video recording: link
Slides:
Introduction & Talk Advancing metadata standards for low carbon energy research, August Wierling, HVL.
Slides:
Overview of the Helmholtz Metadata Collaboration Platform, Wolfgang Süß, KIT.
Slides:
Metadata in materials: The new path forward, Massimo Celino, ENEA.
Chat Notes:
Panel discussion with Q&A
Slides
Briefing for EERAdata paper writing workshop (Who? How? When? What?), Valeria Jana Schwanitz, EERAdata PI.
  • Invitation to comment the draft before the writing workshop
  • Suggested read on the history of metadata

Metadata - Shaping Knowledge from Antiquity to the Semantic Web by Richard Gartner

There is also a list of interesting references.

Join us and collaborate interactively![edit]

Becoming a co-author[edit]

Workshop participants are invited to become co-authors of the planned community paper.

How to join the team writing the community paper? Briefing is provided during the preparatory workshop (also recorded). Includes recommendation for read and watch aheads.

The community paper addresses 3 key questions:

  • Q1: How to align mental models of those searching for data with navigation along metadata? What is specific to the energy domain?
  • Q2: What consequences for the construction of domain-specific metadata follow? and how can they be dynamically updated?
  • Q3: What are the recommendations to the low carbon community?

The envisaged output of the paper is:

  • A community-reviewed draft for a low carbon energy ontology in support of metadata,
  • Metadata suggestion cards for low carbon energy researchers,
  • Recommendations on how to proceed.

Invitation to comment the draft before the writing workshop, e.g.:

  • Review of key questions,
  • Suggestions of literature, illustrative cases, metadata perspectives & issues,
  • Suggestions to any section (e.g., Introduction, Method, Discussion of Results, Outlook)
  • Adding snippets of texts,

Any other idea is welcome!

Using above prior comments & contributions provided by 23.11.2020, the EERAdata team will revise the draft before the workshop starts. During the workshop, the discussion & writing of the paper continues in small writing teams. A prior registration to teams is therefore recommended. However, teams can change anytime. The EERAdata team is strongly interested in inviting a large group of contributing authors. It is not only in the interest of the community project as such, but it increases the credibility of the main output - the draft (or road towards) a low carbon energy ontology in support of metadata”.

Co-authorship is granted to active participants of the EERAdata workshop as well as to other contributors to the community paper.

EERAdata wiki - How to?[edit]

The easiest way is you simply start writing and editing. You can install an easy editor by doing what is described in the picture to the right.

How to install an easy editor

Consult the User's Guide for information on using the wiki software.

EERAdata on github[edit]

Link here and share your thoughts and issues! Note, work in progress.

EERAdata @ Research Gate[edit]

Link here and share your thoughts and issues! Note, work in progress.

Just have fun![edit]

Read before playing: This is a little memory game where players need to identify triples, that is three tiles that belong to one data set. All of them relate to low carbon energy. Some tiles have the form of a picture, others show metadata descriptions (in various formats) and some are even sound files. So, the game is to have fun, while learning about metadata. Try out if you can find all triples. Note if you have opened 3 tiles and they do not match, all will automatically close. Just as you know it from a real-life memory game. Link

Agenda and notes[edit]

Builds on read aheads. Online talks and discussions. Space for interaction with participants after each presentation. Moderated discussion of collected comments.

30th November Framing workshop (Part 1): Online hackathon to write a community paper
Link to the recording of the first part of the session: Link1
Link to the recording of the second part of the session Link2
10.00-10.15 Welcome and introduction to the first draft of the planned paper: “Advancing metadata for low carbon energy research - current state and call for action”, Valeria Jana Schwanitz & August Wierling, HVL.
10.15-10.30 Discussion of the structure, collection of comments and notes
10.30-10.45 Creation of writing teams in break out groups. Preferences are collected beforehand if possible.
10.45-15.00 Work in writing teams (break out groups in zoom). The work-flow, coffee and lunch breaks etc. are organized by the groups themselves.
15.00-16.00 Wrap up day 1 of the framing workshop: Reports from the work done in break out groups, collection of feedback and comments from all. Suggestions of questions to be investigated in use case workshops.
1st December Use case “EU policy and energy research taxonomies”

Link to the recording of the session: Link

10:00am - 2:00pm The workshop focuses on metadata, exploring how to link ontologies used for policy-making and those used in research domains. The workshop will use the software "GitMind" to work on the ontology, incorporating suggestions from workshop participants. The participants will form one of the writing teams for the community paper. Link to workshop page with agenda, notes and results
10:00 - 11:00 Intro to the session
  • Welcome and introduction to the topic
  • Present gitmind, create logins
  • Present the 4 ontology islands
  • Share the ontology
  • Present the tasks
11:00-13:00 Time for individual work

(incl. break/Lunch time). Space for discussion as requested.

12:45 - Collecting of contributions

(e.g., gitmind maps elaborating on the 4 ontology islands)

13:00-14:00 Discuss results with everybody


2nd December Use case “Buildings efficiency”

Link to the recording of the session: upcoming

10:00-12:00 (GMT+1) The aim of this session is to build up on discussions that took place in 1st EERADATA workshop together with the participation of the invited experts, utilizing the experience related to the FAIR assessment in order to: a) Highlight main issues with FAIRness, b) Jointly come up with possible remedies and solutions, and c) Identify pointers to metadata standardization, if and how it can be achieved in Buildings Efficiency domain. Link to workshop page with agenda, notes and results
10.00-10.10 Introduction: What does the “Buildings Efficiency” use case of EERADATA aim? Mehmet Efe Biresselioğlu, IUE
10.10-10.25 Standardized Flexibility: FAIR data principles in H2020 projects - ECHOES and ENCHANT - databases, Jens Olgard Dalseth Røyrvik, NTNU Samfunnsforskning
10.25-10.40 Achieving Fair Principles: An Overview of the Hotmaps Project and Hotmaps Building Stock EU28 Dataset, Simon Pezzutto, EURAC
10.40-10.55 openENTRANCE CS1 residential energy demand response: The importance of nomenclature and data transparency for open source energy system model platform, Ryan O’Reilly, Energieinstitut an der JKU
10.55-11.10 Exceed database: Lessons learned and continuous improvement, Daniele Antonucci, EURAC
11.10-11.15 Short Break
11.15-12.00 Panel Discussion with Q&A
3rd December Use case “Metadata user stories”

Link to the recording of the session: Link

10:00-12:00 (GMT+1) The goal of this workshop is to understand how users find, request, use and also provide data. The workshop connects to functional specifications needed for developing the EERAdata platform and other infrastructure in support of data FAIRification work flows. Participants will have the possibility to comment on first mock-ups of the platform and try out some functionalities. For example, using the ontologies produced during UC “EU policy and energy research taxonomies” and participating in virtual discussions on the EERAdata Community Forum. Link to workshop page with agenda, notes and results Link to zoom meeting: upcoming.
10:00-10:15 Welcome & Introduction

Aims and intended role of the platform in the community (Manfred Paier, AIT)

10:15-10:30 Platform elements

Status quo (initial mock-ups, Community Forum) (Astrid Unger, AIT)

10:30-11:30 User stories

Experiences from UC4, UC1, other Use Cases and Gap Analysis

1. Searching for data

2. Linking existing datasets

General discussion

11:30-11:45 Implications for the data model

Top-down metadata, bottom-up metadata (Michael Barber, AIT)

11:45-12:00 Way forward

for designing and implementing the platform (Manfred Paier, AIT)

4th December Art workshop “Collaborative mosaic about FAIR data and Sustainable Development Goals”
9-15 The aim is to create a mosaic which reflects the views of different researchers and data stakeholders on the relevance of FAIR data and the achieving of the Sustainable Development Goals. Participants are ask to contribute their retrospective from the year 2030. The artist Barbara Bellier will connect the individual pieces creating the collaborative mosaic.
4th December EERAdata participates in International FAIR Convergence 2020 Symposium
15-17 M4M Workshops: Making domain-relevant machine-actionable metadata at scale. There is no FAIR data without machine-actionable metadata. Semantically-rich, domain-relevant, machine-actionable metadata is fundamental for FAIR interoperation of data, algorithms and compute resources. However, FAIR-enabling metadata schema often do not yet exist for many communities and need to be created de novo by the relevant domain experts. In other cases there may be already existing metadata schema but they are difficult to find, assess and reuse. In an effort to support convergence of metadata components, the GO FAIR initiative has launched a systematic and scalable approach to the creation, registration and reuse of machine-actionable metadata called Metadata for Machines (M4M) Workshops. M4M workshops bring together representatives from given scientific communities and FAIR metadata experts to jointly define standards for the metadata that they wish to use in practice. We have found scientists to be highly motivated to develop and use metadata standards through our M4M workshops, and tooling such as the CEDAR Workbench to be effective in the collaborative development of new metadata templates. Thirteen M4M Workshops that have been delivered and in this Session, the most recent workshops (for VODAN Africa, DeiC and ZonMW) will be reviewed. This will be followed by a focused discussion on a proposed M4M workshop for EERA Data Consortium to be run early in 2021.

The agenda, roughly, is:

  • to use the first hour to recap recent M4M accomplishments (VODAN Africa, DeiC, ZonMW COVID Program), and
  • to use the second hour to scope plans for an M4M workshop with the EERA Data Consortium
7th December Framing workshop (Part 2): Online hackathon to write a community paper

Link to the recording of the first part of the session: Link1 Link to the recording of the second part of the session: upcoming

10.00-10.45 Welcome and input from use case workshops presented by use case leaders (see above).
10.45-11.00 Guidance for writing teams
11.00-11.15 Creation of writing teams in break out groups (continuation of day 1).
11.15-15.00 Work in break out groups. The work-flow, coffee and lunch breaks etc. are organized by the groups themselves.
15.00-16.00 Wrap up day 2 of the framing workshop: Reports from the work done in break out groups. Collection of feedback and comments from all. Guidance on how to finalize the paper collaboratively after the workshop.

Metadata knowledge base[edit]

Domain_specific_classification_schemes