Contents for analysing data: Difference between revisions

From EERAdata Wiki
Jump to navigation Jump to search
Access restrictions were established for this page. If you see this message, you have no access to this page.
Created page with "This step analyses the data to support the As a preparation of the FAIRification itself, this step also includes checking if some FAIR features already are existing in the..."
 
No edit summary
 
(6 intermediate revisions by the same user not shown)
Line 1: Line 1:
This step analyses the data to support the  
A proper understanding of the data is essential for carrying out any FAIRification activity. If the data are own data or coming from an in-house activity, such an understanding may come easily. But if the data are provided by a third party, a detailed analysis might be necessary.  This step analyses the data to support the FAIRification step. Issues to be considered are:


* How are the data organized? Do the data meet the intended formatting? Are data missing?


As a preparation of the FAIRification itself, this step also includes checking if some FAIR features already  
* Are some FAIR features already existing in the data such as persistent identifiers? If the data are extensive, running a (semi-) automatic FAIR assessment tool is helpful.
are existing in the data such as persistent identifiers. If the data are extensive, running  
 
Here is a list of (semi-) automatic tools:
The second step is to analyze the data to prepare for subsequent FAIRification (e.g., improving
 
interoperability) and is within the pre-FAIRification phase of the workflow. This process may include:
* [https://fairsharing.github.io/FAIR-Evaluator-FrontEnd/#!/ FAIR Evaluation Services]
1) investigating the data in whatever form(s) it is available (specified in Step 1) and checking whether
 
both the data representation (format) and the meaning of the data elements (the data semantics) are clear
* [https://www.fairsfair.eu/f-uji-automated-fair-data-assessment-tool F-UJI Automated FAIR Data Assessment Tool]
and unambiguous, and 2) checking whether the data already contain FAIR features, such as persistent
 
unique identifiers for data elements [14] (FAIR principle F1 [1]) by e.g., using FAIRness assessment tooling
* [https://fair-checker.france-bioinformatique.fr/ FAIR-Checker]
[2, 3, 4]. It is evident that this step is tightly connected with Step 1 since e.g., selecting a relevant subset
of the data and defining driving user questions(s) are highly relying on being familiar with the data.

Latest revision as of 15:11, 17 November 2022

A proper understanding of the data is essential for carrying out any FAIRification activity. If the data are own data or coming from an in-house activity, such an understanding may come easily. But if the data are provided by a third party, a detailed analysis might be necessary. This step analyses the data to support the FAIRification step. Issues to be considered are:

  • How are the data organized? Do the data meet the intended formatting? Are data missing?
  • Are some FAIR features already existing in the data such as persistent identifiers? If the data are extensive, running a (semi-) automatic FAIR assessment tool is helpful.

Here is a list of (semi-) automatic tools: