Technical Report: CSVM Ecosystem

Abstract : The CSVM format is derived from CSV format and allows the storage of tabular like data with a limited but extensible amount of metadata. This approach could help computer scientists because all information needed to uses subsequently the data is included in the CSVM file and is particularly well suited for handling RAW data in a lot of scientific fields and to be used as a canonical format. The use of CSVM has shown that it greatly facilitates: the data management independently of using databases; the data exchange; the integration of RAW data in dataflows or calculation pipes; the search for best practices in RAW data management. The efficiency of this format is closely related to its plasticity: a generic frame is given for all kind of data and the CSVM parsers don't make any interpretation of data types. This task is done by the application layer, so it is possible to use same format and same parser codes for a lot of purposes. In this document some implementation of CSVM format for ten years and in different laboratories are presented. Some programming examples are also shown: a Python toolkit for using the format, manipulating and querying is available. A first specification of this format (CSVM-1) is now defined, as well as some derivatives such as CSVM dictionaries used for data interchange. CSVM is an Open Format and could be used as a support for Open Data and long term conservation of RAW or unpublished data.
Liste complète des métadonnées

Littérature citée [2 références]  Voir  Masquer  Télécharger

https://hal.archives-ouvertes.fr/hal-00730995
Contributeur : Frédéric Rodriguez <>
Soumis le : mardi 11 septembre 2012 - 16:21:36
Dernière modification le : vendredi 18 janvier 2019 - 15:18:02
Document(s) archivé(s) le : vendredi 16 décembre 2016 - 11:50:04

Fichier

CSVM_ecosystem_1.22b.pdf
Fichiers produits par l'(les) auteur(s)

Identifiants

  • HAL Id : hal-00730995, version 1
  • ARXIV : 1209.2946

Collections

Citation

Frédéric Rodriguez. Technical Report: CSVM Ecosystem. 31 pages including 2p of Annex. 2012. 〈hal-00730995〉

Partager

Métriques

Consultations de la notice

254

Téléchargements de fichiers

229