MANULEX: a grade-level lexical database from French elementary school readers.

Abstract : This article presents MANULEX, a Web-accessible database that provides grade-level word frequency lists of nonlemmatized and lemmatized words (48,886 and 23,812 entries, respectively) computed from the 1.9 million words taken from 54 French elementary school readers. Word frequencies are provided for four levels: first grade (G1), second grade (G2), third to fifth grades (G3-5), and all grades (G1-5). The frequencies were computed following the methods described by Carroll, Davies, and Richman (1971) and Zeno, Ivenz, Millard, and Duvvuri (1995), with four statistics at each level (F, overall word frequency; D, index of dispersion across the selected readers; U, estimated frequency per million words; and SFI, standard frequency index). The database also provides the number of letters in the word and syntactic category information. MANULEX is intended to be a useful tool for studying language development through the selection of stimuli based on precise frequency norms. Researchers in artificial intelligence can also use it as a source of information on natural language processing to simulate written language acquisition in children. Finally, it may serve an educational purpose by providing basic vocabulary lists.
Type de document :
Article dans une revue
Behavior Research Methods Instruments and Computers, Psychonomic Society, 2004, 36 (1), pp.156-66
Liste complète des métadonnées

Littérature citée [31 références]  Voir  Masquer  Télécharger

https://hal.archives-ouvertes.fr/hal-00733549
Contributeur : Liliane Sprenger-Charolles <>
Soumis le : lundi 24 septembre 2012 - 11:46:52
Dernière modification le : vendredi 26 janvier 2018 - 01:02:22
Document(s) archivé(s) le : vendredi 16 décembre 2016 - 16:41:57

Fichier

Lete_Sprenger-Charolles_Cole_M...
Fichiers produits par l'(les) auteur(s)

Identifiants

  • HAL Id : hal-00733549, version 2
  • PUBMED : 15190710

Collections

Citation

Bernard Lété, Liliane Sprenger-Charolles, Pascale Colé. MANULEX: a grade-level lexical database from French elementary school readers.. Behavior Research Methods Instruments and Computers, Psychonomic Society, 2004, 36 (1), pp.156-66. 〈hal-00733549v2〉

Partager

Métriques

Consultations de la notice

269

Téléchargements de fichiers

634