Back to home

Credits and attributions

CardyFlash is possible thanks to a global community of researchers and developers who share their data and tools openly.

wordfreq library

CC BY-SA 4.0

Our app uses the wordfreq library to assess the frequency and difficulty of the words generated by AI.

wordfreq data is distributed under the Creative Commons Attribution-ShareAlike 4.0 license (CC BY-SA 4.0). In accordance with this license, we acknowledge and thank Robyn Speer and the project's contributors for their invaluable work.

Main data sources

This app's lexical knowledge base draws on several high-quality open corpora:

OpenSubtitles & SUBTLEX

To capture natural spoken language, we include frequency data from OpenSubtitles.

Special SUBTLEX notice

We use the frequency lists from SUBTLEX-US, SUBTLEX-UK, among others, created by Marc Brysbaert et al.

“It is a core policy of the SUBTLEX authors that this data be freely available for any purpose. CardyFlash does not sell this raw downloadable data; we provide an AI-based management and generation tool that uses these frequencies to improve the study experience. We thank the Centre for Reading Research at Ghent University for their contribution.”

Commitment to free software

“Knowledge has no owner, but the effort to share it deserves all our respect.”