FreeDict is built on open and open-source data. This page credits every dictionary and dataset we use, with its license.
Definitions, etymology and translations are derived from English Wiktionary and its sister-language editions, available under the Creative Commons Attribution-ShareAlike 4.0 licence. Content derived from Wiktionary remains under that licence.
Licence: CC BY-SA 4.0
· en.wiktionary.org
CEFR level data comes from the CEFR-J Wordlist, copyright Tono Laboratory at Tokyo University of Foreign Studies. It is made available for research and commercial use without charge, on condition the dataset is cited.
Licence: Free for research and commercial use, with citation
· www.cefr-j.org
Wiktionary data is imported via Kaikki.org, Tatu Ylonen's machine-readable extraction of Wiktionary produced by the Wiktextract project. The underlying content is Wiktionary's and carries its CC BY-SA licence.
Licence: CC BY-SA 4.0
· kaikki.org
Higher CEFR levels (C1 and C2) come from the Octanove Vocabulary Profile by Octanove Labs, available under the Creative Commons Attribution-ShareAlike 4.0 licence. It is a separate dataset from the CEFR-J Wordlist and carries a different licence.
Licence: CC BY-SA 4.0
· github.com
Synonym, antonym and semantic-relation data comes from the Open English WordNet, which builds on Princeton WordNet. Available under the Creative Commons Attribution 4.0 licence.
Licence: CC BY 4.0
· github.com
Some example sentences come from the Tatoeba Project, a collection of sentences contributed by volunteers. Available under the Creative Commons Attribution 2.0 France licence, with a portion released under CC0 1.0.
Licence: CC BY 2.0 FR
· tatoeba.org
Some historical definitions and etymologies come from the 1913 Webster's Revised Unabridged Dictionary, which is in the public domain and available via Project Gutenberg.
Licence: Public domain
· www.gutenberg.org
Some bilingual dictionaries are built from WikDict, which compiles translation data from Wiktionary. The underlying content carries Wiktionary's CC BY-SA licence.
Licence: CC BY-SA 4.0
· www.wikdict.com
Word frequency data comes from wordfreq by Robyn Speer. The frequency data is available under the Creative Commons Attribution-ShareAlike 4.0 licence and is itself built from several corpora, which wordfreq asks be credited: the Google Books Ngram Viewer, the SUBTLEX word lists by Marc Brysbaert and colleagues, and OpenSubtitles.
Licence: CC BY-SA 4.0 (data)
· github.com
A large part of FreeDict is written in-house: the learner layer on word pages, the comparison articles, grammar guides and vocabulary lists. This material is our own and is not derived from the third-party sources listed on this page.
Licence: FreeDict original content
· freedict.com
Everything on FreeDict that is not derived from the sources above — the plain-English usage notes, the comparison articles, grammar guides, vocabulary lists, quiz content and editorial definitions — is written in-house and is © FreeDict.com, all rights reserved.
You are welcome to quote briefly with a link. Republishing substantial parts of it, or using it to build or populate another dictionary, is not permitted without written permission — the open licences above apply to the source data, not to our editorial. If you want to use something, ask us.