Data Credits & Licenses

HSK Nest is built on outstanding open datasets. This page credits them as their licenses require — and because they deserve it.

HSK vocabulary

The HSK 1–9 and frequency word lists are derived from complete-hsk-vocabulary by Yanis Zafirópulos, used under the MIT License.

Chinese–English dictionary

Word-entry suggestions for Chinese lists are powered by CC-CEDICT, © the CC-CEDICT editors and contributors, used under the Creative Commons Attribution-ShareAlike 4.0 International License (CC BY-SA 4.0). Our build of the dictionary data has been modified from the original: entries are trimmed, and for some words the order of senses is curated so the most useful everyday meaning is shown first to learners. These modifications are shared under the same license in the open-source repository.

Example sentences

Chinese–English example sentences come from Tatoeba (via the manythings.org sentence pairs), used under the CC-BY 2.0 (France) license. Each sentence stores its individual attribution (sentence IDs and contributor usernames), preserved verbatim in the app's database and data files.

Fonts

Interface and code text use Geist and Geist Mono by Vercel, used under the SIL Open Font License 1.1.

Application code

HSK Nest itself is open source under the GNU AGPL-3.0 license — source code on GitHub.