TRANSPARENCY

Data sources and editorial policy

StrokeSheet combines openly licensed language data with our own interface, learning workflows, and quality checks. This page explains what the data can—and cannot—tell you.

Last reviewed: September 7, 2026

Primary reference sources

CC-CEDICT

CC BY-SA 4.0

Chinese–English dictionary readings and definitions used as reference data.

View upstream source →

Make Me a Hanzi and Hanzi Writer Data

See the upstream Arphic Public License attribution and project terms.

Character stroke paths, medians, and decomposition-related reference data used by stroke-order displays.

View upstream source →

krmanik/HSK-3.0

CC BY-SA 4.0

New HSK vocabulary dataset used for level browsing and study references.

View upstream source →

What StrokeSheet adds

We organise the reference data into searchable character, Pinyin, HSK, radical, component, and stroke-count pages; build worksheet and review tools; and write practical learning guides. Automated checks catch missing or internally inconsistent fields, but they do not turn every source entry into an independently verified linguistic claim.

Editorial boundaries

Meanings are concise study references, not exhaustive translations. A character may have multiple readings or meanings depending on the word and context. We do not invent unsupported word origins, example sentences, rankings, or historical explanations. When reliable information is unavailable, we omit the claim or keep the page out of search indexing.

HSK scope

HSK pages are study aids based on the cited dataset. Learners should consult current official examination guidance for registration requirements, test format, and the syllabus used by their exam centre.

Report an error

Email info@strokesheet.com with the page URL, the field you believe is wrong, your proposed correction, and a reliable reference when possible. We review corrections before changing shared data.