TRANSPARENCY
Data sources and editorial policy
StrokeSheet combines openly licensed language data with our own interface, learning workflows, and quality checks. This page explains what the data can—and cannot—tell you.
Last reviewed: September 7, 2026
Primary reference sources
CC-CEDICT
CC BY-SA 4.0Chinese–English dictionary readings and definitions used as reference data.
View upstream source →Make Me a Hanzi and Hanzi Writer Data
See the upstream Arphic Public License attribution and project terms.Character stroke paths, medians, and decomposition-related reference data used by stroke-order displays.
View upstream source →Hanzi Writer
MITOpen-source library used to render and animate Chinese character stroke order.
View upstream source →krmanik/HSK-3.0
CC BY-SA 4.0New HSK vocabulary dataset used for level browsing and study references.
View upstream source →What StrokeSheet adds
We organise the reference data into searchable character, Pinyin, HSK, radical, component, and stroke-count pages; build worksheet and review tools; and write practical learning guides. Automated checks catch missing or internally inconsistent fields, but they do not turn every source entry into an independently verified linguistic claim.
Editorial boundaries
Meanings are concise study references, not exhaustive translations. A character may have multiple readings or meanings depending on the word and context. We do not invent unsupported word origins, example sentences, rankings, or historical explanations. When reliable information is unavailable, we omit the claim or keep the page out of search indexing.
HSK scope
HSK pages are study aids based on the cited dataset. Learners should consult current official examination guidance for registration requirements, test format, and the syllabus used by their exam centre.
Report an error
Email info@strokesheet.com with the page URL, the field you believe is wrong, your proposed correction, and a reliable reference when possible. We review corrections before changing shared data.