NTJ Japanese Learning Answer Dataset — August 2026 NTJ 日語學習作答彙總資料集 2026 年 8 月 Version 1.1.0 · snapshot 2026-08-14 · temporal coverage 2026-08-02/2026-08-13 56,613 answer events from 1,053 anonymous learner identifiers. Canonical page: https://www.nihongotojapan.com/research/japanese-learning-data-2026 LICENCE CC-BY-4.0 — https://creativecommons.org/licenses/by/4.0/ CC BY 4.0 covers the data only: the aggregate figures on this page and the two downloadable files (japanese-learning-2026.json and japanese-learning-2026.csv). The charts produced by Nihongo to Japan (SVG/PNG) and the text of the report are not covered and remain rights-reserved. The data is released under CC BY 4.0: you may copy, redistribute, adapt and build upon it, including commercially, provided you give appropriate credit and indicate if changes were made. ⚠️ CC BY is irrevocable — anyone who receives the data under these terms keeps those rights even if the publisher changes the terms later. ATTRIBUTION (copy this) Nihongo to Japan, Japanese Learning Answer Dataset — August 2026, v1.1.0, CC BY 4.0, https://www.nihongotojapan.com/research/japanese-learning-data-2026 PERMITTED WITHOUT ASKING - Cite the statistics in papers, research reports, teaching materials and journalism, including commercial publications. - Re-analyse the downloadable aggregate figures and produce your own charts (those charts are yours). - Adapt, remix, transform and redistribute in any form, including redistributing the data files wholesale or including them in your own data product. - Use in classrooms and educational settings without asking first. REQUESTED — these are requests, NOT licence conditions - Carry the sampling limitations alongside the figures (self-selected observational sample, 12 days only, not representative). This is a request, not a licence condition — but quoting the numbers without them publishes a misleading statistic. - Link back to this page so readers can see the methodology and the version. - Say "learner identifiers" rather than "people"; the identifier count is an upper bound on the number of people. - State the version (e.g. v1.1.0, snapshot 2026-08-14) so the figures you cited stay identifiable after later updates. NOT COVERED BY THIS LICENCE - The charts produced by Nihongo to Japan (public/images/research/*.svg and *.png): not covered by CC BY — please contact us before reproducing them. Making your own charts from the data needs no permission. - The text of the report and of this website: not covered by CC BY; rights reserved. - The site name, logo and brand marks: this licence grants no rights to use them. HOW TO READ THE DATA RESPONSIBLY - Every figure is an observed value within this dataset, not an inference about any population. - The sample is self-selected and observational; it is not a representative survey. - The levels use different question sets and there is no IRT equating, so the data cannot be used to compare the objective difficulty of JLPT levels. - "learner identifiers" are anonymous IDs that cannot be merged across devices; the count is an upper bound on the number of people, not a headcount. Contact: https://www.nihongotojapan.com/research/japanese-learning-data-2026#contact This file is generated by scripts/build-research-factpack.mjs — do not edit it by hand.