INDEPENDENT RESEARCH EDITIONLANGUAGE · EDUCATION · WORLDAbout this atlas
A world of Japanese.
Mapped, measured, read.
The Kotoba Gazette
Seven survey waves
2006—2024
KOTOBA ATLAS / NEWSPAPER EDITIONJF survey data · 2025 country profiles
THE WORLD / JAPANESE-LANGUAGE EDUCATION

Japanese, beyond Japan.

Where a language is taught tells one story.
What learners read tells another.

SURVEY WAVES / 2006—2024
2024 · 7 / 7
01 / THE ATLASCOUNTRY/AREA-LEVEL SURVEY DATA
SURVEY 2024
世界の日本語THE WORLD / GLOBAL DISTRIBUTION
Height: square-root scale · Click a column or country · Scroll to zoom

Explore countries & areas

COUNTRY / AREALEARNERS ↓SCRIPT BACKGROUNDTEXTBOOK COVERAGE
03 / THROUGH TIME

The changing world of Japanese.

Compare countries, areas, regions and measures on one plot. Add selections below, then toggle individual series in the legend.

Reading this comparison

Hover or focus a point to inspect a value.

Year-by-year categories, sources and comparison limits

These are repeated surveys of institutions, not a fixed panel of the same schools or learners. Changes can reflect education activity, coverage and response. Japan is outside scope. Country names are linked across renamed entities; an absent country-year is missing, never an assumed zero. Geometry uses a single present-day base map, not historical borders.

Common stage view: Primary + secondary, higher education, and education outside schools. Native primary/secondary splits are retained where the source reports them. They are never estimated from a combined total. The multiple-stage category in 2009/2012 remains separate; stage-specific growth is not directly comparable with later redistributed learner counts.

Teacher counts: counting practices changed across waves, especially 2021→2024. The 2024 report counts a teacher in each educational stage taught within an institution, unlike the institution total collected in 2021. The teacher chart marks these breaks with dashed segments and suppresses growth percentages across a definition break.

Other survey categories: ownership, degree awards, teacher training, motivations, implementation problems/status and online teaching use different questionnaires or availability. This atlas links each year's original tables and documents their availability; it does not merge these categories into a fabricated comparable panel. Textbook mentions and cover previews remain a separate 2025 profile snapshot.

04 / COMPARATIVE STUDIES

Patterns across borders.

Compare education counts or ratios, learning settings and encoded textbook evidence. The saved default uses the strongest tested configuration on a documented cohort.

Features · availability follows the survey year

Exploratory groups, not validated learner types. Region and script background are inputs, so geographical separation is partly built in. Curation status may group countries by data availability; deselect it to check sensitivity.

How the mixed-data clustering works

Choose deterministic k-medoids or average-linkage clustering over a group-weighted mixed dissimilarity. Each selected block has equal weight; numeric counts and ratios use log(1+x) and within-cohort range scaling. A ratio is missing when its denominator is missing or zero. Ratio blocks and count blocks can be selected separately; using both can repeat scale-related information.

Textbook features: one binary presence per normalized title-family label per area, regardless of repeated stage records. Only reported-use evidence enters the encoding; historical/publication mentions and explicitly identified digital resources are excluded. Labels are Unicode/space normalized, not inferred ISBN matches. The vocabulary retains labels documented in at least two cohort areas. One-hot uses binary weights; TF–IDF uses binary TF × (log((1 + reviewed areas)/(1 + document frequency)) + 1), followed by L2 normalization. Cosine distance contributes one feature block, so a large vocabulary does not overwhelm numeric ratios. Empty or unreviewed vectors are missing, not assumed absence. Cluster cards show the strongest mean encoded features, not a mean title count.

Evaluated default: candidate feature sets, both encodings, both algorithms and k = 2…8 are compared on the same complete-case 2024 cohort. Selection balances silhouette and five deterministic 80% subsample stability checks (adjusted Rand index), subject to minimum-cluster-size and maximum-cluster-share safeguards. This is the best tested configuration under these criteria, not an objectively true partition. Region, script background and review status are excluded from default selection to avoid groups driven directly by those labels. They remain optional exploratory inputs. The evaluation and assignments are saved in the open data.

Missing groups are omitted pairwise and weights renormalized; a run with no comparable fields for a pair is rejected. The medoid minimizes within-group distance. Silhouette measures separation under the chosen distance, not pedagogical validity. Historical waves disable contemporary textbook features. Group numbers are recomputed per wave, not persistent identities.

Reference: mixed-type dissimilarities and missing values ↗
05 / THE OPEN DATA DESK

Follow every record back to its source.

Download the evidence behind this edition: original title wording, PDF pages, education settings and review coverage. The survey panel spans 2006–2024; textbook evidence is a separate 2025 profile snapshot.

A publication mention is not adoption. A local percentage is not a national rate. Title families do not identify a specific edition or establish readability.

Repository ↗

SOURCE NOTES / 01

Sources and measurement.

Historical panel: Seven survey waves from 2006 through 2024. Sources and year-specific categories are available in “Through time”. Maps, tables, statistics and clustering follow the selected year. The 2025 textbook profile layer is available only beside the latest survey snapshot; it is never backfilled into historical survey rows.

Population denominators: 2024 uses the population column printed in JF’s regional tables (149 values; Kosovo is missing). Those footnotes cite UN data available in January 2025 and Taiwan’s December 2024 population. Most underlying population reference dates are unspecified, not necessarily 2024. Areas absent from those tables and earlier waves retain annual external sources. Population ratio trends have a denominator-source break in 2024. Each country’s source link identifies the relevant PDF page.

2024 survey: country/area totals for learners, teachers, institutions and educational stages, imported from corrected table 1-1a (March 2026). The 204 table rows include zero counts; 143 report learners. These counts describe institutional education abroad, not all Japanese learners. Japan is outside the scope.

2025 country profiles: 1,299 page-linked material observations across 123 country/area profiles, with 166 profiles reviewed within documented scope. Exact source titles, setting, pages and evidence type are retained. Reported use, curriculum material, historical use and publication mentions are distinguished; a mention is not current adoption. Sources unavailable for this review are marked explicitly. The textbook section is the main scope; adjacent digital-resource coverage varies and is documented per profile. This is not an exhaustive bibliography of every PDF. Download review coverage.

Kanji / non-kanji areas: an explicit atlas convention, not a Japan Foundation variable. The default kanji group comprises China, Taiwan, Hong Kong and Macao. Other surveyed areas are grouped as non-kanji for this comparison. An optional switch includes South Korea, reflecting the alternative convention that includes historical Hanja literacy. Japan is outside the overseas survey. Singapore and other multilingual areas can include Chinese-literate learners; neither category assigns an individual learner's L1 or kanji knowledge. Do not interpret this geography as measured reader background. The active scheme and category are included in CSV exports.

Cover previews: 11 verified title-family covers, displayed where the audited title matches a verified cover family, sourced from publisher or textbook developer pages and credited in the enlarged preview. A cover represents one identifiable volume/edition of a title family; it does not establish which edition or translation is used in a country. Missing or ambiguous matches are shown as “Cover not verified.” Cover copyright remains with the respective rights holders.

The map uses a logarithmic color scale for survey counts. In the textbook layer, the selected title is highlighted where a curated 2025 profile mentions it; other curated profiles and unreviewed areas are distinguished. Missing mentions do not establish absence of use. Optional 3D columns add a square-root height scale for the active metric; the textbook layer instead highlights one distinct title family at a time, with equal-size location markers. Shared scale, enabled by default, uses the maximum over all seven survey waves. Turning it off uses the selected wave maximum. Both stay fixed under geographical filtering. Zero values and uncurated textbook fields have no column; consult the country table to distinguish them. Columns are illustrative anchored markers, not extruded territorial boundaries. Small-area marker positions may be slightly offset for visibility. Filters affect the map and table; headline totals always describe the full survey. Small areas absent from the simplified map remain searchable in the table. Boundaries are illustrative and do not express a position on territorial status.

Source: The Japan Foundation / 出典:国際交流基金(加工して作成)。Independent research visualization. This newspaper edition is a separate design of Kotoba Atlas; it is not affiliated with a news publisher or the Japan Foundation. Survey panel retrieved 7 October 2026; material evidence reviewed 8 October 2026.

SELECTED OBSERVATION

SeriesYearValueSince previous waveSource