Languages: Asia and South America #71

Open
opened 2026-09-06 20:18:31 +00:00 by tiagoagueda · 1 comment
Owner

Decision

milestone for 0.4.0 - all languages from Asia and South America

Shape

Asia. Mandarin Chinese (simplified and traditional are two catalogues, not one),
Hindi, Bengali, Urdu, Japanese, Korean, Vietnamese, Thai, Indonesian, Malay, Tagalog,
Tamil, Telugu, Marathi, Gujarati, Punjabi, Persian, Hebrew, Kazakh, Uzbek, Burmese, Khmer,
Lao, Nepali, Sinhala.

South America. Spanish and Portuguese already arrived with 0.2.0 and cover most of the
continent, but they are the European spellings: pt-BR and the Latin American Spanish
variants are separate catalogues, because "ficheiro" and "arquivo" are not the same word to
the person reading them. Beyond those: Quechua, Guaraní, Aymara.

What is new here and was not in 0.3.0:

  1. Right-to-left again, for Hebrew, Persian and Urdu — but #67 will have done the layout,
    so these are catalogues.
  2. Scripts that need real font coverage: CJK, Devanagari, Thai, Khmer, Burmese. The same
    test as 0.3.0, on a larger scale: a glyph that does not draw is a box on somebody's CV.
  3. Languages with no plural distinction at all (Chinese, Japanese, Korean, Vietnamese,
    Thai): nplurals=1, which the tooling has to accept as readily as Arabic's six.
  4. Word wrapping and line breaking for CJK and Thai, in the browser and in WeasyPrint.
    Neither breaks on spaces, and a CV that wraps in the wrong place looks careless.
  5. Regional variants as first-class catalogues, which the language picker has to present
    without turning into a list nobody can scan.

Classification

Enhancement. Not breaking.

Depends on

#67 for right-to-left. Sequenced after the African languages, so the per-language checks
that 0.3.0 establishes — plurals, formats, fonts, theme titles — are a routine by the time
this many catalogues arrive at once.

## Decision > milestone for 0.4.0 - all languages from Asia and South America ## Shape **Asia.** Mandarin Chinese (simplified and traditional are two catalogues, not one), Hindi, Bengali, Urdu, Japanese, Korean, Vietnamese, Thai, Indonesian, Malay, Tagalog, Tamil, Telugu, Marathi, Gujarati, Punjabi, Persian, Hebrew, Kazakh, Uzbek, Burmese, Khmer, Lao, Nepali, Sinhala. **South America.** Spanish and Portuguese already arrived with 0.2.0 and cover most of the continent, but they are the European spellings: **pt-BR** and the Latin American Spanish variants are separate catalogues, because "ficheiro" and "arquivo" are not the same word to the person reading them. Beyond those: Quechua, Guaraní, Aymara. **What is new here and was not in 0.3.0:** 1. **Right-to-left again**, for Hebrew, Persian and Urdu — but #67 will have done the layout, so these are catalogues. 2. **Scripts that need real font coverage**: CJK, Devanagari, Thai, Khmer, Burmese. The same test as 0.3.0, on a larger scale: a glyph that does not draw is a box on somebody's CV. 3. **Languages with no plural distinction at all** (Chinese, Japanese, Korean, Vietnamese, Thai): `nplurals=1`, which the tooling has to accept as readily as Arabic's six. 4. **Word wrapping and line breaking** for CJK and Thai, in the browser and in WeasyPrint. Neither breaks on spaces, and a CV that wraps in the wrong place looks careless. 5. **Regional variants as first-class catalogues**, which the language picker has to present without turning into a list nobody can scan. ## Classification Enhancement. Not breaking. ## Depends on #67 for right-to-left. Sequenced after the African languages, so the per-language checks that 0.3.0 establishes — plurals, formats, fonts, theme titles — are a routine by the time this many catalogues arrive at once.
tiagoagueda added this to the 0.4.0 milestone 2026-09-06 20:18:31 +00:00
Author
Owner

Re-scoped by #269 and moved to 0.5.0.

Its Asian catalogues are the font-gated tier: Han, Devanagari, Bengali, Tamil and Thai all
wait on #74, and Mandarin waits with them despite having more speakers than anything else on
the list — a page of tofu boxes is worse than the same page in English.

Its right-to-left members — ur, fa, he — move to Tier 3 (0.6.0) with the rest of the
RTL work, for the same reason ar does.

Quechua, Guaraní and Aymara are Latin-script and blocked by nothing; whoever starts Tier 1
may pull them forward rather than leave them waiting on fonts they do not need.

Re-scoped by #269 and **moved to 0.5.0**. Its Asian catalogues are the font-gated tier: Han, Devanagari, Bengali, Tamil and Thai all wait on #74, and Mandarin waits with them despite having more speakers than anything else on the list — a page of tofu boxes is worse than the same page in English. Its right-to-left members — `ur`, `fa`, `he` — move to Tier 3 (0.6.0) with the rest of the RTL work, for the same reason `ar` does. Quechua, Guaraní and Aymara are Latin-script and blocked by nothing; whoever starts Tier 1 may pull them forward rather than leave them waiting on fonts they do not need.
tiagoagueda modified the milestone from 0.4.0 to 0.5.0 2026-09-17 19:44:14 +00:00
Sign in to join this conversation.
No milestone
No project
No assignees
1 participant
Notifications
Due date
The due date is invalid or out of range. Please use the format "yyyy-mm-dd".

No due date set.

Reference
Postulo/postulo#71
No description provided.