Skip to content

Chapter 12Katakana Orthography and the Extended Small-Kana Codas

The katakana writing of Ainu — the syllabary mapping, the extended small-kana coda letters, the modern normative standard, and historical kana systems.

12.1 Katakana in Ainu writing

Katakana functions alongside the Latin phonemic transcription as one of two co-primary scripts for Hokkaido Ainu. Most community publications, pedagogical materials, and digital media produced in Hokkaido use katakana as their default form, while the Latin transcription defined in Chapter 11 (The Latin Phonemic Transcription) prevails in the academic literature and in this grammar. Both scripts represent the same underlying phonological system; the correspondence between them is described in Chapter 11 (The Latin Phonemic Transcription) and summarized in the coda correspondence table below.

The katakana syllabary handles the open (CV) core of the Ainu syllable with standard Japanese kana. The structural challenge lies in the closed (CVC) syllable: Ainu allows a wide coda-consonant inventory — /p t k s m n r w y h/ — most of which Japanese kana provides no dedicated symbol for in coda position. The solution is a set of reduced-size katakana characters placed after the preceding CV mora to represent the coda consonant. The modern normative form of this system derives from the 1994 Hokkaido Utari Kyōkai textbook アコㇿイタㇰ Hokkaidō Utari Kyōkai (1994), whose title (Akor Itak, 'our language') already illustrates the convention: ㇿ for coda /r/ and ㇰ for coda /k/. Nakagawa's extensive grammar follows the same assignments throughout Nakagawa (2024).

12.2 The standard CV syllabary

For open syllables, Ainu uses the inherited katakana inventory without structural modification. The five vowels are ア /a/, イ /i/, ウ /u/, エ /e/, オ /o/; the consonant-plus-vowel mora grid follows the Japanese kana table, but the phonemic values of several cells diverge from modern Japanese reading conventions. ツ corresponds to Ainu /tu/ and チ to /ti/ — reflecting the phonemically straightforward /t/ plus the respective vowel, without the Japanese phonetic shifting of /tu/ to [tsɯ] or /ti/ to [tɕi] Nakagawa (2024). Similarly, ス = /su/ without the Japanese fronting of /s/ before /u/; the allophonic conditions on /s/ and /c/ in Ainu are described in Chapter 16 (The Consonant Inventory and Its Phonetic Realization).

The affricate /c/ and its katakana representation carry an orthographic history. Kindaichi's early grammars and text editions introduced contracted kana (yōon: チャ, チュ, チェ) for /ca/, /cu/, /ce/, treating Ainu /c/ as parallel to Japanese /tʃ/ Kindaichi & Chiri (1936). This practice conflicted with the writing habits of Ainu speakers themselves, who transcribed /c/ without the contracted syllable convention: Kirikae documents that writers such as Sunazawa, Yamamoto, and Nabewasawa consistently avoided yōon for /c/ Kirikae (1997), and Nakagawa interprets this uniformity as reflecting a shared native-speaker phonological intuition Nakagawa (2006). The normative standard dropped yōon for /c/ Hokkaidō Utari Kyōkai (1994); community discussion has confirmed the consensus aynu-corpora Discord (2023–2026) (nukopoli, aynu-corpora Discord, 2024-11-17) ‹consensus›. The full convergence history belongs to Chapter 15 (Orthographic Standardization and the Word-Division Problem).

The Ainu coda inventory and phonotactic constraints on coda position are defined in Chapter 21 (Syllable Structure, Phonotactics, and Word-Edge Constraints); the phonemes themselves are described in Chapter 16 (The Consonant Inventory and Its Phonetic Realization).

12.3 The extended small-kana coda system

A coda consonant in Ainu katakana is written as a small (subscript-sized) katakana character placed immediately after the kana of the preceding mora. These characters occupy the Katakana Phonetic Extensions block, U+31F0–U+31FF, added to the Unicode Standard in version 3.2 (2002) specifically to support Ainu orthography. The block encodes sixteen positions, each derived from the syllable kana in the same consonant row as the coda value it encodes. The table below gives the standard coda assignments; variants and special cases are discussed in the subsections that follow.

Latin codaKatakanaUnicode position(s)Notes
kU+31F0 (small ku)
sU+31F1 (small si)standard; ㇲ (U+31F2, small su) found in some older or variant texts
tU+31F3 (small to)
nU+30F3 (full-size N)full-size ン; the small ン (𛅧 U+1B167) has been proposed but has limited font support — discussed in the /n/ subsection below
pㇷ゚U+31F7 + U+309A (two code points)composite: small hu (ㇷ) plus combining semi-voiced mark (゚); no pre-composed single character exists — discussed in the /p/ subsection below
mU+31FA (small mu)
rㇻ ㇼ ㇽ ㇾ ㇿU+31FB–U+31FF (small ra ri ru re ro)standard system selects by preceding vowel: ㇻ after /a/, ㇼ after /i/, ㇽ after /u/, ㇾ after /e/, ㇿ after /o/ — discussed in the /r/ subsection below
wU+30A6 (full-size U)no dedicated small kana; full ウ follows the preceding mora in glide sequences
yU+30A4 (full-size I)no dedicated small kana; full イ follows the preceding mora in glide sequences
hㇵ ㇶ ㇷ ㇸ ㇹU+31F5–U+31F9 (small ha hi hu he ho)member selected by preceding vowel; phonemic status of coda /h/ is disputed — discussed in the /h/ subsection below

12.3.1 The bilabial coda ㇷ゚

Coda /p/ has no pre-composed Unicode character in the Katakana Phonetic Extensions block. The standard encoding combines small hu (ㇷ, U+31F7) with the combining katakana-hiragana semi-voiced sound mark (゚, U+309A), yielding the two-code-point grapheme cluster ㇷ゚. This parallels the pre-composed semi-voiced syllable kana of Japanese (パ = U+30D1; プ = U+30D7, etc.), but no pre-composed equivalent exists for the small form. Text-processing operations that count code points rather than grapheme clusters will report ㇷ゚ as two characters; software producing Ainu katakana must normalize the sequence at input and search time. Fonts that support the full Katakana Phonetic Extensions block generally render the combination correctly aynu-corpora Discord (2023–2026) (toracatman, aynu-corpora Discord, 2024-03-28).

12.3.2 The /r/ coda series

Coda /r/ is distributed across five small kana — ㇻ ㇼ ㇽ ㇾ ㇿ (small ra, ri, ru, re, ro; U+31FB–U+31FF) — one for each of the five Ainu vowels. The standard convention selects the member whose vowel matches the vowel of the preceding syllable: the title アコㇿイタㇰ illustrates ㇿ (small ro) for the coda /r/ of kor, following the /o/ of Hokkaidō Utari Kyōkai (1994).

The five-way choice is not universally observed. The Upopoy National Ainu Museum uses full-size ル (U+30EB) for all coda /r/ positions, without the small-kana form aynu-corpora Discord (2023–2026) (thegodofneet, aynu-corpora Discord, 2024-11-16). Some pedagogical and digital materials use ㇽ (small ru) invariably regardless of preceding vowel ‹contested›. The morphophonological behavior of coda /r/ — its assimilation triggers and the echo-vowel phenomenon — is described in Chapter 28 (Assimilation and Cluster Simplification: Coda /r/, Nasals, and Clusters); the echo vowel mechanism associated with the affiliative suffix is treated in Chapter 45 (Morphophonology of the Affiliative Suffix and Echo/Copy Vowels).

12.3.3 Coda /n/ and the small-ン question

Coda /n/ is written with full-size ン (U+30F3), the same character as the Japanese moraic nasal. The borrowing is typographically convenient but phonologically imperfect: Japanese ン is an independent mora that contributes its own weight to the prosodic system, whereas Ainu coda /n/ is a non-moraic appendix in the same CVC syllable as its preceding vowel Nakagawa (2024); Shiraishi (2022). To mark this difference, Unicode 12.0 (2019) added a small ン, 𛅧 (U+1B167, KATAKANA LETTER SMALL N), in the Small Kana Extensions block (U+1B130–U+1B16F); it has also been advocated in community discussions as the typographically appropriate form for the Ainu coda nasal aynu-corpora Discord (2023–2026) (toracatman, aynu-corpora Discord, 2023-12-08). Font support for U+1B167 remains sparse as of the mid-2020s and the character is not in common use; full-size ン remains the practical standard ‹corpus-suggested›.

12.3.4 Coda /h/ in katakana

The five small h-row kana (ㇵ ㇶ ㇷ ㇸ ㇹ; U+31F5–U+31F9; small ha, hi, hu, he, ho) are encoded for coda /h/ notation and follow the same vowel-agreement principle as the /r/ series: ㇵ after /a/, ㇶ after /i/, ㇷ after /u/, ㇸ after /e/, ㇹ after /o/. In practice, coda /h/ is the most orthographically variable member of the Hokkaido system: the phonemic status of word-final /h/ remains disputed Nakagawa (2024), and many katakana-based texts either omit final /h/ entirely or mark it only where affixation causes its resurfacing as a full onset consonant. The phonemic-status argument is treated in Chapter 20 (The Laryngeals: Glottal Stop, /h/, and the Final-h Question).

12.3.5 Glide codas /w/ and /y/

The glides /w/ and /y/ in coda position are written with the corresponding vowel kana — ウ and イ respectively — at full size and without a small-kana reduction. These sequences arise in the surface forms that descriptive grammars analyze as coda-glide syllables (ay, uy, oy, aw, iw, ew and their cognates), rather than as true diphthongs with a phonemically distinct glide nucleus Nakagawa (2024). Whether the underlying segment is a vowel, a phonemic glide, or an epenthetic transition is a phonological question treated in Chapter 22 (Glides, Vowel Hiatus, and the Diphthong Question). No dedicated small-kana form exists for either position.

12.4 Historical development

The earliest sustained katakana transcriptions of Ainu, from the Meiji period onward, used standard Japanese kana with no provision for coda consonants. Kindaichi Kyōsuke, whose grammars and yukar editions from the 1930s established the classical transcriptional tradition in Ainu linguistics Kindaichi & Chiri (1936), rendered Ainu in full-size kana and introduced yōon contractions (チャ, ショ, etc.) for palatalized syllables. Coda consonants in this framework were typically indicated by writing the following syllable's onset kana in the same position or were left unmarked, producing representations that obscured syllable structure Kirikae (1997).

The development of a systematic small-kana coda notation emerged from the work of Ainu speakers and linguists in the postwar decades. Kirikae documents writing practices that predate the Unicode standardization, including variant assignments for individual coda positions Kirikae (1997). Nakagawa traces the parallel development of Ainu-authored phonemic transcription traditions, situating the kana system in the broader history of Ainu-language literacy Nakagawa (2006). The 1994 Hokkaido Utari Kyōkai textbook consolidated the current standard assignments and provided their widest early distribution Hokkaidō Utari Kyōkai (1994); Ōno (2022). Unicode 3.2 (2002) then gave the system stable, interoperable code points across platforms and operating systems.

Font coverage followed at varying rates. As of the mid-2020s the Katakana Phonetic Extensions block is supported by BIZ UDゴシック/UDMinchō, MS Gothic/Minchō, Meiryo, and Yu Gothic/Minchō; some educational fonts such as UDデジタル教科書体 do not cover it aynu-corpora Discord (2023–2026) (toracatman, aynu-corpora Discord, 2024-03-28). The toc-chapter for multi-script rendering and font interoperability is Chapter 13 (Cyrillic and Multi-Script Rendering).

12.5 Latin–katakana correspondence and round-tripping

The two scripts represent the same phonological system and conversion between them is in principle mechanical, but several coda positions admit variant katakana realizations that require a convention choice. Coda /s/ may appear as ㇱ or ㇲ; coda /r/ may use any of the five members of the ra-series or the full-size ル; coda /h/ may be absent or written with any of the five h-series members; and the punctuation and spacing conventions differ between the 「,.」 style of (Hokkaidō Utari Kyōkai 1994) and the 「、。」 style more common in Nakagawa's materials Nakagawa (2024). These ambiguities mean that Latin-to-katakana conversion requires explicit convention declarations, while katakana-to-Latin conversion is more consistently deterministic. The word-spacing and boundary conventions that further complicate round-tripping are addressed in Chapter 15 (Orthographic Standardization and the Word-Division Problem).

From a typological perspective, the katakana system for Ainu sits in an unusual structural position: the combination of a full CV baseline kana plus a subordinate coda kana within a single syllable unit resembles an abugida — a semi-syllabic system in which a base character carries an inherent vowel and a secondary mark modifies the coda — more than a pure syllabary or an alphabet. This characterization has been proposed in community discussion aynu-corpora Discord (2023–2026) (nukopoli, aynu-corpora Discord, 2024-03-30) ‹contested›; no formal typological classification has been assigned to Ainu katakana in the published linguistic literature.

References cited in this chapter

aynu-corpora Discord (2023–2026) ·Hokkaidō Utari Kyōkai (1994) ·Kindaichi & Chiri (1936) ·Kirikae (1997) ·Nakagawa (2006) ·Nakagawa (2024) ·Ōno (2022) ·Shiraishi (2022)