agree with #553. Also, Guangyun gives fanqie 防錽 for 笵, which appears to be more immediately equivalent to the character, ⿱竹汜, for which SAT-09407 is the in the fanqie given in the evidence.
Should probably not be unified with 𥐦 (U+25426) because the phonetic components, 己 (kỷ) and 已 (dĩ) are very different. UTC-03297 would be non-cognate with U+25426 in Vietnamese.
As you can see both in the GĐNHV example above and in this image from KCHN, Vietnam uses both forms. Not sure it's a good idea to unify, even if there is semantic overlap
The shapes are too different to recognize as identical. If we are going to arbitrarily equate simplified components based on interchangeability, then we should apply that across the board, including 馬/马, 金/钅, etc.
Oppose Unification
The point is not that we can derive correspondences, we can similarly derive correspondences from 馬 to 马, 鳥 to 鸟, etc. But, if we are going to say that because we can derive correspondences between glyphs that on the surface look quite different, then we should start using stronger unification that includes traditional and simplified. I don't think people want that, so the same treatment should be applied to simplified forms in languages other than Chinese used in the PRC.
The original source reference for V2-7A3B is Vũ Văn Kính, "Tự Điển Chữ Nôm", p. 272, shown in the image below. As you can see, the phonetic is 迭 (điệt). So, the current shape of U+28540 is incorrect. Unification will be acceptable if we change the shape of U+28540 to VN-F2173.
Unihan data, the ORT Attributes predictor, and most other candidates in WS2024 give 8. It would be better to be consistent.
Total Stroke Count
Given the variations across geographies and font designs, and the fact that unification precludes most shape-based determination of attributes, CJKJRG / IRG originally chose to use the Kangxi values, the most common denominator in dictionaries used by the CJKV countries. This avoided a lot of fruitless debate. Kangxi is 9 strokes, but as you point out, that later changed. I'm fine with either 8 or 9, but we should be consistent moving forward and change the ORT tools to support our decision. Otherwise, maybe we should just stop using TS.
秩 is the phonetic and 刀 the semantic. I don't see how radical 93 is appropriate here. If anything the secondary radical, taken from 秩, should be 115 (禾)
The attributes predictor tool gives 8 for the stroke count. https://hc.jsecs.org/irg/ws2021/app/attributes-predictor.php?ids=%E2%BF%B1亡目务&radical=109.0
The Kangxi dictionary cites a similar passage from "博雅" using what appears to be a variant of SAT-10655 : 鎢錥謂之銼𨰠(U+28C20)。U+28C20 uses 羸, with 羊 in place of 虫. Is there a clearer image?
The evidence shown for the Thổ / Tày language of Lạng Sơn and Cao Bằng is actually a simplified form, ⿰身⿱𫩠彐. Is the implication that this is unifiable with the full form ⿰身當?
The analysis in the dictionary says that this character is composed of "tiêu" (髟) and "đề" (提). The evidence looks like 提 on the bottom. So we conclude that 提 is correct.
The correct form is ⿰昏及 VN-F221C. The character is used in compounds such as "mập mờ" = loose, vague, dim, ambiguous, etc. So, the semantic is 昏, not 昬. We can change the glyph or replace with VN-F221C.
Takeuchi, p. 314 also has this character, with 昏, noting that it is possibly an error for 𥄫. "昏と及の組合せ。又は、𥄫の誤字?"
An alternative solution would be to change the glyph to 展. This is the only instance we have with this design. All other examples of this character used as a phonetic in the "Giúp Đọc" have 展.
Glyph design
All the other glyphs in Nom Na Tong use 展, which is clearly the phonetic, so we will modify the font. Note, the current stroke count is already that for 展, so no need to change.
There are 21 Vietnamese characters with 叕 as an immediate constituent. The distribution of the stroke shape in question is about half and half. We will investigate the issues with normalization.
Nom Na Tong and other Nôm fonts, such as Han-Nom Minh and Han-Nom Kai use 𥝢 for most of the characters shown above and some others:
Chars with 𥝢 in Nom Na Tong: 棃犂黎瓈𥗍𨛫㰀嚟𠠍𤂱𤑬
The only exceptions we can find are U+853E and VN-F03D1.
䄪 is not necessarily standard. Other dictionaries show 𥝢 for U+68C3 and U+853E, as in the entries below from Taberd.
Changing the glyphs of U+853E and VN-F03D1 will be the least disruptive and conform to Vietnamese usage. We can do the horizontal extension after we change U+853E.
We need to discuss attributes for the abbreviated component 𫇦 U+2B1E6. For most characters that use this, the radical is 140 with TC = 6 strokes. It would be good to follow that convention, with rad. 151 as secondary. Following that scheme, even with RS=151 as primary, SC=6, TC = 13, and FS = 2.
IRG Working Set 2024v1.0
Source: Lee COLLINS
Date: Generated on 2026-07-27
Unification
Showing 19 comments.
The evidence above suggest this is a unifiable variant of 䏻 (U+43FB), which is in turn a variant of 能
麚 (U+9E9A)
Semantic identity implied by parallel definition in other texts of Shuowen. Add new UCV for equivalence of ⿰𡰥⿱コ又 and ⿰𡰥㕛 as components.
Semantic and shape similarity suggest unification with𧵍 (U+27D4D)
鼽 (U+9F3D)
Agree with unification 𮔔 (U+2E514) . We will withdraw this.
The original source reference for V2-7A3B is Vũ Văn Kính, "Tự Điển Chữ Nôm", p. 272, shown in the image below. As you can see, the phonetic is 迭 (điệt). So, the current shape of U+28540 is incorrect. Unification will be acceptable if we change the shape of U+28540 to VN-F2173.
Attributes
Showing 133 comments.
Warning: Undefined array key 0 in /home/jsecs/www/hc/irg/ws2024/app/DBCharacters.php on line 328
163′.9.2
Evidence
Showing 32 comments.
Takeuchi, p. 314 also has this character, with 昏, noting that it is possibly an error for 𥄫. "昏と及の組合せ。又は、𥄫の誤字?"
Glyph Design & Normalization
Showing 17 comments.
Nom Na Tong and other Nôm fonts, such as Han-Nom Minh and Han-Nom Kai use 𥝢 for most of the characters shown above and some others:
Chars with 𥝢 in Nom Na Tong: 棃犂黎瓈𥗍𨛫㰀嚟𠠍𤂱𤑬
The only exceptions we can find are U+853E and VN-F03D1.
䄪 is not necessarily standard. Other dictionaries show 𥝢 for U+68C3 and U+853E, as in the entries below from Taberd.
Changing the glyphs of U+853E and VN-F03D1 will be the least disruptive and conform to Vietnamese usage. We can do the horizontal extension after we change U+853E.
We will add 𥝢 as a normalization rule for 䄪.
Entries from Taberd showing the use of 𥝢
P. 699
P. 260
Other
Showing 8 comments.
Data for Unihan
Showing 17 comments.
This is based on the passage from《古今圖書集成》shown in the evidence for WS2021:03752. Kangxi has 觜觿,大龜也.
Submitter Request
Showing 1 comments.