Please wait while loading

IRG Working Set 2024v2.0

Source: Lee COLLINS
Date: Generated on 2026-07-27

Show Deleted | Show comments from version: 1.0 2.0 3.0 4.0 5.0 | Show comments with status: Show All New Only Unresolved Only
The Image/Source column is displayed as it was in WS2024 v2.0. The character may have a different status in the latest working set.

Unification

Showing 20 comments.

SnImage/SourceComment TypeDescription
02382
02382
犬 94.9.5
SAT-09479
TS 12 · IDS
Oppose Unification
There is no evidence of any semantic relation between this character and 𬌼 (U+2C33C). Also, the shapes are not identical. No unification without additional evidence establishing a relationship.
03054
03054
网 122.3.2
SAT-09536
TS 9 · IDS
Oppose Unification
Though there is a structural relationship, the overall result of unifying with an enclosing element is completely counter intuitive.
00112
00112
人 9.4.3
SAT-10237
TS 6 · IDS 𠆢
Oppose Unification
Agree with #8122. It's always useful to have unusual components encoded for purposes of illustration.
02861
02861
竹 118.5.1
UTC-03402
TS 11 · IDS
Unification
02858
竹 118.5.1
GCW-00187
TS 11 · IDS
Need to resolve unification to WS2024-02858 GCW-00187
03680
03680
足 157.14.1
UTC-03479
TS 21 · IDS 𧾷𠪨
Unification
[ Unresolved from v1.0 ]
Agree with #83
00398
00398
口 30.5.5
VN-F007E
TS 8 · IDS
Oppose Unification
[ Unresolved from v1.0 ]
The Vietnamese is clearly not cognate with 𠲙 (U+20C99) and the shapes are not identical. Better to keep separate.
02090
02090
水 85.8.2
VN-F02DB
TS 11 · IDS
Oppose Unification
[ Unresolved from v1.0 ]
As you can see both in the GĐNHV example above and in this image from KCHN, Vietnam uses both forms. Not sure it's a good idea to unify, even if there is semantic overlap

02144
02144
水 85.11.4
VN-F02E4
TS 14 · IDS
Unification
U+30736
This should have been unified with 𰜶 (U+30736) and withdrawn
02640
02640
目 109.16.5
VN-F03B9
TS 21 · IDS
Unification
Agree with unification
02669
02669
石 112.6.4
VN-F03C2
TS 11 · IDS
Unification
Agree with unification
02868
02868
竹 118.7.4
VN-F040A
TS 13 · IDS 𣲭
Unification
U+6ED7
The meaning is "drain dry", which is similar to 滗 (U+6ED7), “xế” is a native word, so this a case where a variant of 滗 was borrowed for its meaning. Unification should be appropriate.
03243
03243
艸 140.3.1
VN-F04D4
TS 7 · IDS
Unification
Still pending unification.
03329
03329
艸 140.13.1
VN-F04F7
TS 17 · IDS
Unification
U+8548
蕈 (U+8548). Same semantic, very similar shape
03637
03637
足 157.5.2
VN-F0573
TS 12 · IDS 𧾷
Unification
Agree with unification based on shape.
03878
03878
金 167.7.3
VN-F05D5
TS 15 · IDS
Unification
U+91C6
U+91C7
Agree with unification. 釆 (biện) is often used for the phonetic 'thai', 'hai', more properly written 采 (thái)
01489
01489
心 61.10.3
VN-F1FB8
TS 14 · IDS
Oppose Unification
[ Unresolved from v1.0 ]
The shapes are too different to recognize as identical. If we are going to arbitrarily equate simplified components based on interchangeability, then we should apply that across the board, including 馬/马, 金/钅, etc.
Oppose Unification
[ Unresolved from v1.0 ]
The point is not that we can derive correspondences, we can similarly derive correspondences from 馬 to 马, 鳥 to 鸟, etc. But, if we are going to say that because we can derive correspondences between glyphs that on the surface look quite different, then we should start using stronger unification that includes traditional and simplified. I don't think people want that, so the same treatment should be applied to simplified forms in languages other than Chinese used in the PRC.
01730
01730
日 72.4.1
VN-F1FF7
TS 8 · IDS
Oppose Unification
[ Unresolved from v1.0 ]
The shapes are too different to recognize as identical.
01437
01437
心 61.4.1
VN-F1FFC
TS 7 · IDS
Oppose Unification
[ Unresolved from v1.0 ]
The shapes are too different to recognize as identical.
02054
02054
气 84.9.4
VN-F2173
TS 13 · IDS
Unification
[ Unresolved from v1.0 ]
U+8FED
U+28540
The original source reference for V2-7A3B is Vũ Văn Kính, "Tự Điển Chữ Nôm", p. 272, shown in the image below. As you can see, the phonetic is 迭 (điệt). So, the current shape of U+28540 is incorrect. Unification will be acceptable if we change the shape of U+28540 to VN-F2173.

Attributes

Showing 164 comments.

SnImage/SourceComment TypeDescription
03601
03601
贝 154′.2.2
GCA-J0023
TS 6 · IDS
FS
[ Unresolved from v1.0 ]
FS=3
04602
04602
鸟 196′.10.1
GCA-J0159
TS 15 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
11
Total Stroke Count
[ Unresolved from v1.0 ]
16
04621
04621
卤 197′.8.2
GCA-T0002
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
3
03267
03267
艸 140.7.4
GCW-00202
TS 11 · IDS
IDS
Agree with logic for IDS changed noted above.
03822
03822
辵 162.12.1
GCW-00228
TS 15 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
16
03844
03844
阜 170.9.2
GCW-00229
TS 11 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
12
Radical
[ Unresolved from v1.0 ]
170
03841
03841
阜 170.15.2
GCW-00230
TS 17 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
18
Radical
[ Unresolved from v1.0 ]
170
03992
03992
邑 163.6.1
GCW-00241
TS 8 · IDS
Radical
[ Unresolved from v1.0 ]
163
Total Stroke Count
[ Unresolved from v1.0 ]
9
04396
04396
魚 195.3.5
GCW-00247
TS 14 · IDS
FS
[ Unresolved from v1.0 ]
2
04448
04448
鱼 195′.3.5
GCW-00248
TS 11 · IDS
FS
[ Unresolved from v1.0 ]
2
04609
04609
鸟 196′.11.5
GCW-00261
TS 16 · IDS
FS
[ Unresolved from v1.0 ]
2
04017
04017
邑 163.16.2
GCW-00278
TS 18 · IDS
Radical
[ Unresolved from v1.0 ]
163
Total Stroke Count
[ Unresolved from v1.0 ]
19
04271
04271
马 187′.13.1
GDM-00489
TS 16 · IDS
FS
[ Unresolved from v1.0 ]
2
04485
04485
鱼 195′.7.5
GPGLG-4004
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
2
03993
03993
邑 163.6.1
GXM-00476
TS 8 · IDS
Radical
[ Unresolved from v1.0 ]
163
Total Stroke Count
[ Unresolved from v1.0 ]
9
03999
03999
邑 163.8.3
GXM-00477
TS 10 · IDS
Radical
[ Unresolved from v1.0 ]
163
03609
03609
贝 154′.10.3
GZ-0082105
TS 14 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=11, TC = 15. The ORT Attributes predictor gives 7 strokes for 成. The unihan data is inconsistent. Personally I think 6 is more intuitive.
04170
04170
風 182.2.3
GZ-0341301
TS 11 · IDS
FS
[ Unresolved from v1.0 ]
2
03606
03606
贝 154′.6.4
GZ-1111302
TS 10 · IDS
FS
[ Unresolved from v1.0 ]
FS=2
04144
04144
革 177.8.3
GZ-2672204
TS 17 · IDS
FS
[ Unresolved from v1.0 ]
4
03702
03702
車 159.4.1
GZ-2912503
TS 11 · IDS
FS
[ Unresolved from v1.0 ]
FS=4
00744
00744
口 30.17.4
GZ-2972401
TS 20 · IDS
Radical
radical is 31 (囗 U+56D7)
04569
04569
鳥 196.15.3
GZ-3842501
TS 26 · IDS
IDS
[ Unresolved from v1.0 ]
⿰外𬷨
04148
04148
韋 178.9.4
GZ-4901401
TS 19 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
10
Total Stroke Count
[ Unresolved from v1.0 ]
19
04168
04168
頁 181.15.3
GZHSJ-0005
TS 24 · IDS
IDS
[ Unresolved from v1.0 ]
⿳⿴𦥑同冖頁, 𦥑 produces the expected number of strokes.
04649
04649
黽 205.13.2
GZHSJ-0021
TS 26 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
14
Total Stroke Count
[ Unresolved from v1.0 ]
27
FS
[ Unresolved from v1.0 ]
1
03838
03838
邑 163.8.5
GZHSJ-0038
TS 10 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
11
04281
04281
骨 188.9.4
GZHSJ-0053
TS 18 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
19
03816
03816
辵 162.11.2
GZHSJ-0070
TS 14 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
15
03836
03836
阜 170.7.1
GZHSJ-0076
TS 9 · IDS 𦔮
Total Stroke Count
[ Unresolved from v1.0 ]
10
Radical
[ Unresolved from v1.0 ]
170
03815
03815
辵 162.11.1
GZHSJ-0125
TS 14 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
15
04640
04640
黃 201.6.2
GZHSJ-0145
TS 17 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
18, based on IRGN2221 which suggests 12 strokes for 黃/黄
03832
03832
邑 163.4.3
GZHSJ-0162
TS 6 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
7
02769
02769
示 113.14.2
KC-10153
TS 19 · IDS
Residual Stroke Count
13
Total Stroke Count
18
04236
04236
香 186.18.1
KC-10181
TS 27 · IDS
FS
[ Unresolved from v1.0 ]
2 (based on IRGN 954AR)
03746
03746
車 159.10.1
KC-10197
TS 17 · IDS
FS
[ Unresolved from v1.0 ]
FS=2
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=11, TS=18
03879
03879
金 167.7.4
KC-10260
TS 15 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
8
Total Stroke Count
[ Unresolved from v1.0 ]
16
03545
03545
言 149.25.2
SAT-09174
TS 32 · IDS
FS
[ Unresolved from v1.0 ]
FS=5
03757
03757
車 159.11.5
SAT-09336
TS 18 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=10, TS=17
03903
03903
金 167.11.3
SAT-09597
TS 19 · IDS 𣐂
FS
[ Unresolved from v1.0 ]
2
04645
04645
黑 203.10.2
SAT-09617
TS 21 · IDS
FS
[ Unresolved from v1.0 ]
5
Total Stroke Count
[ Unresolved from v1.0 ]
22
04632
04632
鹿 198.15.4
SAT-09630
TS 26 · IDS 鹿
FS
[ Unresolved from v1.0 ]
2
04215
04215
食 184.14.1
SAT-09712
TS 23 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
22
03527
03527
言 149.8.4
SAT-09780
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
FS = 3
03859
03859
里 166.7.2
SAT-09832
TS 11 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
4
FS
[ Unresolved from v1.0 ]
1
03636
03636
足 157.5.2
SAT-10051
TS 12 · IDS
FS
[ Unresolved from v1.0 ]
FS=1
03843
03843
邑 163.17.1
SAT-10079
TS 20 · IDS 𦾔
FS
[ Unresolved from v1.0 ]
2
04283
04283
骨 188.14.5
SAT-10195
TS 24 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
15
Total Stroke Count
[ Unresolved from v1.0 ]
25
04063
04063
雨 173.10.2
T13-3F42
TS 18 · IDS
FS
[ Unresolved from v1.0 ]
1
04083
04083
雨 173.13.3
T13-3F4F
TS 21 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
14
Total Stroke Count
[ Unresolved from v1.0 ]
22
04084
04084
雨 173.13.3
T13-3F51
TS 21 · IDS
FS
[ Unresolved from v1.0 ]
4
02096
02096
水 85.8.4
TB-7C28
TS 11 · IDS
IDS
[ Unresolved from v1.0 ]
More compact IDS: ⿰氵𱣅
02888
02888
竹 118.10.3
TB-7D2D
TS 16 · IDS
FS
4
03853
03853
酉 164.12.4
TB-7E3E
TS 19 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
12
Total Stroke Count
[ Unresolved from v1.0 ]
19
04059
04059
雨 173.10.1
TB-7E5E
TS 18 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
11
Total Stroke Count
[ Unresolved from v1.0 ]
19
04570
04570
鳥 196.17.2
TB-7E79
TS 28 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
15
Total Stroke Count
[ Unresolved from v1.0 ]
26
04633
04633
鹿 198.16.1
TB-7E7B
TS 27 · IDS 鹿
Residual Stroke Count
[ Unresolved from v1.0 ]
15
Total Stroke Count
[ Unresolved from v1.0 ]
26
04248
04248
馬 187.6.4
TC-6576
TS 16 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
8
Total Stroke Count
[ Unresolved from v1.0 ]
18
03964
03964
門 169.9.5
TC-6625
TS 17 · IDS 𤕰
Residual Stroke Count
[ Unresolved from v1.0 ]
8
Total Stroke Count
[ Unresolved from v1.0 ]
16
04548
04548
鳥 196.8.4
TC-6836
TS 19 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
7
Total Stroke Count
[ Unresolved from v1.0 ]
18
04000
04000
阜 170.8.3
土 32.8.5
TC-743B
TS 11 · IDS 𨸼
FS
[ Unresolved from v1.0 ]
4
04195
04195
食 184.5.3
TD-3463
TS 13 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
4
Total Stroke Count
[ Unresolved from v1.0 ]
12
04395
04395
魚 195.3.3
TD-584A
TS 14 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
4
Total Stroke Count
[ Unresolved from v1.0 ]
15
03630
03630
走 156.8.1
TD-647C
TS 15 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=9, TC=16 (IRGN2221)
02756
02756
示 113.10.3
TD-7B2B
TS 15 · IDS
FS
4
04152
04152
音 180.9.1
TE-297C
TS 18 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
8
Total Stroke Count
[ Unresolved from v1.0 ]
17
03934
03934
金 167.20.3
TE-7D7B
TS 28 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
21
Total Stroke Count
[ Unresolved from v1.0 ]
29
03632
03632
走 156.8.5
UK-30090
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
FS=2
03824
03824
辵 162.12.3
UK-30195
TS 16 · IDS
FS
[ Unresolved from v1.0 ]
4
04445
04445
鱼 195′.2.5
UK-30227
TS 10 · IDS
FS
[ Unresolved from v1.0 ]
3
04201
04201
食 184.7.1
UK-30268
TS 16 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
15
FS
[ Unresolved from v1.0 ]
Unihan data, the ORT Attributes predictor, and most other candidates in WS2024 give 8. It would be better to be consistent.
03881
03881
金 167.7.5
UK-30292
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
1
03755
03755
車 159.11.2
UK-30348
T13-3E7A
TS 18 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=12, TS=19
04196
04196
食 184.6.1
UK-30458
TS 15 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
14
Total Stroke Count
[ Unresolved from v1.0 ]
Unihan data, the ORT Attributes predictor, and most other candidates in WS2024 give 8. It would be better to be consistent.
Total Stroke Count
[ Unresolved from v1.0 ]
Given the variations across geographies and font designs, and the fact that unification precludes most shape-based determination of attributes, CJKJRG / IRG originally chose to use the Kangxi values, the most common denominator in dictionaries used by the CJKV countries. This avoided a lot of fruitless debate. Kangxi is 9 strokes, but as you point out, that later changed. I'm fine with either 8 or 9, but we should be consistent moving forward and change the ORT tools to support our decision. Otherwise, maybe we should just stop using TS.
03561
03561
豆 151.4.5
UK-30559
TS 11 · IDS
FS
[ Unresolved from v1.0 ]
FS=3
04216
04216
食 184.14.1
UK-30582
TS 23 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
22
03615
03615
贝 154′.15.3
UK-30630
TS 19 · IDS
FS
[ Unresolved from v1.0 ]
FS=2
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=14, TC=20, based on Attributes Predictor, which counts the extra dot in 署
03611
03611
贝 154′.12.3
UK-30644
TS 16 · IDS
FS
[ Unresolved from v1.0 ]
FS=4
04009
04009
阜 170.11.3
UK-30801
TS 14 · IDS
FS
[ Unresolved from v1.0 ]
4
03866
03866
金 167.6.1
UTC-03274
TS 14 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
8
FS
[ Unresolved from v1.0 ]
2
Total Stroke Count
[ Unresolved from v1.0 ]
16
03884
03884
金 167.8.1
UTC-03282
TS 16 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
9
Total Stroke Count
[ Unresolved from v1.0 ]
17
03907
03907
金 167.11.4
UTC-03285
TS 19 · IDS
FS
[ Unresolved from v1.0 ]
3
02762
02762
示 113.11.4
UTC-03292
TS 15 · IDS
Total Stroke Count
16
02758
02758
示 113.11.1
UTC-03293
TS 15 · IDS
Total Stroke Count
16
02742
02742
示 113.8.1
UTC-03294
TS 12 · IDS
Total Stroke Count
13
04508
04508
鱼 195′.10.1
UTC-03463
TS 13 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
18
04605
04605
鸟 196′.10.1
UTC-03470
TS 15 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
9
Total Stroke Count
[ Unresolved from v1.0 ]
14
04618
04618
鸟 196′.16.1
UTC-03476
TS 21 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
17
Total Stroke Count
[ Unresolved from v1.0 ]
22
04262
04262
馬 187.19.2
UTC-03481
TS 28 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
29
02346
02346
宀 40.9.3
牛 93.8.4
VN-F0069
TS 12 · IDS
Radical
[ Unresolved from v1.0 ]
Agree with swap
01319
01319
工 48.15.5
VN-F0170
TS 18 · IDS
Total Stroke Count
The IRG Attributes Predictor counts 巨 as 5 strokes, Unihan has 4. We should discuss and document the stroke count we are going to use and fix the ORT if we decide it's 4. Otherwise keep TC=18.
02253
02253
火 86.10.3
戈 62.10.3
VN-F01DF
TS 14 · IDS
Radical
Swapping is fine. But maybe we should consider 196 鳥, since it means 'crow'. We chose 86 based on Kangxi and Unihan, but if there is no need to follow those, 'bird' is the best semantic.
01146
01146
尸 44.4.3
戶 63.4.3
VN-F01E2
TS 8 · IDS
Radical
[ Unresolved from v1.0 ]
Agree with #6756
01736
01736
日 72.4.4
VN-F025A
TS 8 · IDS
IDS
The IDS proposed above seems confusing. 㓁 is a variant of rad. 122 and always appears above. If we merely want to reduce the # of strokes, U+5197 would be better since it can have the shape ⿱冖儿.
01802
01802
日 72.11.5
VN-F0266
TS 15 · IDS
Residual Stroke Count
The current values come from the ORT Attributes Predictor. Unihan gives contradictory values for U+52D9: TC= 10, but RSUnicode = 19.9 (total 11)
Residual Stroke Count
IRGN 954AR counts 攵 and variants (including 夂?) as 4, so we need to agree on the composition of 务 (U+52A1) and 務.
Residual Stroke Count
IRGN2171 suggests 9 for U+6544 敄.
01813
01813
日 72.13.5
VN-F0269
TS 17 · IDS
Residual Stroke Count
The ORT Attributes Predictor gives 13 strokes for U+9115 and U+9109. So keep the current RS values until we publish an agreed upon value for IRG work.
01832
01832
曰 73.10.2
VN-F0280
TS 14 · IDS
Radical
Keep 73 as secondary, since it is the semantic.
02144
02144
水 85.11.4
VN-F02E4
TS 14 · IDS
FS
IRGN 954AR says FS=4
02519
02519
疒 104.6.4
VN-F037D
TS 10 · IDS
Total Stroke Count
11
00281
00281
刀 18.10.3
禾 115.7.3
VN-F03EC
TS 12 · IDS
Radical
[ Unresolved from v1.0 ]
秩 is the phonetic and 刀 the semantic. I don't see how radical 93 is appropriate here. If anything the secondary radical, taken from 秩, should be 115 (禾)
02843
02843
穴 116.14.5
VN-F03FD
TS 19 · IDS
FS
1
02986
02986
糸 120.9.4
VN-F045E
TS 15 · IDS
FS
[ Unresolved from v1.0 ]
5
03452
03452
虫 142.18.4
VN-F0525
TS 24 · IDS
IDS
Agree with #8307. Also true for WS2004-01856, WS2024-01989, and WS2024-02195.
03619
03619
赤 155.12.1
VN-F056C
TS 17 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
SC=12, TC=19
03620
03620
赤 155.12.1
VN-F056D
TS 17 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
TS=19
03682
03682
足 157.15.2
VN-F0597
TS 22 · IDS 𧾷
FS
IRGN954AR shows 乙 (5) , but the ORT Attributes predictor gives 2. 2 seems better since it is the first stroke in
FS
Above comment should end with 賞.
04166
04166
頁 181.12.5
阜 170.18.2
VN-F060F
TS 20 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
21
04416
04416
魚 195.10.3
VN-F0677
TS 21 · IDS
FS
[ Unresolved from v1.0 ]
ORT Attributes predictor gives FS=3, which is correct?

04240
04240
馬 187.4.5
VN-F07F9
TS 14 · IDS
IDS
[ Unresolved from v1.0 ]
⿱⿻又丷馬
02772
02772
禸 114.8.4
八 12.10.4
VN-F085E
TS 12 · IDS
Total Stroke Count
Agree with #8371
03805
03805
辵 162.9.2
VN-F109A
TS 13 · IDS 𠀐
Residual Stroke Count
[ Unresolved from v1.0 ]
8
Total Stroke Count
[ Unresolved from v1.0 ]
12
04191
04191
飞 183′.12.3
VN-F1F88
TS 15 · IDS
Radical
183.1 is the correct radical, but the ORT is showing the full variant
01833
01833
曰 73.12.2
VN-F1FAC
TS 16 · IDS
Radical
Keep 73 as secondary.
03203
03203
自 132.16.1
VN-F1FBD
TS 26 · IDS
Residual Stroke Count
Needs update to 22
02989
02989
糸 120.10.2
VN-F1FBF
TS 16 · IDS
Radical
add 61 as 2nd radical, SC=12, FS=2
01158
01158
尸 44.13.4
戶 63.13.4
VN-F1FDE
TS 17 · IDS
Radical
[ Unresolved from v1.0 ]
Agree with #6757
01379
01379
廾 55.18.1
VN-F1FED
TS 21 · IDS
Radical
Radical for 弄 (U+5F04) is 55. Should be kept as secondary.
04153
04153
音 180.11.3
VN-F1FFF
TS 20 · IDS
Radical
The IRG needs to have an intelligble policy on assignment of radicals. We originally based it on semantic, then the policy seems to have switched to "most intuitive". 180 is "intuitive"; 108 is semantic. Which is it to be?
00959
00959
土 32.15.4
宀 40.15.1
VN-F200A
TS 18 · IDS
Radical
Note that 立 is the phonetic, the semantic is 塞. 立 is fine as primary radical if we are no longer basing it on the semantic.
02611
02611
目 109.8.4
VN-F200F
TS 13 · IDS
Residual Stroke Count
[ Unresolved from v1.0 ]
The attributes predictor tool gives 8 for the stroke count. https://hc.jsecs.org/irg/ws2021/app/attributes-predictor.php?ids=%E2%BF%B1亡目务&radical=109.0
03008
03008
糸 120.13.3
VN-F2036
TS 19 · IDS
Radical
Agree with 2nd radical
04549
04549
鳥 196.8.4
VN-F20BC
TS 16 · IDS 𫠓
Radical
Agree with comment #10413
03083
03083
羊 123.12.5
VN-F2112
TS 18 · IDS
Total Stroke Count
[ Unresolved from v1.0 ]
18 is correct. According to the Attributes Predictor and Unihan data, 羊 is 6, not 7. Both give 10 strokes for 羞.

Evidence

Showing 24 comments.

SnImage/SourceComment TypeDescription
03792
03792
辵 162.4.3
GCW-00226
TS 8 · IDS
Evidence
[ Unresolved from v1.0 ]
Evidence 3 suggests that this is interchangeable with 赼 (U+8D7C)
02990
02990
糸 120.10.4
KC-10159
TS 16 · IDS
New evidence
Also found in Vietnamese, VN-F0466, with reading "súc"

00380
00380
口 30.3.3
UK-30330
TS 6 · IDS
New evidence
[ Unresolved from v1.0 ]
This was standardized / normalized in the "Kho Chữ Nôm Mã Hoà" as V+60888
04085
04085
雨 173.13.3
UTC-03243
TS 21 · IDS 𪫕
New evidence
Evidence from a printed version of the same Manyōshū poem in an edition edited by Tsuru Hisashi and Moriyama Takashi showing emendation of UTC-03243 found in the Nishi Honganji manuscript to 霺 based on Ōya and Kyoto University manuscripts.

03009
03009
糸 120.13.5
UTC-03246
TS 19 · IDS 𭆤
New evidence
Printed evidence from Moriyama Takashi and Tsuru Hisashi eds, “Man'yōshū”, ISBN 4-273-00019-9, p 338:

01511
01511
心 61.13.3
V4-4858
TS 17 · IDS
New evidence
Missing evidence # from VVK:

01671
01671
手 64.14.3
V4-496B
TS 17 · IDS
New evidence
The missing evidence:

03451
03451
虫 142.18.2
谷 150.17.2
V4-5377
TS 24 · IDS
New evidence
Correct image for evidence #1

01010
01010
大 37.10.1
VN-F0122
TS 13 · IDS
New evidence
Evidence from Kho Chữ Hán Nôm Mã Hoá p. 584

01488
01488
心 61.10.3
VN-F01C2
TS 13 · IDS
Evidence
"Giúp đọc Nôm và Hán Việt" is currently the only evidence we have. But based on the analysis given in that dictionary, "Hv tâm quải", which means that it's composed of the Hán Việt characters 忄 and 挂, the glyph should be ⿰忄 挂.
02155
02155
水 85.12.4
VN-F02EA
TS 16 · IDS
New evidence
From Vũ Văn Kính, "Đại Tự Điển Chữ Nôm", p. 481, showing composition of VN-F02EA from băng 氷 + giá 這

02881
02881
竹 118.9.3
VN-F040C
TS 15 · IDS
New evidence
Correct evidence for #2 above

03152
03152
肉 130.7.5
VN-F04A6
TS 11 · IDS
New evidence
Correct evidence on p. 1421

03145
03145
肉 130.6.5
VN-F04A8
TS 10 · IDS
New evidence
From "Kho Chữ Hán Nôm Mã Hoá", p. 499

00323
00323
十 24.8.4
八 12.8.1
VN-F07DD
TS 10 · IDS
New evidence
[ Unresolved from v1.0 ]
Kieu1866, p. 23b:

01484
01484
心 61.10.1
VN-F1F94
TS 13 · IDS
Evidence
Agree with #7850. Will submit separate images in future.
00340
00340
厂 27.8.4
VN-F2002
TS 10 · IDS
Evidence
This is currently the only example we have of this character. The component ⿻沈丶 is thought to derive from a simplified form of 㴷 (đắm: shipwrecked, see TĐCNTD p. 341), where 耽 has been reduced to the form V+60779 shown below:



The more common form is V+607C5 in the above, also shown here from the same source as VN-F2002:



An appropriate normalization would be V+607C5.
03597
03597
貝 154.13.2
VN-F2019
TS 20 · IDS
New evidence
ĐTĐCN p. 136

03007
03007
糸 120.13.3
VN-F202F
TS 19 · IDS 𤠰
New evidence
From BTCN

02520
02520
疒 104.6.5
VN-F2069
TS 11 · IDS
Evidence
I agree that the phonetic doesn't make sense. We would expect 卯 (mão), making this equivalent to U+24D60. TĐCNT is currently the only source we have, but will try to find more.
00144
00144
人 9.9.2
VN-F2079
TS 11 · IDS 𱽗
Evidence
We recently found that this character is the orignal source given for V4-407A, on p. 517 of "Góp phần nghiên cữu văn hoá Việt Nam"



V4-407A is currently encoded as 𫢠 U+2B8A0. One solution would be to move V4-407A to WS2024:00144 and change kIRG_VSource for U+2B8A0 to VN-2B8A0.
04577
04577
鸟 196′.2.3
几 16.5.3
VN-F2080
TS 7 · IDS
Evidence
The reading is shown in the transliteration that on the page that follows (in green). The note explains that the original text is in Vietnamese, so no translation is given. "phượng" is the legendary bird typically translated as "phoenix". It is more commonly written: 鳯.

03083
03083
羊 123.12.5
VN-F2112
TS 18 · IDS
New evidence
Correct evidence from Takeuchi:

03020
03020
糸 120.23.5
VN-F231F
TS 28 · IDS
New evidence
From BTCN

Glyph Design & Normalization

Showing 23 comments.

SnImage/SourceComment TypeDescription
03217
03217
舛 136.0.1
一 1.7.2
SAT-09449
TS 8 · IDS
Glyph design
We need to resolve the issue of design
04142
04142
革 177.5.1
SAT-09987
TS 14 · IDS
Glyph design
[ Unresolved from v1.0 ]
Please consider normalizing the form 戹 to be more in keeping with the design found in other Japanese fonts.
04541
04541
鳥 196.6.1
UTC-03116
VN-F0696
TS 17 · IDS
Glyph design
We are considering changing the design so that the left side looks more like the 戎 in UTC-03116
01886
01886
木 75.7.3
VN-F028C
TS 11 · IDS
Glyph design
There are more than 40 glyphs using the same design in the NomNaTong font. It would be a significant effort to change them all. We would need to better understand the rationale for this design before making such a change.
02155
02155
水 85.12.4
VN-F02EA
TS 16 · IDS
Glyph design
The evidence is contradictory. The glyph has 永, but the structural analysis shows "băng giá", which in chữ Hán is 氷這. The character means "frost", so it has been corrected to use "ice" instead of "eternal".
02628
02628
目 109.11.4
VN-F03B4
TS 16 · IDS
Normalization
The combination of simplified 门 with 舀 is anomalous. We can consider 阎 to be a normalized form.
02685
02685
石 112.8.5
VN-F03C5
TS 13 · IDS
Glyph design
[ Unresolved from v1.0 ]
There are 21 Vietnamese characters with 叕 as an immediate constituent. The distribution of the stroke shape in question is about half and half. We will investigate the issues with normalization.
02710
02710
石 112.11.4
VN-F03CD
TS 16 · IDS
Glyph design
Agree, will fix in next font update.
02713
02713
石 112.12.3
VN-F03D1
TS 17 · IDS
Glyph design
[ Unresolved from v1.0 ]
The best solution is 1, modify VN-F03D1 to use 𥝢.

Nom Na Tong and other Nôm fonts, such as Han-Nom Minh and Han-Nom Kai use 𥝢 for most of the characters shown above and some others:

Chars with 𥝢 in Nom Na Tong: 棃犂黎瓈𥗍𨛫㰀嚟𠠍𤂱𤑬



The only exceptions we can find are U+853E and VN-F03D1.

䄪 is not necessarily standard. Other dictionaries show 𥝢 for U+68C3 and U+853E, as in the entries below from Taberd.

Changing the glyphs of U+853E and VN-F03D1 will be the least disruptive and conform to Vietnamese usage. We can do the horizontal extension after we change U+853E.

We will add 𥝢 as a normalization rule for 䄪.

Entries from Taberd showing the use of 𥝢

P. 699


P. 260
Glyph design
We have an update to this glyph for the change proposed above.
03139
03139
肉 130.6.1
VN-F04A1
TS 10 · IDS
Glyph design
Unfortunately a large number of the characters with element 戎 (U+620E, nhung) are designed this was in the reference font. This includes 戎 itself. We'll have to consider how to migrate these.
03145
03145
肉 130.6.5
VN-F04A8
TS 10 · IDS
Glyph design
Agree. The glyph should be composed with 奸, not 㚥. Will change it.
03193
03193
肉 130.15.4
VN-F04BE
TS 19 · IDS
Glyph design
The glyph already has that general shape, can you provide more detail?
03621
03621
赤 155.12.1
VN-F056E
TS 19 · IDS
Normalization
[ Unresolved from v1.0 ]
The glyph has been normalized based on the given analysis for this entry: "xích hùng", i.e. 赤雄.
03990
03990
邑 163.4.3
VN-F060A
TS 7 · IDS
Normalization
[ Unresolved from v1.0 ]
We can consider changing, but the last published standard had already normalized to the current shape:

00216
00216
八 12.6.3
VN-F066F
TS 8 · IDS
Glyph design
[ Unresolved from v1.0 ]
We will address this.
04560
04560
鳥 196.11.2
VN-F06D3
TS 22 · IDS
Glyph design
Re. #8603, please provide more detail concerning the revision.
01659
01659
手 64.12.4
VN-F0B52
TS 15 · IDS
Glyph design
The phonetic, "dan" argues for U+67EC. Here is another analysis (Vũ Văn Kính, "Tự điễn chứ Nôm" p. 225) showing that the traditional and simplified forms both contain U+67EC, read "lan", as phonetic.

01563
01563
手 64.4.4
VN-F0CBC
TS 7 · IDS 𠬠
Glyph design
The element on the right is a simplification of the characters 沒 / 没, read "một", through these steps 没 > 𠬛 > 𠬠 or 𱥺 > 𠬠. There are 2 basic forms, 𠬠 and 𰰝. This is documented in the character definition shown in the image below from TĐCNTD p. 802



Below is an example of VN-F0CBC from "Lục Vân Tiên" showing a form somewhat between 𠬠 and 𰰝



Historically, there are many examples of 𰰝, but the current trend is to standardize on 𠬠, as shown in this the "BẢNG CHỮ HÁN NÔM CHUẨN THƯỜNG DÙNG" http://www.hannom-rcv.org/NS/bchnctd%20300623.pdf
03805
03805
辵 162.9.2
VN-F109A
TS 13 · IDS 𠀐
Glyph design
Agree to design change
01011
01011
大 37.10.1
VN-F1F13
TS 13 · IDS
Glyph design
About the comment #8629. I assume that the suggestion is to change the 亅 in the top 可 to 丨. That's a reasonable suggestion, but there are at least 6 other V-source characters already encoded (𣘁 U+24819, etc.) that use the current design. Since that's a majority, the lesser impact solution would be to normalize the remaining characters 哥 U+54E5, 歌 U+6B4C, U+2BC04, and VN-F176E (in WS2021) to use 亅.
03560
03560
谷 150.7.3
VN-F1FDB
TS 14 · IDS
Glyph design
Both variants are found in Vietnamese, TĐCNDG entry shown below has U+79C3 秃. In the NomNaTong font there are 5 glyphs composed with U+79C3 秃 and 5 composed with U+79BF 禿. Of the characters with V-Source references, if we normalize to 禿, we would also want to change 𥟉 U+257C9 / V3-3531 and 𥟹 U+257F9 / V2-7F31. If we normalize to 秃, we would only change U+22B33 / VN-22B33

03334
03334
艸 140.13.3
VN-F1FFE
TS 17 · IDS
Glyph design
The glyph should look like 𮚃(U+2E683) in the NomNaTong font, with a slash. We will change the glyph and the IDS.

Other

Showing 16 comments.

SnImage/SourceComment TypeDescription
03562
03562
豆 151.5.1
GCW-00224
TS 12 · IDS 𫇦
Other
[ Unresolved from v1.0 ]
We need to discuss attributes for the abbreviated component 𫇦 U+2B1E6. For most characters that use this, the radical is 140 with TC = 6 strokes. It would be good to follow that convention, with rad. 151 as secondary. Following that scheme, even with RS=151 as primary, SC=6, TC = 13, and FS = 2.
02563
02563
白 106.9.1
KC-10128
TS 14 · IDS
Other
FS needs update
04296
04296
鬼 194.1.1
T13-3F6A
TS 8 · IDS
Other
[ Unresolved from v1.0 ]
Should consider a new radical number for this ⿱田儿 variant of 194
02828
02828
穴 116.4.4
TC-7A3E
TS 9 · IDS
Other
Reading and shape suggest variant of 穷 (U+7A77), simplification of 窮 (U+7AAE)
02556
02556
白 106.4.1
TC-7A46
TS 9 · IDS
Other
Reading 'hào' and shape suggest variant of 昊 (U+660A)
02945
02945
米 119.9.4
TD-6227
TS 15 · IDS
Other
No ⿰米婁 has been encoded.
04599
04599
鸟 196′.9.2
UK-30053
TS 14 · IDS
Other
[ Unresolved from v1.0 ]
The wrong glyph is marked in Evidence 1, but the correct one is next to it.
03493
03493
見 147.7.1
UK-30724
TS 14 · IDS
Other
ORT reports 3 for FS, but I believe that wrong since the traditional phonetic for 吞 is 天, the first stroke of which is 1. This does not appear to be structured with the variant 呑 (U+5451), whose first stroke is 3.
03250
03250
艸 140.5.3
UTC-03295
TS 8 · IDS
Other
[ Unresolved from v1.0 ]
What is the justification for labeling this as similar to U+31FC3?
Other
Evidence # 3 for UTC-03292, which has 逃入清化 (he fled into Thanh Hoá), parallels the phrase 奔清⿱花一 above and suggests that this character is a variant of 化 (U+5316, read hoá). Thanh Hoá is more commonly written 清化.
02913
02913
竹 118.15.1
VN-F0423
TS 21 · IDS
Other
Both characters, U+23813 and VN-F0423 mean "a type of bamboo". The major sources, BTCN, ĐTĐCN, GĐNHV, KCHN, and Takeuchi, all show the form VN-F0423, with 竹, appropriately, as the radical. Since VN-F0423 appears to be the correct form, if we were to unify these, Vietnam would request changing the representative glyph for U+23813 to be that of VN-F0423.
03331
03331
艸 140.13.2
VN-F04F6
TS 17 · IDS 𠣜
Other
It's possible that the shape in #8394 was intended, but the only evidence we have shows the current glyph design. Here is KCHN

03544
03544
言 149.23.2
貝 154.23.
VN-F1E12
TS 30 · IDS
Other
[ Unresolved from v1.0 ]
second radical should have FS=4
02803
02803
禾 115.9.4
VN-F1E76
TS 14 · IDS
Other
[ Unresolved from v1.0 ]
Re #1148, Prof. Hồng uses U+3004 to indicate a sense that is different from the primary sense of the character being defined.
02531
02531
疒 104.9.3
VN-F2049
TS 14 · IDS 𢚩
Other
SC and TC need to be updated as noted above.
04025
04025
雨 173.1.3
VN-F208C
TS 9 · IDS 丿
Other
'phiét' means 'lightning' in the Tày language. It's more commonly written 𲉷 (U+32277). 丿 , read 'phiệt' in Vietnamese, is the phonetic element.

Data for Unihan

Showing 22 comments.

SnImage/SourceComment TypeDescription
02981
02981
糸 120.8.4
SAT-09364
TS 14 · IDS
Semantic variant
縮 (U+7E2E)
02400
02400
犬 94.14.3
SAT-09470
TS 17 · IDS 𡼡
Semantic variant
U+736F
02554
02554
白 106.3.3
SAT-09837
TS 8 · IDS
Semantic variant
U+7680
03081
03081
羊 123.10.1
SAT-10168
TS 16 · IDS
Semantic variant
羨 (U+7FA8)
02962
02962
米 119.16.4
SAT-10224
TS 22 · IDS 𢋎
Semantic variant
䊳 (U+42B3)
02969
02969
糸 120.6.3
SAT-10235
TS 12 · IDS 𠆢
Semantic variant
𥿳 (U+25FF3)
02293
02293
火 86.18.2
SAT-90071
TS 22 · IDS
Semantic variant
𤑼 (U+2447C)
02409
02409
玉 96.4.2
TB-7C51
TS 8 · IDS
Semantic variant
U+73C3
02964
02964
糸 120.3.1
TC-7A23
TS 9 · IDS
Semantic variant
Based on evidence in #7176, this is a variant of 紊 (U+7D0A), also read みだる.
03449
03449
虫 142.16.4
UTC-03372
TS 22 · IDS 𠫓
Semantic variant
蠃 (U+8803). Also variant of WS2024:UTC-03376
03444
03444
虫 142.15.4
UTC-03376
TS 21 · IDS
Semantic variant
蠃 (U+8803). The reading "つぶら" suggests a semantic relationship with the variant 螺 (つぶ), both a kind of conch.
03257
03257
艸 140.6.3
UTC-03392
TS 10 · IDS 𦘱
Semantic variant
藤 (U+85E4)
01324
01324
己 49.18.2
VN-F0177
TS 21 · IDS
Simp variant
U+2AA71
02663
02663
石 112.5.4
GZ-0532301
TS 9 · IDS
Trad variant
U+255AD
02867
02867
竹 118.7.4
VN-F0408
TS 13 · IDS 𫔭
Trad variant
U+25CD0
03331
03331
艸 140.13.2
VN-F04F6
TS 17 · IDS 𠣜
Trad variant
繭 (U+7E6D)
03081
03081
羊 123.10.1
SAT-10168
TS 16 · IDS
Unihan data
kFanqie 涎箭
02738
02738
示 113.7.2
TC-5D33
UTC-03242
TS 12 · IDS
Unihan data
kJapanese ほ
03496
03496
見 147.14.2
UTC-03252
TS 21 · IDS
Unihan data
kJapanese みる
03449
03449
虫 142.16.4
UTC-03372
TS 22 · IDS 𠫓
Unihan data
kJapanese ほらがい
03257
03257
艸 140.6.3
UTC-03392
TS 10 · IDS 𦘱
Unihan data
kJapanese トウ ふじ
03597
03597
貝 154.13.2
VN-F2019
TS 20 · IDS
Unihan data
kVietnamese buôn