文字図譜
ライセンス:CC BY-SA 4.0 · GitHub · 互助認字(中文) · 漢字泉 #集字厨ノ天下(日本語)
字種
文字
出典: cjkvi-variants · derived-ids · hng-basic-data · mj-shrink-map · opencc · unihan · wikidata · yitizi
- cjkvi-variants CJKVI Variants Database (cjkvi-variants); PD (https://github.com/cjkvi/cjkvi-variants/blob/e4f1da248c9737a243f9930b5dc497cef5d5ae16/cjkvi-variants.txt); 川幡太一, CJKVI Variants Database (cjkvi-variants)
- derived-ids Predicted component variants (derived, not attested): up to two substitutions of data/vocab/han-component-variants.tsv made in a character's BabelStone IDS; Glyph Atlas, CC BY-SA 4.0 (LICENSE-DATA)
- hng-basic-data 漢字字体規範史データセット(HNG)基本データセット; CC-BY-SA-4.0 (https://github.com/chise/hng-basic-data/blob/e2174a30844b8100c34af1c0dbe1e301f186883e/README.md); 漢字字体規範史データセット保存会『漢字字体規範史データセット(HNG)』基本データセット, CC BY-SA 4.0
- mj-shrink-map MJ縮退マップ; CC-BY-SA-2.1-JP (https://moji.or.jp/wp-content/mojikiban/oscdl/MJShrinkMap.1.2.0.json); 文字情報基盤 MJ縮退マップ(IPA), CC BY-SA 2.1 JP
- opencc Open Chinese Convert (OpenCC) dictionaries; Apache-2.0 (https://github.com/BYVoid/OpenCC/blob/2939943bd6f4d459b46d7fdcf07a885ab3f01761/LICENSE); OpenCC contributors, Apache License 2.0
- unihan Unicode Han Database (Unihan), Unihan_Variants.txt; Unicode-3.0 (https://www.unicode.org/license.txt); Copyright Unicode, Inc.; Unicode License v3
- wikidata Wikidata, property P5475 (CJKV variant character); CC0-1.0 (https://www.wikidata.org/wiki/Wikidata:Licensing); Wikidata contributors, CC0 1.0
- yitizi Yitizi 異體字 (nk2028); CC0-1.0 (https://github.com/nk2028/yitizi/blob/60e232c40d6076af07679151c008349363e699ae/LICENSE); nk2028, Yitizi, CC0 1.0
出典: honkoku-ruby · unihan-kjapanese · wiktionary-ja
- honkoku-ruby 振り仮名 of the みんなで翻刻 transcriptions, counted in data/vocab/ruby-spellings.tsv (みんなで翻刻(https://honkoku.org/)翻刻データ, CC BY-SA 4.0, revision be63dc2); crowd transcriptions with 1.5 errors or spelling inconsistencies per 100 characters in an expert audit, transcribers are asked to type modern forms, and the markup guide does not say whether a 振り仮名 is on the page. Only spellings that table attests in at least two documents are listed. The statement is the project's: the spelling with the reading the table keys it by, which the locator names; the forms as typed are in that table's `ruby` column. Readings without their dakuten are included for など (なと), まで (まて) and ごとし (ことし, ことく), with the iteration mark for ただ (たゞ), and left out for ばかり (はかり), which also spells the verb はかる (計り, 量り); すなわち, なお, ゆえ and もって, typed in modern kana, count for すなはち, なほ, ゆゑ and もつて, ごとく for ごとし, and なを and ゆへ, Edo-period spellings, for なほ and ゆゑ. Spellings of another word with the same reading are left out: 直, 治, 愈, 癒, 修 and 療 for なほす, 名 for 名を, 待, 蟶 and 馬刀 for まて. So are 他事 and 別 for ほか, each of whose documents are copies of one text; 却 for さて, which both documents read as half of 却説(さてまた); and 麦介, which sits inside garbled 割書 markup. Under the later words, the main ones are: 股, 叉, 岐 and 俣 (the noun また), 待 and 俟 (まつ, as in またず), 全 (half of 全く); 正, 直, 糺, 忠 and 爛 (ただす, ただれ); 但, whose readings are mostly ただし, and ただ 啻, whose two documents are copies of one book; 虎列 (コレラ), and 這 for これ, whose second document reads it as half of 這箇; 好 and 嗜 (このむ); 園 for その; 今年, 今茲 and 本年 for ことし; 殺 and 転 for ころ; 試, 溜 and 例 for ため. Readings that are half of a compound (往古, 以来, 当時) are not listed one by one. もつとも is not counted: Wiktionary gives 最も and 尤も as two words.
- unihan-kjapanese Unicode Han Database (Unihan) 18.0.0, Unihan_Readings.txt, field kJapanese, whose data UAX #38 (revision 41) attributes to the 文字情報基盤 database of CITPC; Unicode-3.0 (https://www.unicode.org/license.txt); Copyright Unicode, Inc.; Unicode License v3. A reading list, Provisional in UAX #38: it does not name the word or the sense.
- wiktionary-ja Wiktionary 日本語版; CC-BY-SA-4.0 (https://ja.wiktionary.org/wiki/Wiktionary:著作権); Wiktionary 日本語版の執筆者 (revision in the locator). An open wiki, not a reviewed dictionary.
ひらがなカタカナ漢字ハングル口訣記号
60字形 · ここに60件 · コーパスから60件
字形大図譜829件確認済み · 要修正162件