文字図譜
ライセンス:CC BY-SA 4.0 · GitHub · 互助認字(中文) · 漢字泉 #集字厨ノ天下(日本語)
字種
文字
異体字
出典: cjkvi-variants · derived-ids · hng-basic-data · mj-shrink-map · opencc · unihan · wikidata · yitizi
- cjkvi-variants CJKVI Variants Database (cjkvi-variants); PD (https://github.com/cjkvi/cjkvi-variants/blob/e4f1da248c9737a243f9930b5dc497cef5d5ae16/cjkvi-variants.txt); 川幡太一, CJKVI Variants Database (cjkvi-variants)
- derived-ids Predicted component variants (derived, not attested): up to two substitutions of data/vocab/han-component-variants.tsv made in a character's BabelStone IDS; Glyph Atlas, CC BY-SA 4.0 (LICENSE-DATA)
- hng-basic-data 漢字字体規範史データセット(HNG)基本データセット; CC-BY-SA-4.0 (https://github.com/chise/hng-basic-data/blob/e2174a30844b8100c34af1c0dbe1e301f186883e/README.md); 漢字字体規範史データセット保存会『漢字字体規範史データセット(HNG)』基本データセット, CC BY-SA 4.0
- mj-shrink-map MJ縮退マップ; CC-BY-SA-2.1-JP (https://moji.or.jp/wp-content/mojikiban/oscdl/MJShrinkMap.1.2.0.json); 文字情報基盤 MJ縮退マップ(IPA), CC BY-SA 2.1 JP
- opencc Open Chinese Convert (OpenCC) dictionaries; Apache-2.0 (https://github.com/BYVoid/OpenCC/blob/2939943bd6f4d459b46d7fdcf07a885ab3f01761/LICENSE); OpenCC contributors, Apache License 2.0
- unihan Unicode Han Database (Unihan), Unihan_Variants.txt; Unicode-3.0 (https://www.unicode.org/license.txt); Copyright Unicode, Inc.; Unicode License v3
- wikidata Wikidata, property P5475 (CJKV variant character); CC0-1.0 (https://www.wikidata.org/wiki/Wikidata:Licensing); Wikidata contributors, CC0 1.0
- yitizi Yitizi 異體字 (nk2028); CC0-1.0 (https://github.com/nk2028/yitizi/blob/60e232c40d6076af07679151c008349363e699ae/LICENSE); nk2028, Yitizi, CC0 1.0
出典: honkoku-ruby · unihan-kjapanese · wiktionary-ja
- honkoku-ruby 振り仮名 of the みんなで翻刻 transcriptions, counted in data/vocab/ruby-spellings.tsv (みんなで翻刻(https://honkoku.org/)翻刻データ, CC BY-SA 4.0, revision be63dc2); crowd transcriptions with 1.5 errors or spelling inconsistencies per 100 characters in an expert audit, transcribers are asked to type modern forms, and the markup guide does not say whether a 振り仮名 is on the page. Only spellings that table attests in at least two documents are listed. The statement is the project's: the spelling with the reading the table keys it by, which the locator names; the forms as typed are in that table's `ruby` column. Readings without their dakuten are included for など (なと) and まで (まて) and left out for ばかり (はかり), which also spells the verb はかる (計り, 量り); すなわち and なお, typed in modern kana, count for すなはち and なほ, and so does なを, an Edo-period spelling of なほ. Spellings of another word with the same reading are left out: 直, 治, 愈, 癒, 修 and 療 for なほす, 名 for 名を, 待, 蟶 and 馬刀 for まて. So are 他事 and 別 for ほか, each of whose documents are copies of one text; 却 for さて, which both documents read as half of 却説(さてまた); and 麦介, which sits inside garbled 割書 markup. もつとも is not counted: Wiktionary gives 最も and 尤も as two words.
- unihan-kjapanese Unicode Han Database (Unihan) 18.0.0, Unihan_Readings.txt, field kJapanese, whose data UAX #38 (revision 41) attributes to the 文字情報基盤 database of CITPC; Unicode-3.0 (https://www.unicode.org/license.txt); Copyright Unicode, Inc.; Unicode License v3. A reading list, Provisional in UAX #38: it does not name the word or the sense.
- wiktionary-ja Wiktionary 日本語版; CC-BY-SA-4.0 (https://ja.wiktionary.org/wiki/Wiktionary:著作権); Wiktionary 日本語版の執筆者 (revision in the locator). An open wiki, not a reviewed dictionary.
ひらがなカタカナ漢字ハングル口訣記号
11字形 · コーパスから11件
字形大図譜717件確認済み · 要修正160件