73 releases (39 breaking)

new 0.41.0 Apr 13, 2025
0.40.1 Mar 27, 2025
0.38.1 Nov 30, 2024
0.32.2 Jun 29, 2024
0.3.2 Feb 20, 2020

#2236 in Text processing

Download history 2962/week @ 2024-12-22 2423/week @ 2024-12-29 3980/week @ 2025-01-05 4118/week @ 2025-01-12 2850/week @ 2025-01-19 3239/week @ 2025-01-26 3850/week @ 2025-02-02 3547/week @ 2025-02-09 3948/week @ 2025-02-16 5714/week @ 2025-02-23 5763/week @ 2025-03-02 4877/week @ 2025-03-09 5038/week @ 2025-03-16 4940/week @ 2025-03-23 8642/week @ 2025-03-30 6102/week @ 2025-04-06

25,072 downloads per month
Used in 20 crates (via lindera)

MIT license

140KB
3K SLoC

Lindera IPADIC

License: MIT Crates.io

Dictionary version

This repository contains mecab-ipadic.

Dictionary format

Refer to the manual for details on the IPADIC dictionary format and part-of-speech tags.

Index Name (Japanese) Name (English) Notes
0 表層形 Surface
1 左文脈ID Left context ID
2 右文脈ID Right context ID
3 コスト Cost
4 品詞 Major POS classification
5 品詞細分類1 Middle POS classification
6 品詞細分類2 Small POS classification
7 品詞細分類3 Fine POS classification
8 活用形 Conjugation type
9 活用型 Conjugation form
10 原形 Base form
11 読み Reading
12 発音 Pronunciation

User dictionary format (CSV)

Simple version

Index Name (Japanese) Name (English) Notes
0 表層形 surface
1 品詞 Major POS classification
2 読み Reading

Detailed version

Index Name (Japanese) Name (English) Notes
0 表層形 Surface
1 左文脈ID Left context ID
2 右文脈ID Right context ID
3 コスト Cost
4 品詞 POS
5 品詞細分類1 POS subcategory 1
6 品詞細分類2 POS subcategory 2
7 品詞細分類3 POS subcategory 3
8 活用形 Conjugation type
9 活用型 Conjugation form
10 原形 Base form
11 読み Reading
12 発音 Pronunciation
13 - - After 13, it can be freely expanded.

API reference

The API reference is available. Please see following URL:

Dependencies

~12–24MB
~387K SLoC