kuromoji.js

JavaScript implementation of Japanese morphological analyzer. This is a pure JavaScript porting of Kuromoji.

Updates to original package

rewrited with es6, promise
removed the zlib
supports loading dicts from url

Usage

Build dictionary data from IPADIC:

npm run build-dict

Build the module:

npm run build

Usage

You can tokenize sentences with only 5 lines of code. If you need working examples, you can see the files under the demo or example directory.

Node.js

Load this library as follows:

var kuromoji = require("./kuromoji");

You can prepare tokenizer like this:

kuromoji.builder({ dicPath: "path/to/dictionary/dir/" }).build()
    .then((tokenizer) => {
        // tokenizer is ready
        const path = tokenizer.tokenize("すもももももももものうち");
        console.log(path);
    }).catch(err => console.log(err));

Browser

You only need the dist/kuromoji.js and dict/*.dat files

<script src="url/to/kuromoji.js"></script>

In your JavaScript:

kuromoji.builder({ dicPath: "/url/to/dictionary/dir/" }).build(function (err, tokenizer) {
    // tokenizer is ready
    var path = tokenizer.tokenize("すもももももももものうち");
    console.log(path);
});

API

The function tokenize() returns an JSON array like this:

[ {
    word_id: 509800,          // 辞書内での単語ID
    word_type: 'KNOWN',       // 単語タイプ(辞書に登録されている単語ならKNOWN, 未知語ならUNKNOWN)
    word_position: 1,         // 単語の開始位置
    surface_form: '黒文字',    // 表層形
    pos: '名詞',               // 品詞
    pos_detail_1: '一般',      // 品詞細分類1
    pos_detail_2: '*',        // 品詞細分類2
    pos_detail_3: '*',        // 品詞細分類3
    conjugated_type: '*',     // 活用型
    conjugated_form: '*',     // 活用形
    basic_form: '黒文字',      // 基本形
    reading: 'クロモジ',       // 読み
    pronunciation: 'クロモジ'  // 発音
  } ]

(This is defined in src/util/IpadicFormatter.js)

Name		Name	Last commit message	Last commit date
Latest commit History 180 Commits
demo		demo
example		example
scripts		scripts
src		src
test		test
.codeclimate.yml		.codeclimate.yml
.eslintignore		.eslintignore
.eslintrc.js		.eslintrc.js
.gitignore		.gitignore
.node-version		.node-version
CHANGELOG.md		CHANGELOG.md
LICENSE-2.0.txt		LICENSE-2.0.txt
NOTICE.md		NOTICE.md
README.md		README.md
debug.js		debug.js
package-lock.json		package-lock.json
package.json		package.json
webpack.config.js		webpack.config.js
webserver.js		webserver.js

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

kuromoji.js

Updates to original package

Usage

Directory

Usage

Node.js

Browser

API

About

Releases

Packages

Languages

mirigana/kuromoji.js

Folders and files

Latest commit

History

Repository files navigation

kuromoji.js

Updates to original package

Usage

Directory

Usage

Node.js

Browser

API

About

Resources

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages