🏆 #2,519 overall#99 of 148 in Tokenization & Preprocessing🔥 active this week

ongjin /garu

초경량 한국어 형태소 분석기 — 1MB 모델로 브라우저에서 실행(WASM). Ultra-lightweight Korean morphological analyzer in the browser. npm: garu-ko

$ git clone https://github.com/ongjin/garu.git
GitHub social preview for ongjin/garu
Stars
177
177
Forks
9
9
Language
Python
License
MIT
Created
Mar 27, 2026
0.5 years old
Last push
Sep 21, 2026
🔥 this week

Categories

GitHub topics

More in Tokenization & Preprocessing

Tokenization & Preprocessing

daac-tools/vibrato🔥 active

🎤 vibrato: Viterbi-based accelerated tokenizer

🧠 Natural Language Processing
RustApache-2.0updated Sep 19, 2026
GitHub ↗★ 423⑂ 26
Tokenization & Preprocessing

ku-nlp/jumanpp

Juman++ (a Morphological Analyzer Toolkit)

🧠 Natural Language Processing
C++Apache-2.0updated Apr 17, 2026
GitHub ↗★ 413⑂ 47
Tokenization & Preprocessing

daac-tools/vaporetto

🛥 Vaporetto: Very accelerated pointwise prediction based tokenizer

🧠 Natural Language Processing
RustApache-2.0updated Jul 20, 2026
GitHub ↗★ 299⑂ 12
Tokenization & Preprocessing

ikawaha/kagome🔥 active

Self-contained Japanese Morphological Analyzer written in pure Go

GoMITupdated Sep 19, 2026
GitHub ↗★ 983⑂ 60
Tokenization & Preprocessing

lovit/soynlp

한국어 자연어처리를 위한 파이썬 라이브러리입니다. 단어 추출/ 토크나이저 / 품사판별/ 전처리의 기능을 제공합니다.

🧠 Natural Language Processing
PythonOtherupdated Mar 10, 2026
GitHub ↗★ 993⑂ 183
Tokenization & Preprocessing

open-korean-text/open-korean-text

Open Korean Text Processor - An Open-source Korean Text Processor

🧠 Natural Language Processing
ScalaApache-2.0updated Mar 12, 2024
GitHub ↗★ 669⑂ 95

Data from GitHub · snapshot Sep 24, 2026