Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
nomadamas avatar

Korean Character Count

  • 2.8k installs
  • 6.5k repo stars
  • Updated July 27, 2026
  • nomadamas/k-skill

korean-character-count deterministically counts Korean graphemes, lines, and bytes using default or neis profiles via a Node helper script.

About

Korean Character Count deterministically counts Korean text for self-introduction essays, applications, and free-form fields where one character matters, without LLM estimation. The default profile uses Intl.Segmenter with ko grapheme granularity for character counts, UTF-8 Buffer.byteLength for bytes, and newline sequences CRLF, LF, CR, U+2028, and U+2029 counted as single line breaks with empty strings returning zero lines. The neis profile shares grapheme and line rules but counts Korean graphemes as 3 bytes, ASCII as 1 byte, and newlines as 2 bytes for school record compatibility. A Node 18 helper script korean_character_count.js accepts text, file, or stdin input with json or text output formats. The skill never trims or normalizes input arbitrarily and reports which profile produced results. Submission workflows for NEIS or school records should select neis only when that contract is required.

  • Deterministic Intl.Segmenter grapheme counting without LLM guesses.
  • default and neis byte profile contracts documented.
  • Line counting rules for CRLF and Unicode line separators.
  • CLI helper for text, file, and stdin with json output.
  • No arbitrary trim or normalize on submitted text input.

Korean Character Count by the numbers

  • 2,824 all-time installs (skills.sh)
  • +126 installs in the week ending Jul 28, 2026 (Skillselion tracking)
  • Ranked #185 of 3,301 Productivity & Planning skills by installs in the Skillselion catalog
  • Security screen: LOW risk (skills.sh audit)
  • Data as of Jul 28, 2026 (Skillselion catalog sync)
At a glance

korean-character-count capabilities & compatibility

Capabilities
intl.segmenter ko grapheme character counting · utf 8 and neis byte profile calculations · newline aware line counting rules · cli text file and stdin input support · json and text output format selection · profile contract reporting in responses
Use cases
translation · copywriting
Pricing
Free
From the docs

What korean-character-count says it does

LLM이 글자 수를 눈대중으로 예측하면 재현성이 없다.
SKILL.md
입력을 임의로 trim/정규화하지 않고
SKILL.md
npx skills add https://github.com/nomadamas/k-skill --skill korean-character-count

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs2.8k
repo stars6.5k
Security audit3 / 3 scanners passed
Last updatedJuly 27, 2026
Repositorynomadamas/k-skill

How do I count Korean essay characters and bytes exactly for form limits without LLM estimation?

Count Korean text deterministically with grapheme, line, and byte contracts for essays and form character limits.

Who is it for?

Users validating Korean self-introduction or application text against strict character or byte limits.

Skip if: Skip for non-Korean text-only counting or semantic editing of essay content.

When should I use this skill?

User asks to count Korean characters, bytes, or lines for essays, NEIS, or form limits.

What you get

Grapheme, line, and byte counts with stated profile contract from korean_character_count.js output.

  • grapheme count results
  • utf-8 byte totals
  • hangul segment breakdown

By the numbers

  • Requires Node.js 18 or newer for Intl.Segmenter

Files

SKILL.mdMarkdownGitHub ↗

한국어 글자 수 세기

What this skill does

자기소개서, 지원서, 자유서술형 폼처럼 글자 수 제한이 중요한 한국어 텍스트를 대상으로 LLM 추정 없이 결정론적으로 카운트한다.

  • 기본 글자 수: Intl.Segmenter 기반 Unicode extended grapheme cluster
  • 줄 수: CRLF, LF, CR, U+2028, U+2029 를 줄바꿈 1회로 계산
  • 기본 byte 수: UTF-8 실제 인코딩 길이
  • 호환 프로필: neis byte 규칙

When to use

  • "이 자기소개서 1000자 넘는지 정확히 세줘"
  • "이 텍스트를 UTF-8 byte 기준으로 계산해줘"
  • "줄 수랑 byte 수도 같이 알려줘"
  • "한글/영문/이모지 섞인 문장을 추정 말고 코드로 세줘"

Why this skill exists

  • 글자 수 제한은 1자 차이도 민감하다.
  • LLM이 글자 수를 눈대중으로 예측하면 재현성이 없다.
  • 이 스킬은 입력을 임의로 trim/정규화하지 않고, 문서화된 계약으로만 센다.

Contracts

default profile

  • characters: Intl.Segmenter("ko", { granularity: "grapheme" })
  • bytes: Buffer.byteLength(text, "utf8")
  • lines:
  • empty string => 0
  • non-empty => 줄바꿈 시퀀스 수 + 1
  • CRLF2줄바꿈이 아니라 1줄바꿈으로 센다.

neis profile

  • characters: default 와 동일
  • lines: default 와 동일
  • bytes:
  • 한글 grapheme => 3B
  • ASCII grapheme => 1B
  • Enter/줄바꿈 시퀀스 => 2B
  • 그 외 문자는 UTF-8 byte 길이로 fallback

Prerequisites

  • node 18+
  • 설치된 skill payload 안에 scripts/korean_character_count.js helper 포함
  • 별도 API 키 없음

Workflow

1. 텍스트를 직접 받거나 파일/STDIN으로 읽는다. 2. node scripts/korean_character_count.js 로 결정론적 카운트를 실행한다. 3. 필요한 프로필(default/neis)과 출력 형식(json/text)을 고른다. 4. 결과를 그대로 반환하고, 어떤 계약으로 셌는지 함께 알려준다.

CLI examples

node scripts/korean_character_count.js --text "가나다"
node scripts/korean_character_count.js --text $'첫 줄\r\n둘째 줄🙂'
node scripts/korean_character_count.js --text $'첫 줄\n둘째 줄🙂' --profile neis --format text
node scripts/korean_character_count.js --file ./essay.txt --profile default
cat essay.txt | node scripts/korean_character_count.js --stdin --profile neis

Response policy

  • 추정하지 말고 helper 결과를 그대로 쓴다.
  • 어떤 profile로 셌는지 함께 보여준다.
  • 기본값이 필요하면 default profile을 사용한다.
  • 제출처가 NEIS/학교생활기록부 같은 별도 계약을 요구할 때만 neis 를 쓴다.

Done when

  • 글자 수, 줄 수, byte 수가 함께 반환된다.
  • defaultneis 계약 차이가 문서에 명시된다.
  • node scripts/korean_character_count.js --help 가 동작한다.
  • 혼합 한국어/영문/공백/개행/emoji 입력에 대한 테스트가 있다.

Notes

  • Unicode grapheme clusters: https://www.unicode.org/reports/tr29/
  • WHATWG Encoding Standard: https://encoding.spec.whatwg.org/
  • Node Buffer.byteLength: https://nodejs.org/api/buffer.html

Related skills

FAQ

default or neis profile?

Use default for UTF-8 bytes; use neis only when school or NEIS byte rules require the special mapping.

Does it estimate character counts?

No. It runs the helper script and returns documented contract results without LLM guessing.

How are lines counted?

Empty string is zero lines; otherwise newline sequences count once with CRLF treated as one break.

Is Korean Character Count safe to install?

skills.sh reports 3 of 3 security scanners passed. Review the Security Audits panel on this page before installing in production.

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.