Text unicode normalization check
text-unicode-normalization-check · version 1.0.0 · Text · free, no key needed
Check if a text string is in various Unicode normalization forms (NFC, NFD, NFKC, NFKD).
Use when you need to: text unicode normalization check · check unicode normalization · is text nfc.
Supported
- text unicode normalization check
- check unicode normalization
- is text nfc
Not supported
- normalize text
- fix encoding
- detect encoding
Behavior
- Checks the provided string against the four standard Unicode normalization forms.
- Returns true for a form if the string is identical to its normalized version in that form.
- Assumes the input is valid UTF-16 (JavaScript strings).
Input
text(string, required): max length 1048576
Output
is_nfc(boolean, required)is_nfd(boolean, required)is_nfkc(boolean, required)is_nfkd(boolean, required)
Limits
- max text bytes: 1048576
Example
Request input:
{
"text": "é"
}
Response:
{
"result": {
"is_nfc": true,
"is_nfd": false,
"is_nfkc": true,
"is_nfkd": false
}
}
How to call it
MCP
Connect https://computefirst.net/mcp (setup), then call execute with:
{
"id": "text-unicode-normalization-check",
"version": "1.0.0",
"input": {
"text": "é"
}
}
HTTP (no key)
curl -X POST https://computefirst.net/v1/tools/text-unicode-normalization-check/versions/1.0.0/execute \
-H "Content-Type: application/json" \
-d '{"text":"é"}'
The machine-readable contract is at /v1/tools/text-unicode-normalization-check/versions/1.0.0.
CLI
node cli.mjs run text-unicode-normalization-check 1.0.0 --input input.json --base-url https://computefirst.net
Get the client at /clients/cli/.
Related tools
- Text NFC: Normalize Unicode text to NFC (canonical composition).
- Text indentation detect: Detect whether a text uses tabs or spaces for indentation by counting lines.
- Text NFD: Normalize Unicode text to NFD (canonical decomposition).
- Text prefix suffix check: Exact prefix and/or suffix boolean checks using JS string startsWith/endsWith.
- Text slice codepoints: Slice a string by Unicode code-point indexes, not UTF-16 code units.
- Text apply unified patch: Apply a one-file unified patch to text at original line addresses with exact hunk matches.