Multilingual editor
Review Chinese and Latin boundaries without ASCII-only splitting.
Choose zh or en, compare CJK punctuation, decimals and actual candidate text.
Can correct editorial assumptions instead of treating segmentation as grammar proof.
Inspect sentence or line candidates with language-aware segmentation, word/scalar/grapheme lengths, source offsets and complete JSON review reports.
A quick decision brief for this specific tool
Text source
Complete JSON report
A dedicated interactive panel for this data format or operation
Choose your path
Inspect locale-sensitive sentence candidates or explicit line records, compare word/scalar/grapheme lengths and trace complete segments to original UTF-16 ranges.
Review Chinese and Latin boundaries without ASCII-only splitting.
Choose zh or en, compare CJK punctuation, decimals and actual candidate text.
Can correct editorial assumptions instead of treating segmentation as grammar proof.
Find segments worth reviewing without a universal length grade.
Choose words/scalars/graphemes, set a review threshold and compare complete candidates.
Threshold markers prompt contextual review, not an accessibility verdict.
Trace long or punctuation-free records without losing text.
Switch sentence/line boundaries, inspect UTF-16 ranges, sort longest first and export 220 candidates.
Original segment numbers and complete text survive sorting and preview limits.
Outputs and checklists are planning aids. Review the linked current authorities and the records, terms, instructions, and requirements that apply to your exact situation before a consequential decision.
Tools you might need next
Measure UTF-16 units, Unicode scalars, grapheme clusters and UTF-8 bytes, inspect exact offsets and check an optional length budget.
Convert text to sentence case. Capitalize after periods. For proper sentences. Free online text tool — process content instantly in your browser with no
Estimate locale-sensitive sentence segments and expose token, line, paragraph, preview, and fallback context without a grammar or clarity verdict.
Sentence mode uses this runtime’s Intl.Segmenter for a supported language tag. Line mode uses explicit Unicode line records. Ambiguous text can differ across runtime dictionaries.
Count word-like segments, scalars or graphemes after excluding boundary Unicode whitespace. Report original source ranges and complete untrimmed segment text separately.
Mean, median, minimum and maximum describe selected lengths. A strict greater-than threshold is your review marker, not a universal readability or accessibility standard.
Updated: August 2026
Review Chinese and Latin boundaries without ASCII-only splitting. Choose zh or en, compare CJK punctuation, decimals and actual candidate text.
Find segments worth reviewing without a universal length grade. Choose words/scalars/graphemes, set a review threshold and compare complete candidates.
Trace long or punctuation-free records without losing text. Switch sentence/line boundaries, inspect UTF-16 ranges, sort longest first and export 220 candidates.
Read the actual content and check reader needs; length alone proves neither.
Line mode deliberately treats records as candidates, including punctuation-free fragments.
Inspect locale-sensitive sentence candidates or explicit line records, compare word/scalar/grapheme lengths and trace complete segments to original UTF-16 ranges.