Multilingual interface designer
Keep emoji families and accented characters intact in short labels.
Paste mixed Unicode text, set a grapheme budget and inspect retained characters.
No partial surrogate or combining sequence is emitted.
Shorten text at complete Unicode grapheme boundaries. Include the marker in a character or byte budget, keep either end or both, then preview and export.
A quick decision brief for this specific tool
Original text
Shortened text
A dedicated interactive panel for this data format or operation
Choose your path
Shorten text at complete grapheme boundaries under a marker-inclusive grapheme, UTF-16 or UTF-8 budget while retaining the chosen ends.
Keep emoji families and accented characters intact in short labels.
Paste mixed Unicode text, set a grapheme budget and inspect retained characters.
No partial surrogate or combining sequence is emitted.
Fit a preview within an explicit UTF-8 budget.
Choose UTF-8 bytes, use a short marker, inspect original/output sizes and test a too-large marker.
Output stays within budget or returns a clear refusal.
Compare beginning, ending and middle excerpts without losing source.
Choose marker positions, clear and undo source, restore the prior successful run and export.
Source remains available and the marker is included in the stated limit.
Outputs and checklists are planning aids. Review the linked current authorities and the records, terms, instructions, and requirements that apply to your exact situation before a consequential decision.
Tools you might need next
Decode binary bytes with strict UTF-8 or ASCII validation. Choose token grouping and BOM handling, inspect every byte, and export the complete decoded text.
Decode complete Morse tokens with explicit slash or line-break word boundaries. Refuse unknown signals, inspect character rows, and export the full transcript.
Repeat whole text blocks with exact copy counts, literal separators and optional numbering. Preserve whitespace, review bounded previews, then copy or export.
Intl.Segmenter supplies the current runtime’s grapheme boundaries. UTF-16 and UTF-8 budgets still cut only between whole graphemes; unsupported segmentation fails explicitly.
The budget ranges from 0 to 100,000 selected units and includes the marker. End keeps a prefix; Start keeps a suffix; Middle gives the prefix the rounded-up half of remaining capacity, then fills available space.
No normalization or whitespace trimming is applied. Runtime Unicode versions may differ, and character counts do not guarantee pixel width, readability or a platform-specific limit.
Updated: August 2026
Keep emoji families and accented characters intact in short labels. Paste mixed Unicode text, set a grapheme budget and inspect retained characters.
Fit a preview within an explicit UTF-8 budget. Choose UTF-8 bytes, use a short marker, inspect original/output sizes and test a too-large marker.
Compare beginning, ending and middle excerpts without losing source. Choose marker positions, clear and undo source, restore the prior successful run and export.
The marker consumes part of the budget; it is not appended beyond the stated maximum.
A grapheme is not a fixed visual width or a word. Check the final layout and meaning before publishing.
Create a bounded excerpt without splitting an emoji or combining sequence. Select how to count, where to place the marker and which part of the original to retain.