We use cookies.This website uses essential cookies to operate core features. With your consent, we also use analytics cookies to understand traffic and improve the service. For more details, see our .
Was this tool helpful to use?
Your feedback helps us make it better
Count the frequency of characters or phrases in your text. Supports custom N-gram lengths and character types.
Please enter text and configure options to start analysis
Overview
Understand what the tool solves, how it works, and the boundaries of its data.
Paste text to count one-character units or chunks from one to five characters long. The analyzer ranks the most frequent units and shows their counts and percentages in a chart and table. It is a character-frequency tool: it does not count words, identify an encrypted message, or solve a cipher.
For chunk lengths above one, the text is split into consecutive, non-overlapping blocks starting at the beginning. For example, HELLO with a length of two is grouped as HE and LL; the final O is too short and is omitted. This differs from overlapping bigrams, which would also include EL and LO.
The analyzer retains ASCII English letters, plus the optional digits and whitespace. Accented letters, non-Latin scripts, punctuation and emoji are filtered out. This limitation matters when analyzing multilingual text: a result may represent only a subset of the original input.
Complete chunks = floor(filtered character count ÷ selected chunk length)
Chunk percentage = chunk count ÷ complete chunks × 100%
Percentages are rounded to two decimal places. The displayed table is limited to the 50 highest-count chunks, while percentages use all complete chunks. As a result, adding only the visible percentages may give less than 100%. Ties may appear in either order.
Guide
Follow the workflow and verify inputs and outputs with practical examples.
Use a sample whose character set matches the analyzer’s ASCII-based coverage. The result updates as the text or options change.
Select one for individual characters or two through five for fixed, non-overlapping chunks.
Keep ignore-case on to combine uppercase and lowercase forms. Turn on numbers or whitespace only when those characters should be part of the sample.
Use the table for the top 50 ranked chunks. Remember that omitted low-frequency chunks still contribute to the percentage denominator.
Q&A
Find concise answers to common questions and confusing cases.
No. It advances by two characters at a time, producing fixed non-overlapping chunks. A short remainder at the end is discarded.
The current character filter keeps ASCII English letters and optionally digits and whitespace. Other Unicode characters are removed before counting.
The table shows at most the 50 most frequent chunks, but each percentage is calculated using every complete chunk in the filtered input.
Notes
Review scope, result limitations, and important precautions before use.
This is a frequency counter for a limited set of characters, not a word tokenizer, language detector or cipher solver. Because unsupported characters are filtered before chunking, results from multilingual text can omit substantial portions of the input. Use a Unicode-aware analyzer if the full character repertoire matters.
Related
Discover related tools, collections, and available API capabilities.