Character & Word Counter
Count characters, words, sentences, paragraphs and UTF-8 bytes as you type, with reading time and a token estimate for AI context limits.
About these counts
Detailed statistics
Most frequent words
Import from File
Drag & drop a file here
or
Supports .txt, .csv, .log, .json and other text files
Counting happens entirely in your browser as you type. Your text is never uploaded, logged or stored.
What It Does
This counts what is actually in your text, live as you type. Characters, words, sentences, paragraphs and lines are the obvious part, and most counters stop there. The numbers that tend to be missing elsewhere are the ones that cause real problems: the UTF-8 byte count, which is what a database column or a protocol field is genuinely limited by, and a proper grapheme count, so an emoji registers as the one character a person sees rather than the two units a computer stores. Alongside those it estimates reading and speaking time, gives a rough token count for fitting text into an AI context window, and shows which words you have leaned on most. Everything updates as you type and nothing is sent anywhere.
When to Use It
- You are writing a meta description or a social post with a hard character limit and need to know exactly where you stand as you edit.
- A form or database field keeps rejecting text that looks short enough, and you suspect the limit is counted in bytes rather than characters.
- You are checking whether a document will fit inside a language model's context window before pasting it in.
- An article needs a reading-time label and you want the figure without installing a plugin.
- You are editing prose and want to see which words you have repeated too often.
Worked Examples
The quick brown fox jumps over the lazy dog.
Plain ASCII, so the character and byte counts are identical — nine words, one sentence, 44 of each. This is the baseline case where every counter agrees.
Café naïve résumé 🔒
The interesting case. Four visible items but noticeably more bytes, because each accented letter takes two bytes in UTF-8 and the emoji takes four. This gap is exactly what breaks a field limit that was specified in characters but enforced in bytes.
First paragraph here.
Second paragraph follows. It has two sentences.
Shows the structural counts working: two paragraphs, three sentences, and a line count that differs from the paragraph count because of the blank line between them.
Features
How to Use
1. Type or paste your text into the box — the counts update immediately. 2. The headline figures show characters, words, sentences and bytes. 3. Expand the detailed statistics for reading time, averages and the most frequent words. 4. Watch the byte count if you are working against a database or protocol limit. 5. Click Clear to empty the box and start again.
Common Mistakes
- Assuming a character limit means characters. Database columns and many APIs count bytes, so text with accents, emoji or non-Latin script hits the limit sooner than its character count suggests.
- Trusting a counter that reports code units. Most report string length, which counts an emoji as two and can count a single flag or family emoji as eight — a real problem for anything with a hard limit.
- Treating an AI token estimate as exact. Tokenisation depends on the model's vocabulary, and code or non-English text produces far more tokens per character than ordinary prose.
- Forgetting that trailing whitespace counts. A stray newline at the end of a paste is a real character, and it is enough to push text one over a limit for no visible reason.
- Comparing word counts across tools and expecting them to match. Different applications treat hyphens, contractions and numerals differently, so a small discrepancy is normal.