UTF-8 Byte Counter
Count the number of UTF-8 bytes and characters in a text to determine the exact storage size and character length at a glance. GDPR-compliant.
Your inputs are processed in your browser and are not transmitted to our servers. Note: third-party resources (e.g. advertising and analytics from Google/Cloudflare) and an optional PayPal donation link may transfer data when loading or when clicked. Browser extensions or plugins may be able to read content that is visible in the input fields.
The result will appear here …
How to use this tool (video)
This video is hosted on YouTube. When you play it, data may be sent to Google.
UTF-8 Byte Counter
The UTF-8 byte counter counts the number of UTF-8 bytes and characters in a text. UTF-8 is a variable-width encoding where different characters require different numbers of bytes. ASCII characters use 1 byte, Latin umlauts use 2 bytes, and emojis use 4 bytes. The counter shows you the exact byte count and character count. The converter works locally in your browser.
How it works
UTF-8 uses variable byte counts depending on the character. ASCII characters in the range 0 to 127 require exactly 1 byte. Letters with diacritics such as ä, ö, ü, or é require 2 bytes. Characters from the Greek or Cyrillic scripts also require 2 bytes. Emojis and characters from Asian writing systems require 4 bytes. For example, the word "Hällo" has 5 characters and requires 6 bytes because ä needs 2 bytes. An emoji-heavy text can require significantly more bytes than characters.
What the tool can and cannot do
The counter can determine the number of UTF-8 bytes and characters in a text. It shows the exact byte count per character and the total sum. However, it cannot determine the actual file size on disk because filesystem overhead is not considered. This tool is a pure text analysis utility.
Why local processing in the browser
All calculations happen locally in your browser. Your text inputs are never sent to a server. There is no upload and no server log files. You need neither an account nor a registration. After the page loads, the counter works offline too. There is no server that could be attacked and nothing is transmitted over the network. You can analyze texts as often as you like.
Frequently asked questions
How does UTF-8 work?
UTF-8 is a variable-width encoding. ASCII characters use 1 byte, Latin umlauts use 2 bytes, and emojis use 4 bytes. This makes UTF-8 backwards compatible with ASCII.
Do I need to be online?
No, after the page has loaded, the counter works offline. The required assets are cached.
Are my data uploaded?
No, your text inputs are never sent to a server. Everything is processed locally in your browser.
Can the counter determine file sizes on disk?
No, the counter only determines the UTF-8 byte count of text. Filesystem overhead is not considered.
What happens with spaces and punctuation?
Spaces and most punctuation marks require 1 byte in UTF-8. Only certain special characters require more.
Is the tool free to use?
Yes, the UTF-8 byte counter is a free tool that runs directly in your browser. There are no hidden costs.