टेक्स्ट
Character & Word Counter
रीयल-टाइम टेक्स्ट स्टैटिस्टिक्स: कैरेक्टर्स (स्पेस के साथ और बिना), शब्द, लाइनें, वाक्य, पैराग्राफ और बाइट्स में साइज़ (UTF-8)।
This tool counts characters, words, lines, sentences, and paragraphs in real time as you type or paste text, and also shows the size in bytes under UTF-8 encoding.
What gets counted
- Characters with and without spaces — shown separately so you can check against limits like meta descriptions or tweets.
- Words — split on spaces and the punctuation between them.
- Sentences and paragraphs — split on terminal punctuation (. ! ?) and blank lines, respectively.
- UTF-8 byte size — useful when a limit is measured in bytes rather than characters (non-Latin text and emoji take more than one byte each).
Common uses
- Checking a meta description or title's length before publishing so it fits the search result limit.
- Keeping text within a character limit for an assignment, bio, or product description.
- Getting quick text stats without opening a text editor with a word-count panel.
Things to keep in mind
One visual "character" (like an emoji with a skin-tone modifier or a country flag) can be made of several Unicode code points and count differently depending on the counting method.
Social media and form limits sometimes count bytes or grapheme clusters rather than characters — check against this tool's byte counter if that matters.
इस टूल के बारे में लेख: कैरेक्टर और शब्द गिनना: यूनिकोड इसे जटिल क्यों बनाता है
अक्सर पूछे जाने वाले प्रश्न
वास्तव में कौन-से आंकड़े गणना किए जाते हैं?
टूल स्पेस सहित और बिना स्पेस के अक्षर, शब्द, लाइनें, वाक्य, पैराग्राफ, और UTF-8 एन्कोडेड बाइट्स में टेक्स्ट का आकार गिनता है।
बाइट काउंट अक्षर काउंट से अलग क्यों होता है?
UTF-8 में, एक सामान्य लैटिन अक्षर एक बाइट लेता है, लेकिन एक सिरिलिक अक्षर, एक्सेंट वाला अक्षर, या इमोजी दो से चार बाइट्स ले सकते हैं, इसलिए कई मल्टी-बाइट कैरेक्टर्स वाले टेक्स्ट का बाइट काउंट उसके कैरेक्टर काउंट से काफ़ी ज़्यादा होगा।
क्या मेरा टेक्स्ट विश्लेषण के लिए सर्वर पर भेजा जाता है?
नहीं, सारी गणनाएं आपके ब्राउज़र में JavaScript में रीयल टाइम में होती हैं, कुछ भी ऑनलाइन नहीं भेजा जाता।
बिना स्पेस वाली भाषाओं के लिए शब्द गिनती अशुद्ध क्यों होती है?
चीनी, जापानी या थाई भाषा में परंपरागत रूप से शब्दों को स्पेस से अलग नहीं किया जाता — सीमा व्याकरण और संदर्भ से तय होती है। ऐसी भाषाओं के लिए स्पेस पर आधारित गिनती या तो पूरे टेक्स्ट को एक "शब्द" गिन देती है, या अर्थहीन हो जाती है — इसके लिए अलग सेगमेंटेशन एल्गोरिदम चाहिए।
एक इमोजी कभी-कभी कई कैरेक्टर के बराबर क्यों गिना जाता है?
स्किन टोन मॉडिफ़ायर या देश के झंडे वाले इमोजी यूनिकोड में कई कोड पॉइंट से बने होते हैं जो एक विज़ुअल ग्रैफ़ीम क्लस्टर में जुड़ते हैं। कोड पॉइंट पर आधारित नैव गिनती ऐसे विज़ुअली एक कैरेक्टर को दो, तीन या ज़्यादा गिन सकती है।