Text Processing
Count, clean, compare, and transform text locally in your browser.
Available Tools (55)
← Back to all toolsText Statistics & Analysis
Word Counter
Count words, characters, sentences, and paragraphs in real time. Calculate reading and speaking time for essays, articles, and speeches.
Readability Score Calculator
Readability Score Calculator: Uses an English syllable heuristic plus word, sentence, and character counts to calculate Flesch Reading Ease, Flesch-Kincaid Grade Level, and ARI. Non-English text needs a language-specific method.
Multilingual Stop Words Remover
Multilingual Stop Words Remover: Matches words against standard NLTK English stop words corpus and removes matches.
Text Word Frequency & Keyword Extractor
Text Word Frequency & Keyword Extractor: Tokenizes words, normalizes case, and aggregates occurrence counts sorted descending.
Local Extractive Text Summarizer
Local Extractive Text Summarizer: Scores sentences using TextRank graph algorithm based on lexical overlap and term significance.
Reading & Speaking Time Estimator
Reading & Speaking Time Estimator: Calculates reading time based on 225 WPM (silent reading) and speaking time based on 130 WPM (presentation pace).
Repeated Phrase & Cliché Finder
Repeated Phrase & Cliché Finder: Extracts n-grams (length 2-5) and counts identical contiguous phrase matches.
Line & Character Density Counter
Line & Character Density Counter: Scans line breaks (\r\n and \n) and categorizes empty versus populated rows.
Duplicate Word Highlighter
Duplicate Word Highlighter: Scans for identical contiguous words across spaces and punctuation boundaries.
Text Sentiment Polarity Checker
Text Sentiment Polarity Checker: Tokenizes text against AFINN-165 sentiment lexicon and aggregates valence scores.
Tool category
CJK & Latin Character Counter
CJK & Latin Character Counter: Scans codepoints against CJK Unified Ideographs, Hiragana, Katakana, and Latin script Unicode blocks.
Traditional & Simplified Chinese Converter
Traditional & Simplified Chinese Converter: Applies contextual word-level phrase dictionary matching (OpenCC standard) rather than simplistic single-character substitution.
Chinese Pinyin Translator
Chinese Pinyin Translator: Queries authoritative Mandarin phonetic dictionary and applies tonal diacritics.
Japanese Romaji & Kana Converter
Japanese Romaji & Kana Converter: Maps Hiragana codepoints (U+3040-U+309F) to Katakana (U+30A0-U+30FF) and Hepburn Romaji transcription.
Morse Code Translator & Audio Player
Morse Code Translator & Audio Player: Maps each alphanumeric character to standard dots and dashes (S = ... O = ---).
Braille Text Translator
Braille Text Translator: Maps ASCII letters to 6-dot Braille cell Unicode codepoints (U+2800 to U+283F).
Emoji Shortcode & Codepoint Converter
Emoji Shortcode & Codepoint Converter: Queries emoji shortcode dictionary and replaces tokens with multi-byte Unicode surrogate characters.
Unicode Glyph & Codepoint Inspector
Unicode Glyph & Codepoint Inspector: Extracts Unicode scalar value, UTF-8 byte encoding, UTF-16 surrogate values, and formal character name.
Reverse Text & Mirror Generator
Reverse Text & Mirror Generator: Reverses Unicode grapheme sequence while preserving multi-byte surrogate pairs.
Zalgo Glitch Text Generator
Zalgo Glitch Text Generator: Attaches random combining diacritical marks (U+0300 to U+036F) above, below, and through character stems.
Tool category
Chinese-English Spacing Beautifier
Chinese-English Spacing Beautifier: Applies Paranoid text spacing rules, inserting U+0020 between CJK ideographs and Western alphanumeric glyphs.
Full-width & Half-width Converter
Full-width & Half-width Converter: Subtracts 0xFEE0 offset from full-width Unicode characters (U+FF01 to U+FF5E) to map to standard ASCII.
Invisible Characters Cleaner
Invisible Characters Cleaner: Scans for non-printing Unicode codepoints (U+200B, U+200C, U+200D, U+FEFF, U+00AD) and strips them.
Mojibake & Encoding Repair Assistant
Mojibake & Encoding Repair Assistant: Interprets byte sequence through Latin-1 / Windows-1252 table and re-encodes as UTF-8.
Whitespace & Empty Lines Cleaner
Whitespace & Empty Lines Cleaner: Strips leading/trailing line whitespace and collapses multiple consecutive empty lines into a single blank line.
Tab to Space & Space to Tab Converter
Tab to Space & Space to Tab Converter: Replaces tab characters (\t) with corresponding number of space characters (U+0020).
Join Lines with Custom Delimiter
Join Lines with Custom Delimiter: Splits on line breaks and joins elements with the specified delimiter string.
Text Word Wrap & Line Formatter
Text Word Wrap & Line Formatter: Wraps lines on word boundaries without breaking mid-word when possible.
Paragraph Re-wrapper & De-hyphen
Paragraph Re-wrapper & De-hyphen: Removes soft hyphenation (`para- \n graph` -> `paragraph`) and joins broken lines into continuous paragraphs.
Text Justifier & Column Aligner
Text Justifier & Column Aligner: Distributes additional whitespace evenly between word boundaries to reach exact width.
Text Number Padding & Leading Zeros
Text Number Padding & Leading Zeros: Prepends pad characters until each string matches target minimum width.
Tool category
Markdown to Plain Text Stripper
Markdown to Plain Text Stripper: Parses Markdown tokens, strips header hashes, bold/italic markers, and extracts raw link text.
Markdown Realtime Previewer & HTML Exporter
Markdown Realtime Previewer & HTML Exporter: Parses CommonMark / GFM AST and renders sanitized HTML markup.
HTML to Markdown Converter
HTML to Markdown Converter: Parses HTML DOM elements, converting headings, emphasis tags, and anchor links into standard Markdown syntax.
LaTeX Math Equation Renderer
LaTeX Math Equation Renderer: Tokenizes LaTeX mathematical grammar and generates accessible MathML and SVG vector representations.
Subtitle Time Shift & Sync (SRT/VTT)
Subtitle Time Shift & Sync (SRT/VTT): Parses timecodes into milliseconds, adds 1,500ms offset, and reformats into standard SRT time strings.
SRT ↔ WebVTT Subtitle Converter
SRT ↔ WebVTT Subtitle Converter: Prepends `WEBVTT` header line, converts timestamp millisecond comma delimiters into periods.
Subtitle to Plain Dialogue Extractor
Subtitle to Plain Dialogue Extractor: Strips sequence index numbers, timestamp timecode lines, and blank lines, joining dialogue into running prose.
BBCode to Markdown & HTML
BBCode to Markdown & HTML: Maps BBCode tags ([b], [i], [url], [code], [img]) to corresponding Markdown elements.
CSV Delimiter Auto-Detector & Re-formatter
CSV Delimiter Auto-Detector & Re-formatter: Detects semicolon delimiter, parses quoted values, and reformats into RFC 4180 comma-separated syntax.
Tool category
Text PII Anonymizer & Redactor
Text PII Anonymizer & Redactor: Scans text for regular expression entity patterns (emails, phone numbers, IPv4 addresses, and proper names).
Text Word & Character Diff
Text Word & Character Diff: splits both texts while retaining whitespace tokens and lists tokens absent from the other side; it is not a Myers or aligned diff.
Zero-Width Steganography Tool
Zero-Width Steganography Tool: Converts binary bytes of the secret string into alternating zero-width spaces (U+200B) and zero-width non-joiners (U+200C).
Distinct Lines Extractor
Distinct Lines Extractor: Preserves first occurrence of each line while removing all subsequent duplicates.
Multi-Rule Find & Replace Matrix
Multi-Rule Find & Replace Matrix: Executes sequential or simultaneous regex token replacements according to defined rule table.
Shuffle Text Lines Randomly
Shuffle Text Lines Randomly: Applies Fisher-Yates shuffle algorithm with CSPRNG entropy from crypto.getRandomValues().
Email Address Extractor & Deduplicator
Email Address Extractor & Deduplicator: Applies RFC 5322 compliant email regex, normalizes to lowercase, and removes duplicate occurrences.
URL Link Extractor
URL Link Extractor: Extracts HTTP/HTTPS URLs matching RFC 3986 URI syntax rules.
IP Address Extractor
IP Address Extractor: Extracts valid IPv4 dotted-quad addresses and IPv6 colon-hexadecimal strings.
Tool category
Caesar Cipher & ROT13 Decoder
Caesar Cipher & ROT13 Decoder: Shifts each Latin alphabetical letter backward by 13 positions modulo 26.
Vigenère Cipher Tool
Vigenère Cipher Tool: Shifts each letter using repeating modular key values (L=11, E=4, M=12, O=14, N=13).
Base32 Encoder & Decoder
Encodes UTF-8 bytes using RFC 4648 Base32: A–Z and 2–7, with = padding. Decoding reverses the 5-bit groups.
Base58 Encoder & Decoder
Uses the Bitcoin Base58 alphabet without 0, O, I or l. Leading zero bytes are preserved as leading 1 characters.
Ascii85 / Base85 Converter
Uses ASCII85: four UTF-8 bytes become five base-85 digits. A full group of four zero bytes is abbreviated as z; partial final groups are shorter.
Quoted-Printable Decoder & Encoder
Encodes every UTF-8 byte as =hh hexadecimal and adds soft line breaks so encoded lines do not exceed 76 characters. Decoding restores bytes and removes soft line breaks.
