Hapus karakter khusus, emoji, simbol, angka, aksen, dan sampah tak terlihat dalam satu alat browser pribadi — tanpa merusak skrip non-Latin.
Preset
Khusus & simbol
Huruf & angka
Tak terlihat & terkendali
Kutipan & Unicode
Aturan Anda sendiri
Urutan pemrosesan sudah diperbaiki: normalisasi → tidak terlihat → kontrol → tanda kutip → aksen → emoji → kelas karakter → aturan khusus → ASCII → rapi.
Mengapa alat ini berbeda
Penghapusan aksen yang memahami aksara — bahasa Hindi, Arab, Ibrani, dan Thailand tetap mempertahankan cirinya dan tidak terkoyak
Emoji muncul sebagai satu kesatuan, sehingga keluarga, bendera, dan warna kulit tidak meninggalkan sisa yang tidak terlihat
Simbol dan tanda baca adalah saklar yang terpisah dan bukan satu wadah “karakter khusus” yang tidak jelas
Pertahankan atau hapus karakter Anda sendiri, sehingga tanda hubung atau garis bawah dapat bertahan jika dibersihkan
Keempat bentuk normalisasi Unicode dijelaskan dalam bahasa yang sederhana
Pemrosesan dalam browser pribadi dalam 16 bahasa
Cara menghapus karakter khusus
Hapus karakter khusus, hapus tanda baca dan simbol secara bersamaan, pertahankan huruf, angka, dan spasi. Jika Anda hanya menginginkan salah satunya, gunakan sakelar Hapus tanda baca dan Hapus simbol yang terpisah.
Hapus karakter non-alfanumerik
Hapus penyimpanan non-alfanumerik hanya huruf dan angka dari bahasa apa pun. Matikan Pertahankan spasi untuk menciutkan semuanya menjadi satu string yang tidak terputus, yang berguna untuk ID dan kode.
Hapus angka dari teks
Hapus angka demi angka di setiap skrip, termasuk angka Arab-India dan Devanagari, bukan hanya 0 hingga 9.
Hapus huruf dari teks
Hapus huruf, hapus huruf dari setiap skrip, sisakan angka, tanda baca, dan spasi. Pasangkan dengan Hapus tanda baca untuk mengeluarkan angka bersih dari pasta yang berantakan.
Hapus emoji dari teks
Hapus emoji menghapus setiap emoji sebagai satu kesatuan yang lengkap. Emoji keluarga, bendera negara, atau lambaikan tangan dengan warna kulit adalah satu kelompok, jadi tidak ada yang tertinggal untuk merusak teks Anda nanti.
Hapus simbol dari teks
Hapus simbol yang menargetkan karakter matematika, mata uang, dan tanda seperti plus, sama dengan, dolar, dan euro. Tanda baca seperti koma dan titik tetap ada kecuali Anda juga menghapus tanda baca.
Hapus aksen dari teks
Hapus aksen mengubah kafe menjadi kafe dan Ελλάδα menjadi Ελλαδα. Sengaja melewatkan Devanagari, Arab, Ibrani, Thai, dan Cyrillic, karena pada aksara tersebut tandanya merupakan bagian dari huruf, bukan hiasan.
Konversikan kutipan cerdas
Kutipan pintar menjadi tanda kutip lurus menggantikan tanda kutip keriting dan apostrof dengan tanda kutip biasa, seperti yang diharapkan oleh kode, file CSV, dan sistem lama. Kutipan langsung ke pintar melakukan kebalikannya untuk penerbitan. Tanda hubung dan elipsis ke ASCII juga meratakan tanda hubung en, tanda hubung em, dan elipsis karakter tunggal.
Hapus karakter yang tidak terlihat
Mode aman menghilangkan spasi lebar nol dan tanda urutan byte. Mode agresif juga menghilangkan joiner, yang dibutuhkan oleh beberapa teks Arab dan India, jadi gunakanlah hanya jika Anda tahu teks Anda tidak memerlukannya. Menghapus tanda arah akan menghapus penggantian dua arah dan tanda hubung lunak.
Normalisasikan teks Unicode
NFC menyusun huruf dan tanda menjadi karakter tunggal dan merupakan pilihan paling aman untuk penyimpanan dan pencarian. NFD memisahkan mereka. NFKC dan NFKD melangkah lebih jauh dan melipat kemiripan, sehingga font, pengikat, dan angka yang dilingkari menjadi teks biasa.
Hapus karakter kontrol
Hapus karakter kontrol, hapus byte nol dan kode kontrol tak kasat mata lainnya yang merusak database dan ekspor. Pertahankan tab dan jeda baris diaktifkan secara default sehingga tata letak Anda tetap bertahan.
Some text problems are not about how many spaces you have — they are about which characters survived the copy. Microsoft Word inserts curly quotes that break JSON. Design tools export em dashes and ellipsis characters plain-text editors mishandle. Emoji render as boxes in legacy systems. OCR output scatters section signs and soft hyphens through otherwise readable sentences. Zero-width joiners from RTL paste break search indexes silently.
The WriteWithin Text Character Cleaner targets character classes, Unicode normalization, and custom keep/remove rules — all in your browser, with no upload. Use it when the Whitespace Cleaner has already fixed spacing but symbols, emoji, or invisible control codes still cause trouble.
What this tool removes and keeps
Special, non-alphanumeric, numbers, and letters
Remove special characters clears punctuation and symbols together while keeping letters, numbers, and spaces. For finer control, toggle Remove punctuation and Remove symbols separately — commas and full stops are punctuation; plus, equals, and currency signs are symbols.
Remove non-alphanumeric (keep letters and numbers only) strips everything else from any script. Turn off Keep spaces to collapse the result into one unbroken string — handy for IDs, slugs, and matching keys. Remove numbers drops digits in every numeral system, not just 0–9. Remove letters strips letters from every script, leaving digits and punctuation if those options stay off.
Emoji clusters and symbols
Remove emojis deletes each emoji as a complete unit. Family emoji, country flags, and skin-tone modifiers include zero-width joiners internally; the cleaner removes the whole cluster so no invisible leftovers break your text later. Remove symbols targets math, currency, and sign characters such as plus, equals, dollar, and euro while leaving punctuation unless you remove that too.
Accents and script-aware behavior
Remove accents turns café into cafe and Ελλάδα into Ελλαδα using decomposition that understands which scripts treat marks as decoration. It deliberately skips Devanagari, Arabic, Hebrew, Thai, and Cyrillic, where marks are part of the letter — blind stripping would corrupt meaning. This script-aware approach is something generic “remove diacritics” tools often get wrong.
Smart quotes, dashes, and typographic punctuation
Smart quotes to straight replaces curly quotes and apostrophes with plain ASCII — what code, CSV, and older databases expect. Straight quotes to smart does the reverse for publishing. Straighten dashes converts en dashes, em dashes, and the single-character ellipsis to ASCII hyphen and three dots.
Invisible and control characters
Safe invisible mode removes zero-width spaces and byte-order marks. Aggressive mode also strips joiners Arabic and Indic text may need — use only when you know your content is unaffected. Remove direction marks clears bidirectional overrides and related controls that scramble mixed RTL/LTR paste. Remove control characters drops null bytes and other invisible codes that break databases; Keep tabs and line breaks stays on by default so layout survives.
Unicode normalization: NFC, NFD, NFKC, NFKD
NFC composes letters and combining marks into single characters — the safest default for storage and search. NFD splits them apart. NFKC and NFKD go further, folding compatibility characters so fancy fonts, ligatures, and circled numbers become plain text. Choose NFKC when importing user-generated content into strict ASCII systems; choose NFC for general multilingual storage on WriteWithin.
Custom keep and remove characters
Keep characters lists symbols that must survive a aggressive cleanup — hyphens, underscores, dots for filenames. Remove characters lists exact code points or literals to strip regardless of other toggles. Pair Keep alnum only with Keep characters -_. for filename-safe output without guessing which symbol class each mark belongs to.
Presets for common jobs
Plain text — NFC, safe invisible removal, straight quotes, straight dashes, control cleanup, tidy spaces
The Whitespace Cleaner collapses repeated spaces, replaces NBSP, converts tabs, fixes PDF line breaks, and optionally removes blank lines. It does not remove emoji, punctuation classes, or smart quotes. When Word paste shows “correct” spacing but still fails a JSON parser, the problem is usually curly quotes or em dashes — character-level, not space-level.
Character Cleaner includes Tidy spaces for light collapse after stripping symbols, but for heavy spacing work — PDF joins, NBSP normalization, show-invisible preview — use Whitespace Cleaner first. Typical chain: Whitespace Cleaner → Character Cleaner → Line Tools for list dedupe.
Add Keep characters -_. if underscores or dots must remain.
Turn on ASCII only when the destination filesystem requires it.
Enable Tidy spaces to collapse gaps left after symbol removal.
Copy and rename files or create slug fields in your CMS.
Common mistakes to avoid
Removing accents on Arabic or Hindi text — the tool skips meaningful scripts, but aggressive invisible mode can still harm joiners; prefer safe invisible on multilingual paste.
Using Remove special when you only mean emoji — toggle Remove emojis alone to keep punctuation.
Expecting PDF line repair here — hard line breaks are a Whitespace Cleaner job.
NFKC on display text you want to look pretty — NFKC folds stylistic variants; use NFC for human-facing copy.
Forgetting Keep tabs and line breaks — turning it off flattens structured logs; disable only when you truly need one line.
Benefits of script-aware character cleaning
Emoji removed as whole clusters — no invisible joiner leftovers
Accent stripping respects scripts where marks carry meaning
Separate switches for symbols, punctuation, numbers, and letters
All four Unicode normalization forms with plain-language presets
Private in-browser processing for sensitive drafts and client data
Real-world use cases
JSON and API payloads: straight quotes, no control characters, NFC normalization. E-commerce SKU cleanup: letters and numbers only, symbols off. Social post prep for SMS: emoji off, accents optional. OCR and PDF paste: combine with Whitespace Cleaner, then strip section signs and odd symbols here. Filename batches: ASCII safe preset with custom keep characters for dots and hyphens in version numbers.
Common questions
Will removing accents break Hindi, Arabic, or Thai? No. Accent removal targets Latin and Greek decoration marks. Devanagari, Arabic, Hebrew, Thai, and Cyrillic marks are left intact.
What is the difference between symbols and special characters? Symbols are math, currency, and signs like + = $ €. Punctuation is , . ! ? and quotes. Remove special clears both; you can also toggle each class.
Does emoji removal leave broken leftovers? No. Clusters including skin tones, flags, and families remove as one unit.
Which Unicode form should I use? NFC for general storage; NFKC when folding lookalikes and compatibility characters for strict systems.
Is my text uploaded? No. Character cleaning runs entirely in your browser.
Pertanyaan yang sering diajukan
Apakah menghilangkan aksen akan merusak teks Hindi, Arab, atau Thailand?
Tidak. Penghapusan aksen hanya menargetkan bahasa Latin dan Yunani, yang mana tandanya adalah hiasan. Pada Devanagari, Arab, Ibrani, Thailand, dan Cyrillic tandanya merupakan bagian dari huruf, jadi dibiarkan saja.
Apa perbedaan antara simbol dan karakter khusus?
Simbolnya adalah matematika, mata uang, dan karakter tanda seperti + = $€. Tanda baca adalah tanda seperti , . ! ? dan kutipan. Remove special characters clears both at once, and you can also toggle each one separately.
Apakah penghapusan emoji meninggalkan sisa yang rusak?
Tidak. Emoji gabungan seperti keluarga, bendera, dan warna kulit dihapus sebagai satu kesatuan, sehingga tidak ada penggabung lebar nol atau pemilih variasi yang tertinggal.
Formulir normalisasi Unicode mana yang harus saya gunakan?
NFC adalah default aman untuk menyimpan dan membandingkan teks. NFD memisahkan huruf dari tandanya. NFKC dan NFKD juga melipat kemiripan, mengubah font mewah, pengikat, dan angka yang dilingkari menjadi karakter biasa.
Apakah teks saya diunggah?
Tidak. Pembersihan karakter sepenuhnya dilakukan di browser Anda. Teks tidak pernah meninggalkan perangkat Anda.