๐Ÿ“š General & Other

Matching Digits Is Not as Simple as 0-9

removes lines containing ASCII digits. Whether that covers the data depends on where the data came from.

grep -v '[0-9]' removes lines containing ASCII digits. Whether that covers the data depends on where the data came from.

Unicode has many digits

\d in a Unicode-aware engine matches Arabic-Indic digits, Devanagari digits and other numeral systems โ€” not just 0-9. Python's re with a string pattern does this by default; grep with a plain bracket range does not. On international content the two disagree, and the disagreement is silent.

Full-width digits look identical

Text from Japanese or Chinese sources may contain full-width forms that render like ordinary digits but occupy different code points. [0-9] misses them entirely. Normalising to NFKC folds them to ASCII first, which is usually what you want.

Digits in percent-encoding

Every percent-escape in a URL contains hex digits, so %20 in a path makes any URL with an encoded space match a digit filter. If the list is URLs, decode before filtering, or accept that encoded characters will pull rows out for reasons unrelated to the numbers you meant.

Try it: Remove Lines With Digits on SeoWolf's Notepad.