feat: add rapidfuzz, minting Text Processing > Fuzzy Matching

The industry's fuzzy string matching answer (web-verified: the
production recommendation over thefuzz — same API, MIT license, C++
speed — and preferred over textdistance for string metrics). 181.7M
downloads/month (pepy), 4.1K stars, pushed 2026-08. Sole obvious
choice; the subcategory label rides this commit so it is never empty,
completing the displacement of textdistance.

Co-Authored-By: Claude <noreply@anthropic.com>
This commit is contained in:
Vinta Chen
2026-08-16 15:09:32 +08:00
co-authored by Claude
parent 043cb6d377
commit db9c262342
+2
View File
@@ -857,6 +857,8 @@ _Libraries for parsing and manipulating plain texts._
- [charset-normalizer](https://github.com/jawah/charset_normalizer) - Universal character encoding detector, the default of the requests ecosystem.
- [chardet](https://github.com/chardet/chardet) - Python character encoding detector.
- [ftfy](https://github.com/rspeer/python-ftfy) - Makes Unicode text less broken and more consistent automagically.
- Fuzzy Matching
- [rapidfuzz](https://github.com/rapidfuzz/RapidFuzz) - Rapid fuzzy string matching using various string metrics, with a C++ core.
- General
- [difflib](https://docs.python.org/3/library/difflib.html) - (Python standard library) Helpers for computing deltas.
- [pyfiglet](https://github.com/pwaller/pyfiglet) - An implementation of figlet written in Python.