I want to use string similarity functions to find corrupted data in my database.
I came upon several of them:
- Jaro,
- Jaro-Winkler,
- Levenshtein,
- Euclidean and
- Q-gram,
I wanted to know what is the difference between them and in what situations they work best?