This project is a text-matching tool that uses the Levenshtein distance algorithm to find the similarity between sets of strings. Built with Python and Streamlit, this tool allows users to upload CSV files or manually input strings for comparison. Users can also fine-tune the matching criteria using automatic or manual sliders.
Learn more about Levenshtein distance on Wikipedia
Clone the repository:
git clone https://github.com/dvonpasecky/fuzzy-matcher.git
cd fuzzy-matcherInstall the required packages:
pip install -r requirements.txt- Choose "Upload CSV" from the sidebar.
- Upload your CSV file. The file should contain two columns of strings to be compared.
- Choose "Manual Input" from the sidebar.
- Manually enter strings in the columns that appear.
Choose either the "Automatic Slider" or the "Manual Slider" to adjust the matching criteria.
Toggle the "Case Sensitive" option on for case-sensitive matching.
Download the filtered results as a CSV file.
Pull requests are welcome. For major changes, please open an issue first to discuss what you would like to change.
You can access the hosted webapp here.
