Polyleven Alternatives

Similar projects and alternatives to polyleven based on common topics and language

distlib

2 20 4.4 C polyleven VS distlib

Distance related functions (Damerau-Levenshtein, Jaro-Winkler , longest common substring & subsequence) implemented as SQLite run-time loadable extension. Any UTF-8 strings are supported.
SymSpell

16 3,037 5.8 C# polyleven VS SymSpell

SymSpell: 1 million times faster spelling correction & fuzzy search through Symmetric Delete spelling correction algorithm
InfluxDB

www.influxdata.com featured

Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
Java String Similarity

0 2,654 0.0 Java polyleven VS Java String Similarity

Implementation of various string similarity and distance algorithms: Levenshtein, Jaro-winkler, n-Gram, Q-Gram, Jaccard index, Longest Common Subsequence edit distance, cosine similarity ...
RapidFuzz

11 2,348 9.2 C++ polyleven VS RapidFuzz

Rapid fuzzy string matching in Python using various string metrics
lev

4 4 0.0 C polyleven VS lev

Levenshtein distance function as C Extension for Python 3 (by duranbe)
SaaSHub

www.saashub.com featured

SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better polyleven alternative or higher similarity.

Suggest an alternative to polyleven

polyleven reviews and mentions

Posts with mentions or reviews of polyleven. We have used some of these posts to build our list of alternatives and similar projects.

Spellcheck and Levenshtein distance
1 project | /r/learnmachinelearning | 15 Nov 2022

polyleven is the fastest Levenshtein distance library I've been able to find. It also has a threshold parameter which can be used to speed up the calculations. That being said, I've had a lot more success speeding up the processing of large text datasets by converting the words to a vector space (using e.g. word2vec) then calculating euclidean distance, which is much faster than calculating Levenshtein distance (assuming you are using vectorized operations). The fastest solution would probably be to use approximate nearest neighbor search (see for example the faiss library), but again you'll have to embed your words in a vector space and you'll need to decide if this is viable for your use case.

Stats

Basic polyleven repo stats

Mentions

Stars

Activity

10.0

Last Commit

over 1 year ago

fujimotos/polyleven is an open source project licensed under MIT License which is an OSI approved license.

The primary programming language of polyleven is C.

Popular Comparisons

polyleven

Polyleven Alternatives

Similar projects and alternatives to polyleven based on common topics and language

distlib

SymSpell

InfluxDB

Java String Similarity

RapidFuzz

lev

SaaSHub

polyleven reviews and mentions

Stats

Popular Comparisons