A Corpus of 21st Century Scots Texts

Intro a b c d e f g h i j k l m n o p q r s t u v w x y z Texts Writers Statistics Top200 Search Compare

Levenshtein Distance

Enter a word to find nearest neighbouring words, for example ahint

- basic concord - pre-sorted concord - post-sorted concord - map and chronology - chronogrid - fine-grain concord -

Similar words to hærsts in Corpus

Levenshtein Double Levenshtein SoundEx MetaPhone Manually curated
hærsts (0) - 1 freq
hærst (1) - 1 freq
hæst (2) - 1 freq
hairsts (2) - 7 freq
hærth (2) - 1 freq
horsis (3) - 3 freq
hoasts (3) - 8 freq
firsts (3) - 1 freq
horss (3) - 22 freq
herst (3) - 6 freq
hurts (3) - 17 freq
heests (3) - 2 freq
heysts (3) - 1 freq
haerts (3) - 19 freq
haarst (3) - 1 freq
hirsty (3) - 2 freq
æst (3) - 1 freq
hoarses (3) - 5 freq
kïsts (3) - 1 freq
dær's (3) - 1 freq
tærs (3) - 1 freq
ært (3) - 3 freq
herts (3) - 119 freq
bursts (3) - 19 freq
mæst (3) - 1 freq
hærsts (0) - 1 freq
hærst (2) - 1 freq
hairsts (4) - 7 freq
hærth (4) - 1 freq
hæst (4) - 1 freq
fært (6) - 1 freq
hairst's (6) - 1 freq
horses (6) - 113 freq
hearts (6) - 32 freq
bursts (6) - 19 freq
fæsis (6) - 1 freq
mæst (6) - 1 freq
høst (6) - 3 freq
hairst (6) - 188 freq
hirst (6) - 2 freq
huirst (6) - 1 freq
bærn's (6) - 1 freq
hysts (6) - 1 freq
herts (6) - 119 freq
he'rts (6) - 4 freq
hosts (6) - 13 freq
hïts (6) - 1 freq
hairts (6) - 62 freq
herst (6) - 6 freq
hurts (6) - 17 freq
SoundEx code - H623
hairst - 188 freq
horsed - 3 freq
harkit - 10 freq
hairstyle - 2 freq
haircut - 22 freq
herst - 6 freq
hairstin - 14 freq
hairsts - 7 freq
hirstlin - 2 freq
huirst - 1 freq
hirst - 2 freq
harrassed - 2 freq
hairsted - 1 freq
horse-tradin - 1 freq
harrased - 1 freq
hairstit - 4 freq
hairset - 3 freq
hairst-rig - 1 freq
horchata - 2 freq
horchatería - 1 freq
hirsty - 2 freq
horse-drawn - 1 freq
hærst - 1 freq
hærsts - 1 freq
hairst-blinks - 1 freq
hairst's - 1 freq
hairstless - 1 freq
haircuts - 2 freq
haarst - 1 freq
hairst-time - 1 freq
harestanes - 1 freq
hairst-taest - 1 freq
hairst-gowd - 1 freq
hairster - 1 freq
hairstpark - 1 freq
hairsters - 4 freq
hairst-moose - 1 freq
herkit - 1 freq
harassit - 1 freq
harrowgate - 1 freq
€˜hairst - 1 freq
hirsute - 1 freq
hairstyles - 2 freq
harassed - 1 freq
heresdaibhi - 19 freq
hoorsaday - 1 freq
hoursdrive - 1 freq
hryqcdn - 1 freq
harighotra - 1 freq
harrisdistiller - 1 freq
MetaPhone code - RSTS
roosts - 1 freq
rests - 11 freq
restis - 1 freq
wrists - 10 freq
rest's - 1 freq
ruists - 1 freq
recedes - 1 freq
resets - 1 freq
resides - 1 freq
hærsts - 1 freq
rusts - 1 freq
rusty's - 1 freq
roasts - 1 freq
reists - 1 freq
roasties - 1 freq
HÆRSTS
Time to execute Levenshtein function - 0.314568 milliseconds
The Levenshtein distance is the number of characters you have to replace, insert or delete to transform one word into another, its useful for detecting typos and alternative spellings
Time to execute Double Levenshtein function - 0.689304 milliseconds
In a stroke of genius, this runs the Levenshtein function twice, once without vowels and adds the distance together, giving double weight to consonants.
Time to execute SoundEx function - 0.030134 milliseconds
Soundex is a phonetic algorithm for indexing names by sound, as pronounced in English. The goal is for homophones to be encoded to the same representation so that they can be matched despite minor differences in spelling.
Time to execute MetaPhone function - 0.073534 milliseconds
Metaphone is a phonetic algorithm, published by Lawrence Philips in 1990, for indexing words by their English pronunciation.[1] It fundamentally improves on the Soundex algorithm by using information about variations and inconsistencies in English spelling and pronunciation to produce a more accurate encoding, which does a better job of matching words and names which sound similar.
Time to execute Manually curated function - 0.000864 milliseconds
Manual Curation uses a lookup table / lexicon which has been created by hand which links words to their lemmas, and includes obvious typos and spelling variations. Not all words are covered.