oliverguhr/spelling-correction-multilingual-base

🤗 Hugging Face sourcetext-generationmit248M params990 MBsafetensors✓ 7 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo oliverguhr/spelling-correction-multilingual-base ./model-folder
Needs a seeder →

This is an experimental model that should fix your typos and punctuation. If you like to run your own experiments or train for a different language, take a look at the code.

Model description

This is a proof of concept spelling correction model for English and German.

Intended uses & limitations

This project is work in progress, be aware that the model can produce artefacts. You can test the model using the pipeline-interface:

from transformers import pipeline

fix_spelling_pipeline = pipeline("text2text-generation",model="oliverguhr/spelling-correction-multilingual-base")
def fix_spelling(text, max_length = 256):
    return fix_spelling_pipeline("fix:"+text,max_length = max_length)

print(fix_spelling_pipeline("can we mix the languages können wir die sprachen mischen",max_length=2048))