You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

mBART50 Fine-tuned for Sinhala Spell Correction

Model Details

Research

This model was developed as part of our research on spell correction using pre-trained language models for low-resource languages.

Limitations

  • The model is specifically fine-tuned for Sinhala spell correction and may not generalize to other languages.
  • Performance may vary for informal text, social media content, code-mixed text, and text containing uncommon or domain-specific vocabulary.
  • The model may occasionally modify text that is already correctly spelled.

Citation

If you use this model in your research or applications, please cite our paper:

A. Gunathilake, N. Karunarathna, T. Bandaranayake, S. Ranathunga, N. de Silva and N. Jayatilleke, "LMSpell: Spell Correction with Pre-Trained Language Models," 2026 Moratuwa Engineering Research Conference (MERCon), Moratuwa, Sri Lanka, 2026, pp. 503-508, doi: 10.1109/MERCon71835.2026.11691371.

@INPROCEEDINGS{11691371,
  author={Gunathilake, Akesh and Karunarathna, Nadil and Bandaranayake, Tharusha and Ranathunga, Surangika and de Silva, Nisansa and Jayatilleke, Nevidu},
  booktitle={2026 Moratuwa Engineering Research Conference (MERCon)},
  title={LMSpell: Spell Correction with Pre-Trained Language Models},
  year={2026},
  pages={503-508},
  doi={10.1109/MERCon71835.2026.11691371}
}
Downloads last month
-
Safetensors
Model size
0.6B params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for lm-spell/mbart50-ft-ssc

Finetuned
(320)
this model

Dataset used to train lm-spell/mbart50-ft-ssc