Skip to content

Model and approach research for Tibetan spell-checking model #399

Description

@tenzinyonten

Description:
We want to build a spell-checking model to correct errors in Tibetan text sourced from OCR outputs, PDFs, and epubs. Before implementation, we'll research existing models and approaches for Tibetan (or similar low-resource language) spell-checking and OCR error correction, weigh their tradeoffs, and recommend the best one to move forward with.

Sub Task

  • Research and compare existing models/approaches for Tibetan spell-checking and OCR error correction.
  • Write a blog post recommending a model/approach to move forward with.

Reviewer

Metadata

Metadata

Assignees

Labels

No labels
No labels

Type

No type

Fields

Priority

None yet

Projects

Status
Done

Milestone

No milestone

Relationships

None yet

Development

No branches or pull requests

Issue actions