flexudy/t5-base-multi-sentence-doctor
Flexudy/t5-base-multi-sentence-doctor is machine learning model.
About flexudy/t5-base-multi-sentence-doctor
The current version of the model was only trained on 150K sentences from the tatoeba dataset . The datasets are available in the data folder (where sentence_doctor_dataset_300K is a larger dataset with 100K sentences for each language). We might release a version trained on more data . The model works on English, German and French text. The model is based on sentences that were extracted with OCR software or text extractors. It attempts to reconstruct sentences based on the its context (sourrounding text). The task is pretty straightforward: reconstruct the "intended" sentence, and its context, reconstruct the sentence . We are telling the model to repair the,