Team Ai
Datasetpublic

armvectores/handwritten_text_detection

Handwritten text detection dataset Data domain The blanks were provided by youth organization "Armenian Club" (telegram, instagram ), Russia Moscow. The text on blanks was written during dictation "Teladrutyun" in 2018 The blanks were labeled by Amir and Renal during research project in HSE MIEM Dataset info Contains labeled dictations blanks in YOLO format 91 image in total, 73 (80%) for train and 18 (20%) for test No image alignment or any… See the full description on the dataset page: https://huggingface.co/datasets/armvectores/handwritten_text_detection.

sourceHugging Facemitupdated 2y agoView on Hugging Face
7likes316downloads
Dataset Card

Handwritten text detection dataset

Data domain

The blanks were provided by youth organization "Armenian Club" (telegram, instagram ), Russia Moscow.

The text on blanks was written during dictation "Teladrutyun" in 2018

The blanks were labeled by Amir and Renal during research project in HSE MIEM

Dataset info

Contains labeled dictations blanks in YOLO format

91 image in total, 73 (80%) for train and 18 (20%) for test

No image alignment or any preprocess

Resolution 1320x1020, 96 dpi

How to use

1) clone repo

git clone https://huggingface.co/datasets/armvectores/handwritten_text_detection
cd handwritten_text_detection

2) use data.yaml for training

from ultralytics import YOLO

model = YOLO('yolov8n.pt')
model.train(data='data.yaml', epochs=20)

Data sample

<img src="blank_sample.png" width="700" />