Team Ai
Modelpublic

omoured/YOLOv10-Document-Layout-Analysis

sourceHugging Faceagpl-3.0updated 2y agoView on Hugging Face
100likes
Model Card

๐Ÿค— Live Demo here: https://huggingface.co/spaces/omoured/YOLOv10-Document-Layout-Analysis

<!-- ABOUT THE PROJECT -->

About ๐Ÿ“‹

The models were fine-tuned using 4xA100 GPUs on the Doclaynet-base dataset, which consists of 69103 training images, 6480 validation images, and 4994 test images.

<p align="center"> <img src="https://github.com/moured/YOLOv10-Document-Layout-Analysis/raw/main/images/samples.gif" height="320"/> </p>

Results ๐Ÿ“Š

ModelmAP50mAP50-95Model Weights
YOLOv10-x0.9240.740Download
YOLOv10-b0.9220.732Download
YOLOv10-l0.9210.732Download
YOLOv10-m0.9170.737Download
YOLOv10-s0.9050.713Download
YOLOv10-n0.8920.685Download

Codes ๐Ÿ”ฅ

Check out our Github repo for inference codes: https://github.com/moured/YOLOv10-Document-Layout-Analysis

References ๐Ÿ“

  1. 1.YOLOv10
BibTeX
@article{wang2024yolov10,
  title={YOLOv10: Real-Time End-to-End Object Detection},
  author={Wang, Ao and Chen, Hui and Liu, Lihao and Chen, Kai and Lin, Zijia and Han, Jungong and Ding, Guiguang},
  journal={arXiv preprint arXiv:2405.14458},
  year={2024}
}
  1. 1.DocLayNet
@article{doclaynet2022,
  title = {DocLayNet: A Large Human-Annotated Dataset for Document-Layout Analysis},  
  doi = {10.1145/3534678.353904},
  url = {https://arxiv.org/abs/2206.01062},
  author = {Pfitzmann, Birgit and Auer, Christoph and Dolfi, Michele and Nassar, Ahmed S and Staar, Peter W J},
  year = {2022}
}

Contact

LinkedIn: https://www.linkedin.com/in/omar-moured/