AI Tools.

Search

object detection models

2 models · ranked by HuggingFace downloads

PP-DocLayoutV3_safetensors

by PaddlePaddle

PP-DocLayoutV3 is PaddleOCR's third-generation document layout detection model, converted to safetensors format for HuggingFace compatibility. It performs object detection to identify layout regions — text blocks, tables, figures, formulas, headings — in document images using a transformer-based backbone. The model is a building block in PaddleOCR's full document parsing pipeline.

847,545 ↓ · 39 ♡

table-transformer-detection

by microsoft

A DETR-based object detection model from Microsoft Research trained to locate tables in document images. It is the detection stage in a two-step pipeline — a separate structure recognition model then parses the detected table's rows and columns.

585,407 ↓ · 428 ♡