Glossary › ocr-free
GLOSSARY
ocr-free
appears in 1 paper titles
Definition
Describes document models that map pixels straight to an answer or a structured record, with no separate text-recognition stage in front. The appeal is that OCR errors cannot propagate and that you avoid maintaining per-language recognition engines; Donut and Pix2Struct are the standard examples. The cost is resolution — reading small print takes many image tokens — and a much larger appetite for training data.