Why OpenAI Models Struggle with PDF Extraction(and Why Gemini Fairs Much Better)medium.com 5 pointskapitalx1 year ago2 commentsSaveHideCopy link On HNComments−ban-lan-gen1yDoes this only apply to PDFs with many images? If it is mostly text and table, could it just extract plain text?−kapitalxOP1yThis applies to PDFs with lots of text. The smaller the text on the page, the more impacted it is.
Comments
Does this only apply to PDFs with many images? If it is mostly text and table, could it just extract plain text?
This applies to PDFs with lots of text. The smaller the text on the page, the more impacted it is.