Загрузка видео...
Не удалось загрузить видео
GPT-4o is bad at processing PDF documents. Whoever tells you otherwise is not living in the real world. In 2024, people fill out forms using pen and paper. Try to answer questions from those forms using modern models, and you'll be disappointed. I recorded a video to show you... show more
196,847 просмотров • 1 год назад •via X (Twitter)
Комментарии: 10

I second this. I used to upload scanned docs to it, but better results came when I extracted text from scan using Google lense and send that to gpt as input.

That's the way.

Convert the PDF to an image, and it works as good as any other document you upload. Click Bait!

I'd love for this to work, but it doesn't. Feel free to try yourself.

I see you play modern day influencer role under the disguise of being machine learning engineer. Ah. I see.

An anonymous account accusing me of disguising my identity. Irony is dead.

AI still has a long way to go in mastering practical document processing.

It does. That's where we come in: instead of hoping for the model to work, we can make it easy to find the information it needs.

Had a go this morning at getting ChatGPT to scrape a basic web site. It was hopeless, seemingly if the HTML tags are not so descriptive that a child of 5 could scrape them it struggles

Anything you can do to preprocess the data and make it easier to consume, goes a long way.
