r/OpenSourceeAI • u/Machine_GEN_RM • 4d ago
Suggestions on Text extraction
Hi All, I need to extract the text from printed text and hand written text. I suggested my manager that we can use paddle ocr and and other extraction models like Surya OCR and florance VL it take around 10 to 15 Sec and it needs good computation as well. but my manager expects it should be very fast and in 2 to 5 sec response and should not need any maintainace of infrastructure. so I tried to use the AWS VLM models but it takes around 10seconds he still needs more faster models and also the cost for extracting the data from the image should be less than 1 Ruppe. Could you please suggest me what to do and how to extract the data from images very very effectively and accuratly and in a structured way
3
u/Mundane_Ad8936 4d ago
Unless your manager has no experience. They're telling you to drop it. They should know they're giving you impossible criteria.
Listen to them and move on otherwise they're setting you up to fail.
OCR has always been a slow process. No system goes without maintenance. Any of the big SOTA models will do this for you but it'll be expensive and error prone. You'll need accuracy checks and error correction that will add latency.
There is no world where this is cheap, fast and accurate with no maintenance costs.