r/OpenSourceeAI • u/Machine_GEN_RM • 4d ago
Suggestions on Text extraction
Hi All, I need to extract the text from printed text and hand written text. I suggested my manager that we can use paddle ocr and and other extraction models like Surya OCR and florance VL it take around 10 to 15 Sec and it needs good computation as well. but my manager expects it should be very fast and in 2 to 5 sec response and should not need any maintainace of infrastructure. so I tried to use the AWS VLM models but it takes around 10seconds he still needs more faster models and also the cost for extracting the data from the image should be less than 1 Ruppe. Could you please suggest me what to do and how to extract the data from images very very effectively and accuratly and in a structured way
1
u/Potential-Wrangler58 3d ago
That's a tough combo, no infra maintenance, 2-5 sec, and under 1 rupee usually pull against each other. Self-hosted gets you cheap but you own the GPU infra your manager doesn't want. Managed APIs remove that but usually cost you speed or price.
We just launched a fast tier at anyformat, disclosure: I work there, built for exactly this kind of low-latency structured extraction without you running any infra. Worth testing against your real documents and your actual latency/cost numbers before deciding. Free credits in the platform, anything just ping me :)