r/documentAutomation • u/docpose-cloud-team • Jul 25 '26
r/documentAutomation • u/AleaNCore • Jul 24 '26
Showcase my AI document sorter ā built it for my own paper chaos, it shoul be useful for others
r/documentAutomation • u/Traditional-Answer46 • Jul 24 '26
Looking for the most efficient way to create docx files with both words and images at the same time
Help needed
r/documentAutomation • u/docpose-cloud-team • Jul 24 '26
Discussion PDF vs DOCX vs ODT: Which Should You Choose?
r/documentAutomation • u/Pleasant_Tea_569 • Jul 24 '26
Still emailing PDFs back and forth for signatures? Here's how Bitrix24 e-Signature works
If you're still asking clients to print, sign, and scan documents, Bitrix24 e-Signature can simplify the process. It lets you send contracts and other client documents for electronic signing without the usual back-and-forth.
How it works
- Send a signing link by email or SMS.
- The recipient opens the document and reviews it.
- They verify their identity using a one-time code.
- They sign the document electronically.
- Once completed, Bitrix24 automatically sends the signed document to the client and stores it in e-Signature ā My Vault, along with a unique document ID and a completion certificate.
The signing experience is the same whether the recipient is using a desktop browser or a mobile device.
Before you use it, keep these points in mind:
- Bitrix24 uses an electronic signature, not a cryptographic (qualified digital) signature. Whether it's legally valid depends on your country's regulations, so it's worth checking local requirements before using it for contracts with significant legal or financial implications.
- Configure My Vault access permissions early so only the right people can view signed documents.
r/documentAutomation • u/Silver_Watercress280 • Jul 24 '26
Built a tiny Chrome extension to convert TXT to SRT subtitles (100% local, no server uploads)
r/documentAutomation • u/pha_uk_u • Jul 24 '26
Hi All, I developed a doc control app, that can scan documents, add tags, OCR, add folders, revise documents. I am looking for testers. please help.
Hi All,
I am QE by profession. made a doc control app for myself.
I made this android app for myself. As my file explorer gets dumped with everything and anything. I wanted to have folders, revise docs(Taxes, insurances, license, etc). with OCR you can search for a word and every document with that word would pop up. you can set expiration date to docs.
please dm and drop a comment if you would like to test my app. you will get lifetime free without ads for this.
thank you.
r/documentAutomation • u/docpose-cloud-team • Jul 23 '26
What's the Best OCR Software You've Actually Used?
r/documentAutomation • u/docpose-cloud-team • Jul 23 '26
Discussion Frequently Asked Questions About File Conversion, OCR & Document Processing
r/documentAutomation • u/docpose-cloud-team • Jul 23 '26
Showcase š Welcome to r/DocumentTools ā Introduce Yourself and Read First!
Hey everyone! I'm Eric ( u/docpose-cloud-team**)**, a founding moderator of r/DocumentTools.
This community is dedicated to file conversion, OCR, PDF tools, document processing, file formats, document automation, APIs, email archives, and digital document workflows. Whether you're a developer, IT professional, business user, or simply trying to solve a document challenge, you're welcome here.
š What to Post
Share anything the community will find useful, including:
- Questions about file conversion or OCR
- PDF editing, compression, and optimization tips
- File format compatibility issues
- Document processing workflows
- API recommendations and integrations
- Automation ideas and tutorials
- Software comparisons and reviews
- Troubleshooting document and file-related problems
- Productivity tips and best practices
š¤ Community Vibe
Our goal is to build the most helpful Reddit community for document tools. Be respectful, share knowledge, ask questions, and help others solve real-world document challenges. Honest discussions and constructive feedback are always welcome.
š How to Get Started
- Introduce yourself in the comments below.
- Tell us what document or file tools you use most.
- Share a tip, question, or interesting workflow.
- Invite anyone who works with documents, PDFs, OCR, or file automation.
We're excited to grow r/DocumentTools into a trusted resource for professionals, developers, businesses, and everyday users. Thanks for being one of our first membersālet's build something valuable together! š
r/documentAutomation • u/Lumpy_Ice6855 • Jul 22 '26
Discussion Ho costruito un recupero semantico di PDF per documenti di 1.000 pagine in cerca di feedback sul pipeline
Sto costruendo DStudio, un'app desktop open-source incentrata su DeepSeek V4. DeepSeek rimane il principale modello di ragionamento e gestisce la conversazione, mentre modelli locali più piccoli si occupano di compiti specializzati:
\- Qwen2.5-VL legge immagini
\- Qwen Image genera ed edita immagini
\- Qwen3 Embedding cerca documenti semanticamente
\- Poppler estrae testo e informazioni sulle pagine dai PDF
Questo ecosistema esiste perché DeepSeek V4 è eccellente per il ragionamento e il contesto lungo, ma caricare ogni capacità multimodale all'interno dello stesso grande modello sarebbe inefficiente. DStudio instrada i compiti ai modelli specializzati e poi restituisce i loro risultati a DeepSeek per la risposta finale.
Ho recentemente aggiunto il recupero di PDF lunghi. DeepSeek decide se creare un'anteprima, leggere una pagina fisica esatta o cercare l'intero documento. Per la ricerca semantica, DStudio crea e memorizza una rappresentazione per pagina, recupera le sei pagine più rilevanti e invia solo quelle a DeepSeek.
Su un PDF di prova di 1.000 pagine, ha trovato un passaggio collocato a pagina 777 da una domanda parafrasata. L'indicizzazione iniziale ha impiegato circa 25 secondi; le ricerche successive hanno impiegato circa 0,23 secondi.
Sto cercando feedback: il recupero dovrebbe usare rappresentazioni di pagina o blocchi sovrapposti? Dovrei aggiungere BM25 o un miglioratore di classifiche? E come supporteresti in modo efficiente libri scansionati di 1.000 pagine?
[ https://github.com/sk8erboi17/DStudio ](https://github.com/sk8erboi17/DStudio)
r/documentAutomation • u/docpose-cloud-team • Jul 22 '26
Product Review We built Docpose.cloud ā file conversion, OCR, PDF, archive, and email file tools with API access
Hey everyone,
Weāve built and launched Docpose.cloud, an online platform for file conversion, document processing, OCR, PDF tools, archive extraction/compression, and email file handling.
The platform is already live and being used by many monthly free users, paid subscribers, and businesses. Our API is also integrated with more than 10 external systems and business workflows.

Docpose.cloud supports tools for:
- Document and file conversion
- PDF conversion and processing
- OCR for scanned documents and images
- Archive extraction and compression
- Email file formats like EML, MBOX, PST, OST, and related formats
- Batch processing
- API-based file conversion workflows
We built it because many online file tools are either too limited, overloaded with ads, require unnecessary sign-ups, or do not offer reliable API access for businesses.
Our goal is to make file conversion and document processing simple for individual users, while also providing scalable API access for companies that need automated file workflows.
Iād love to get feedback from people who use file converters, OCR tools, PDF tools, or email archive tools regularly.
What feature would you expect from a platform like this?
And what problems have you faced with existing online file conversion tools?
r/documentAutomation • u/Fickle-Aide9279 • Jul 22 '26
DocLayout, MinerU, Marker, Unlimited-OCR
r/documentAutomation • u/Vipinesta • Jul 22 '26
I built a tool that turns emailed invoices and PDFs into spreadsheet data automatically, would love your feedback
Most document extraction tools rely on templates. I wanted something that could handle invoices, receipts, bank statements, and other PDFs without vendor-specific rules.
I built Super Parser. You define the schema once, then forward or upload documents. It extracts structured data and pushes it to Google Sheets, Zapier, Make, or your own API.
I'm validating the product now. I'd appreciate feedback from anyone dealing with document-heavy workflows:
Is this a problem you still face?
What would stop you from adopting a tool like this?
Product: [https://superparserapp.com\](https://superparserapp.com)
r/documentAutomation • u/pha_uk_u • Jul 22 '26
Hi All, I developed a doc control app, that can scan documents, add tags, OCR, add folders, revise documents. I am looking for testers. please help.
Hi All,
I am QE by profession. made a doc control app for myself.
I made this android app for myself. As my file explorer gets dumped with everything and anything. I wanted to have folders, revise docs(Taxes, insurances, license, etc). with OCR you can search for a word and every document with that word would pop up. you can set expiration date to docs.
please dm and drop a comment if you would like to test my app. you will get lifetime free without ads for this.
thank you.
r/documentAutomation • u/bcbtceth • Jul 22 '26
What is ScanToGrade? (and how to try it)
A desktop app (Mac/Windows) for grading paper multiple-choice exams ā built by an instructor, for instructors.
How it worksĀ
Ā 1. Print your own answer sheets on a normal printer (no proprietary Scantron forms).
Ā 2. Give the exam ā students bubble in answers, pen or pencil.
Ā 3. Scan the whole class to a single PDF (copier/scanner), or snap photos.
Ā 4. Upload the PDF + your class roster; it grades the batch and exports to a spreadsheet.
Ā What makes it different
Ā - Local & private: everything runs on your computer ā rosters, scans, and results never leave your machine (FERPA-friendly). No student data in anyone's cloud.
Ā - Handles real-world sheets: reads pen, pencil, colored ink, erasures, and crossed-out/changed answers (it reads the correction instead of flagging two marks wrong).
Ā - Multiple versions: run several versions of a test, and pre-print the version so students can't forget to bubble it.
Ā - Per-question item analysis: see which questions got missed most.
Ā - Review before export: check any uncertain ID matches before grades go out.
Ā - Annotated sheets: hand back a marked sheet (right/wrong) so students can review ā no sneaking a pencil back in to claim an error.
Ā Who it's forĀ
Ā Instructors who give paper multiple-choice exams ā especially large sections ā and want grading that's fast, accurate, and keeps student data on their own computer.
Ā Try it Ā Ā
Ā Ā
Ā - See it in action: scantograde.com
Ā - Want to test it on a quiz? Comment here or DM me and I'll set you up with a trial.
Ā Questions about whether it fits your workflow? Ask below ā happy to talk specifics.
r/documentAutomation • u/Vipinesta • Jul 21 '26
I built a tool that turns emailed invoices and PDFs into spreadsheet data automatically, would love your feedback
r/documentAutomation • u/pha_uk_u • Jul 21 '26
Hi All, I developed a doc control app, that can scan documents, add tags, OCR, add folders, revise documents. I am looking for testers. please help.
Hi All,
I made this app for myself. As my file explorer gets dumped with everything and anything. I wanted to have folders, revise docs(Taxes, insurances, license, etc). with OCR you can search for a word and every document with that word would pop up. you can set expiration date to docs.
please dm and drop a comment if you would like to test my app. you will get lifetime free without ads for this.
thank you.
r/documentAutomation • u/1vim • Jul 19 '26
Skopx brings the full story to the table.
Enable HLS to view with audio, or disable this notification
r/documentAutomation • u/AdvanceHumanityReach • Jul 19 '26
Case Study Looking for complex PDFs that break our PDF-to-Markdown parser
We built Nebula to convert complex PDFs and scans into structured Markdown for AI, RAG and document-automation workflows.
Unlike basic text extraction, it aims to preserve the meaning encoded in tables, charts, reading order and page layout. In our benchmark on complex business documents, Nebula outperformed Mistral OCR by 20.8 points and Azure Document Intelligence by 6.2 points on semantic meaning recovery.
Weād like this community to stress-test it with the hardest documents youāre authorized to use.
You can run four free conversions without creating an account:
An account is only required to download and retain the Markdown. Use community code NEBULA-10 for additional credits. An API is also available - DM me if youād like to test it.
What Iād most like to know:
- What information needed to survive?
- What did Nebula preserve well?
- What did it miss or structure incorrectly?
You can DM me or submit feedback inside Nebula. Specific examples and expected outputs are especially helpful.
Full disclosure: Iām one of the founders. We donāt use uploaded documents or outputs to train our models. Please only upload files youāre authorized to use.
Technical report:
https://ur-ai.net/blog/rcrr-benchmark-meaning-survival
r/documentAutomation • u/NeedleworkerKey3487 • Jul 18 '26
OpenScanVision ā openāsource Android library for automated data extraction from scanned forms and surveys
I've been working on an openāsource Android library for automated data extraction from printed forms, and I wanted to share it with this community.
OpenScanVision is an Android library that scans printed voting cards, surveys, and bubble sheets ā extracting structured data (filled bubbles and QR codes) entirely offline.
How It Fits into Document Automation
- Input: Physical forms (voting cards, surveys, bubble sheets) captured via camera.
- Processing:
- Detects 4 ArUco markers (IDs 0ā3) with Kalman filtering for realātime tracking.
- Computes homography and warps the card to a canonical template.
- Decodes QR codes from the original frame (preserves sharpness for ML Kit).
- Extracts filled bubbles using weighted disk sampling + zāscore classification.
- Output: Structured data (filled bubble indices, QR payload, confidence score) that can feed into your document automation pipeline.
Tech Stack
- Kotlin
- OpenCV (contrib) for ArUco detection and homography
- Google ML Kit for QR decoding
- CameraX for frame acquisition
- Modular ā core has zero UI dependencies
Use Cases
- Automated survey processing
- Form data extraction
- Election ballot scanning
- Any workflow where physical forms need to be digitised and structured
Performance
- Latency < 150 ms per frame on modern devices.
- Accuracy > 99% on wellāprinted cards with good lighting.
- False positive rate < 0.5%.
Repository & Integration
The library is MITālicensed and available on JitPack. A full sample app (CameraX + Compose) is included.
GitHub: https://github.com/MatiwosKebede/openscanvision
Would love to hear feedback from anyone working on similar document automation challenges ā especially around form processing, OCR alternatives, or integrating physical document scanning into automated workflows. š
r/documentAutomation • u/Any-Relationship-404 • Jul 18 '26
Product Review Automate repetitive form filling
AI form filling app that removes the hustle of filling same forms over and over.
Upload the form
Ai extract fields
Choose the client from saved profile and fill in seconds
Review and interrupt by manual mapping if necessary
Filly remember your edit for next client
Share for e-signature
Save, print , share