AI-powered document intelligence system using OCR, NLP, Named Entity Recognition, and FastAPI for structured information extraction from scanned documents.
-
Updated
Jul 19, 2026 - Python
AI-powered document intelligence system using OCR, NLP, Named Entity Recognition, and FastAPI for structured information extraction from scanned documents.
Learned visual-token pruning for Donut document OCR. 65% cross-attention KV memory cut for -0.26 word-recall points; learned selection beats random by up to +34 points. Pre-registered analysis, executable verifiers, documented negative results.
Research-oriented Document AI using LayoutLMv3 + EasyOCR for multimodal document understanding, entity extraction, and robustness evaluation.
Layout-aware financial-document intelligence: TATR table extraction, OCR content, DocLayNet layout, FUNSD relations, and table-grounded RAG QA. Honest subset evaluation; artifact-backed Gradio demo.
To associate your repository with the funsd topic, visit your repo's landing page and select "manage topics."