Property Title Intelligence & Verification

BhumiClear

Automated preliminary Indian property title screening engine. Features multilingual OCR extraction across Indian scripts, deed chain verification, timeline gap detection, tax mutation validation, and advocate briefing pack generation.

OCR
Multi-Script Engine
LangGraph
Chain Verification
Patta/EC
Discrepancy Check
Dossier
Advocate Briefing

Screening Workflow Pipeline

Automated legal document ingestion and cross-reference verification.

1

Document Vault & Multi-Script OCR

Tesseract + EasyOCR + Gemini Vision ensemble extracts text, stamp paper seals, and registration stamps across English, Tamil, Telugu, Kannada, Hindi, and Marathi.

2

Property Identity Dossier

Cross-document correlation of Survey Numbers, Sub-division, Patta, Village, Taluk, Boundaries (Four-boundaries verification), and party name variations.

3

Chronology & Timeline Gaps

Constructs the full chain of ownership across 30+ years. Detects unrecorded partitions, unregistered power of attorneys (PoA), and registration date lags.

4

Advocate Briefing Pack

Generates a structured, concise legal briefing document highlighting missing EC/Patta certifications, boundary mismatches, and specific inquiry questions for title certification.

Architecture & Stack

Engineered in Python with LangGraph and Streamlit.

Layer Component Responsibility
Workflow Engine LangGraph 1.2, Python 3.11 Stateful DAG execution for legal validation and cross-doc checking
Vision & OCR Tesseract, EasyOCR, Poppler, Gemini Vision High-accuracy multi-lingual document extraction and quality scoring
User Interface Streamlit 1.40 Interactive intake, document viewer, and timeline explorer
Portal Verification SSRF-hardened HTTP client Reachability testing against state land-records portals