|
EstiScan
CASE STUDY
This isn’t “hand it all to a VLM on full-auto.” Deterministic methods handle coordinates & dimensions, AI handles meaning & Japanese, and risk areas are verified by humans. We compute the takeoff from traceable structured data, not from estimates.
Left to a VLM alone, dimensions become “estimates,” and miscounts and hallucinations propagate into the takeoff (= cost).
A pixel-only TIFF leaves lines, text, and parts unreadable by machines. With a VLM alone, dimensions become estimates.
Miscounts and hallucinations grow into cost errors through unit-price × quantity takeoff.
You can’t trace or audit “why that value.” You can’t explain the basis for the cost afterward.
▣ Deterministic= coordinates & dimensions (mm) ◆ AI= meaning & Japanese 👤 Human= review of risk areas
Preprocessing & vectorization convert pixels into line segments with start/end points and coordinates. Dimensions are fixed geometrically, not estimated by AI.
Built around commercially usable OSS such as OpenCV / DeepLSD / PaddleOCR / Detectron2 / Qwen2.5-VL, avoiding vendor lock-in.
Click to experience the actual processing flow (mock data) right in your browser.