Media & documents
OCR + vLLM evidence
Ground image answers in both OCR and model interpretation.
Ground image answers in both OCR and model interpretation:
Arka reports what OCR says, what the visual/vLLM model says, and a combined
answer that identifies agreement, conflicts, and uncertainty. It does not
silently treat a model guess as text evidence. Natural language requests are
supported.