Files
nash_su 7c544b504d plan: multimodal image extraction + indexing
Spec for adding image extraction from PDF/PPTX/DOCX, vision-LLM
captioning, and indexing of captions through the existing RAG
pipeline. Phased delivery; Phase 5 (multimodal embedding) deferred.

See plans/multimodal-images.md for the full design, current-state
audit, open questions, and per-phase implementation notes.
2026-04-27 16:43:33 +08:00
..