vision-augmentImage & Video ProcessingAI & Machine LearningCaoMeiYouRenAlicenseAqualityAmaintenanceEnables non-vision LLMs to understand images, extract text via OCR, and parse documents through a unified MCP interface, with local-first processing and optional OpenAI-compatible channels. Updated a month ago (2026-08-11 18:41 UTC)3MIT