MCP server that gives AI chat real vision by reading images, files, logs, and zip archives via local OCR in any language, with privacy-first processing.
Enables text-only coding models to read images, PDFs, presentations, spreadsheets, and other non-text files through a single analyze_media tool, combining local document extraction, OCR, and optional vision models with clear evidence labeling.
A streamlined MCP server for XMP metadata embedding with beautiful formatting and smart filename indicators, enabling metadata embedding, reading, validation, and report generation for lifestyle, product, and orbit schemas.
MCP server that enables Codex to run image generation and inspection as durable jobs, returning bounded text results and references instead of image bytes to keep context manageable.
MCP server for MarkItUp's AI image-annotation pipeline. Generate polished marketing-visual variations of any screenshot, regenerate, AI outpaint, and remove backgrounds —
powered by Claude analysis + Gemini rendering.
Enables text-only agents to process images by accepting image files, base64 data, or URLs, sending them to multimodal models, and returning structured text results via MCP.
MCP server for tinify.ai image optimization. AI-powered upscaling, resizing/cropping, compression, and SEO filename & alt text generation — all in one tool.
A local MCP server that provides image processing tools including resizing, cropping, format conversion, compression, rotation, flipping, thumbnailing, watermarking, effects, placeholder generation, and overlaying.
Enables intelligent multi-provider image generation through OpenAI and Google Gemini APIs with automatic provider selection, support for reference images, real-time data grounding, and conversational refinement.
Enables MCP-compatible AI agents to generate background images and videos through the museav platform, and to perform local image post-processing such as background removal, upscaling, watermark removal, and compression, returning file paths instead of base64.
Enables video and audio processing through FFmpeg, supporting format conversion, compression, trimming, audio extraction, frame extraction, video merging, and subtitle burning through natural language commands.
An MCP server for AI-powered image generation using Google Gemini models with intelligent automatic model selection, supporting multiple resolutions and aspect ratios.