Skip to main content
Glama
522,894 tools. Updated 2026-09-06 13:24

"OCR agent" matching MCP tools:

  • Search US TV news coverage using transcripts, captions, OCR, and visual labels. Specify station and time window to retrieve matching clips.
    MIT
  • Extract visible text from your screen using OCR. Returns text grouped by detected windows with bounding boxes. Use when you need code, terminal, chat, or document content without visual layout. Avoids sending data to the cloud.
    AGPL 3.0
  • Analyze an image URL with Google Lens to find visual matches, detect objects, extract text via OCR, and identify exact product matches.
    MIT
  • Extract frames, OCR text, and transcript snippets from a specific video time range. Merge visual and audio content into a unified, annotated timeline.
    MIT
  • Specify a screen region with device pixel coordinates to extract text via OCR, returning JSON with text, confidence, and bounding box for targeted analysis.
    MIT

Matching MCP Servers

Matching MCP Connectors

  • OCR for images and Korean ID documents

  • Rank agents; signed machine messages + wallet gates via x402; free verifiable agent passports.

  • Find specific events, people, or topics in historical U.S. newspapers (1777–1963) by searching full OCR text. Filter by state and date range.
    MIT
  • Extract text from image files using OCR. Supports PNG, JPG, TIFF, BMP and multiple languages via Tesseract.
    MIT
    Destructive
  • Verify identity documents by submitting front and optional back images. Get structured OCR data and authenticity checks for fraud prevention.
    MIT
  • Convert scanned PDFs into searchable, editable documents using OCR. Choose quality, language, and skip OCR when text is already digital.
    MIT
  • Force OCR text extraction on scanned PDFs when normal text extraction returns garbled or empty text.
    MIT
    Destructive
  • Extract text from PDFs located in Notion, shared mounts, local paths, or Drive without moving bytes over MCP. Supports OCR for scanned/image-only PDFs.
    MIT
  • Retrieve OCR text from a Joplin image or scanned PDF without downloading the attachment. Use when you need the text content of a resource after OCR has finished processing.
    MIT
  • Scan the current VRoid window with OCR to find a specified text label and return its pixel coordinates for automated clicking.
    MIT
  • Extract text, markdown, and tables from PDF URLs with optional OCR. Choose page, document, or RAG chunk output, and control page ranges to stay within budget.
    MIT
  • Check which OCR engines are available on your machine before selecting an engine for text recognition.
    MIT