Skip to main content
Glama

PDF Analyzer MCP Server

The PDF Analyzer MCP Server gives AI agents the ability to read and analyze PDF documents, enabling document Q&A through natural conversations.

Supports multiple LLM providers: Google Gemini, Anthropic Claude, and OpenAI on their direct APIs, plus Google Vertex AI and Anthropic on Vertex AI for service-account auth. Choose your preferred provider and model during setup.

macOS / Linux:

curl -fsSL https://raw.githubusercontent.com/IntelligentElectron/pdf-analyzer/main/install.sh | bash

Windows (PowerShell):

irm https://raw.githubusercontent.com/IntelligentElectron/pdf-analyzer/main/install.ps1 | iex

Why use the native installer:

  • No dependencies — standalone binary, no Node.js required

  • Auto-updates — checks for updates on startup

  • Signed binaries — macOS binaries are notarized by Apple

Platform

Install Directory

macOS

~/Library/Application Support/pdf-analyzer/

Linux

~/.pdf-analyzer/

Windows

%LOCALAPPDATA%\pdf-analyzer\

Update

The server checks for updates on startup. To update manually:

pdf-analyzer --update

Related MCP server: PDF RAG MCP Server

Alternative: Install via npm

For developers who prefer npm:

npm install -g @intelligentelectron/pdf-analyzer

Or use with npx (no installation required):

npx @intelligentelectron/pdf-analyzer --help

Requires Node.js 20+.

To update:

npm update -g @intelligentelectron/pdf-analyzer

Setup

After installing, run the interactive setup to choose your provider, model, and enter your API key:

pdf-analyzer --setup

You'll be prompted to choose from:

Provider

Fast Model

Flagship Model

Get API Key

Google Gemini

Gemini 3 Flash

Gemini 3.1 Pro

Google AI Studio

Anthropic Claude

Claude Sonnet 4.6

Claude Opus 4.7

Anthropic Console

OpenAI GPT

GPT-5.4 Mini

GPT-5.4

OpenAI Platform

Claude Opus 4.6 is offered alongside 4.7 as the previous flagship. The Vertex AI providers offer the same Gemini and Claude models, and authenticate with a service account JSON key file instead of an API key.

You can re-run --setup at any time to switch providers or models.

Connect the MCP with your favorite AI tool

After setup, connect the MCP to your AI agent of choice.

Claude Code

Install Claude Code, then run:

claude mcp add --scope user pdf-analyzer -- pdf-analyzer

OpenAI Codex

Install OpenAI Codex, then run:

codex mcp add pdf-analyzer -- pdf-analyzer

Usage

Once connected, ask your AI assistant to analyze any PDF:

  • "Analyze /path/to/document.pdf and summarize the key points"

  • "What tables are in this PDF? Extract the data from table 2"

  • "Compare the findings in sections 3 and 5 of this report"

The server accepts:

  • Local file paths: /Users/name/docs/report.pdf

  • URLs: https://example.com/document.pdf

Supported Platforms

Platform

Binary

macOS (Universal)

pdf-analyzer-darwin-universal

Linux (x64)

pdf-analyzer-linux-x64

Linux (ARM64)

pdf-analyzer-linux-arm64

Windows (x64)

pdf-analyzer-windows-x64.exe

Running as a hosted server

Setting PORT starts the server over Streamable HTTP instead of stdio, serving MCP at /mcp, a direct POST /analyze REST endpoint, and GET /health.

See deploy/README.md for deploying it to Cloud Run, including the provider and auth matrix, the IAM roles each provider needs, and how to reach the private service.

Documentation

See docs/architecture.md for how the server is put together.

See CONTRIBUTING.md for development guidelines.


About

Created by Valentino Zegna

This project is hosted on GitHub under the IntelligentElectron organization.

License

Apache License 2.0 - see LICENSE

Available Tools

1 tool
analyze_pdfA

Analyze a PDF document using AI. Provide an absolute file path, URL, cached file URI (from a previous response, Google only), or array of cached file URIs (from a previous chunked response, Google only) and a list of questions to ask about the PDF content. With the Google provider, returns a cached_uris array that can be reused for subsequent queries on the same document.

ParametersJSON Schema
NameRequiredDescriptionDefault
queriesYesArray of questions to ask about the PDF
pdf_sourceYesPDF source: absolute local file path, URL, cached file URI from a previous response (Google only), or array of cached file URIs from a previous chunked response (Google only)

TDQS

A3.9/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full burden. It mentions the caching behavior for Google but fails to describe the primary output (analysis results). The agent does not know what the tool returns beyond an optional cached_uris array.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two concise sentences, each serving a purpose. The description is front-loaded with the main action, followed by input specifics and a note on output. No wasted words.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness3/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the simplicity (2 params, no output schema, no annotations), the description covers input types and caching but omits the primary output format, error handling, or any limitations. It is partially complete.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, but the description adds value by clarifying provider-specific details (cached URIs for Google only) and contextualizing the pdf_source options beyond the schema's description.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states 'Analyze a PDF document using AI' with specific verb and resource, and distinguishes input types. It is precise and leaves no ambiguity about the tool's function.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

While no sibling tools exist, the description provides clear context on acceptable input formats (file path, URL, cached URI) and hints at provider-specific usage (Google caching). It lacks explicit when-not-to-use or alternatives, but the context is sufficient.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Tool Schema Changelog

Recent tool additions, removals, and schema changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. 1 tool updatev1.2.4
    • First observedanalyze_pdf

TDQS

A3.8/5.0
Disambiguation5/5

With only one tool, there is no possibility of ambiguity. The tool 'analyze_pdf' stands alone with a clear purpose.

Naming Consistency5/5

The single tool 'analyze_pdf' follows a clear verb_noun pattern, which is standard and consistent.

Tool Count2/5

A single tool for a PDF analyzer is too few. While the tool itself is non-trivial (AI analysis), the server lacks any additional operations like extraction, merging, or conversion, making it feel incomplete for typical PDF tasks.

Completeness2/5

The server only offers AI-based analysis of PDFs. There are obvious gaps such as missing CRUD operations for PDF content, text extraction, or file manipulation, which limits its usefulness.

Maintenance

ActivityMaintained
ResponsivenessSyncing

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Connectors

Related MCP Servers

  • F
    license
    Not graded
    quality
    D
    maintenance
    Enables AI assistants to search and query PDF documents through a local RAG system with vector embeddings. Provides semantic document search capabilities while keeping all data stored locally without external dependencies.
    -
  • A
    license
    Not graded
    quality
    D
    maintenance
    Enables intelligent search and question-answering over PDF documents using semantic similarity and keyword search. Supports OCR for scanned PDFs, persistent vector storage with ChromaDB, and maintains source tracking with page numbers.
    6
    MIT
  • A
    license
    Not graded
    quality
    B
    maintenance
    Enables AI agents and users to process documents through natural language, supporting PDF operations like text extraction, redaction, splitting, form filling, annotations, and content search.
    275
    61
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    Enables AI applications to read and process PDF files with intelligent file search, text extraction, image processing, and optional OCR support for scanned documents.
    MIT

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/IntelligentElectron/pdf-analyzer'

If you have feedback or need assistance with the MCP directory API, please join our Discord server