TianshangScribe
TianshangScribe
面向开发者、CLI 自动化和 AI 代理的跨平台 Office 文档处理工具。可创建、编辑、模板填充和转换 Word(.docx)、Excel(.xlsx)和 PowerPoint(.pptx)文档,支持 LaTeX 风格标记、原生 OMML 数学公式以及模板引擎({{placeholders}}、{{#each}} 循环、{{#if}} 条件)。附带一个 MCP 服务器,提供 7 个工具(create、edit、fill template、convert、extract、validate、compare),支持 stdio、SSE 和 Streamable HTTP 传输,并带有 bearer-token 认证和速率限制。
警告:API 不稳定——可能发生破坏性变更
本项目处于 1.0 之前(0.x)阶段。CLI 选项、MCP 工具签名、模板语法和输出格式尚未冻结,可能随时变更,恕不另行通知。 兼容性承诺:任何破坏性变更将至少提前一个版本在 CHANGELOG 中公布,并附有迁移指南。 生产环境使用时,请锁定特定版本,并在升级前查看 CHANGELOG。
安装
pip install tianshang-scribe
# Or from source:
git clone https://github.com/Tianshang301/TianshangScribe.git
cd TianshangScribe
pip install -e ".[dev]"Linux 部署
Docker(推荐用于通过 Streamable HTTP 运行 MCP 服务器):
git clone https://github.com/Tianshang301/TianshangScribe.git
cd TianshangScribe
docker compose up -d
# Streamable HTTP MCP Server at http://localhost:8080/mcp
# (override transport / auth / rate limits via TIANSHANG_SCRIBE_* env vars).deb 包(Debian / Ubuntu):
# Download from GitHub Releases
sudo dpkg -i tianshang-scribe_0.7.1_all.deb
tianshang-scribe --helppipx(隔离的 CLI):
pipx install tianshang-scribe
tianshang-scribe --help需要 Python 3.10+ · python-docx · openpyxl · python-pptx · typer · rich · lxml
Related MCP server: docx-forge-mcp
快速开始
# Create a Word document
tianshang-scribe -w --create -a "Hello World" -o hello.docx
# Replace text (--regex for regex mode)
tianshang-scribe input.docx -r "old" --replace-new "new" -o output.docx
# LaTeX markup with nesting
tianshang-scribe -w --create --latex-style \
-s "font=Times New Roman,size=14" \
-a "\bfseries{\itshape{bold italic}} \fontsize{24}{Heading} \color{FF0000}{red}" \
-o styled.docx
# Math formulas —auto-converted to native Word OMML
tianshang-scribe -w --create \
--math "x = \frac{-b \pm \sqrt{b^2 - 4ac}}{2a}" \
--math "\sum_{i=0}^{n} i^2" \
-o formulas.docx
# Template filling (JSON / CSV / YAML →{{placeholder}})
tianshang-scribe template.docx -t data.json -o filled.docx
# Convert to PDF (office2pdf ~2MB, or LibreOffice fallback)
tianshang-scribe input.docx --topdf -o output.pdf
# MCP Server —stdio mode (Claude Code / Cursor)
python -m tianshang_scribe.mcp.server
# MCP Server —SSE mode (Dify / Coze / FastGPT)
python -m tianshang_scribe.mcp.server --transport sse --port 8080
# Excel: import CSV, sort, export JSON
tianshang-scribe -e --create --from-csv data.csv --sort "A1:A10 asc" --to-json -o out.json
# Excel: add formula, protect workbook
tianshang-scribe budget.xlsx --formula "B10 =SUM(B2:B9)" --protect "p@ss" -o protected.xlsx全局选项
参数 | 说明 |
| 输入文档路径(使用 |
| 处理 Word 文档 |
| 处理 Excel 工作簿 |
| 处理 PowerPoint 演示文稿 |
| 输出文件路径 |
| 允许覆盖已有文件 |
| 输出为 PDF |
| 从标准输入读取 |
| 写入标准输出 |
当省略 -w/-e/-p 时,将根据输入文件的扩展名推断文档类型。
操作
选项 | 说明 | 示例 |
| 创建空白文档 |
|
| 添加文本 |
|
|
|
|
| 查找并替换 |
|
| 删除内容 |
|
| 清除内容 / 格式 / 链接 |
|
| 修改内容 |
|
| 设置样式 |
|
| 模板填充 |
|
| 提取数据( |
|
| 设置属性 |
|
| 启用 LaTeX 解析 | |
| 添加数学公式(Word) |
|
| 数学解析方言(office/mathtype) |
|
| OMML 数学字体(默认 Cambria Math) |
|
| 以 MathType OLE 对象(MTEF)嵌入 |
|
| 添加标题(Word) |
|
| 正则表达式模式 | 与 |
| 合并文件 |
|
| 拆分文档(仅 Excel: |
|
| 添加批注(Word)/ 演讲者备注(PPT) |
|
| 添加表格(Word) |
|
| 添加图表(Excel) |
|
| 批处理模式 |
|
| 批处理的 glob 模式 |
|
| 计划任务 SQLite 数据库路径 |
|
| 注册计划任务 |
|
| 移除计划任务 |
|
| 列出计划任务 |
|
| 立即运行计划任务 |
|
| 运行到期的计划任务 |
|
| 在沙箱中运行脚本 |
|
| 从标准输入读取 | |
| 写入标准输出 |
Word 专用选项
选项 | 说明 | 示例 |
| 添加标题 |
|
| 添加数学公式 |
|
| 启用 LaTeX 标记 | |
| 生成目录 |
|
| 插入分节符 |
|
| 设置页眉 |
|
| 设置页脚 |
|
| 文字水印 |
|
| 转换为 Markdown |
|
| 转换为 HTML |
|
Excel 专用选项
选项 | 说明 | 示例 |
| 添加工作表 |
|
| 删除工作表 |
|
| 重命名工作表 |
|
| 设置列宽 |
|
| 设置行高 |
|
| 设置单元格公式 |
|
| 导入 CSV 数据 |
|
| 排序区域 |
|
| 添加图表 |
|
| 设置密码 |
|
| 移除密码 |
|
| 导出为 CSV | |
| 导出为 JSON | |
| 导出为 HTML |
LaTeX 风格标记
在 --add 内容中嵌入以下标记。使用 --latex-style 启用。支持嵌套。
语法 | 效果 |
| 粗体 |
| 斜体 |
| 小型大写字母 |
| 下划线 |
| 罗马体(衬线) |
| 无衬线体 |
| 等宽体 |
| 指定字体 |
| 字号(磅) |
| 颜色(十六进制) |
| 居中对齐 *— |
| 左对齐 *— |
| 右对齐 *— |
| 行距 *— |
| 缩进 *— |
| 插入标题 |
| 分页 |
| 插入图片 |
*— 段落级格式(创建新段落)。
字体配置
命令 | 效果 |
| 默认西文字体 |
| 默认中文字体 |
| 无衬线字体 |
| 中文无衬线字体 |
| 等宽字体 |
| 中文等宽字体 |
Word OOXML 原生区分 w:ascii(西文)和 w:eastAsia(中文)字体,可在混合文字中自动切换字体。
数学公式
LaTeX 数学公式可通过 --math 转换为原生 Word OMML(Office Math Markup Language)格式。该转换器是一个手写的递归下降解析器(expression → term → factor → atom),作用于嵌套的不可变令牌树(fraction、root、N-ary、sub/sup、accent、styled、delimiter 令牌),并通过一张 O(1) 命令表进行分派,该命令表包含预编译的正则表达式和零拷贝参数切片。使用 --math-font "Times New Roman" 可将公式渲染为 MathType 风格的衬线字体,而不是 Word 默认的 Cambria Math(<m:mathPr><m:mathFont>)。--math-style mathtype 可切换 LaTeX 解析方言以获得 MathType 兼容。使用 --math-mtef 可将公式作为真正的 MathType OLE 对象(MTEF 二进制格式)嵌入——这可由旧版 MathType(6.x 及更早版本)编辑,也与 --extract math 读取回的格式相同。输出在所有版本之间保持字节级稳定(受 golden-snapshot 回归测试套件保护)。
支持语法
分类 | 命令 |
分数 |
|
根式 |
|
上标/下标 |
|
求和/积分 |
|
极限 |
|
具名函数 |
|
希腊字母 |
|
符号 |
|
关系 |
|
箭头 |
|
重音 |
|
括号 |
|
数学字体 |
|
数学排版
遵循主流数学期刊标准(AMS、Elsevier、Springer):
内容 | 样式 | 示例 |
单字母变量 | 斜体 |
|
数字 | 正体 |
|
具名函数 | 正体 |
|
小写希腊字母 | 斜体 |
|
大写希腊字母 | 正体 |
|
自动识别
--add 文本中的命令即使没有 $...$ 包裹也会自动识别为数学公式:
带参数:
\frac\sqrt\sum\int\prod\lim重音:
\hat{x}\bar{x}\vec{x}等一元运算符:
\sin\cos\tan\log\ln等纯文本中的
H_{2}O和m^{2}会变成 Unicode 下标/上标(H₂O / m²)
样式语法
--style 使用逗号分隔的键值对:
--style "font=Times New Roman,size=14,bold,italic,color=FF0000,align=center"键 | 别名 | 值 | 说明 |
|
| 字体名 | 西文字体 |
|
| 字体名 | CJK 字体 |
|
| pt | 字号 |
| flag | 粗体 | |
| flag | 斜体 | |
| flag | 下划线 | |
|
|
| 十六进制颜色 |
|
|
| 对齐方式 |
布尔键(bold italic underline)只要存在即为 True。
模板填充
支持 JSON、CSV 和 YAML 数据源。替换文档中的 {{placeholder}}。嵌套对象可用点号表示法展开。循环可迭代列表值。条件语句可显示/隐藏块。
{
"name": "John Doe",
"date": "2026-07-28",
"user": { "city": "Beijing" },
"show": true,
"paid": false,
"items": [
{ "product": "Widget", "price": "10" },
{ "product": "Gadget", "price": "20" }
]
}{{name}} → John Doe
{{user.city}} → Beijing
{{#each items}} → repeats the block for each item
{{product}}: {{price}}
{{/each}}
{{#if show}} → shown only when show is truthy
Confidential content
{{/if}}
{{#if role=admin}} → shown only when role equals "admin"
Admin dashboard
{{/if}}
{{#unless paid}} → shown only when paid is falsy
Payment required
{{/unless}}Excel 功能
功能名称 | CLI 选项 |
工作表管理 |
|
列/行调整 |
|
公式 |
|
数据导入 |
|
数据导出 |
|
排序 |
|
图表 |
|
保护 |
|
PPT 功能
功能名称 | 说明 |
幻灯片管理 | 添加、删除、重新排列幻灯片( |
版式 | 按名称或索引应用幻灯片版式( |
演讲者备注 | 添加演示者备注( |
数学公式 |
|
切换效果 | 设置幻灯片切换效果——淡入、推入、擦除等( |
导出 | 将幻灯片保存为图片( |
媒体压缩 | 压缩图片( |
保护 | 设置/移除密码( |
退出码
代码 | 含义 |
| 成功 |
| 一般错误 |
| 参数错误 |
| 未实现 |
MCP 服务器
天匠Scribe 包含 MCP(Model Context Protocol)服务器——AI 智能体可以创建、编辑、填充模板、转换,以及从 Office 文档中提取数据。
Quick Connect
stdio(Claude Code, Cursor):
{"mcpServers": {"tianshang-scribe": {
"command": "python", "args": ["-m", "tianshang_scribe.mcp.server"]
}}}SSE(Dify, Coze, FastGPT):
python -m tianshang_scribe.mcp.server --transport sse --host 0.0.0.0 --port 8080{"mcpServers": {"tianshang-scribe": {
"url": "http://localhost:8080/sse", "transport": "sse"
}}}Tools (7)
Tool | Description |
| 使用结构化内容块创建 .docx / .xlsx / .pptx |
| 对现有文档执行替换、删除、修改、样式和添加操作 |
| 用数据填充 |
| 在格式之间转换(docx↔pdf/md/html,xlsx↔csv/json) |
| 提取元数据、全文或文档结构 |
| 在填充前预检模板占位符与数据是否匹配 |
| 比较两个 .docx 文件的段落级差异 |
Capabilities
Feature | Detail |
协议 | MCP 2024-11-05 · stdio + SSE · JSON-RPC 2.0 |
资源 |
|
提示词 | 5 个内置工作流模板( |
进度 | 在 PDF 转换和长操作期间发送 |
响应 | 多类型 |
模式 | 所有参数均支持 |
生产环境(仅 SSE)
# With authentication
TIANSHANG_SCRIBE_AUTH_TOKEN="secret" \
python -m tianshang_scribe.mcp.server --transport sse --host 0.0.0.0 --port 8080
# Health check
curl http://localhost:8080/health
# {"status":"ok","version":"0.7.1","uptime_seconds":3600,"active_sessions":3,"tools_available":7}
# CORS whitelist
python -m tianshang_scribe.mcp.server --transport sse --cors-origins "https://coze.com,https://dify.ai"端点:GET /health · GET /sse · POST /message?session_id=X
完整文档:docs/mcp/README.md。
python tests/integration/mcp/mcp_stdio_smoke.py # 9/9 quick tests (stdio)
python tests/integration/mcp/test_sse.py # 3/3 SSE transport tests
python tests/integration/mcp/mcp_agent_sim.py # 11-scenario Agent simulation架构
src/
└── tianshang_scribe/ # importable package (tianshang_scribe.*)
├── cli/ # Typer CLI entry
│ ├── main.py # Command parsing & dispatch
│ └── global_opts.py # File path / type inference
├── core/ # Document engine abstraction
│ ├── document.py # DocumentABC unified interface
│ ├── word_engine.py # Word engine (python-docx)
│ ├── excel_engine.py# Excel engine (openpyxl)
│ └── ppt_engine.py # PPT engine (python-pptx)
├── rendering/ # Style & formula rendering
│ ├── styles.py # TextStyle dataclass
│ ├── latex_parser.py # LaTeX markup parser
│ ├── math_omml.py # LaTeX →OMML math converter
│ └── template.py # Template filling engine
├── transform/ # Format conversion
│ └── pdf.py # PDF export (office2pdf + LibreOffice)
├── mcp/ # MCP Server (official mcp SDK 2.x)
│ ├── server.py # build_server + entry (stdio / SSE / Streamable HTTP)
│ ├── transport.py # transport wiring + ASGI middleware
│ ├── schemas.py # pydantic models + as_dict
│ ├── auth.py # Bearer token auth
│ ├── rate_limit.py # token bucket rate limiting
│ ├── metrics.py # Prometheus-style metrics
│ ├── security.py # read-only / destructive classification
│ ├── prompts.py # 5 prompt workflows
│ ├── tools/ # 7 Agent tools
│ │ ├── _registry.py # tool registry (schemas auto-derived)
│ │ ├── create.py / edit.py / template.py / convert.py
│ │ ├── validate.py / compare.py
│ └── errors.py # structured error codes + fixes
└── utils/ # Utility functions
└── file_utils.py技术栈
组件 | 技术 |
CLI | Typer + Rich |
Word | python-docx |
Excel | openpyxl |
PPT | python-pptx |
数学 | 手写递归下降解析器 → OMML XML(不可变令牌树、命令分派表) |
模板 | 自定义引擎({{placeholder}}, {{#each}}, {{#if}}) |
office2pdf(约 2MB Rust 二进制,零依赖)+ LibreOffice 后备方案 | |
质量 | pytest(936 项测试)· ruff · mypy |
构建 EXE
pip install pyinstaller
pyinstaller --onefile --name tianshang-scribe --hidden-import openpyxl.cell._writer --hidden-import openpyxl.cell.read_only --hidden-import openpyxl.styles --hidden-import openpyxl.chart --hidden-import openpyxl.comments src/tianshang_scribe/cli/main.py
# dist/tianshang-scribe.exe (~35 MB)演示
python -m demo.generate_demos
# demo/demo_word.docx —LaTeX + math + TOC + watermark
# demo/demo_excel.xlsx —CSV import + formulas + chart + protection
# demo/demo_ppt.pptx —slides + notes + transitions + math formulasCLI 合规性测试:
python demo/test_cli.py开发
git clone https://github.com/Tianshang301/TianshangScribe.git
cd TianshangScribe
pip install -e ".[dev]"
pytest tests/ -v # Run tests
ruff check src/tianshang_scribe/ tests/ # Lint
mypy src/tianshang_scribe/ # Type check许可证
Apache-2.0
Available Tools
12 toolsanalyze_excel_dataARead-onlyIdempotent
Analyze an Excel workbook without touching it: per-sheet row/column counts, headers, inferred column types (numeric min/max/mean, categorical values), null counts, sample rows, and duplicate-row detection. Read-only — never modifies the input file. After analysis use edit_excel_workbook to apply fixes, or extract_document_data for raw text and metadata.
| Name | Required | Description | Default |
|---|---|---|---|
| options | No | ||
| input_path | Yes |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already declare read-only, non-destructive, and idempotent behavior. The description reinforces this with 'never modifies the input file' and 'Read-only', adding explicit confirmation. It also discloses the scope of analysis (per-sheet counts, headers, inferred types), which goes beyond annotations by describing what the tool inspects, but doesn't detail edge cases like formula evaluation or large file behavior.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is concise, with the key purpose and read-only guarantee front-loaded. It lists specific outputs in a compressed list and then mentions alternatives in a final sentence. No wasted words, and every sentence contributes to understanding.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool has a detailed output schema (not shown in detail but indicated as 'true'), the description doesn't need to explain return values. With only 2 parameters (one required), the description covers what the tool does and its safety profile. The only gap is not specifying which options from the shared ToolOptions are relevant, but for a read-only analysis tool, most options (like backup, restore) are likely inapplicable, and the description's clarity about read-only behavior implicitly signals that. Slight incompleteness for an agent unfamiliar with the options structure.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The schema has 0% description coverage, and the only functional parameter is 'input_path' which is self-explanatory. The options parameter is a complex nested object with many fields for other tools, but the description doesn't clarify which options apply to this tool. Even with low schema coverage, the input_path is obvious, and the options are likely irrelevant to this read-only analysis, so the description adds no extra meaning beyond what the parameter name suggests.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool's purpose: analyzing an Excel workbook without modifying it, and lists specific analysis outputs (row/column counts, headers, column types, null counts, sample rows, duplicate detection). It distinguishes itself from sibling tools by explicitly mentioning edit_excel_workbook and extract_document_data as alternatives.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description gives clear guidance on when to use the tool: for analysis before applying fixes or extracting text. It names alternatives (edit_excel_workbook for fixes, extract_document_data for raw text), but doesn't provide explicit when-not-to-use conditions, though the alternative mentions serve as implicit exclusions.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
compare_documentsA
Compare two Word (.docx) documents and report paragraph-level differences: additions, removals, and changes with paragraph indices. Note: comparison and snapshots are ONLY supported for Word (.docx); Excel (.xlsx) and PowerPoint (.pptx) comparison is not yet available (you will receive an UNSUPPORTED_FORMAT error). Also manages document snapshots via options.action: "snapshot" records path_a state, "list_snapshots" lists recorded snapshots, and "restore" writes a snapshot back to path_b (snapshot store default: ~/.tianshang-scribe/snapshots/). The compare mode never modifies its inputs; snapshot/restore write to the snapshot store. To read a single document, use extract_document_data.
| Name | Required | Description | Default |
|---|---|---|---|
| path_a | Yes | Path to the first document. | |
| path_b | Yes | Path to the second document. | |
| options | No | Tool options. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description explicitly states that compare mode never modifies inputs, while snapshot/restore write to the snapshot store, and restore overwrites path_b. It also mentions backup and dry_run options, providing full transparency about side effects despite annotations not indicating these behaviors.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is well-structured, starting with the core purpose, then detailing snapshot actions, format limitations, and an alternative. Every sentence adds necessary information without redundancy, maintaining a balance between completeness and brevity.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool's complexity (multiple actions, options, and side effects), the description covers all essential aspects: supported formats, error conditions, alternative tools, default behaviors, and side effects. The presence of an output schema further completes the picture, so nothing critical is missing.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The schema already provides descriptions for all parameters (100% coverage). The tool description adds valuable context, such as the default snapshot_dir and the meaning of each action, enriching the semantic understanding beyond the schema alone.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool's primary function: comparing two Word documents for paragraph-level differences. It explicitly names the resource (two .docx files) and the actions (additions, removals, changes). It also distinguishes itself from a sibling tool (extract_document_data) and specifies format limitations, ensuring unambiguous purpose.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description provides explicit guidance on when to use the tool (for comparing two documents) and when not to (unsupported formats like .xlsx/.pptx), and offers an alternative for single-document reading (extract_document_data). It also explains the snapshot actions and side effects, making usage conditions clear.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
convert_documentA
Convert a document between formats while preserving structure where possible: Word (.docx) to PDF/Markdown/HTML, Excel (.xlsx) to PDF/CSV/JSON/HTML, PowerPoint (.pptx) to PDF. Writes the converted file to output_path (default: .). PDF conversion requires office2pdf or LibreOffice to be installed. To read content for analysis instead, use extract_document_data.
| Name | Required | Description | Default |
|---|---|---|---|
| options | No | Tool options. | |
| input_path | Yes | Path to the source document. | |
| output_path | No | Output file path. | |
| target_format | Yes | Target output format: - "pdf": PDF document (requires office2pdf or LibreOffice) - "csv": Comma-separated values (Excel only) - "json": JSON array of rows (Excel only) - "html": HTML table (Excel) / styled HTML (Word) - "md"/"markdown": Markdown (Word only) |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description discloses that the tool writes a converted file to output_path, and notes a dependency for PDF conversion. However, it does not mention overwrite behavior, idempotency, or failure modes. Since annotations provide no hints (all false), the description bears the burden, and it provides some but not comprehensive behavioral transparency.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is two sentences with no fluff, front-loading the core purpose and including necessary caveats and alternatives.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The description covers the main functionality, dependencies, and provides an alternative tool for reading content. It does not describe the options parameter (shared across tools) or return values, but the output schema exists, and the schema covers the options. Given the complexity, it is reasonably complete, though it could mention overwrite semantics.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The description adds meaning beyond the schema by specifying which target formats are valid per source type (e.g., CSV/JSON only for Excel, Markdown only for Word). It also clarifies the default output_path behavior. This goes beyond the schema's generic descriptions.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the verb 'Convert', the resource 'document between formats', and enumerates specific source-target combinations (Word, Excel, PowerPoint to various formats). It also distinguishes itself from the sibling tool extract_document_data by explicitly saying to use that for reading content instead.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description provides explicit when-not-to-use guidance: 'To read content for analysis instead, use extract_document_data.' It also mentions the prerequisite for PDF conversion (office2pdf or LibreOffice), which guides the agent on environment requirements. This is clear context.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
create_excel_workbookA
Create a NEW Excel workbook from typed sheet specs: each sheet carries name, headers, rows, cell formulas, freeze panes, number_format, conditional_format, data_validation, and column widths. Writes a new .xlsx at output_path and overwrites any file already there. For targeted changes to an existing workbook use edit_excel_workbook; for read-only inspection use analyze_excel_data.
| Name | Required | Description | Default |
|---|---|---|---|
| sheets | Yes | Worksheets to build (in order). | |
| options | No | Tool options. | |
| metadata | No | Optional document properties (title/author/...). | |
| output_path | Yes | Output .xlsx path to create. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description warns that the tool 'overwrites any file already there,' which is important behavioral context. However, the annotation destructiveHint=false contradicts this warning, since overwriting an existing file is a destructive side effect. Per the rubric, a description that contradicts annotations scores 1.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is composed of three purposeful sentences, with the core creation intent first, the overwrite warning second, and usage routing third. The enumeration of sheet properties is somewhat redundant with the schema but still contributes to quick understanding.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The description covers core behaviors including creation, overwriting, and sibling routing. It acknowledges the diversity of sheet configuration through the property list and explains the difference between creating versus editing. The main gap is the contradiction with the destructive annotation, which could mislead agents.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so the schema already documents each parameter. The description repeats the sheet property names but does not add deeper meaning or usage context beyond the schema. A baseline 3 is appropriate.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states a specific verb and resource: 'Create a NEW Excel workbook' from typed sheet specs, and lists the specific capabilities (headers, rows, formulas, freeze panes, etc.). It also implicitly differentiates from siblings by focusing on creation versus modification or analysis.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Explicit guidance is provided: use edit_excel_workbook for targeted changes to an existing workbook and analyze_excel_data for read-only inspection. This tells the agent exactly when to choose this tool versus the alternatives.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
create_office_documentA
Create a NEW Word (.docx), Excel (.xlsx), or PowerPoint (.pptx) file from a structured content list of typed blocks (heading, paragraph, formula, table, image), with LaTeX-style markup (\bfseries{bold}) and math formulas (\frac{a}{b}). Writes to output_path (default: a temp file) and persists. To edit an existing file, use edit_office_document. Excel blocks may carry: sheet_name, cell+formula, freeze, chart_type+chart_data_range, number_format, conditional_format, data_validation, hyperlink, named_range. PPT blocks may carry: slide_index, slide_layout, notes, transition, chart_type+chart_data, rows (table), path (picture), fill/line (shape). Multiple PPT text/table/chart blocks stack onto one slide unless slide_index is set.
| Name | Required | Description | Default |
|---|---|---|---|
| style | No | Global document style. | |
| format | Yes | Document format: - "docx": Word document — reports, letters, contracts, proposals - "xlsx": Excel workbook — spreadsheets, data tables, charts - "pptx": PowerPoint — slides, presentations, pitch decks | |
| content | Yes | Ordered list of content blocks. | |
| options | No | Tool options. | |
| metadata | No | Document metadata (title, author, etc.). | |
| output_path | No | Output file path. | |
| template_data | No | Key-value pairs to fill {{placeholder}} in content. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Reveals concrete behaviors: persists output, default temp file, PPT block stacking overwrites slides unless slide_index provided. This goes well beyond the annotations. Could mention whether it overwrites existing output_path files (destructiveHint=false but not stated) — that's the main gap.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is moderately long but front-loaded with the core action ('Create a NEW Word, Excel, or PowerPoint file') and switches to the sibling tool early. The format-specific details (Excel vs PPT block properties) are useful but slightly dense; still, they are organized and valuable. Minor redundancy with the schema ('action' default, 'output_path' default) but not bloated.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Covers most end-to-end context: when to use, what formats are supported, how blocks map to formats, output behavior, and the sibling tool for editing. Given 7 parameters and nested blocks, it could add a short example or note about template_data interplay, but it's already quite complete. The block types in description map well to schema fields (cell, formula, rows, style, template_data).
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Although schema coverage is 100% (7 parameters fully described), the description adds significant semantic value: it explains the relationship between blocks and formats (Excel-only properties like cell/formula, PPT-only like slide_index/transition), clarifies 'Writes to output_path (default: a temp file) and persists', and emphasizes that unlisted-block properties are typed. The schema defines the shape; the description explains usage semantics (stacking, default behavior, format applicability), which is exactly what an agent needs.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description states a specific action ('Create a NEW Word, Excel, or PowerPoint file') with a clear verb ('create'), object ('office document'), and scope (docx/xlsx/pptx). It explicitly says 'To edit an existing file, use edit_office_document', which differentiates it from the editing sibling. The description also names the three supported formats and details per-format capabilities (Excel blocks with formulas/conditional formatting, PPT blocks with slides/transitions), so an agent can unambiguously tell what this tool produces and how it differs from edit_office_document.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Gives clear context (create vs edit, format selection) and explicitly mentions the alternative tool. Missing explicit 'when NOT to use' exclusions beyond the edit case, but the guidance is sufficient.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
create_presentationA
Create a NEW PowerPoint deck from typed slide specs: layout, title, bullets, positioned text boxes, tables, charts, pictures, speaker notes, and transitions. Writes a new .pptx at output_path and overwrites any file already there. To change an existing deck slide by slide instead, use edit_presentation.
| Name | Required | Description | Default |
|---|---|---|---|
| slides | Yes | Slides to build (in order). | |
| options | No | Tool options. | |
| metadata | No | Optional document properties (title/author/...). | |
| output_path | Yes | Output .pptx path to create. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotation Contradiction: The description explicitly states the tool 'overwrites any file already there,' which is a destructive side effect on the output path, yet annotations declare destructiveHint=false. This creates conflicting signals for the agent about whether an existing file may be lost. The overwrite warning is transparent, but the annotation contradicts it, so the score must drop to 1 per the rubric.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
Three sentences with no filler. The first sentence states the primary purpose, the second exposes the overwrite side effect, and the third names the sibling alternative. Information is front-loaded and every sentence earns its place.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The description is largely complete: it names the tool's purpose, enumerates the supported slide components, discloses the overwrite behavior, and points to the alternative for editing. It does not describe the return value, but the context indicates an output schema exists. The only true gap is the annotation contradiction, which is addressed separately; otherwise, the description provides enough context for an agent to invoke the tool correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so the schema already documents the parameters and nested slide structures. The description adds general context about output_path and the kinds of slide content supported, but it does not clarify ambiguous schema details such as the generic ToolOptions block (which references compare_documents actions unrelated to presentation creation). Baseline 3 is appropriate because the schema carries the parameter-semantic burden.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description states a specific verb and resource ('Create a NEW PowerPoint deck') and enumerates the supported slide components (layout, title, bullets, positioned text boxes, tables, charts, pictures, notes, transitions). It explicitly distinguishes itself from edit_presentation by saying it creates from scratch rather than changing an existing deck slide by slide.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description gives a clear condition for use: create a new .pptx from typed slide specs. It explicitly routes to a sibling tool: 'To change an existing deck slide by slide instead, use edit_presentation.' This gives the agent a direct when-to-use/when-not-to-use rule.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
edit_excel_workbookADestructive
Edit an existing .xlsx workbook with typed operations: write_cell, set_formula, freeze_panes, add_chart, conditional_format, data_validation, add_table, sort, add_sheet, set_range_style, and number_format. Rewrites the file — when output_path is omitted the INPUT FILE IS OVERWRITTEN IN PLACE, so pass output_path or options {"backup": true} to keep a .bak copy. To build a new workbook use create_excel_workbook; to inspect one first use analyze_excel_data.
| Name | Required | Description | Default |
|---|---|---|---|
| options | No | Tool options (dry_run, backup). | |
| input_path | Yes | Path to the existing .xlsx workbook. | |
| operations | Yes | Typed Excel edit operations applied in order. | |
| output_path | No | Output path (defaults to the input file). |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
While the annotations already indicate destructiveHint=true, the description adds critical behavioral specifics: rewriting the file, overwriting in place when output_path is omitted, and the backup option. It also mentions the dry_run option for validation without writing, providing a comprehensive picture of side effects.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is well-structured: it opens with the core action, lists the operations compactly, then provides essential behavioral warnings and sibling alternatives. It is concise yet comprehensive, with no unnecessary fluff, and the key information (overwrite, backup) is prominently placed.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The tool is a mutation tool with output schema present, so return values need no explanation. The description covers all necessary context: how to perform edits, optional output path, backup behavior, dry-run validation, and how to avoid overwriting. It is complete for an agent to invoke the tool correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
All parameters have descriptions (schema coverage 100%), and the ExcelEditOp description in the schema explicitly lists which fields are meaningful for each action (e.g., 'write_cell': cell, value, sheet_name, style, is_formula). This gives clear parameter semantics beyond the terse per-field descriptions, ensuring the agent understands how to use each field.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool edits an existing .xlsx workbook and lists the specific typed operations (write_cell, set_formula, etc.), distinguishing it from creation and analysis tools. The verb 'edit' and resource '.xlsx workbook' make the purpose explicit.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description provides explicit usage guidance by naming sibling tools: 'To build a new workbook use create_excel_workbook; to inspect one first use analyze_excel_data.' This tells the agent when to use this tool versus alternatives, and the mention of overwriting and backup options further clarifies usage scenarios.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
edit_office_documentADestructive
Legacy general-purpose editor kept for backward compatibility; for Excel use edit_excel_workbook and for PowerPoint use edit_presentation instead of this wide operation model. Still supports replace/delete/modify/style/add/clear on any type plus the Excel/PPT capability actions. Rewrites the file — when output_path is omitted the INPUT FILE IS OVERWRITTEN IN PLACE, so pass output_path or options {"backup": true} for a .bak copy. New documents belong to create_office_document.
| Name | Required | Description | Default |
|---|---|---|---|
| options | No | Tool options. | |
| input_path | Yes | Path to the existing document. | |
| operations | Yes | List of edit operations applied in order. | |
| output_path | No | Output path (defaults to the input file). |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided to indicate read-only or destructive nature meaningful. The description only states it edits Word documents without describing side effects, reversibility, or whether it overwrites the source file, which is a significant gap for an edit tool.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
A single concise sentence that states the tool's purpose and names alternatives. Clear and efficient.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The description gives purpose and alternative routing but lacks behavioral details like output format, mutation of input, or return value. With no destructive/read-only hints beyond the false flags, the agent relies on the schema for specifics.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The schema provides detailed per-operation parameter descriptions with 100% coverage illustration. The description itself adds little parameter detail, relying on the schema, which is acceptable given the coverage. Slight deduction because the description does not summarize the key operation fields.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly identifies the tool as an editor for office documentsaisôcienta and specifies it can still edit Word documents, while naming the sibling tools for Excel and PPT. An agent can readily distinguish it from related tools.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
Explicitly instructs to use edit_excel_workbook or edit_presentation instead for Excel/PPT, leaving no ambiguity about when this tool should be used.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
edit_presentationADestructive
Edit an existing .pptx with typed operations: add_slide, add_text, replace_text, add_table, add_chart, add_picture, add_shape, apply_layout, set_transition, and add_notes. Rewrites the deck — when output_path is omitted the INPUT FILE IS OVERWRITTEN IN PLACE, so pass output_path or options {"backup": true} for a .bak copy. To generate a new deck use create_presentation.
| Name | Required | Description | Default |
|---|---|---|---|
| options | No | Tool options (dry_run, backup). | |
| input_path | Yes | Path to the existing .pptx presentation. | |
| operations | Yes | Typed PowerPoint edit operations applied in order. | |
| output_path | No | Output path (defaults to the input file). |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already mark destructiveHint=true, but the description adds crucial context: rewriting the deck and in-place overwrite when output_path is omitted. It also discloses the mitigation (output_path or backup option), which goes well beyond the structured annotation and is vital for safe invocation.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is two sentences with no wasted words. It front-loads the purpose, lists operations efficiently, then delivers the critical overwrite warning and sibling pointer. Every sentence earns its place.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
The description covers the most important context for a mutation tool: the destructive overwrite behavior and how to prevent it. Since an output schema exists and the input schema documents the per-operation fields, nothing essential is missing for an agent to call the tool correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so the baseline is 3. The description lists the operation names but does not add parameter-level meaning beyond what the schema already provides; the schema itself documents per-action field mapping in the PptEditOp description. Therefore the description neither compensates nor degrades parameter semantics.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description uses a specific verb and resource: 'Edit an existing .pptx with typed operations' and enumerates the exact supported operations. It also distinguishes itself from the sibling create_presentation by explicitly saying 'To generate a new deck use create_presentation', so an agent can tell them apart.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
It explicitly states when to use the tool (editing an existing .pptx) and names the alternative for new decks: 'To generate a new deck use create_presentation'. It also gives essential usage guidance around output_path and the backup option, telling the agent to pass output_path or options {'backup': true} to avoid overwriting the input file.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
extract_document_dataARead-onlyIdempotent
Read data from a document: metadata (author/title/etc.), text (plain text plus a block count), or structure (paragraphs/sheets/slides). Read-only — never modifies the input file. To compare two documents, use compare_documents.
| Name | Required | Description | Default |
|---|---|---|---|
| mode | No | What to extract: metadata, text, or structure. | metadata |
| options | No | Tool options. | |
| input_path | Yes | Path to the source document. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
The description restates the read-only and non-destructive behavior that the annotations already provide (`readOnlyHint: true`, `destructiveHint: false`, `idempotentHint: true`), so it adds little new behavioral disclosure. The mode-specific output hints (e.g., "plain text plus a block count") offer minor added context, but no deeper risks, auth needs, or side effects are described.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is three short clauses, front-loaded with the core action, and every sentence earns its place. It conveys purpose, read-only safety, and a sibling alternative with no wasted words.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the presence of an output schema and strong annotations, the description is largely complete: it states what the tool extracts, confirms it is non-destructive, and offers a comparison alternative. The only notable gap is the confusing `options` parameter, which is not addressed, but the schema's 100% coverage mitigates this.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
The schema already covers all parameters with 100% description coverage, so the baseline is 3. The description adds useful meaning to the `mode` parameter by giving examples (author/title, paragraphs/sheets/slides). However, it does not clarify the `options` parameter, whose schema references compare_documents-specific actions, so the description doesn't fully disambiguate all parameters.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description uses a specific verb and resource: "Read data from a document," and enumerates the three extraction modes (metadata, text, structure) with concrete examples. It also distinguishes itself from the sibling `compare_documents` tool by saying "To compare two documents, use compare_documents."
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
It clearly states when to use the tool: when you need to read metadata, text, or structure. It gives an explicit alternative for a different use case (comparison), and the read-only declaration makes it clear this is not for editing or modifying documents.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
fill_templateA
Fill {{key}} placeholders in a template document with data, expanding {{#each list}} loops and {{#if}}/{{#unless}} conditions. Writes a NEW file to output_path (default: _filled.); the input template is left unchanged. Run validate_template first to catch missing keys early.
| Name | Required | Description | Default |
|---|---|---|---|
| data | Yes | Key-value data to fill placeholders. Supports nested objects. | |
| options | No | Tool options. | |
| output_path | No | Output file path. | |
| template_path | Yes | Path to the template document containing placeholders. |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Beyond the annotations (which only mark non-readOnly, non-destructive), the description adds key behavioral context: it writes a NEW file, leaves input unchanged, and expands specific template constructs. This helps the agent understand side effects. It doesn't mention whether it can overwrite an existing output_path, but overall it adds meaningful behavior beyond annotations.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is two sentences, front-loaded with the core verb and resource, then details output behavior and a recommended prerequisite. Every sentence contributes unique information, with no padding. It's concise and well-structured for quick scanning.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool has an output schema (not detailed here) and rich annotations, the description covers the main functionality, output behavior, and a usage tip. It doesn't elaborate on options like dry_run or backup, but those are already in the schema. The description is sufficient for an agent to decide when and how to call it, though it could mention how it handles nested data structures more explicitly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema coverage is 100% (all parameters are described in the schema). The description adds value by specifying the default output_path behavior ('<template>_filled.<ext>'), which the schema does not mention. It also indirectly explains how data interacts with loops/conditions, though that's not parameter-specific. This raises the score above baseline from schema.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description clearly states the tool fills placeholders in a template document, expanding loops and conditions, and writes a new file. It distinguishes itself from siblings like edit_office_document by explicitly noting it creates a new file and leaves the input unchanged. The verb 'fill' plus specific placeholder syntax and condition handling make the purpose unambiguous.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description advises running validate_template first to catch missing keys early, providing a clear sequencing guideline. It does not explicitly mention when not to use this tool versus alternatives like edit_office_document or convert_document, but the 'new file' behavior implicitly contrasts with in-place modification. The guidance is adequate but lacks explicit alternatives and exclusions.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
validate_templateARead-onlyIdempotent
Validate that all {{placeholder}} variables, {{#each}} loops, and {{#if}}/{{#unless}} conditions in a template can be resolved against data, reporting missing keys. Read-only — never modifies the file. Call this BEFORE fill_template to catch missing keys early.
| Name | Required | Description | Default |
|---|---|---|---|
| data | Yes | Key-value data to validate against placeholders. | |
| template_path | Yes | Path to the template document (.docx/.xlsx). |
Output Schema
| Name | Required | Description |
|---|---|---|
No output parameters | ||
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
Annotations already declare readOnlyHint=true, idempotentHint=true, and destructiveHint=false, so the safety profile is covered. The description adds useful behavioral context by specifying what constructs are validated and that it reports missing keys, going slightly beyond the structured annotations without contradicting them.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is two sentences with no filler. It front-loads the core purpose, states the safety guarantee, and gives actionable usage guidance. Every sentence earns its place.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool has rich annotations, a complete input schema, an output schema, and a clear relationship to fill_template, the description covers purpose, timing, and safety sufficiently. No critical context is missing for an agent to select and invoke this tool correctly.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, with both parameters already described clearly in the input schema: data is 'Key-value data to validate against placeholders' and template_path is 'Path to the template document (.docx/.xlsx).' The description does not add significant parameter-level detail beyond what the schema provides, so the baseline of 3 is appropriate.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description uses a specific verb ('Validate') with a clear resource ('template') and enumerates exactly what is checked ({{placeholder}} variables, {{#each}} loops, {{#if}}/{{#unless}} conditions). It also states the outcome ('reporting missing keys') and distinguishes itself from fill_template by name.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
The description explicitly says 'Call this BEFORE fill_template to catch missing keys early,' giving clear when-to-use guidance and naming the related alternative tool. It also clarifies the read-only nature, reinforcing that this is a pre-flight check rather than a mutation operation.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections. Dates show when Glama detected each change.
7 tool updates
v0.8.0- Added
analyze_excel_data - Added
create_excel_workbook - Changed
create_office_document16 fields changed- added
Input schema / $defs / ContentBlock / properties / cellAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target cell reference, e.g. \"A1\" (Excel). Used by formula/hyperlink/write_cell.", + "title": "Cell" +} - added
Input schema / $defs / ContentBlock / properties / chart_dataAdded value: +{ + "anyOf": [ + { + "items": { + "items": {}, + "type": "array" + }, + "type": "array" + }, + { + "type": "null" + } + ], + "default": null, + "description": "PPT chart data: first row series names, then [category, *values] rows.", + "title": "Chart Data" +} - added
Input schema / $defs / ContentBlock / properties / chart_data_rangeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel chart data range, e.g. \"Sheet1!A1:B10\".", + "title": "Chart Data Range" +} - added
Input schema / $defs / ContentBlock / properties / chart_typeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Chart type: bar, line, pie, area, doughnut, scatter (Excel/PPT).", + "title": "Chart Type" +} - added
Input schema / $defs / ContentBlock / properties / conditional_formatAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel conditional format spec, e.g. \"B2:B100=color_scale\" or \"C1:C5=cell_is:greaterThan:20\".", + "title": "Conditional Format" +} - added
Input schema / $defs / ContentBlock / properties / data_validationAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel data validation spec, e.g. \"C2:C50=list:yes,no\" or \"B1:B10=whole:1:100\".", + "title": "Data Validation" +} - added
Input schema / $defs / ContentBlock / properties / formulaAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel formula string, e.g. \"=SUM(B1:B10)\". Requires cell.", + "title": "Formula" +} - added
Input schema / $defs / ContentBlock / properties / freezeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel freeze panes cell, e.g. \"A2\".", + "title": "Freeze" +} - added
Input schema / $defs / ContentBlock / properties / hyperlinkAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "URL to hyperlink the cell given by `cell` to (Excel).", + "title": "Hyperlink" +} - added
Input schema / $defs / ContentBlock / properties / named_rangeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel named range spec, e.g. \"MyRange=A1:B2\".", + "title": "Named Range" +} - added
Input schema / $defs / ContentBlock / properties / notesAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "PPT speaker notes text for the target slide.", + "title": "Notes" +} - added
Input schema / $defs / ContentBlock / properties / number_formatAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel number format spec, e.g. \"A1:A10=0.00%\".", + "title": "Number Format" +} - added
Input schema / $defs / ContentBlock / properties / sheet_nameAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target worksheet name (Excel). Routes write/formula/style to this sheet.", + "title": "Sheet Name" +} - added
Input schema / $defs / ContentBlock / properties / slide_indexAdded value: +{ + "anyOf": [ + { + "type": "integer" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target slide index (0-based) for PPT content (tables/charts/text). None = current/last slide.", + "title": "Slide Index" +} - added
Input schema / $defs / ContentBlock / properties / slide_layoutAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "PPT slide layout name or index (applied on slide creation).", + "title": "Slide Layout" +} - added
Input schema / $defs / ContentBlock / properties / transitionAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "PPT slide transition name, e.g. \"fade\".", + "title": "Transition" +}
- Added
create_presentation - Added
edit_excel_workbook - Changed
edit_office_document29 fields changed- changed
Input schema / $defs / EditOperation / properties / action / descriptionPrevious value: -"Operation type:\n- \"replace\": find old_text, replace with new_text (regex optional)\n- \"delete\": remove text matching target\n- \"modify\": find old_text, replace with new_text\n- \"style\": apply a style string to the whole document\n- \"add\": append text at a column (Excel)\n- \"clear\": clear document content"New value: +"Operation type. Only the listed fields are meaningful per action:\n- \"replace\": old_text, new_text, regex (all doc types)\n- \"delete\": target, regex (all doc types)\n- \"modify\": old_text, new_text (all doc types)\n- \"style\": style, apply_all (all doc types)\n- \"add\": text, column (Excel append; Word/PPT append paragraph)\n- \"clear\": (no fields)\n- \"write_cell\": cell, text (value), sheet_name, style, is_formula (Excel)\n- \"set_formula\": cell, formula, sheet_name (Excel)\n- \"freeze_panes\": range (Excel)\n- \"add_chart\": chart_type + chart_data_range (Excel) OR chart_type + chart_data (PPT), sheet_name (Excel)\n- \"conditional_format\": conditional_format OR range + cell_is opts (Excel)\n- \"data_validation\": data_validation OR range (Excel)\n- \"add_table\": rows (first row = header), slide_index (PPT)\n- \"add_picture\": path, slide_index (PPT)\n- \"add_shape\": slide_index, fill, line, shape_type (PPT)\n- \"apply_layout\": slide_index, layout (PPT)\n- \"set_transition\": slide_index, transition (PPT)\n- \"add_notes\": slide_index, notes (PPT)\n- \"add_slide\": layout (PPT, optional)\n- \"sort\": range, key_columns, orders, order (Excel)\n- \"add_sheet\": sheet_name (Excel)\n- \"set_range_style\": range, style (Excel)\n- \"number_format\": number_format (\"RANGE=FORMAT\", e.g. \"A1:A10=0.00%\")\nUnused fields for an action are ignored." - changed
Input schema / $defs / EditOperation / properties / action / enumPrevious value: -[ - "replace", - "delete", - "modify", - "style", - "add", - "clear" -]New value: +[ + "replace", + "delete", + "modify", + "style", + "add", + "clear", + "write_cell", + "set_formula", + "freeze_panes", + "add_chart", + "conditional_format", + "data_validation", + "add_table", + "add_picture", + "add_shape", + "apply_layout", + "set_transition", + "add_notes", + "add_slide", + "sort", + "add_sheet", + "set_range_style", + "number_format" +] - added
Input schema / $defs / EditOperation / properties / cellAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target cell reference, e.g. \"A1\" (Excel).", + "title": "Cell" +} - added
Input schema / $defs / EditOperation / properties / cf_typeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Conditional format type (color_scale/data_bar/cell_is/formula).", + "title": "Cf Type" +} - added
Input schema / $defs / EditOperation / properties / chart_dataAdded value: +{ + "anyOf": [ + { + "items": { + "items": {}, + "type": "array" + }, + "type": "array" + }, + { + "type": "null" + } + ], + "default": null, + "description": "PPT chart data for add_chart.", + "title": "Chart Data" +} - added
Input schema / $defs / EditOperation / properties / chart_data_rangeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel chart data range for add_chart.", + "title": "Chart Data Range" +} - added
Input schema / $defs / EditOperation / properties / chart_typeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Chart type for add_chart.", + "title": "Chart Type" +} - added
Input schema / $defs / EditOperation / properties / dv_typeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Data validation type (list/whole/decimal/date/text_length).", + "title": "Dv Type" +} - added
Input schema / $defs / EditOperation / properties / fillAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Shape fill color hex (add_shape).", + "title": "Fill" +} - added
Input schema / $defs / EditOperation / properties / formulaAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Excel formula string (set_formula).", + "title": "Formula" +} - added
Input schema / $defs / EditOperation / properties / formula1Added value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Data validation formula1 (e.g. \"yes,no\").", + "title": "Formula1" +} - added
Input schema / $defs / EditOperation / properties / formula2Added value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Data validation formula2 (e.g. upper bound).", + "title": "Formula2" +} - added
Input schema / $defs / EditOperation / properties / is_formulaAdded value: +{ + "anyOf": [ + { + "type": "boolean" + }, + { + "type": "null" + } + ], + "default": null, + "description": "write_cell only: true stores text as a formula (must start with \"=\"), false forces a literal string even when it starts with \"=\", omitted keeps the automatic behaviour.", + "title": "Is Formula" +} - added
Input schema / $defs / EditOperation / properties / key_columnsAdded value: +{ + "anyOf": [ + { + "items": { + "type": "integer" + }, + "type": "array" + }, + { + "type": "null" + } + ], + "default": null, + "description": "0-based sort key columns (sort).", + "title": "Key Columns" +} - added
Input schema / $defs / EditOperation / properties / layoutAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Slide layout name/index (apply_layout).", + "title": "Layout" +} - added
Input schema / $defs / EditOperation / properties / lineAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Shape line color hex (add_shape).", + "title": "Line" +} - added
Input schema / $defs / EditOperation / properties / notesAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Speaker notes text (add_notes).", + "title": "Notes" +} - added
Input schema / $defs / EditOperation / properties / number_formatAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Number format spec \"RANGE=FORMAT\" (number_format).", + "title": "Number Format" +} - added
Input schema / $defs / EditOperation / properties / orderAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Single sort order asc/desc (sort).", + "title": "Order" +} - added
Input schema / $defs / EditOperation / properties / ordersAdded value: +{ + "anyOf": [ + { + "items": { + "type": "string" + }, + "type": "array" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Per-key sort orders asc/desc (sort).", + "title": "Orders" +} - added
Input schema / $defs / EditOperation / properties / pathAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Image path (add_picture).", + "title": "Path" +} - added
Input schema / $defs / EditOperation / properties / rangeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Cell range for freeze_panes/conditional_format/data_validation/add_chart (Excel).", + "title": "Range" +} - added
Input schema / $defs / EditOperation / properties / rowsAdded value: +{ + "anyOf": [ + { + "items": { + "items": {}, + "type": "array" + }, + "type": "array" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Table rows for add_table (PPT). First row is the header.", + "title": "Rows" +} - added
Input schema / $defs / EditOperation / properties / shape_typeAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Autoshape type (add_shape), e.g. rectangle/oval/arrow/line.", + "title": "Shape Type" +} - added
Input schema / $defs / EditOperation / properties / sheet_nameAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target worksheet (Excel).", + "title": "Sheet Name" +} - added
Input schema / $defs / EditOperation / properties / slide_indexAdded value: +{ + "anyOf": [ + { + "type": "integer" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Target slide index (PPT add_table/add_chart).", + "title": "Slide Index" +} - changed
Input schema / $defs / EditOperation / properties / text / anyOfPrevious value: -[ - { - "type": "string" - }, - { - "type": "null" - } -]New value: +[ + { + "type": "string" + }, + { + "type": "integer" + }, + { + "type": "number" + }, + { + "type": "boolean" + }, + { + "type": "null" + } +] - changed
Input schema / $defs / EditOperation / properties / text / descriptionPrevious value: -"Text to add (add)."New value: +"Text to add (add) or scalar value to write (write_cell; formulas start with \"=\")." - added
Input schema / $defs / EditOperation / properties / transitionAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Transition name (set_transition).", + "title": "Transition" +}
- Added
edit_presentation
7 tool updates
- Changed
compare_documents4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "compare_documentsDictOutput", + "type": "object" +}
- Changed
convert_document4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "convert_documentDictOutput", + "type": "object" +}
- Changed
create_office_document4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "create_office_documentDictOutput", + "type": "object" +}
- Changed
edit_office_document4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "edit_office_documentDictOutput", + "type": "object" +}
- Changed
extract_document_data4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "extract_document_dataDictOutput", + "type": "object" +}
- Changed
fill_template4 fields changed- added
Input schema / $defs / ToolOptions / properties / actionAdded value: +{ + "anyOf": [ + { + "enum": [ + "compare", + "snapshot", + "list_snapshots", + "restore" + ], + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Sub-operation for compare_documents: \"compare\" (default) diffs two documents; \"snapshot\" records path_a state; \"list_snapshots\" lists recorded snapshots; \"restore\" writes a snapshot back to path_b.", + "title": "Action" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_dirAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Directory to store/read document snapshots (default: ~/.tianshang-scribe/snapshots/).", + "title": "Snapshot Dir" +} - added
Input schema / $defs / ToolOptions / properties / snapshot_idAdded value: +{ + "anyOf": [ + { + "type": "string" + }, + { + "type": "null" + } + ], + "default": null, + "description": "Snapshot identifier used by the restore action.", + "title": "Snapshot Id" +} - changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "fill_templateDictOutput", + "type": "object" +}
- Changed
validate_template1 field changed- changed
Output schema / (root)Previous value: -nullNew value: +{ + "additionalProperties": true, + "title": "validate_templateDictOutput", + "type": "object" +}
7 tool updates
v0.3.0- First observed
compare_documents - First observed
convert_document - First observed
create_office_document - First observed
edit_office_document - First observed
extract_document_data - First observed
fill_template - First observed
validate_template
TDQS
Multiple tools have unclear boundaries: create_office_document can produce Excel and PowerPoint files, making its purpose overlap directly with create_excel_workbook and create_presentation. Similarly, the legacy edit_office_document duplicates the typed editors, and extract_document_data overlaps with analyze_excel_data for reading. The long cross-referencing descriptions try to compensate, but an agent could easily pick the wrong tool for a create-xlsx or create-pptx task.
The server follows a consistent verb_noun snake_case convention throughout (create_*, edit_*, analyze_*, extract_*, convert_*, compare_*, fill_*, validate_*), making the intent of each tool predictable. Minor deviations exist: compare_documents is plural while others are singular, and objects alternate granularity (excel_workbook vs office_document vs presentation vs template) rather than using one unifying noun. Overall the pattern is easy to learn and apply.
At 12 tools, the server sits comfortably in the well-scoped 3-15 range for a document-processing domain covering Word, Excel, PowerPoint, and templates. The count is substantial enough to cover the full workflow without feeling bloated or sparse.
The server covers the document lifecycle well: create (excel/office/presentation), read (extract/analyze), update (edit variants), plus convert, compare, snapshots, and template fill/validate. Major gaps include the lack of delete/remove operations (e.g., no remove_sheet or delete_slide) and comparison that only supports Word, with Excel/PPT hitting UNSUPPORTED_FORMAT errors. These are workable gaps rather than dead ends.
Maintenance
Related MCP Connectors
Document-to-Markdown MCP server — convert PDF, Office and HTML into LLM-ready Markdown.
Generate PDF/DOCX/XLSX/PPTX from templates+JSON. Convert Office/HTML/MD to PDF. Universal templating
Generate PDF, Word (.docx) and PowerPoint (.pptx) documents from Markdown over MCP.
Use your own Word templates to convert Markdown → DOCX/PDF/HTML from any MCP-compatible AI.
Related MCP Servers
- AlicenseBqualityDmaintenanceA universal MCP server for document processing, conversion, and automation. Handle PDF, DOCX, HTML, Markdown, and more through a unified API and toolset.1333139MIT
- AlicenseAqualityDmaintenanceMCP server for Word document (.docx) creation and manipulation — the production-grade document automation tool for AI agents.963MIT
- AlicenseAqualityBmaintenanceMCP server for reading, writing, editing, formatting, and exporting Microsoft Office documents (Word, Excel, PowerPoint) via stdio JSON-RPC, with 47 tools and cross-platform support.47MIT
- AlicenseCqualityDmaintenanceA unified MCP server for document processing that enables creating, editing, and converting Word documents (DOCX), PDFs, Markdown, and images, with support for templates, formatting, and batch operations.100MIT
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/Tianshang301/TianshangScribe'
If you have feedback or need assistance with the MCP directory API, please join our Discord server