Skip to main content
Glama

Ask Pipeworx — Grounded

ask_pipeworx_grounded
Read-onlyIdempotent

Hallucination-resistant answer mode for high-stakes reads. Same routing as ask_pipeworx — picks the right tool from 5,798 across 1517 sources, fills arguments, fetches the data — then EXTRACTS the answer using ONLY what the tool result contains. Returns {answer, evidence (verbatim quote), confidence, source, fetched_at, refusal_reason:null} on success, OR an explicit refusal {answer:null, refusal_reason:"not_in_source"|"no_tool_match"|"tool_error"|"data_truncated"|"llm_error"} when the data doesn't directly answer. Use whenever an answer will be quoted, cited, or acted on, and the agent must not invent facts (financial verdicts, legal claims, medical lookups, public statements). Costs one extra LLM call vs ask_pipeworx — prefer ask_pipeworx for casual lookups.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
qNoAlias for question.
textNoAlias for question.
inputNoAlias for question.
queryNoAlias for question.
promptNoAlias for question.
questionYesYour question in natural language. Accepts query, q, prompt, text, input as aliases.

Schema Changelog

Changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. First observed

TDQS

A4.7/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Even though annotations already mark this as readOnly/idempotent/non-destructive, the description goes beyond them: it discloses the grounded extraction behavior, the exact success and refusal return shapes, the refusal_reason enum, and the fact that it will refuse rather than guess. It also discloses the hidden operational cost of an extra LLM call. These are valuable behavioral traits not visible in annotations.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is dense but every sentence contributes: purpose, routing, behavior, return/refusal contract, usage, and tradeoff against ask_pipeworx. The most important qualifier ('Hallucination-resistant answer mode for high-stakes reads') is front-loaded, and the rest follows logically without repetition.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

With no output schema, the description fully compensates by specifying the return object and refusal object in detail. It covers when to use, when not to use, key behavioral constraints, the cost implication, and the exact failure modes. An agent has everything needed to invoke it correctly.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100% and the one functional parameter, question, is fully documented with aliases. The description adds no semantic detail about how to phrase the question or what makes a good query. Under the baseline rule for high schema coverage, 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description opens with a precise value proposition: 'Hallucination-resistant answer mode for high-stakes reads.' It clearly distinguishes this tool from ask_pipeworx by stating it extracts answers using ONLY the tool result and returns explicit refusals when the data doesn't directly answer. This gives an agent a clear, specific verb+resource+behavior profile.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description explicitly states when to use it — 'whenever an answer will be quoted, cited, or acted on, and the agent must not invent facts' — and when not to, with 'prefer ask_pipeworx for casual lookups.' It even names the alternative tool and explains the cost tradeoff (one extra LLM call). This is ideal routing guidance.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.9/5.0
Disambiguation2/5

Multiple tools overlap heavily: ask_pipeworx, ask_pipeworx_beta (currently identical), and ask_pipeworx_grounded share the same routing core, while bet_research, polymarket_edges, polymarket_arbitrage, and polymarket_fill_risk all target prediction-market edge detection. The descriptions are detailed, but an agent must read a lot of nuance to avoid selecting the wrong tool.

Naming Consistency3/5

All names are lowercase snake_case and readable, but the convention is mixed: verb_noun names like encode_html and resolve_entity sit alongside bare verbs like remember and forget, and noun-phrase names like entity_profile, recent_changes, and polymarket_fill_risk. The ask_pipeworx_* and polymarket_* families are internally consistent, but there is no single predictable pattern across the whole set.

Tool Count2/5

At 33 tools this exceeds the 25+ threshold for a heavy set. The bloat is especially noticeable because the server is named Htmlentities yet only encode_html and decode_html relate to that purpose; the rest are unrelated Pipeworx research, prediction-market, memory, and subscription utilities.

Completeness4/5

The Pipeworx surface is broadly complete: ask/grounded/deep_research/discover/suggest cover data access, entity_profile/compare_entities/recent_changes/validate_claim cover entity workflows, and subscriptions and memory have create/list/delete lifecycles. Minor gaps such as no update operation for subscriptions or memories are workable, and encode/decode fully covers the literal Htmlentities purpose.