Skip to main content
Glama

Ask Pipeworx Beta

ask_pipeworx_beta
Read-onlyIdempotent

Beta version of ask_pipeworx: identical universal router (same 5,798 tools, same arguments, same response shape) with candidate routing improvements enabled live whenever one is under test. No candidate is active right now (the last was retired on outcome evidence 2026-07-26), so this currently matches ask_pipeworx exactly. Use it exactly like ask_pipeworx when you want the newest routing; results are compared against the stable router to decide what merges. Falls back to nothing — this IS a full working router, just the experimental edge.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
qNoAlias for question.
textNoAlias for question.
inputNoAlias for question.
queryNoAlias for question.
promptNoAlias for question.
questionYesYour question or request in natural language. Accepts query, q, prompt, text, input as aliases.

Schema Changelog

Changes observed during successful MCP inspections. Dates show when Glama detected each change.

  1. Added

TDQS

A4.6/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Beyond the read-only and idempotent annotations, the description discloses the experimental nature, live enablement of candidate routing improvements, the recent retirement of the last candidate, and the comparison process against the stable router. It also explicitly states there is no fallback path, so the agent knows this is a fully functional router. This is rich behavioral context with no contradiction to the annotations.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is four sentences and every sentence adds unique value: identity, current experimental state, usage direction, and clarification that it is not a fallback. The key definition is front-loaded before the beta nuance. The density is appropriate for a tool whose main complexity is its relationship to the stable version.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The description covers identity, current behavior, usage, and experimental comparison, and references the same response shape as ask_pipeworx, which is sufficient given the sibling context. It does not independently describe the exact response format, but no output schema exists and the stable sibling is available as a reference. This is a minor gap.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, with every property described as an alias for the required 'question' parameter, so the schema already carries the parameter semantics. The description only repeats that the tool takes the same arguments as ask_pipeworx and does not add meaningful details beyond what the input schema provides. Baseline 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description immediately identifies the tool as the beta version of ask_pipeworx and as an identical universal router with the same 5,798 tools, arguments, and response shape. This clearly names the operation and differentiates the beta from the stable router via candidate routing improvements. It is easy for an agent to understand what this tool is and how it relates to its closest sibling.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It gives an explicit usage condition: use it exactly like ask_pipeworx when you want the newest routing, and notes that results are compared against the stable router. It also states that no candidate is currently active, so behavior currently matches ask_pipeworx exactly, and clarifies that this is a full working router rather than a fallback. This is clear when-to-use guidance with the stable sibling named as the alternative.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

TDQS

A3.5/5.0
Disambiguation2/5

Several tool clusters have unclear boundaries: ask_pipeworx, ask_pipeworx_beta, ask_pipeworx_grounded, and deep_research all route to the same 5,724 tools with only subtle behavioral differences, and bet_research, polymarket_edges, polymarket_arbitrage, and polymarket_kalshi_spread heavily overlap around prediction-market opportunity discovery. entity_profile, compare_entities, and recent_changes also fan out across the same SEC/news/patent sources, making selection ambiguous for agents.

Naming Consistency2/5

Naming mixes multiple conventions: snake_case verb_noun for odds tools (get_events, list_sports), vendor-prefixed clusters (ask_pipeworx_*, pipeworx_*, polymarket_*), and a few reversed noun-verb names like bet_research. CamelCase is used in ai_visibility_check and generate_llms_txt adds another style. Only the polymarket_* and pipeworx_* families are internally consistent, but the overall pattern is chaotic.

Tool Count2/5

37 tools is well beyond the typical well-scoped server, and the count feels inflated by unrelated meta-tools (suggest_questions, discover_tools, pipeworx_feedback, pipeworx_trending, generate_llms_txt, scan_dependency, remember/recall/forget) that have nothing to do with the server's stated 'Odds Api' purpose. The actual odds surface is only ~5 tools, so the vast majority of the catalog is off-scope padding.

Completeness3/5

The core odds domain is well covered: list_sports, get_events, get_odds, get_event_odds, and get_scores form a coherent lifecycle, plus quota introspection. However, for the server's actual broad-research scope there are noticeable gaps (e.g., no direct single-filing fetch tool despite heavy SEC coverage, a lone npm-dependency tool with no surrounding ecosystem, and no historical/past-odds endpoint), and the heterogeneous domains make completeness uneven.