All scored servers

cityparity: tool descriptions scored

tools/list taken
(UTC)
Scoring rules
v2

Scores, issues and names only

This page shows how cityparity's tool descriptions scored, the issues found and the names of its tools and parameters. The descriptions are the server's own: read them where the copy came from.

This copy comes from Smithery's registry, which keeps each tool's name, description and input schema, but no output schema, annotations or instructions. Points for return values and for marking destructive tools can come out lower than the server's own tools/list would give.

A server with a public URL can be scored as it is today: score from a URL

servers/bissell-skyler-cityparity
Average score
67.2 / 100Cloudy
Tools
6
Near-identical descriptions
0

Across the server

No server-wide issues.

Tools

list_cities

63 / 100Cloudy

Purpose
20 / 20
When to use
0 / 20
Arguments
20 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
5 / 10
  • Major
    Says nothing about when to use it, or when not to.no_usage_context
  • Minor
    No `required` list, so every argument is optional.no_required

get_inbound_tax_regime

65 / 100Cloudy

Purpose
17 / 20
When to use
10 / 20
Arguments
20 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
0 / 10
  • Minor
    No `required` list, so every argument is optional.no_required

rank_cities

65 / 100Cloudy

Purpose
20 / 20
When to use
5 / 20
Arguments
17 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
5 / 10
  • Minor
    Arguments without a description: weights.vacation, weights.childcare, weights.financial, weights.healthcare, weights.safety_net.param_no_description
  • Minor
    No `required` list, so every argument is optional.no_required
  • Minor
    The meaning of the boolean argument scenario.has_partner is unclear.vague_boolean

compare_cities

70 / 100Clear

Purpose
20 / 20
When to use
10 / 20
Arguments
17 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
5 / 10
  • Minor
    Arguments without a description: score_weights.vacation, score_weights.childcare, score_weights.financial, score_weights.healthcare, score_weights.safety_net.param_no_description
  • Minor
    No `required` list, so every argument is optional.no_required

get_city_summary

70 / 100Clear

Purpose
17 / 20
When to use
15 / 20
Arguments
20 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
0 / 10
  • Minor
    No `required` list, so every argument is optional.no_required

get_safety_net

70 / 100Clear

Purpose
17 / 20
When to use
15 / 20
Arguments
20 / 25
Return value
15 / 15
Constraints and side effects
3 / 10
Examples
0 / 10
  • Minor
    No `required` list, so every argument is optional.no_required

What static scoring cannot tell

This score reads only the text of each description and schema: it says what is written, not what a model will do. Short tools whose names already say what they do, such as browser_close, score low here even when models use them well, and in our measurements current models chose the right tool in one step even from one-line descriptions, as long as the names and schemas differed. What the text cannot show, such as what happens over several steps and whether an added sentence helps or misleads, is what an evaluation measures, on the Pro and Team plans.

Is your server listed here and you would rather it were not? Email support@forecall.dev, and we will take it off within 7 days.

All scored servers