Skip to main content
Glama
Christofer566

Technical Answer Validator

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
TAV_HOSTNoHost to bind the service to. Defaults to 127.0.0.1.127.0.0.1
TAV_PORTNoPort to listen on. Defaults to 8080.8080
TAV_API_KEYNoA long random secret at least 24 characters. Required unless TAV_API_KEYS is set.
TAV_API_KEYSNoJSON object mapping each client ID to the generated SHA-256 digest. When set, it takes precedence over TAV_API_KEY.

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Features and capabilities supported by this server

Protocol revision2025-11-25

CapabilityDetails
tools
{
  "listChanged": false
}
prompts
{
  "listChanged": false
}
resources
{
  "subscribe": false,
  "listChanged": false
}
experimental
{}

Tools

Functions exposed to the LLM to take actions

NameDescription
evaluate_answerA

Check an answer against caller-provided concepts, synonyms, and numeric requirements.

Rubric fields: required_concepts (1-50 strings), optional accepted_synonyms keyed by concept, optional numeric_requirements [{value, unit?, tolerance?}], and optional required_count. No question bank is bundled. Review the returned result; it is not an official exam grade.

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources

TDQS

A3.8/5.0

Scored across 1 tool

Disambiguation5/5

With only a single tool, there is no possibility of overlap or misselection. The tool's purpose—validating an answer against caller-supplied concepts, synonyms, and numeric requirements—is unambiguous.

Naming Consistency5/5

The lone tool uses a clear snake_case verb_noun convention (evaluate_answer) that is fully self-consistent. There are no competing names or styles to create inconsistency.

Tool Count3/5

A single tool is at the thin end of the scale given the rubric explicitly treats 1-2 tools as borderline. The validation domain is narrow enough that one monolithic tool is defensible, but there is no granularity for separate concerns.

Completeness4/5

The tool covers the core validation surface well: required concepts, synonyms, numeric requirements with tolerance, and required_count, with a sensible disclaimer about grade authority. Some gaps exist (e.g., no batch evaluation or question-bank management), but agents can work around them.

Maintenance

ActivityMaintained
ResponsivenessNo issues