Skip to main content
Glama

Server Configuration

Describes the environment variables required to run the server.

NameRequiredDescriptionDefault
SWAYSOCKNoThe Sway socket path to connect to for resolving output geometry. If not set, swaymsg uses its default socket.
SESHAT_PIPER_MODELNoPath or name of the piper voice model to use for offline narration. If not set, edge-tts or a default piper voice may be used.

Instructions

Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.

This server publishes no instructions, or was last inspected before Glama recorded them.

Capabilities

Features and capabilities supported by this server

Protocol revision2025-11-25

CapabilityDetails
tools
{}

Tools

Functions exposed to the LLM to take actions

NameDescription
seshat_infoB

Inspect the recorder: active take, stream and recording directories, the streams that will be ingested, the capture outputs, and the binaries narration depends on.

recording_startA

Start recording an active output (or a region within it) into a silent artifact: H.264 MP4 by default, or AV1 WebM, or a constrained GIF fallback. Returns immediately; call recording_stop to finish, then poll recording_status until the phase is completed or failed. The pointer cursor is always included. Narration anchors come from the timeline streams ingested for this take, so the tool server that drives the demonstration has to publish one.

recording_statusB

Report the take lifecycle phase (idle, recording, stopping, processing, narrating, completed, failed) plus live progress or final artifact metadata.

recording_stopA

Gracefully stop the active recording and begin finalizing the requested artifact. Returns the stopping/processing state; poll recording_status for the final artifact path and metadata.

recording_timelineA

Return the take's monotonic event timeline: every event ingested from the take's timeline streams, filtered to the capture window, with a recording-relative t_ms, the emitting source, and a compact payload. Use event ids as narration anchors. A sidecar JSON copy is written next to the artifact on completion.

recording_voiceoverA

Attach a scripted narration track to a completed take: the caller supplies the prose, the server synthesizes speech, aligns segments to timeline anchors (or to an explicit at_ms), and muxes the audio over the existing video stream, optionally burning styled captions. Starts an async 'narrating' phase; poll recording_status. GIF cannot carry audio.

recording_scenesA

Optional, approximate fallback anchors for a completed take: ffmpeg scene-cut timestamps, with optional keyframe OCR via tesseract. Use when no timeline stream was published for the take; scene cuts are secondary evidence, never the sync source.

Prompts

Interactive templates invoked by user choice

NameDescription

No prompts

Resources

Contextual data attached and managed by the client

NameDescription

No resources

TDQS

A3.9/5.0

Scored across 7 tools

Disambiguation5/5

Each tool has a distinct role: start/stop/status for lifecycle, timeline for event anchors, scenes for fallback visual anchors, voiceover for narration, and seshat_info for environment inspection. No two tools appear interchangeable or likely to be confused.

Naming Consistency4/5

Six of seven tools use the recording_ prefix, which is consistent and predictable. The single seshat_info tool breaks the pattern, but the deviation is minor and still readable.

Tool Count5/5

Seven tools is well-scoped for a recording and narration lifecycle. Each tool earns its place with no redundant or bloated operations.

Completeness4/5

The core lifecycle is covered: start, stop, status, timeline, fallback scenes, voiceover, and environment info. Minor gaps like listing or deleting completed artifacts exist, but they are not essential for the stated recording/narration purpose.

Maintenance

ActivityMaintained
ResponsivenessNo issues