Retrieve all served vLLM models and LoRA adapters from a specified inference target, with adapters flagged, for governance-grade AIops monitoring and root-cause analysis.
MIT
x402 LLM proxy + data-enriched analysis (17 sources) + TimesFM predictive IoT intelligence.
Check if a task runs locally vs cloud. Save money on calls that don't need cloud inference.