omni-cache
GitHub管理LLM响应缓存,支持查看统计、清除条目、配置TTL策略及控制语义相似度缓存阈值。
Trigger Scenarios
Install
npx skills add diegosouzapw/OmniRoute --skill omni-cache -g -y
SKILL.md
Frontmatter
{
"name": "omni-cache",
"description": "Manage the LLM response cache. View cache statistics, clear entries, configure TTL policies, and control semantic-similarity caching thresholds."
}
Overview
Manage the LLM response cache. View cache statistics, clear entries, configure TTL policies, and control semantic-similarity caching thresholds.
Authentication
All requests require a valid Bearer token or session cookie. Obtain a token via POST /api/auth/login or configure REQUIRE_API_KEY=false for local development.
Endpoints
GET /api/cache
Get cache statistics
curl https://localhost:20128/api/cache \
-H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache
Clear all caches
curl -X DELETE https://localhost:20128/api/cache \
-H "Authorization: Bearer $OMNIROUTE_TOKEN"
GET /api/cache/stats
Get detailed cache statistics
Returns detailed statistics for all cache layers.
curl https://localhost:20128/api/cache/stats \
-H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache/stats
Clear cache statistics
curl -X DELETE https://localhost:20128/api/cache/stats \
-H "Authorization: Bearer $OMNIROUTE_TOKEN"
Payloads
See the full OpenAPI specification at GET /api/openapi/spec or docs/openapi.yaml for detailed request/response schemas.
Version History
- 1cafd32 Current 2026-07-25 11:47


