context-shunt
GitHub通过调用廉价模型处理大文件读取,避免原始内容进入主上下文以节省Token。适用于跨文件查询事实或理解代码逻辑的场景,而非直接编辑。
Trigger Scenarios
Install
npx skills add alinaqi/maggy --skill context-shunt -g -y
SKILL.md
Frontmatter
{
"name": "context-shunt",
"effort": "low",
"description": "Offload large \/ multi-file reads to a cheap worker model so raw files never enter Claude's context (token savings)",
"when-to-use": "When answering a question that requires reading large or many files, reviewing logs, or scanning generated output — anything you don't need to edit",
"user-invocable": false
}
Context Shunt — Read Cheap, Keep Context Small
Reading big files into context is the most expensive thing an agent does for the
least reasoning value. The shunt hands those reads to a cheap worker model
(bulk-read) that answers a question about the files and returns a compact
summary. The raw bytes never enter this session's context.
This is orthogonal to whole-turn routing (srooter / route-task): those pick the
model for the turn; the shunt trims what a tool call pulls into context when
the turn is legitimately here.
Decision: read raw, shunt, or graph?
- Editing this exact file — read it raw. You need every line; never edit against a summary.
- A fact/answer across large or many files —
bulk-read "<question>" file.... - A code symbol (function/class/route) —
get_code_snippet(qualified_name): free and exact. - Small file (under threshold) you need in full — read it raw.
A shunt answer is for understanding, not for producing a diff.
Usage
bulk-read "how does token refresh work?" src/auth/session.ts src/auth/refresh.ts
bulk-read "which config keys are read at startup?" $(git ls-files 'config/*.yaml')
bulk-read prints structured bullets citing path:line, or
NOT FOUND IN PROVIDED FILES. A token-savings report goes to stderr.
Configuration
Env vars or ~/.claude/shunt.conf (see templates/shunt.conf):
SHUNT—on/offmaster switch for the PreToolUse hook.SHUNT_MIN_LINES— large-read threshold (default 350).SHUNT_MODE—suggest(default) /block/offfor the hook's large-read action.SHUNT_GRAPH_NUDGE—on/offonce-per-session graph nudge.SHUNT_MODEL— worker command (defaultdeepseek --flash; alsogemini-api --flash-lite,qwen3,glm).
The context-shunt-gate PreToolUse hook enforces the thresholds; this skill tells
you when to reach for bulk-read yourself.
Version History
- 2a98228 Current 2026-09-09 08:47


