Context Shunt — Read Cheap, Keep Context Small
SkillFiles & storageOffload large / multi-file reads to a cheap worker model so raw files never enter Claude's context (token savings)
Available today. Use it from your connected AI after setup.
No other account needed.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use the Context Shunt — Read Cheap, Keep Context Small skill
What this skill tells your AI
The instructions your AI receives, as published by alinaqi/maggy in skills/context-shunt/SKILL.md and read by ahel’s review.
Reading big files into context is the most expensive thing an agent does for the
least reasoning value. The shunt hands those reads to a cheap worker model
(bulk-read) that answers a question about the files and returns a compact
summary. The raw bytes never enter this session's context.
This is orthogonal to whole-turn routing (srooter / route-task): those pick the
model for the turn; the shunt trims what a tool call pulls into context when
the turn is legitimately here.
Decision: read raw, shunt, or graph?
- Editing this exact file — read it raw. You need every line; never edit against a summary.
- A fact/answer across large or many files —
bulk-read "<question>" file.... - A code symbol (function/class/route) —
get_code_snippet(qualified_name): free and exact. - Small file (under threshold) you need in full — read it raw.
A shunt answer is for understanding, not for producing a diff.
Usage
bulk-read "how does token refresh work?" src/auth/session.ts src/auth/refresh.ts
bulk-read "which config keys are read at startup?" $(git ls-files 'config/*.yaml')
bulk-read prints structured bullets citing path:line, or
NOT FOUND IN PROVIDED FILES. A token-savings report goes to stderr.
Configuration
Env vars or ~/.claude/shunt.conf (see templates/shunt.conf):
SHUNT—on/offmaster switch for the PreToolUse hook.SHUNT_MIN_LINES— large-read threshold (default 350).SHUNT_MODE—suggest(default) /block/offfor the hook's large-read action.SHUNT_GRAPH_NUDGE—on/offonce-per-session graph nudge.SHUNT_MODEL— worker command (defaultdeepseek --flash; alsogemini-api --flash-lite,qwen3,glm).
The context-shunt-gate PreToolUse hook enforces the thresholds; this skill tells
you when to reach for bulk-read yourself.
Signals
- GitHub stars
- 707
- Forks
- 56
- Last commit
- Sep 2026
Advanced
- Catalog kind
- skill
- Gateway key
context-shunt- Source
- github.com/alinaqi/maggy