Techniques to maximize context window efficiency, reduce latency, and prevent 'lost in middle' issues through strategic masking and compaction. (triggers: *.log, chat-history.json, reduce tokens, optimize context, summarize history, clear output)
Scanned 5/30/2026
Install via CLI
openskills install ComeOnOliver/skillshub---
name: common-context-optimization
description: "Techniques to maximize context window efficiency, reduce latency, and prevent 'lost in middle' issues through strategic masking and compaction. (triggers: *.log, chat-history.json, reduce tokens, optimize context, summarize history, clear output)"
---
## **Priority: P1 (OPTIMIZATION)**
Manage the Attention Budget. Treat context as a scarce resource.
## 1. Observation Masking (Noise Reduction)
**Problem**: Large tool outputs (logs, JSON lists) flood context and degrade reasoning.
**Solution**: Replace raw output with semantic summaries _after_ consumption.
1. **Identify**: outputs > 50 lines or > 1kb.
2. **Extract**: Read critical data points immediately.
3. **Mask**: Rewrite history to replace raw data with `[Reference: <summary_of_findings>]`.
4. **See**: `references/masking.md` for patterns.
## 2. Context Compaction (State Preservation)
**Problem**: Long conversations drift from original intent.
**Solution**: Recursive summarization that preserves _State_ over _Dialogue_.
1. **Trigger**: Every 10 turns or 8k tokens.
2. **Compact**:
- **Keep**: User Goal, Active Task, Current Errors, Key Decisions.
- **Drop**: Chat chit-chat, intermediate tool calls, corrected assumptions.
3. **Format**: Update `System Prompt` or `Memory File` with compacted state.
4. **See**: `references/compaction.md` for algorithms.
## 3. KV-Cache Awareness (Latency)
**Goal**: Maximize pre-fill cache hits.
- **Static Prefix**: strict ordering: System -> Tools -> RAG -> User.
- **Append-Only**: Avoid inserting into the middle of history if possible.
## References
- [Observation Masking Patterns](references/masking.md)
- [Compaction Algorithms](references/compaction.md)
## Anti-Patterns
- **No raw tool dumps**: Mask large outputs immediately after extracting data.
- **No append-only growth**: Compact every 10 turns to preserve intent over dialogue.
- **No middle insertions**: Append-only history maximizes KV cache hits.
No comments yet. Be the first to comment!
Ultra-compressed communication mode. Cuts token usage ~75% by speaking like caveman while keeping full technical accuracy. Supports intensity levels: lite, full (default), ultra, wenyan-lite, wenyan-full, wenyan-ultra. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman. Also auto-triggers when token efficiency is requested.
Adversarial multi-agent planning skill. Self-orchestrates 5 hostile category members (unspecified-low, unspecified-high, deep, ultrabrain, artistry) via team-mode for ruthless cross-critique debate, distills only the defensible insights, then MANDATORILY hands the distilled insight bundle to the `plan` agent for executable plan formalization. Use when planning needs maximum rigor and surfacing of weak assumptions, blind spots, and over-engineering. Triggers: 'hyperplan', 'hpp', '/hyperplan', ...
**Complete production-ready guide for Google Gemini embeddings API** This skill provides comprehensive coverage of the `gemini-embedding-001` model for generating text embeddings, including SDK usage, REST API patterns, batch processing, RAG integration with Cloudflare Vectorize, and advanced use cases like semantic search and document clustering. ---
Interview, source-challenge, verify, save, and ADR-gate fuzzy coding requests into Codex-ready implementation specs. Use when a feature, bugfix, refactor, migration, repo-wide change, or architecture task needs user-verified requirements, source-backed decisions, durable architecture decisions, acceptance criteria, validation commands, rollout notes, saved spec/ADR files, and a Codex execution prompt. Do not use when already fully specified or when the user wants direct implementation now.
Use when a repo needs CodeGraph plus ast-grep for Codex MCP setup, exploration, impact analysis, structural search, or safe refactor planning.