Benchmark compressed context against summary, direct retrieval, and provider-context baselines without enabling product integration
Scanned 9/3/2026
Install to Claude Code
npx -y skills add jmagly/aiwg --skill long-context-bench --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Long Context Bench?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/jmagly-long-context-bench)More formats (shields.io, HTML) on the badges page.
---
namespace: aiwg
name: long-context-bench
platforms: [all]
description: Benchmark compressed context against summary, direct retrieval, and provider-context baselines without enabling product integration
---
# Long-Context Compression Benchmark
Run `aiwg context-bench run <fixture.json>` on real AIWG task families. Record
quality, exact-recovery failures, latency, memory, and provider-realizable
constraints for all four required strategies.
Product integration remains blocked unless compressed skim plus exact recovery
beats the strongest quality baseline without increasing exact-recovery failures.
Preserve weak and failed results in the report.
@implements #2046
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!