Methodology from Anthropic research (Apr 29, 2026) for benchmarking LLM bioinformatics research capabilities.
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill biomysterybench-evaluation --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Biomysterybench Evaluation?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-biomysterybench-evaluation-ai-collection)More formats (shields.io, HTML) on the badges page.
---
name: biomysterybench-evaluation
category: ai_collection
tags: [anthropic, bioinformatics, benchmark, evaluation, science]
---
# BioMysteryBench: Evaluating AI Bioinformatics Capabilities
Methodology from Anthropic research (Apr 29, 2026) for benchmarking LLM bioinformatics research capabilities.
## Core Concept
Benchmark designed to evaluate Claude's ability to perform bioinformatics research tasks — measuring scientific capability in a domain with potential dual-use concerns.
## Methodology
- Series of bioinformatics challenges ranging in difficulty
- Tests ability to analyze biological sequence data, interpret results, and draw conclusions
- Provides framework for measuring AI capabilities in potentially sensitive scientific domains
**Activation**: biomysterybench, bioinformatics, AI science, biology benchmark, research capabilities
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!
Ultra-compressed communication mode. Cuts token usage ~75% by speaking like caveman while keeping full technical accuracy. Supports intensity levels: lite, full (default), ultra, wenyan-lite, wenyan-full, wenyan-ultra. Use when user says "caveman mode", "talk like caveman", "use caveman", "less tokens", "be brief", or invokes /caveman. Also auto-triggers when token efficiency is requested.