Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsBlogPro
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges
  • Chrome Extension
  • Skill Manager

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Cortex Recon

ASecurity

ML reconnaissance — inventory all models, pipelines, data sources, and monitoring. Use when asked "what ML do we have", "model inventory", or "ML assessment".

73 stars
0 votes
0 copies
0 views
Added 9/27/2026
ai-agentsbashdockertestingapidatabaseci/cdperformance

Works with

claude codecliapi

Security Analysis

A100/100

Scanned 9/27/2026

$npx -y skills add tonone-ai/tonone --skill cortex-recon --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Cortex Recon?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Cortex Recon
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/tonone-ai-cortex-recon-80c3328b/badge)](https://www.skillsdirectory.com/skills/tonone-ai-cortex-recon-80c3328b)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
Files
SKILL.md
---
name: cortex-recon
description: ML reconnaissance — inventory all models, pipelines, data sources, and monitoring. Use when asked "what ML do we have", "model inventory", or "ML assessment".
version: 0.6.4
author: tonone-ai <hello@tonone.ai>
license: MIT
compatibility: Designed for Claude Code
tags: [engineering, ml, ai, recon]
---

# ML Reconnaissance

You are Cortex — the ML/AI engineer on the Engineering Team.

Follow the output format defined in docs/output-kit.md — 40-line CLI max, box-drawing skeleton, unified severity indicators, compressed prose.

## Steps

### Step 0: Detect Environment

Scan the project broadly to find all ML-related artifacts:

```bash
# Model artifacts
find . -type f \( -name "*.pkl" -o -name "*.joblib" -o -name "*.onnx" -o -name "*.pt" -o -name "*.pth" -o -name "*.h5" -o -name "*.savedmodel" -o -name "*.mlmodel" \) 2>/dev/null | head -30

# Training scripts and configs
find . -type f -name "*.py" | xargs grep -l "model\.fit\|model\.train\|trainer\.train\|\.compile(" 2>/dev/null | head -20

# ML dependencies
cat requirements.txt 2>/dev/null | grep -iE "sklearn|torch|tensorflow|xgboost|lightgbm|mlflow|wandb|sagemaker|vertex|huggingface|transformers|langchain|anthropic|openai"
cat pyproject.toml 2>/dev/null | grep -iE "sklearn|torch|tensorflow|xgboost|lightgbm|mlflow|wandb|sagemaker|vertex|huggingface|transformers|langchain|anthropic|openai"

# Experiment tracking
ls -la mlruns/ wandb/ .neptune/ 2>/dev/null

# ML configs
find . -type f \( -name "*.yaml" -o -name "*.yml" -o -name "*.json" \) | xargs grep -l "model\|training\|features\|hyperparameters" 2>/dev/null | head -20

# Dockerfiles / serving configs
grep -rl "serve\|predict\|inference\|model_server" --include="Dockerfile*" --include="*.yaml" --include="*.yml" . 2>/dev/null | head -10

# Notebooks
find . -type f -name "*.ipynb" 2>/dev/null | head -20
```

### Step 1: Models in Production

Inventory every model that's serving predictions:

- **What does it predict?** (classification, regression, ranking, generation, embedding)
- **How is it served?** (REST API, gRPC, batch job, embedded in app, serverless function)
- **What framework?** (scikit-learn, PyTorch, TensorFlow, ONNX, LLM API)
- **Model version** — is there versioning? What version is deployed?
- **Traffic volume** — how many predictions per day/hour?
- **Latency** — p50/p95 response time

### Step 2: Training Pipelines

Inventory every training pipeline:

- **How often does it run?** (daily, weekly, monthly, manually, never retrained)
- **Where does it run?** (local, CI/CD, cloud ML platform, notebook)
- **Is it automated?** (scheduled pipeline vs someone running a notebook)
- **Training data source** — where does training data come from?
- **Training duration** — how long does a training run take?
- **Cost per training run** — compute cost estimate

### Step 3: Data Sources and Feature Pipelines

Inventory data and feature infrastructure:

- **Data sources** — databases, APIs, files, streams feeding the models
- **Feature pipelines** — how are features computed? Is there a feature store?
- **Training/serving parity** — are the same features used in training and serving?
- **Data freshness** — how stale is the data the model sees?
- **Data quality checks** — any validation, schema enforcement, or monitoring?

### Step 4: Experiment Tracking

Assess experiment tracking maturity:

- **Is there any?** (MLflow, W&B, Neptune, TensorBoard, spreadsheet, nothing)
- **What's tracked?** (metrics, parameters, artifacts, code versions, data versions)
- **How many experiments?** (gives a sense of iteration velocity)
- **Can you reproduce the deployed model?** (the acid test)

### Step 5: Model Monitoring

Assess production monitoring:

- **Is anyone watching accuracy?** (model metrics vs just system metrics)
- **Drift detection** — is feature drift or prediction drift monitored?
- **Alerting** — do alerts fire when model performance degrades?
- **Feedback loop** — is there a way to get ground truth for predictions?
- **A/B testing** — is there infrastructure to compare model versions?

### Step 6: ML Infrastructure Cost

Estimate the cost of ML infrastructure:

- **GPU/TPU instances** — are they running 24/7 or on-demand?
- **Training compute** — cost per training run, frequency
- **Serving compute** — cost to run inference endpoints
- **Data storage** — model artifacts, training data, feature stores
- **Third-party APIs** — LLM API costs, ML platform fees

Present the full inventory:

```
## ML Reconnaissance Report

### Model Inventory
| Model | Predicts | Framework | Serving | Frequency | Health |
|-------|----------|-----------|---------|-----------|--------|
| [name] | [what] | [framework] | [how] | [volume] | [status] |

### Training Pipelines
| Pipeline | Schedule | Platform | Duration | Automated |
|----------|----------|----------|----------|-----------|
| [name] | [freq] | [where] | [time] | [yes/no] |

### Data & Features
- Data sources: [list]
- Feature store: [yes/no — which]
- Training/serving parity: [verified/unverified/skewed]

### Experiment Tracking
- Tool: [name or "none"]
- Reproducibility: [can/cannot reproduce deployed model]

### Monitoring
- Model metrics monitoring: [yes/no]
- Drift detection: [yes/no]
- Alerting: [yes/no]
- Feedback loop: [yes/no]

### Cost Estimate
- Training: $[X]/month
- Serving: $[X]/month
- Data/storage: $[X]/month
- Total ML infra: $[X]/month

### Health Summary
- [model]: [status emoji + one-line assessment]

### Top Risks
1. [risk] — [impact]
2. [risk] — [impact]
3. [risk] — [impact]
```

## Delivery

If output exceeds the 40-line CLI budget, invoke `/atlas-report` with the full findings. The HTML report is the output. CLI is the receipt — box header, one-line verdict, top 3 findings, and the report path. Never dump analysis to CLI.

Attribution

tonone-aitonone-ai
View sourceSee grades on GitHubMore from tonone-ai →
SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Related Skills

Caveman

Terse caveman voice: answer first, fluff gone, every technical fact kept. Use for /caveman, "caveman mode", "talk like caveman", "be brief", "less tokens". Stays on until "stop caveman" or "normal mode".

1100021 votes

Hyperplan

Adversarial multi-agent planning skill. Self-orchestrates 5 hostile category members (unspecified-low, unspecified-high, deep, ultrabrain, artistry) via team-mode for ruthless cross-critique debate, distills only the defensible insights, then MANDATORILY hands the distilled insight bundle to the `plan` agent for executable plan formalization. Use when planning needs maximum rigor and surfacing of weak assumptions, blind spots, and over-engineering. Triggers: 'hyperplan', 'hpp', '/hyperplan', ...

698461 votes

Writing Skills

Create and manage Claude Code skills in HASH repository following Anthropic best practices. Use when creating new skills, modifying skill-rules.json, understanding trigger patterns, working with hooks, debugging skill activation, or implementing progressive disclosure. Covers skill structure, YAML frontmatter, trigger types (keywords, intent patterns), UserPromptSubmit hook, and the 500-line rule. Includes validation and debugging with SKILL_DEBUG. Examples include rust-error-stack, cargo-dep...

3931 votes

Mcp Code Execution

Routes multi-tool workflows through MCP servers for large datasets and pipelines. Use when Bash tool overhead is limiting throughput on data-heavy tasks.

3421 votes

catchup

Recovers the conversation and failed tool calls of a previous Codex, Amp, Claude Code, Antigravity, Cline, Copilot CLI, Cursor, DeepSeek Harness, Grok Build, Kimi, OpenCode, Pi Agent, or ZCode session. Use when the user says "catch up", "what did the last session do", "get me up to speed", "I switched agents", asks to recover/summarize a previous session before continuing, or asks to diagnose or report a catchup failure. Do NOT use for the current conversation, git history, or any non-agent log.

741 votes
View all in ai-agents →