Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsCommunityBlog
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Ceq Inference

ASecurity

Use when scrutinizing statistical inference for a 《经济学(季刊)》 (China Economic Quarterly, CEQ) manuscript — choosing and justifying the clustering level, handling weak instruments with robust inference, correcting for multiple hypothesis testing, and reporting standard errors that survive a technical reviewer. Default robust SEs are rarely enough at CEQ.

1,052 stars
0 votes
0 copies
0 views
Added 6/4/2026
ai-agentstesting

Security Analysis

A100/100

Scanned 6/4/2026

Install to Claude Code

$npx -y skills add brycewang-stanford/Awesome-Journal-Skills --skill ceq-inference --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Ceq Inference?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Ceq Inference
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/brycewang-stanford-ceq-inference/badge)](https://www.skillsdirectory.com/skills/brycewang-stanford-ceq-inference)

More formats (shields.io, HTML) on the badges page.

Download with Pro
Files
SKILL.md
---
name: ceq-inference
description: Use when scrutinizing statistical inference for a 《经济学(季刊)》 (China Economic Quarterly, CEQ) manuscript — choosing and justifying the clustering level, handling weak instruments with robust inference, correcting for multiple hypothesis testing, and reporting standard errors that survive a technical reviewer. Default robust SEs are rarely enough at CEQ.
---

# 推断细节(ceq-inference)

## 触发时机

- 标准误用了默认稳健,但没说聚类层级理由
- 第一阶段 F 偏低 / 工具偏弱
- 跑了大量子样本/异质性,没做多重检验校正
- 审稿人会逐条挑推断的稿子

## CEQ 视角:推断是技术审稿的主战场

海外训练的审稿人会盯:**聚类层级、弱工具、多重检验、有限样本**。点估计漂亮但推断站不住,照样退。

## 1. 聚类层级(要给理由,不是默认)

- 聚类应与**处理/抽样的层级**一致:处理在省级,就省级聚类(即使样本在个体)。
- 簇数太少(< ~30–50)→ 用 wild cluster bootstrap(Cameron-Gelbach-Miller)。
- 面板同时考虑双向聚类(个体 + 时间)是否必要。
- [ ] 聚类层级与处理分配层级一致,且写明理由
- [ ] 少簇情形用 wild bootstrap

## 2. 弱工具(IV)

- 第一阶段 F 报告;多工具用 Kleibergen-Paap(非同方差稳健)而非简单 F。
- F 偏弱 → 用 **Anderson-Rubin** 等 weak-IV-robust 置信区间,别只靠 t 比。
- 报告 reduced form 与第一阶段,别只给 2SLS。
- [ ] F / KP 报告;弱则给 AR 区间
- [ ] reduced form 与第一阶段都展示

## 3. 多重假设检验

- 跑了多个结果变量 / 多个子组 → Romano-Wolf、List-Shaikh-Xu,或 BH/Bonferroni 校正。
- 异质性"找显著"要预警 p-hacking;最好预先登记或限制切分维度。
- [ ] 多结果/多子组已做 MHT 校正
- [ ] 异质性切分有理论依据,非数据挖掘

## 4. 标准误与有限样本

- DID 现代估计量用其配套(解析或 bootstrap)标准误,别套 TWFE 的。
- 小样本/少处理单位 → 随机化推断(permutation / placebo 分布)。
- [ ] 估计量与标准误匹配
- [ ] 必要时随机化推断

## 反模式

- "标准误聚类到个体层"但处理在省级——典型低估
- 只报 2SLS t 值,不管弱工具
- 报告 20 个异质性里挑出的 2 个显著,不做校正
- 现代 DID 估计量配 TWFE 标准误

## 输出格式

```
【聚类层级】... | 与处理层级一致 □ | 少簇 bootstrap □
【弱工具】F/KP=... | AR 区间 □ | reduced form □
【多重检验】结果数=.. 子组数=.. | 校正方法=..
【标准误-估计量匹配】是 / 否
【缺口】<待补>
【下一步】ceq-mechanism
```

Attribution

brycewang-stanfordbrycewang-stanford
View sourceMore from brycewang-stanford →
SSkills DirectorySkills Directory

Know which skills are safe — weekly.

Best new skills + every skill we flagged as malicious. From the team that scanned 103,619.

Join free

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills DirectorySkills Directory

Know which skills are safe — weekly.

Best new skills + every skill we flagged as malicious. From the team that scanned 103,619.

Join free

Related Skills

Caveman

Ultra-compressed communication mode that cuts output tokens while keeping technical accuracy. Levels: lite, full, ultra and the wenyan variants. Use for /caveman, "caveman mode", "talk like caveman", "be brief" or "less tokens".

1074701 votes

Hyperplan

Adversarial multi-agent planning skill. Self-orchestrates 5 hostile category members (unspecified-low, unspecified-high, deep, ultrabrain, artistry) via team-mode for ruthless cross-critique debate, distills only the defensible insights, then MANDATORILY hands the distilled insight bundle to the `plan` agent for executable plan formalization. Use when planning needs maximum rigor and surfacing of weak assumptions, blind spots, and over-engineering. Triggers: 'hyperplan', 'hpp', '/hyperplan', ...

693161 votes

Mcp Code Execution

Routes multi-tool workflows through MCP servers for large datasets and pipelines. Use when Bash tool overhead is limiting throughput on data-heavy tasks.

3351 votes

catchup

Recovers the conversation and failed tool calls of a previous Codex, Claude Code, Antigravity, Cline, Copilot CLI, Cursor, DeepSeek Harness, Kimi, OpenCode, Pi Agent, or ZCode session. Use when the user says "catch up", "what did the last session do", "get me up to speed", "I switched agents", asks to recover/summarize a previous session before continuing, or asks to diagnose or report a catchup failure. Do NOT use for the current conversation, git history, or any non-agent log.

691 votes

math-skill

A comprehensive mathematical reasoning skill for AI assistants — handles arithmetic to research-level problems with rigorous step-by-step reasoning, systematic verification, and transparent uncertainty handling

381 votes
View all in ai-agents →