Skip to content
Back to skills

Psych Epistemic Situation And Replication Reform

ASecurity

Use before relying on any finding from these fields: what the replication crisis actually established and the genuinely encouraging finding alongside it, the reliability gradient as the practical takeaway, how to read a claim in these fields, why the fields are methodologically hard, and what the reform movement changed in practice. Includes the router for the whole psychology-sociology-cultural-sciences reference.

  • 2 stars
  • 0 votes
  • 0 copies
  • 0 views
  • Added September 19, 2026
ai-agentsgoreact

Works with

  • cli

Security analysis

A100/100

Scanned September 19, 2026

npx -y skills add the-vibey-project/vibey --skill psych-epistemic-situation-and-replication-reform --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Psych Epistemic Situation And Replication Reform?

Add the live security badge to your README. It updates with every re-scan.

Security grade badge for Psych Epistemic Situation And Replication Reform
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/the-vibey-project-psych-epistemic-situation-and-replication-reform/badge)](https://www.skillsdirectory.com/skills/the-vibey-project-psych-epistemic-situation-and-replication-reform)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
SKILL.md
---
name: psych-epistemic-situation-and-replication-reform
description: "Use before relying on any finding from these fields: what the replication crisis actually established and the genuinely encouraging finding alongside it, the reliability gradient as the practical takeaway, how to read a claim in these fields, why the fields are methodologically hard, and what the reform movement changed in practice. Includes the router for the whole psychology-sociology-cultural-sciences reference."
---

# Psychology, Sociology and Cultural Sciences: The Epistemic Situation, Why These Fields Are Hard, and What Reform Changed

> **Part 1 of 5** of the *Psychology, Sociology and Cultural Sciences* reference (plugin `psychology-sociology-cultural-sciences`), covering §0–§3. Sibling skills: `psych-perception-memory-learning-cognition-and-emotion` (§4–§8), `psych-development-personality-social-clinical-and-neurodiversity` (§9–§15), `psych-sociology-institutions-culture-and-the-weird-problem` (§16–§24), `psych-reference` (§25–§29). Section numbers are shared across the set; a reference written as §N → `skill` points into that sibling skill.
>
> **Currency:** The core findings are decades stable. Two areas are live. See §25 → `psych-reference` for where replication reform actually stands, and the social-media-and-adolescent-mental-health dispute presented as the dispute it is.

> **⚠️ Read §1 first. It is not preamble — it is the single most useful thing in this
> document.**
>
> **In most fields you can learn the findings and treat the epistemics as background.**
> ⚠️ **Here that gets you a head full of confident claims that are false.** **Between
> roughly 2011 and 2020 these fields discovered that a large fraction of their published
> literature did not replicate**, and ⚠️ **much of the popular material about psychology
> was written from the pre-crisis literature and is still in circulation — in
> bestsellers, TED talks, management training and journalism.**
>
> **⚠️ GOTCHA** boxes mark claims that are widely repeated and either false, much weaker
> than advertised, or genuinely contested.
>
> **The three ideas that organize everything below:**
> 1. **⚠️ Knowing WHICH findings survived is more valuable than knowing more findings**
>    (§1–§3).
> 2. **⚠️ The robust results cluster in a recognizable place**: **perception,
>    psychophysics, memory, learning, individual differences, and large-sample
>    demography.** **The fragile ones cluster in social priming, small-N lab studies with
>    surprising outcomes, and anything that made a good headline** (§1.3).
> 3. **⚠️ Humans are far more shaped by situation and structure than intuition allows —
>    and far less infinitely malleable than a century of blank-slate theorizing
>    assumed.** **Both errors are common, and they're symmetrical.**

---

## §0. Routing

| You want... | Go to |
|---|---|
| **⚠️ The epistemic situation — READ FIRST** | **§1** |
| Why these fields are hard | §2 |
| What reform changed | §3 |
| Perception and attention | §4 → `psych-perception-memory-learning-cognition-and-emotion` |
| **Memory** | **§5 → `psych-perception-memory-learning-cognition-and-emotion`** |
| Learning and conditioning | §6 → `psych-perception-memory-learning-cognition-and-emotion` |
| **Cognition, reasoning, bias** | **§7 → `psych-perception-memory-learning-cognition-and-emotion`** |
| Emotion | §8 → `psych-perception-memory-learning-cognition-and-emotion` |
| Development | §9 → `psych-development-personality-social-clinical-and-neurodiversity` |
| **Personality** | **§10 → `psych-development-personality-social-clinical-and-neurodiversity`** |
| Motivation | §11 → `psych-development-personality-social-clinical-and-neurodiversity` |
| **Social psychology** | **§12 → `psych-development-personality-social-clinical-and-neurodiversity`** |
| **Clinical — and what works** | **§13 → `psych-development-personality-social-clinical-and-neurodiversity`** |
| Intelligence | §14 → `psych-development-personality-social-clinical-and-neurodiversity` |
| Neurodiversity | §15 → `psych-development-personality-social-clinical-and-neurodiversity` |
| **Sociological thinking** | **§16 → `psych-sociology-institutions-culture-and-the-weird-problem`** |
| Theory traditions | §17 → `psych-sociology-institutions-culture-and-the-weird-problem` |
| Stratification | §18 → `psych-sociology-institutions-culture-and-the-weird-problem` |
| **Networks** | **§19 → `psych-sociology-institutions-culture-and-the-weird-problem`** |
| Institutions and organizations | §20 → `psych-sociology-institutions-culture-and-the-weird-problem` |
| Deviance and control | §21 → `psych-sociology-institutions-culture-and-the-weird-problem` |
| **Religion and secularization** | **§22 → `psych-sociology-institutions-culture-and-the-weird-problem`** |
| Culture and anthropology | §23 → `psych-sociology-institutions-culture-and-the-weird-problem` |
| **Cross-cultural — WEIRD** | **§24 → `psych-sociology-institutions-culture-and-the-weird-problem`** |
| **What's live** | **§25 → `psych-reference`** |
| Misconceptions | §26 → `psych-reference` |
| Books | §27 → `psych-reference` |
| Quick reference | §28 → `psych-reference` |

---

## §1. ⚠️ The Epistemic Situation

### 1.1 What happened
**⚠️ In August 2015 the Open Science Collaboration published the Reproducibility Project:
Psychology — a 270-author effort that attempted to replicate 100 published psychology
studies.** ⚠️ **About 36% of the replications produced a statistically significant effect,
against 97% of the originals.** **Effect sizes in the replications averaged roughly half
the originals'.**

**⚠️ This was not an isolated result.** **Comparable projects in experimental economics,
cancer biology and other fields found similar or worse.** **Across large-scale replication
efforts, reported rates have generally landed in the 30–70% band**, and ⚠️ **a 2026
analysis reported roughly 50–55% of originally significant claims replicating.**

**⚠️ It also was not primarily about fraud.** **A small number of high-profile fraud cases
occurred, and they are not the mechanism.** **The mechanism was ordinary researchers
following ordinary incentives** (§2).

### 1.2 ⚠️ The genuinely encouraging finding
> **⚠️ GOTCHA — the most important result in this whole area is the one that gets the least
> attention, and it's good news.** **When four labs ran 16 studies under normal working
> conditions but with one change — the researchers PREREGISTERED their hypotheses,
> procedures and analysis plans in advance — and each study was then replicated with
> large samples, ⚠️ 55 of 64 replications found the same effect. An 86% replication rate**,
> against the 30–70% typical of large-scale replication projects.
> **⚠️ This tells you the problem is METHODOLOGICAL, not that human behaviour is
> unstudyable.** **Psychology can produce reliable findings. It reliably produced
> unreliable ones because of how the work was designed, analyzed and published.**

### 1.3 ⚠️ The reliability gradient — the practical takeaway
```
ROBUST — replicated widely, large effects, often decades of data
  ⚠️ Psychophysics and perception · basic memory phenomena · classical and operant
  conditioning · the Big Five factor structure · g and its predictive validity ·
  heritability findings from twin/adoption designs · the strength of weak ties ·
  demographic and stratification patterns · large-sample epidemiology

MODERATE — real but smaller and more context-dependent than advertised
  ⚠️ Most heuristics-and-biases effects (real, but effect sizes and boundary
  conditions were overstated) · attachment · most therapy outcome research ·
  much of developmental psychology · Hofstede-style culture dimensions (§24)

⚠️ FRAGILE OR FAILED — treat any confident claim here with suspicion
  ⚠️ SOCIAL PRIMING as a class · ego depletion · power posing · facial-feedback
  in its strong form · stereotype threat's larger claimed effects · the implicit
  association test as an individual diagnostic or behaviour predictor ·
  much "nudge" research at the effect sizes originally claimed ·
  growth mindset at the effect sizes originally claimed ·
  the strong learning-styles claim (⚠️ this one is simply false, see §26)
```
**⚠️ The pattern is legible**: **the failures concentrate in small-sample laboratory
studies producing surprising, mechanistically vague, media-friendly effects.** **The
survivors tend to be boring, large, measured well, and often predate the era of
publish-or-perish incentives.**

### 1.4 ⚠️ How to read a claim in these fields
```
⚠️ Was it PREREGISTERED? (§1.2 — this is the single strongest signal)
⚠️ Sample size? (N=40 undergraduates predicts nothing)
⚠️ Effect size, not just p<.05? An effect can be real and too small to matter
⚠️ Has it been REPLICATED by an independent lab?
⚠️ Published before ~2015? Then the QRP-era base rate applies
⚠️ Is it correlational being described causally? (§2.4)
⚠️ WEIRD sample generalized to humanity? (§24)
⚠️ Does it appear in a bestseller or TED talk? ⚠️ NOT disqualifying — but popular
   uptake selects for surprisingness, which anti-correlates with robustness
```
**⚠️ And a caution against the opposite error**: **"psychology is all fake" is as wrong as
uncritical acceptance, and it's becoming the more common failure now.** **The robust tier
in §1.3 is genuinely robust, and some of it is among the better-established knowledge in
any science of human beings.**

---

## §2. Why These Fields Are Hard

**⚠️ These are not soft sciences because their practitioners are less rigorous. They are
hard sciences of an intractable object.**

**2.1 Measurement.** ⚠️ **You cannot directly observe a belief, an attitude, a preference
or a mental state.** **You observe self-reports, reaction times, behaviour and physiology,
each an imperfect proxy.** ⚠️ **Construct validity — whether your instrument measures the
thing you named it after — is the field's deepest and least-discussed problem**, and
**"the theory crisis" is the argument that psychology's constructs are often too vague to
generate falsifiable predictions in the first place.**

**2.2 Effect sizes and power.** **Real psychological effects are usually small.**
⚠️ **A small true effect studied with a small sample yields low power, and low-powered
studies that DO reach significance systematically OVERESTIMATE the effect** — **the
winner's curse.** **This is why the replication effect sizes came in at roughly half.**

**2.3 ⚠️ Researcher degrees of freedom.** **Which participants to exclude, when to stop
collecting, which outcome to analyze, which covariates to include, how to transform the
data.** ⚠️ **Each choice is defensible; the combination lets you find significance in
noise without ever intending to cheat.** **Simmons and colleagues demonstrated this can
produce false positives at will; Gelman and Loken called it "the garden of forking
paths," and the crucial point is that it doesn't require p-hacking as a conscious act —
you only have to walk one path and never see the others.**

**2.4 Causation.** ⚠️ **Most social science data is observational, and confounding,
selection and reverse causation are pervasive.** **The credible identification toolkit —
RCTs, natural experiments, instrumental variables, difference-in-differences, regression
discontinuity** — ⚠️ **and where none of them is available, the honest answer is "we have
an association."** **§25.2 → `psych-reference` is a case study in what happens when a field can't run the
experiment.**

**2.5 Publication bias.** ⚠️ **Null results are hard to publish, so the literature is a
biased sample of the research conducted, and meta-analyses of a biased literature inherit
the bias.**

**2.6 The moving target.** ⚠️ **Unlike electrons, people change** — **across cohorts,
technologies and cultures** — **and they respond to being studied (demand
characteristics), and to knowing the theory about them (Hacking's "looping effect").**

---

## §3. What Reform Changed

```
PREREGISTRATION       ⚠️ commit to hypotheses and analysis BEFORE data.
                      Converts exploratory work into honestly-labelled exploration
REGISTERED REPORTS    ⚠️ THE STRONGEST FIX — peer review of the METHOD, with
                      publication guaranteed regardless of outcome. Removes the
                      incentive to find something
OPEN DATA/MATERIALS   verification and reanalysis become possible
LARGER SAMPLES        ⚠️ powered for realistic effect sizes, not for luck
MANY-LABS / MULTI-LAB collaborative replication as a standard format
BETTER STATISTICS     effect sizes and intervals over bare p-values; Bayesian methods
```
**⚠️ Honest assessment of where this stands.** **The reforms work — §1.2 is the evidence.**
**Preregistration rates rose sharply through the late 2010s and sample sizes increased.**
⚠️ **But: preregistration is still not the norm; the reforms are not universally mandated;
they do nothing to retroactively fix the pre-2010s published record; and the incentives
still favour positive, novel, theoretically appealing results.** **One prominent critic's
summary is worth carrying — that in the absence of enforceable field-wide standards,
credibility remains largely a property of individual researchers rather than of the
discipline.**
**⚠️ Practical consequence: judge the paper and the lab, not the field.**

---

# PART II — PSYCHOLOGY

Attribution

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments

Loading comments…