**arXiv ID:** 2507.18681 **Authors:** Manuel de Sousa Ribeiro, Afonso Leote, João Leite **Published:** 2025-07-24T16:30:10Z **Abstract:** Concept probing has recently gained popularity as a way for humans to peek into what is encoded within artificial neural networks. In concept probing, additional classifiers are trained to map the internal representations of a model into human-defined concepts of interest. However, the performance of these probes is highly dependent on the internal represen...
Scanned 9/11/2026
Install to Claude Code
npx -y skills add hiyenwong/ai_collection --skill concept-probing-where-to-find-humandefined-concepts-extended-version --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Concept Probing Where To Find Humandefined Concepts Extended Version?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/hiyenwong-concept-probing-where-to-find-humandefined-concept)More formats (shields.io, HTML) on the badges page.
# Concept Probing: Where to Find Human-Defined Concepts (Extended Version)
**arXiv ID:** 2507.18681
**Authors:** Manuel de Sousa Ribeiro, Afonso Leote, João Leite
**Published:** 2025-07-24T16:30:10Z
**Abstract:**
Concept probing has recently gained popularity as a way for humans to peek into what is encoded within artificial neural networks. In concept probing, additional classifiers are trained to map the internal representations of a model into human-defined concepts of interest. However, the performance of these probes is highly dependent on the internal representations they probe from, making identifying the appropriate layer to probe an essential task. In this paper, we propose a method to automatically identify which layer's representations in a neural network model should be considered when probing for a given human-defined concept of interest, based on how informative and regular the representations are with respect to the concept. We validate our findings through an exhaustive empirical analysis over different neural network models and datasets.
## Skill Description
This skill is generated from the arXiv paper: Concept Probing: Where to Find Human-Defined Concepts (Extended Version) (2507.18681).
## How to Use
[To be filled in by the user or by future automation]
## References
- [arXiv:2507.18681](http://arxiv.org/abs/2507.18681v1)
Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!