All authors
VectorSpaceLab avatar

Claude Skills by VectorSpaceLab

github.com/VectorSpaceLab
6,028 skillsA× 6,018B× 8D× 20 installs1,695 views
Intervention Comparison MatrixA

Compare robustness interventions across clean and shifted metrics using direction-aware evidence signs.

testingpython
0
247
Reduced Recovery EvaluatorA

Assemble executable soft-mode proxy recovery results with metrics, traces, and mechanism checks.

toolspython
0
247
Robustness Benchmark ProtocolA

Validate clean and shifted robustness benchmark protocols with class-overlap and metric-gap contracts.

testingpython
0
247
Detection Metric ProtocolA

Compute AUROC and AUPR detection metrics with paper-faithful score orientations for softmax confidence baselines.

testingpython
0
247
Proxy Recovery HarnessA

Run a bounded mechanism-faithful proxy experiment for maximum softmax OOD detection recovery artifacts.

researchgit
0
247
Softmax Confidence ScoringA

Compute maximum softmax probability detector scores from logits or probabilities for misclassification and OOD detection.

testingpythongit
0
247
Collocation Morphology PreprocessorA

Preprocess English text with WordNet-style collocation matching, tokenization, and inflectional normalization.

toolspython
0
247
Context Sense TaggerA

Select or defer WordNet-style sense pointers using current and previous sentence context.

businesspython
0
247
Lexical Taxonomy SchemaA

Build and validate compact WordNet-style synset and semantic-pointer taxonomies for recovery experiments.

testingpythondatabase
0
247
Semantic Distance EvaluatorA

Estimate WordNet-style semantic distance by traversing typed semantic-pointer graphs.

testingpython
0
247
Taxonomy Recovery HarnessA

Run a reduced executable WordNet taxonomy recovery using generated preprocessing, tagging, and distance skills.

toolspython
0
247
Dpo Loss ObjectiveA

Compute Direct Preference Optimization logistic losses, implicit rewards, and log-ratio diagnostics from policy and reference log probabilities.

testingpythonbash
0
247
Dpo Preference Data ContractA

Normalize pairwise preference records into prompt, chosen, rejected, and metadata fields for Direct Preference Optimization training.

ai-agentspythonbash
0
247
Dpo Reduced Training HarnessA

Run a bounded scalar DPO optimization that records loss, parameter changes, preference accuracy, and mechanism checks for reduced recovery.

testingpythonbash
0
247
Dpo Sequence Logprob AccountingA

Compute response-only sequence log probabilities for DPO by masking prompts and padding before summing token log probabilities.

testingpythonbash
0
247
Ffn Activation Logit AttributionA

Measure whether activated FFN neurons preferentially raise logits for their projected promoted tokens versus controls.

toolspythongit
0
247
Ffn Concept GroupingA

Group top promoted vocabulary tokens into human-readable concept labels with purity and unresolved-token evidence.

toolspythongit
0
247
Ffn Proxy Recovery HarnessA

Run a bounded mechanism-faithful proxy experiment for FFN value-vector concept promotion using generated skills.

researchpythongit
0
247
Ffn Value Vector ExtractionA

Extract and normalize feed-forward output value vectors from transformer-like model weights for concept-promotion analysis.

testingpython
0
247
Ffn Vocabulary ProjectionA

Project FFN value vectors through an LM head to rank promoted tokens for mechanistic concept analysis.

toolspythongit
0
247
Pplm Attribute ObjectivesA

Compute PPLM-style bag-of-words or linear-classifier attribute losses and gradients for controlled generation.

toolspythongit
0
247
Pplm Controlled Generation EvaluationA

Evaluate PPLM controlled-generation proxies with target-mass gain, KL fluency cost, and target consistency checks.

testingpython
0
247
Pplm Fusion GenerationA

Fuse PPLM perturbed and unperturbed token distributions using geometric mixing for fluent controlled decoding.

testingpythongit
0
247
Pplm Perturbation LoopA

Run PPLM-style iterative normalized gradient perturbations with KL regularization and optimizer trace evidence.

toolspythongit
0
247
Rtp Continuation Generation ProtocolA

Create bounded multi-continuation records per prompt for toxic degeneration evaluation with explicit generator metadata.

toolspythonapi
0
247
Rtp Prompt Dataset ProtocolA

Normalize RealToxicityPrompts-style prompt records for toxic and non-toxic prompted generation evaluation.

toolspythonapi
0
247
Rtp Recovery Experiment HarnessA

Run an executable bounded recovery harness that composes prompt normalization, generation, scoring, and aggregation with mechanism checks.

toolspythonapi
0
247
Rtp Toxicity Metric AggregationA

Compute expected maximum toxicity and toxicity probability for RealToxicityPrompts-style continuation sets.

toolspythonapi
0
247
Rtp Toxicity Scoring AdapterA

Attach numeric toxicity scores to generated continuations using Perspective API or declared offline proxy scoring.

toolspythonapi
0
247
Attention Path ExpansionA

Compute frozen-attention direct, head, and virtual-head path contributions for attention-only circuits.

toolsgit
0
247
Circuit Recovery HarnessA

Build and validate bounded mechanism-faithful recovery artifacts for Transformer Circuits proxy experiments.

toolsgit
0
247
Induction Head DetectorA

Detect induction-head copying behavior on repeated token sequences with explicit mechanism checks.

researchgotesting
0
247
Qk Ov Circuit ExpansionA

Expand attention-head QK and OV weights into token-level circuit matrices and copying diagnostics.

toolsgogit
0
247
Residual Logit LensA

Apply logit-lens unembedding and additive residual contribution checks for mechanistic transformer circuit analysis.

developmentpythonbash
0
247
Curriculum RegularizationA

Generate coefficient schedules for curriculum-regularized PINN recovery experiments.

testingpython
0
247
Failure Mode DiagnosticsA

Diagnose PINN failure-mode recovery runs with relative errors and optimizer progress checks.

testing
0
247
Periodic Pde BenchmarkA

Build deterministic periodic PDE benchmarks for PINN failure-mode recovery experiments.

researchpythonreact
0
247
Pinn Residual ObjectiveA

Compute PINN data, boundary, and PDE residual losses for reduced failure-mode experiments.

testingpython
0
247
Reduced Recovery HarnessA

Run bounded reduced PINN recovery experiments that exercise generated benchmark, objective, diagnostics, and curriculum skills.

toolspython
0
247
Deep Ritz Residual Trial NetworkA

Build smooth residual trial networks for Deep Ritz variational PDE objectives.

toolspythonbash
0
247
Deep Ritz Stochastic Quadrature SamplingA

Generate stochastic interior and boundary quadrature samples for Deep Ritz variational PDE training.

toolspythonbash
0
247
Deep Ritz Training RecoveryA

Run bounded Deep Ritz optimization and emit auditable recovery artifacts for variational PDE experiments.

testingpythonbash
0
247
Deep Ritz Variational Energy LossA

Compute Deep Ritz variational energy losses with gradient energy and boundary penalties.

toolspythonbash
0
247
Gated Pinn ArchitectureA

Build a lightweight gated PINN-style scalar model that exposes trainable parameters and gradient paths for reduced recovery experiments.

testingpython
0
247
Gradient Stat AnnealingA

Update PINN data-loss weights from residual and data-fit gradient statistics using the paper's moving-average annealing rule.

testingpython
0
247
Helmholtz Pinn ProblemA

Construct deterministic reduced Helmholtz PINN benchmark data with analytic solution, boundary samples, collocation samples, forcing, and relative L2 scoring.

testingpython
0
247
Pinn Loss DecompositionA

Compute separated Helmholtz PINN residual and boundary losses so gradient imbalance can be diagnosed and corrected.

testingpython
0
247
Reduced Recovery EvaluationA

Run bounded reduced Helmholtz PINN recovery with optimizer updates, lambda traces, relative L2 metrics, and validator-compatible evidence.

tools
0
247
Curvature MemoryA

Maintain bounded positive-curvature L-BFGS correction memory for large-scale quasi-Newton optimization.

toolspython
0
247
Proxy Recovery EvaluationA

Build validator-compatible soft-mode recovery metrics and mechanism checks for L-BFGS proxy experiments.

toolspython
0
247