Skills DirectorySkills Directory
SkillsLearnSecurityCategoriesDocsBlogPro
Sign InSubmit Skill
Skills Directory

Security-tested agent skills for Claude, coding agents, and AI workflows.

Directory

  • Browse Skills
  • All Skills A–Z
  • Claude Skills
  • Claude Code Skills
  • Agent Skills
  • Categories
  • Authors
  • Submit a Skill

Learn

  • Learn Hub
  • Install Claude Skills
  • Write SKILL.md
  • Skills vs MCP
  • Directories Compared

Security

  • Security
  • Methodology
  • Secure Claude Skills
  • Security Badges
  • Chrome Extension
  • Skill Manager

Company

  • About
  • Community
  • Blog
  • API Docs
  • Advertise

2026 Skills Directory. All rights reserved.

ProTermsPrivacyRefunds
Back to skills

Spectral Window Transfer Pretraining

ASecurity

Spectral-window transfer analysis for cross-domain foundation models - pretraining transfer quality is governed by the overlap between the pretrained model's spectral window and the target domain's spectra (infrared-pretrained model falls below untrained control on optical/UV tasks). Use for deciding pretrained-vs-in-domain pretraining for spectroscopy, sensor, or any wavelength/frequency-structured scientific data.

3 stars
0 votes
0 copies
0 views
Added 10/3/2026
researchgo

Security Analysis

A100/100

Scanned 10/3/2026

$npx -y skills add hiyenwong/ai_collection --skill spectral-window-transfer-pretraining --agent claude-code

Installs into .claude/skills of the current project.

Are you the author of Spectral Window Transfer Pretraining?

Add the live security badge to your README — it updates automatically with every re-scan.

Security grade badge for Spectral Window Transfer Pretraining
[![Security: A — Skills Directory](https://www.skillsdirectory.com/api/skills/hiyenwong-spectral-window-transfer-pretraining/badge)](https://www.skillsdirectory.com/skills/hiyenwong-spectral-window-transfer-pretraining)

More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.

Download with Pro
Files
SKILL.md
---
name: spectral-window-transfer-pretraining
description: Spectral-window transfer analysis for cross-domain foundation models - pretraining transfer quality is governed by the overlap between the pretrained model's spectral window and the target domain's spectra (infrared-pretrained model falls below untrained control on optical/UV tasks). Use for deciding pretrained-vs-in-domain pretraining for spectroscopy, sensor, or any wavelength/frequency-structured scientific data.
category: ai_collection
trigger_words: spectral foundation model, transfer learning, domain gap, auroral spectra, masked autoencoder, 1D vision transformer, linear probe, fine-tuning, spectrograph, SpectraFM, SpecFormer, scientific foundation model, pretraining window
---

# Spectral-Window Transfer: From Foundation Models to Scientific Spectra

**Source**: arXiv:2609.31206v1 (2026-09-25) — Le Lain, Cessateur, Lefèvre, cs.LG/physics.space-ph.

## Problem

Scientific instruments (here: auroral spectrographs like ASIS) produce hundreds of thousands of spectra but only hundreds of expert labels. Should you fine-tune an existing pretrained spectral/time-series foundation model, or pretrain in-domain on the unlabeled data?

## Core Finding: Transfer Is Governed by the Spectral Window

Existing pretrained models transfer **according to the overlap between their pretraining spectral window and the target's spectral range**:

- **SpectraFM** (trained in **infrared**) on optical auroral task → **falls below the untrained control** (negative transfer!)
- **SpecFormer** (trained in **optical**) → approaches in-domain pretraining but does not reach it

So for wavelength-structured data, "use the biggest pretrained model" is wrong advice — a mismatched pretraining window is worse than nothing.

## The Winning Recipe (in-domain masked pretraining)

1. Pretrain a **1D Vision Transformer with masked autoencoder** (MAE) on all 223,000 **unlabelled** spectra.
2. Probe frozen representations: they recover the **emission-line intensity ratios physicists actually use** for diagnosing precipitating particles (R² 0.91 vs. 0.77 untrained control) — without any labels.
3. Linear probe: matches classification using 13 hand-designed expert features.
4. Fine-tune: beats the previous supervised auroral classifier on its own benchmark (macro-AP 88.5 vs. 77.8); **+0.159 over training the same architecture from scratch using only 10% of the labels**; attribution confirms the model uses physically meaningful N2+ bands.

## Decision Rule (reusable)

For frequency/wavelength-structured scientific data:

```
If pretrained model's training window ∩ target window ≈ ∅  → negative transfer risk; prefer random init or in-domain pretraining
If partial overlap (adjacent bands)                        → usable but expect a gap vs. in-domain pretraining
If unlabeled target data ≥ ~100× labeled data               → in-domain MAE pretraining is cheap and dominates
```

## Reusable Methodology: Physics-Grounded Pretraining Validation

- **Probe against known physical quantities** (here: emission-line ratios) — if the frozen representation recovers the variables domain scientists use for diagnosis, pretraining learned real structure, not shortcuts.
- **Untrained control is the mandatory baseline** — it quantifies architectural inductive bias vs. learned features; skip it and you cannot detect negative transfer.
- **Attribution check**: confirm the fine-tuned model attends to physically meaningful features (N2+ bands), guarding against spurious correlations.

**Activation**: scientific spectra pretraining, MAE 1D ViT, transfer window analysis, negative transfer detection, label-efficient scientific ML, spectroscopy foundation models.

Attribution

hiyenwonghiyenwong
View sourceSee grades on GitHubMore from hiyenwong →
SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Is this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.

Comments (0)

No comments yet. Be the first to comment!

SSkills DirectorySkills Directory

Ship a skill? Prove it's safe.

Free 120-pattern security scan, letter grade, and an embeddable README badge.

Submit a skill

Related Skills

Competitor Analysis

This skill provides comprehensive analysis of competitor SEO and GEO strategies, revealing what's working in your market and identifying opportunities to outperform the competition.

1823 votes

Deep Research

Universal deep research agent team. 13-agent pipeline for rigorous academic research on any topic. 8 modes: full research, quick brief, paper review, lit-review, fact-check, three-way literature scan, Socratic guided research dialogue, and systematic review with optional meta-analysis. Covers research question formulation, Socratic mentoring, methodology design, systematic literature search, source verification, cross-source synthesis, risk of bias assessment, meta-analysis, APA 7.0 report co...

502942 votes

Paperclip Distill

Use when an operation issue is a Paperclip cursor-window, distill, or backfill — `operationType: "distill"` or `"backfill"` and the body references a Paperclip source bundle for a project or root issue. Turn raw Paperclip activity into a wiki-insightful project page, decisions log, and history note. This skill exists specifically to replace the stiff, datestamp-heavy templated output that the deterministic distiller produces.

953191 votes

Academic Pipeline

Orchestrator for the full academic research pipeline: research -> write -> integrity check -> review -> revise -> re-review -> re-revise -> final integrity check -> finalize. Coordinates deep-research, academic-paper, and academic-paper-reviewer into a seamless 10-stage workflow with mandatory, coverage-bounded integrity checks, two-stage peer review, and auditable quality-assurance artifacts. Triggers on: academic pipeline, research to paper, full paper workflow, paper pipeline, end-to-end p...

502941 votes

Literature Review

Assistance with writing literature reviews by searching for academic sources via Semantic Scholar, OpenAlex, Crossref and PubMed APIs. Use when the user needs to find papers on a topic, get details for specific DOIs, or draft sections of a literature review with proper citations.

6511 votes
View all in research →