All authors
HolobiomicsLab avatar

Claude Skills by HolobiomicsLab

github.com/HolobiomicsLab
12,704 skillsA× 12,683B× 13C× 2D× 60 installs6,245 views
Rnaseq Count Matrix AnalysisA

Use when you have a count matrix (genes × samples) from HTSeq, featureCounts, or transcript abundance quantification (Salmon, kallisto), a sample metadata table with experimental design, and you need to test for differential expression while controlling false discovery rate via independent.

ai-agentsgoexpress
0
15
Salmon Output ParsingA

Use when when you have salmon quant.sf.gz output files from pseudoalignment-based transcript quantification and need to convert transcript-level abundance estimates and counts into gene-level matrices for differential expression analysis with edgeR, DESeq2, or limma-voom.

ai-agentsgoexpress
0
15
Salmon Quantification Output VerificationA

Use when after running salmon quant with the --writeMappings/-z flag to produce SAM output, or when investigating discrepancies between the number of mapped reads reported in quant.sf and the actual number of records written to the output SAM file.

ai-agentsrustc++
0
15
Sam Bam Mapping Record InspectionA

Use when after quantifying the same read set with two versions of a mapping/quantification tool (e.g., C++ salmon 1.11.

ai-agentsrustc++
0
15
Sam Record Parsing And ValidationA

Use when salmon quant is run with the --writeMappings/-z flag and you need to verify that all mapped reads appear in the SAM output file.

ai-agentsrustc++
0
15
Sample Metadata Preparation And AssignmentA

Use when before constructing a DESeqDataSet from any count matrix (whether from tximport, HTSeq, featureCounts, or raw counts). You have sample identifiers (run IDs, file names, or row names) and must link them to condition labels (e.

ai-agentsexpresstesting
0
15
Scanpy Preprocessing Pipeline ExecutionA

Use when you have raw or minimally processed single-cell RNA-seq expression data loaded into an AnnData object (dense, sparse, or Dask-backed array as X), and you need to apply standardized preprocessing transformations (normalization, filtering, PCA) before downstream analysis such as clustering.

ai-agentspythongo
0
15
Scientific Task FormulationA

Use when you encounter a published scientific article or software paper that makes claims about data processing, analysis, or results, but the reproducibility context is unclear, artifacts are scattered, or the connection between claims and outputs is not immediately evident.

ai-agentsgitdocumentation
0
15
Seed Representation Variant ComparisonA

Use when you have observed mapping rate or quantification disagreement (e.g., >0.1% divergence in mapping rate or Pearson r < 0.

ai-agentsrustc++
0
15
Selective Alignment Parameter TuningA

Use when you observe discrepancies in mapping rate or per-transcript quantification between two salmon implementations, or when the default chain-pruning thresholds (orphanChainSubThresh, postMergeChainSubThresh) are leaving a substantial fraction of reads unmapped (e.

ai-agentsrustgo
0
15
Selective Alignment Sensitivity EvaluationA

Use when when comparing mapping outputs between two selective-alignment implementations (e.g., C++ vs. Rust port) on byte-identical reference indices and observing a multi-percentage-point gap in mapping rate or read assignments.

ai-agentsrustgo
0
15
Seurat Workflow Orchestration For ScrnaseqA

Use when you have a raw or Seurat object-backed scRNA-seq expression matrix and need to: (1) stabilize variance across genes with SCTransform normalization, (2) extract feature loadings in reduced dimensionality space via reverse PCA to use as input for GESECA or other coregulation-based enrichment.

ai-agentsgoexpress
0
15
Single Cell Graph Construction NeighborhoodA

Use when you have preprocessed single-cell RNA-seq data (normalized and dimensionality-reduced via PCA) and need to establish cell-to-cell connectivity for trajectory inference, clustering validation, or graph-based visualization.

ai-agentspythongo
0
15
Single Cell Rna Seq NormalizationA

Use when immediately after loading raw single-cell gene expression count matrices (AnnData objects) and before identifying highly variable genes or performing dimensionality reduction.

ai-agentspythonexpress
0
15
Single Cell Rna Seq Quality Control And NormalizationA

Use when when you have a raw or minimally processed scRNA-seq dataset (e.g., a Seurat object loaded from GEO) and need to prepare it for pathway enrichment or coregulation analysis.

ai-agentsgoexpress
0
15
Sparse Matrix Csr Format AssemblyA

Use when when you have computed k-nearest neighbor indices and distances (e.

ai-agentsgotesting
0
15
Sparse Matrix Format HandlingA

Use when your input is an AnnData object with expression matrix X as a sparse scipy matrix or Dask-backed array, and you need to apply preprocessing functions (normalization, PCA, filtering) that could trigger eager materialization.

ai-agentspythonexpress
0
15
Sparse Matrix Verification And ValidationA

Use when after calling squidpy.gr.spatial_neighbors or any graph-building operation that outputs sparse matrices to adata.

ai-agentsrustgo
0
15
Spatial Coordinate Indexing Dense To SparseA

Use when you have computed k-nearest neighbors for spatial coordinates (e.g., via pynndescent or another NN backend) and need to store the resulting adjacency and distance information in a memory-efficient format compatible with downstream graph algorithms.

ai-agentsgotesting
0
15
Spatial Coordinate Integration With Imaging DataA

Use when you have a spatial omics dataset (AnnData object with coordinate columns like 'x', 'y', 'z') and an associated tissue image file (e.g., TIFF, PNG, or HE-stained histology), and you need to extract image-based morphological features (e.

ai-agentsgitapi
0
15
Spatial Enrichment Score ValidationA

Use when after executing a spatial statistics function (e.g., squidpy.gr.sepal) on a spatial transcriptomics dataset in AnnData format, and before using the computed rankings or enrichment scores in downstream analysis.

ai-agentsgoexpress
0
15
Spatial Gene Ranking ComputationA

Use when when working with spatial transcriptomics datasets (e.g., Slide-seq v2, MERFISH) stored in AnnData format and you need to identify genes whose expression shows significant spatial patterns or enrichment within tissue regions.

ai-agentsexpressgit
0
15
Spatial Neighbor Graph ConstructionA

Use when when you have spatial molecular data (e.g., Visium, imaging-based cytometry) stored in an AnnData object with coordinate information in .

ai-agentsgogit
0
15
Spatial Neighbor Graph Properties ValidationA

Use when after calling squidpy.gr.spatial_neighbors on an AnnData object containing spatial coordinates in obsm.

ai-agentsgogit
0
15
Spatial Omics Dataset LoadingA

Use when you have a spatial transcriptomics experiment (e.

ai-agentsexpressgit
0
15
Spatial Statistics InterpretationA

Use when when you have a spatial molecular dataset (e.g., Visium, MERFISH) with categorical cell-type or feature annotations and want to test whether specific categories are preferentially located near or away from each other in tissue space, beyond what random spatial distribution would predict.

ai-agentsgoexpress
0
15
Splicing Matrix NormalizationA

Use when you have transcript-level quantification (TPM or counts from Salmon/kallisto) and need to quantify the inclusion level of specific alternative splicing events (exon skipping, intron retention, alternative splice sites, etc.) in a form suitable for differential splicing analysis across.

ai-agentspythonexpress
0
15
Statistical Comparison P Value ConservationA

Use when you have run the same pathway enrichment analysis (e.

ai-agentsgoexpress
0
15
Statistical Hypothesis Testing Rna SeqA

Use when you have RNA-seq count matrices (from alignment, transcript quantification, or HTSeq-count files) and need to test for differential expression between two or more conditions while controlling for batch effects or other covariates.

ai-agentsgoexpress
0
15
Statistical Precision Comparison Ranked OutputsA

Use when when a statistical method offers a parameter to trade computational cost for precision (e.

ai-agentsgogit
0
15
Statistical Ranking Wilcoxon TestA

Use when when you have leiden or louvain cluster assignments in single-cell data (stored in adata.obs) and need to identify cluster-specific marker genes.

ai-agentspythonexpress
0
15
Temporal Gene Expression DynamicsA

Use when you have time-ordered gene expression data (e.

ai-agentsreactexpress
0
15
Test Failure Diagnosis And Logging AnalysisA

Use when you have modified the Scanpy codebase (e.g., added a feature or bugfix) and need to confirm that all unit and integration tests pass before submitting a pull request, or when a CI workflow fails and you need to reproduce the failure locally to diagnose the root cause.

ai-agentstestinggit
0
15
Trajectory Inference PreprocessingA

Use when you have raw or normalized single-cell RNA-seq expression data stored in an AnnData object (`.h5ad` format) and your analysis goal is to infer developmental or differentiation trajectories.

ai-agentspythongo
0
15
Transcript Abundance Correlation AnalysisA

Use when when validating a new or reimplemented quantification tool against a reference implementation on the same dataset and index, or when investigating whether changes to seed representation, chain pruning thresholds, or other algorithmic parameters affect downstream abundance estimates.

ai-agentsrustgo
0
15
Transcript Abundance Quantification ImportA

Use when you have transcript abundance files (e.g., Salmon quant.sf.gz, kallisto abundance.h5, RSEM .isoforms.results) from a quantification tool and need to construct a gene-level count matrix for DESeq2 analysis.

ai-agentsgoexpress
0
15
Transcript Event ExtractionA

Use when you have a genome annotation GTF file and need to identify all transcript-level alternative splicing events (exon skipping, intron retention, alternative splice sites, mutually exclusive exons, alternative first/last exons) before quantifying their inclusion levels (PSI) across samples or.

ai-agentspythongit
0
15
Transcript Expression Quantification HandlingA

Use when you have transcript abundance estimates from RNA-seq quantification tools (e.

ai-agentsexpressgit
0
15
Transcript Level Abundance ImportA

Use when you have transcript-level quantification files (quant.sf.gz or quant.gz) from salmon, sailfish, kallisto, or oarfish and need to aggregate them into gene-level or transcript-level count, abundance, and length matrices for input to DESeq2, edgeR, or limma-voom.

ai-agentsgoexpress
0
15
Transcript Level Count AggregationA

Use when you have transcript-level abundance and count estimates from salmon, sailfish, kallisto, or oarfish and need gene-level matrices for downstream differential analysis with edgeR, DESeq2, or limma-voom.

ai-agentsgoexpress
0
15
Transcript Quantification Benchmark ComparisonA

Use when you have two implementations of the same quantification method (or major versions) and observe a persistent disagreement in mapped-read counts, per-read alignment agreement, or abundance correlations on the same reference index and read set.

ai-agentsrustgo
0
15
Transcript Quantification IntegrationA

Use when you have transcript quantification output (TPM or raw counts) from a pseudo-aligner (Salmon or kallisto) and an ioe/ioi event definition file from SUPPA2's generateEvents step, and you need to calculate PSI values—the relative inclusion level of alternative splicing events—across multiple.

ai-agentspythonexpress
0
15
Transcript To Gene AggregationA

Use when when you have transcript-level quantification files (e.g., Salmon quant.sf.gz, kallisto abundance.h5, or RSEM .results) and need to construct a gene-level count matrix for DESeq2 differential expression testing.

ai-agentsexpresstesting
0
15
Transcript To Gene MappingA

Use when you have transcript-level quantification files (salmon quant.sf.gz, kallisto, or Sailfish output) and need to perform gene-level differential expression analysis.

ai-agentsexpressgit
0
15
Umap Embedding VisualizationA

Use when after completing PCA and k-nearest neighbor graph construction on preprocessed, log-normalized, highly-variable-gene-filtered single-cell RNA-seq data (stored in an AnnData object), compute UMAP embeddings when you need a 2-D visualization for cluster inspection, cell-type annotation, or.

ai-agentspythongo
0
15
Uncertainty Aware Significance AssessmentA

Use when you have PSI matrices for two or more conditions with replicates per condition, and you need to determine which alternative splicing events show significant changes between conditions while accounting for measurement uncertainty that scales with transcript expression levels.

ai-agentspythonexpress
0
15
Accurate Mass Database SearchA

Use when after peak detection and MS1 feature extraction from FIA-MS, GC-MS, LC-MS(/MS), or CE-MS data, when you need to identify unknown metabolites by matching observed m/z values to a reference database and want to recover HMDB identifiers, molecular formulas, and structural annotations for.

ai-agentspythongo
0
15
Adduct Ion Parent Ion Pairing AnalysisA

Use when when you have binned mass spectrometry imaging peaks and want to understand which detected mass-to-charge ratios represent the same metabolite in different ionization states (parent vs. adduct form).

ai-agentstestinggit
0
15
Annotation Quality Filtering By Cosine SimilarityA

Use when you have in silico annotations (e.g. from GNPS, timaR, or SIRIUS) paired with experimental MS/MS spectra and need to select only the highest-confidence structural matches.

ai-agentsgitdatabase
0
15
Approximate Nearest Neighbor Index ConstructionA

Use when when you have a large collection of reference MS/MS spectra (spectral library) and need to search unknown query spectra against it rapidly, particularly for open modification searching where the modification mass is unknown and candidate space is large.

ai-agentspythongo
0
15