All authors
NVIDIA avatar

Claude Skills by NVIDIA

github.com/NVIDIA
445 skillsA× 401B× 31C× 9D× 2F× 20 installs897 views
Tao Train DinoA

DINO (DETR with Improved DeNoising Anchor Boxes) for 2D object detection. Transformer-based detector with

ai-agentspythongo
0
3,503
Tao Train Fast Foundation StereoA

Real-time stereo depth estimation using FastFoundationStereo (FFS), the distilled bp2 commercial variant of

ai-agentspythongo
0
3,503
Tao Train Foundation StereoB

Stereo depth estimation using FoundationStereo. Predicts disparity maps from stereo image pairs for 3D

ai-agentsgobash
0
3,503
Tao Train Grounding DinoA

Grounding DINO for open-set object detection. Combines DINO-style detection with a BERT text encoder for

ai-agentspythongo
0
3,503
Tao Train Image ClassificationA

PyTorch-based TAO image classification. Supports a wide range of backbones (FAN, EfficientNet, ResNet, etc.)

ai-agentspythonbash
0
3,503
Tao Train Mask Auto EncoderA

Masked Auto-Encoder (MAE) for self-supervised pretraining and fine-tuning. Masks random patches and reconstructs

ai-agentspythonbash
0
3,503
Tao Train Mask Auto LabelA

MAL (Mask Auto-Label) for weakly-supervised segmentation. Produces segmentation masks from minimal annotations

ai-agentspythonbash
0
3,503
Tao Train Mask Grounding DinoA

Mask Grounding DINO for grounded instance segmentation. Extends Grounding DINO with a mask-prediction head for

ai-agentspythonbash
0
3,503
Tao Train Mask2formerA

Mask2Former for universal image segmentation (panoptic, instance, and semantic). Transformer-based with

ai-agentspythongo
0
3,503
Tao Train Metric Learning RecognitionA

Metric-learning recognition (ml-recog) for fine-grained visual recognition. Learns embeddings for

ai-agentspythonrust
0
3,503
Tao Train Nvdinov2A

NVDINOv2 for self-supervised visual representation learning. Trains vision transformers via self-distillation

ai-agentspythongo
0
3,503
Tao Train Nvpanoptix3dA

NVPanoptix3D for panoptic 3D scene reconstruction from posed RGB images. Produces 3D panoptic segmentation

ai-agentspythonrust
0
3,503
Tao Train OcdnetA

OCDNet for scene text detection. Detects arbitrary-oriented text regions in natural images using a

ai-agentspythonbash
0
3,503
Tao Train OcrnetA

OCRNet for scene text recognition. Recognizes text content from cropped text-region images and supports CTC

ai-agentspythonrust
0
3,503
Tao Train OneformerA

OneFormer for universal image segmentation. Unifies panoptic, instance, and semantic segmentation with a

ai-agentspythonrust
0
3,503
Tao Train Optical InspectionA

Optical Inspection for defect detection using Siamese networks. Compares image pairs to detect manufacturing

ai-agentspythonrust
0
3,503
Tao Train PointpillarsA

PointPillars for 3D object detection from LiDAR point clouds. Encodes point clouds into a pseudo-image via a

ai-agentspythongo
0
3,503
Tao Train Pose ClassificationA

Pose classification using ST-GCN (Spatial Temporal Graph Convolutional Network). Classifies skeleton sequences

ai-agentspythongo
0
3,503
Tao Train ReidA

Person re-identification (ReID). Learns discriminative embeddings to match the same person across different

ai-agentspythonrust
0
3,503
Tao Train RtdetrA

RT-DETR (Real-Time DEtection TRansformer) for 2D object detection. Designed for real-time inference with

ai-agentspythongo
0
3,503
Tao Train SegformerA

SegFormer for semantic segmentation. Lightweight transformer-based architecture with hierarchical feature

ai-agentspythongo
0
3,503
Tao Train Single StepA

Standard single-step train/eval/export workflow for any TAO model. Use when training a TAO model on a dataset

ai-agentspythonbash
0
3,503
Tao Train Sparse4dA

Sparse4D for multi-camera temporal 3D object detection and tracking. Uses sparse queries with deformable

ai-agentspythongo
0
3,503
Tao Train Visual ChangenetA

Visual ChangeNet for binary image classification and segmentation in AOI defect detection. Use when training,

ai-agentspythongo
0
3,503
Tao Validate Dataset FormatA

Run `tao-daft validate` to check NVIDIA TAO DAFT datasets for structure, schema, and cross-reference errors. Do

ai-agentspythonrust
0
3,503
Tilegym Improve Cutile Kernel PerfA

Iteratively optimize cuTile kernel performance through systematic profiling, bottleneck analysis, IR comparison, and targeted tuning. Covers tile sizes, occupancy, autotune configs, TMA, latency hints, persistent scheduling, num_ctas, flush_to_zero, and IR-level debugging. Use when asked to "optimize cutile kernel", "improve kernel perf", "tune cutile performance", "make kernel faster", or iteratively benchmark and refine a cuTile GPU kernel in the TileGym project.

ai-agentspythongo
0
3,503
Vss Ask VideoB

Use this skill to ask the VSS agent's video_understanding tool a fresh visual question about a recorded clip. Not for prior tool output, search hits, or metadata-answerable questions.

ai-agentspythonbash
0
3,503
Vss Deploy Dense CaptioningB

Use this skill when deploying standalone RT-VLM dense captioning or calling its REST API (uploads, captions, streams, chat-completions, Kafka). Not for VSS profile deploy or video-search ingestion.

ai-agentsrustgo
0
3,503
Vss Deploy Detection Tracking 2dA

Use this skill when the user wants to deploy, run, debug, tear down, or call the REST API of the RTVI-CV 2D detection / tracking microservice. Trigger when the user says things like 'deploy rtvi-cv', 'start warehouse 2d', 'add a stream', 'check rtvi-cv health', or 'stop the perception container'. Not for VLM, embedding, or analytics — use the matching vss-* skill.

ai-agentspythongo
0
3,503
Vss Deploy Detection Tracking 3dB

Deploy and operate the RTVI-CV-3D microservice as MV3DT (`MODE=mv3dt`): per-camera DeepStream perception plus BEV Fusion over calibrated cameras. Supports the bundled sample dataset, custom video files, and RTSP streams, and chains to `vss-generate-video-calibration` when calibration is missing. Use `vss-deploy-profile` for the full warehouse blueprint and `vss-deploy-detection-tracking-2d` for single-camera 2D detection.

ai-agentsgoshell
0
3,503
Vss Deploy ProfileF

Use to select, configure, deploy, verify, debug, or tear down a VSS profile (base, search, lvs, warehouse, edge). Not for standalone microservices — use the vss-deploy-* skill.

ai-agentsgoshell
0
3,503
Vss Deploy Video EmbeddingA

Use this skill when deploying, operating, or integrating the VSS 3.2 GA RT-Embed Video Embedding microservice. Covers Docker Compose bring-up, GPU and storage prerequisites, the `/v1` REST API (file uploads, text and video embeddings, live RTSP streams, health and metrics), Redis/Kafka/OTel integration, common failure modes, and teardown.

ai-agentsbashdocker
0
3,503
Vss Generate Video CalibrationA

Use to run AutoMagicCalib on local MP4s, RTSP, or the bundled sample dataset, and to deploy vss-auto-calibration when needed. Do not use for non-AMC calibration or runtime analytics.

ai-agentspythongo
0
3,503
Vss Generate Video ReportB

Use this skill when producing a VSS analysis report — Mode A per-clip VLM, Mode B incident-range via video-analytics. Not for standalone video summarization, real-time alerts or ad-hoc Q&A.

ai-agentspythongo
0
3,503
Vss Manage AlertsA

Use for VSS alert workflows — real-time monitoring, Alert-Bridge subscriptions, Slack notifications, incident queries, camera onboarding. Not for non-alert analytics.

ai-agentsgobash
0
3,503
Vss Manage Video Io StorageA

Use to call the VIOS REST API (sensor list, timelines, clip extraction, snapshots, add/delete sensors and streams). Not for VLM inference or search.

ai-agentsshellbash
0
3,503
Vss Query AnalyticsA

Use this skill when reading video-analytics metrics, incidents, alerts, and sensor data via the VA-MCP server (port 9901). Not for live VLM or incident-range narrative reports.

ai-agentsrustshell
0
3,503
Vss Search ArchiveA

Use this skill to run top-level VSS fusion search on archived video, or to ingest video files / RTSP streams for search. Do NOT use for ad-hoc visual Q&A (use vss-ask-video), live captioning (use vss-deploy-dense-captioning), or video summarization and reports (use vss-summarize-video).

ai-agentsgobash
0
3,503
Vss Setup Behavior AnalyticsA

Use to deploy the vss-behavior-analytics service standalone (entrypoint, config-source, optional calibration). Not for the full warehouse deploy.

ai-agentsgobash
0
3,503
Vss Setup Video Analytics ApiA

Use to deploy the vss-video-analytics-api REST service standalone (config-source, data-log bind, Elasticsearch, optional Kafka). Not for full warehouse deploy.

ai-agentsgobash
0
3,503
Vss Summarize VideoB

Use to summarize a recorded video via the LVS summarization microservice (HITL-gated) with a VLM fallback. Not for report generation or live RTSP captioning.

ai-agentsgobash
0
3,503
Warp Compile Time OptimizerA

Use when compile time or startup time is the problem in code that uses Warp: a request to improve, optimize, or cut compile times; an app that is slow to start or stalls at the first wp.launch; seconds of compiling before real work begins; JIT modules recompiling on every run or every CI job. Only applies when the code being optimized uses Warp kernels. Not for steady-state kernel runtime, memory, correctness, building Warp itself from source, or nvcc/C++ build times.

ai-agentspythonc++
0
3,503
Warp Debug GradientsA

Use to diagnose and fix incorrect gradients in differentiable Warp programs. Anything trained, optimized, calibrated, or fit through Warp kernels depends on wp.Tape gradients, so treat any misbehavior of such a workflow as a gradient problem until proven otherwise — use this when training diverges or NaNs, won't train at all, stalls or plateaus above the expected loss, converges to a wrong or biased answer, is worse than a reference implementation, works at small scale but fails at production...

ai-agentspythonrust
0
3,503
Warp EvalA

Evaluate whether an existing hot path is a credible NVIDIA Warp candidate. Use for irregular or spatial queries, particle or geometry simulation, branch-heavy loops, many small launches, host fallbacks, or large intermediates. CPU-only code and absent GPU dependencies are normal unless NVIDIA is prohibited. Exclude required cross-vendor or CPU-only deployment, vendor-lowered dense or NN layers, general Warp API questions, and already-selected Warp kernels. Contribution policy alone is not exc...

ai-agentspythongo
0
3,503
Bionemo Kermt Add Cmim PretrainA

Convert a grover_base checkpoint (encoder-only or encoder + vocab heads) into a hybrid checkpoint by adding a randomly-initialized cMIM decoder + latent_dist, then continue pretraining on the user's corpus as hybrid (vocab + contrast). Effectively kermt-continue-pretrain with a one-time ckpt-conversion step prepended.

ai-agentspythondocker
0
3,503
Bionemo Kermt Continue PretrainA

Continue KERMT pretraining on a custom SMILES corpus with a grover_base, cmim, or hybrid checkpoint. Use a local checkpoint or optionally download a pinned Hugging Face model bundle using HF_TOKEN if configured. Run containerized training and write model bundles, prepared data, logs, and checkpoints to user-selected host directories.

ai-agentspythongo
0
3,503
Bionemo Kermt EmbedA

Extract per-molecule embeddings from any encoder-bearing KERMT checkpoint. Use a local checkpoint or optionally download a pinned Hugging Face model bundle using HF_TOKEN if configured. Run containerized embedding extraction and write model bundles, per-readout .npy embeddings, canonical SMILES, and validity arrays to user-selected host directories.

ai-agentspythongo
0
3,503
Bionemo Kermt FinetuneA

Finetune a pretrained KERMT encoder on a labeled CSV. Validate the checkpoint and data, prepare features, and run containerized training. Use a local checkpoint or optionally download a pinned Hugging Face model bundle using HF_TOKEN if configured. Write model bundles, prepared data, logs, and trained models to user-selected host directories.

ai-agentspythongo
0
3,503
Bionemo Kermt InferA

Run predictions with a finetuned KERMT checkpoint on a SMILES-only CSV. The skill validates that the input ckpt has task FFN heads (refuses pretrain ckpts with a redirect to kermt-finetune), validates the CSV, prepares the data (clean + rdkit_2d features), then launches main.py predict inside the kermt container (blocking, minutes-scale).

ai-agentspythonbash
0
3,503
Bionemo Kermt MonitorA

Check progress for a detached KERMT run (pretrain, finetune, or any kermt_run_detached invocation). Reads run.json, queries docker for container state, tails the pretrain/finetune log, and parses progress lines (epoch, step, val loss).

ai-agentspythongo
0
3,503