Category

Research

Research, evidence gathering, literature, reports, investigation, and synthesis

23,891
skills in category
996
pages available
Security grades appear on each card once the skill has been scanned. Newly imported skills may briefly show without a grade until the backfill job runs.
Open in full browser

Browse research skills

Showing 15,577–15,600 of 23,891 skills

Allava Harnessing Gpt4v Synthesized Data For A Lite Vision Language Model Arxiv 2402 11684v3A

Use this skill when you want to synthesize high-quality instruction data using GPT-4V on diverse image sources for training small VLMs. Avoid it when you cannot afford GPT-4V API calls or are training a large-scale model.

researchgoapi
0
9
Training Compute Optimal Large Language Models Arxiv 2203 15556v1A

Use this skill when you want to know the compute-optimal ratio of model size to training data based on Chinchilla scaling laws. Avoid it when you are not making model/data scaling decisions.

researchgoaws
0
9
Scaling Laws For Neural Language Models Arxiv 2001 08361v1A

Use this skill when you want to understand power-law relationships between model size, data size, and compute for optimal resource allocation. Avoid it when you are not making scaling decisions.

researchgoaws
0
9
Scaling Data Constrained Language Models Arxiv 2305 16264v3A

Use this skill when you want to understand how to optimally train when data is limited and must be repeated, and how to trade off epochs vs model size. Avoid it when you have unlimited unique data and do not need to repeat data.

researchgoaws
0
9
Influence Functions In Deep Learning Arxiv 2002 08484v3A

Use this skill when you want to estimate the influence of individual training examples on model predictions for data debugging and selection. Avoid it when you cannot afford the compute for influence estimation or have too many training examples.

researchgodebugging
0
9
El2n Deep Learning On A Data Diet Arxiv 2107 07075v2A

Use this skill when you want to prune training data using early-epoch error norms (EL2N scores) to select the most informative examples. Avoid it when you cannot compute early training predictions or all data is equally important.

researchgoperformance
0
9
Data Mixing Laws Optimizing Data Mixtures By Predicting Language Modeling Performance Arxiv 2403 16952v2A

Use this skill when you want to predict LLM performance from data mixing ratios and optimize the mix without expensive full-scale training. Avoid it when you have a single data source or cannot run ablation experiments.

researchgoaws
0
9
Beyond Neural Scaling Laws Beating Power Law Scaling Via Data Pruning Arxiv 2206 14486v2A

Use this skill when you want to beat standard scaling laws by pruning low-quality or redundant data points using perplexity or EL2N scores. Avoid it when you do not have compute for data quality scoring or your data is already curated.

researchgoaws
0
9
Badge Batch Active Learning By Diverse Gradient Embeddings Arxiv 1906 03671v2A

Use this skill when you want to select diverse, uncertain samples for annotation using gradient embeddings that capture both uncertainty and diversity. Avoid it when simple uncertainty sampling is sufficient or you cannot compute gradients.

researchgo
0
9
What Makes For Good Visual Tokenizers For Large Language Models Arxiv 2305 12223v2A

Use this skill when you want to understand which visual tokenizer properties matter most for VLM performance. Avoid it when you have already chosen your visual encoder.

researchgoperformance
0
9
Vila On Pre Training For Visual Language Models Arxiv 2312 07533v4A

Use this skill when you want systematic ablation insights on VLM pretraining choices: frozen vs unfrozen LLM, interleaved vs paired data, text mixing. Avoid it when you have a well-established pretraining recipe.

researchgoperformance
0
9
T5 Exploring The Limits Of Transfer Learning With A Unified Text To Text Transformer Arxiv 1910 10683v4A

Use this skill when you want to understand the C4 text cleaning methodology and its impact on LLM pretraining quality. Avoid it when you are not cleaning web text or have a different cleaning pipeline.

researchgoperformance
0
9
Scaling Instruction Finetuned Language Models Arxiv 2210 11416v5A

Use this skill when you want to understand how to scale instruction fine-tuning across number of tasks, model size, and chain-of-thought data for maximum benefit. Avoid it when you are doing single-task fine-tuning.

researchgoaws
0
9
Pali 3 Smaller Faster Stronger Arxiv 2310 09199v2A

Use this skill when you want insights on training a smaller but stronger VLM through improved data quality, SigLIP encoder, and efficient training recipe. Avoid it when you are not optimizing VLM training efficiency.

researchgoperformance
0
9
Openclip An Open Source Implementation Of Clip Arxiv 2212 07143v2A

Use this skill when you want to train CLIP models on open datasets (LAION) with reproducible training recipes and systematic data ablations. Avoid it when you are using pre-trained CLIP models and do not need to train your own.

researchgoperformance
0
9
Open Vocabulary Object Detection Using Captions Arxiv 2011 10678v2A

Use this skill when you want to train an object detector using image captions as weak supervision instead of bounding box annotations. Avoid it when you have abundant bounding box annotations.

researchgoperformance
0
9
Multimodal Neurons In Artificial Neural Networks Arxiv 2103 01002v1A

Use this skill when you want to understand how CLIP learns multimodal neurons that respond to both visual and textual concepts for debugging data quality. Avoid it when you do not need interpretability or representation analysis.

researchgodebugging
0
9
Multimodal Learning With Transformers A Survey Arxiv 2206 06488v2A

Use this skill when you need a comprehensive survey of multimodal learning with transformers covering fusion strategies, pretraining, and applications. Avoid it when you need a specific method rather than a broad overview.

researchgoperformance
0
9
Mm1 Methods Analysis And Insights From Multimodal Llm Pre Training Arxiv 2403 09611v3A

Use this skill when you want to understand optimal data mixing recipes for multimodal LLM pretraining across interleaved, image-text, and text-only data. Avoid it when you only have one data type or cannot run ablation studies to tune the mix.

researchgoperformance
0
9
Llava Onevision Easy Visual Task Transfer Arxiv 2408 03326v2A

Use this skill when you want a unified training recipe for single-image, multi-image, and video understanding with curated data at each stage. Avoid it when you only need single-image understanding.

researchgoperformance
0
9
Lima Less Is More For Alignment Arxiv 2305 11206v1A

Use this skill when you want to align an LLM with just 1,000 carefully curated instruction examples, demonstrating that quality trumps quantity. Avoid it when you have abundant alignment data or need maximum diversity.

researchgo
0
9
Label Noise Sgd Provably Prefers Flat Global Minimizers Arxiv 2106 06530v4A

Use this skill when you want to understand how label noise in training data affects optimization and generalization, particularly for noisy web data. Avoid it when you do not deal with noisy labels.

researchgo
0
9
Improved Baselines With Visual Instruction Tuning Arxiv 2310 03744v2A

Use this skill when you want to curate a high-quality academic-task-oriented instruction dataset for VLM fine-tuning. Avoid it when you need only synthetic conversation data or cannot access the academic VQA datasets.

researchgoperformance
0
9
Idefics2 An 8b Parameters Multimodal Model Arxiv 2405 02246v2A

Use this skill when you want an efficient 8B VLM training recipe with curated data combining web interleaved, paired, and instruction data. Avoid it when you need a larger model or have a different data strategy.

researchgoperformance
0
9