few-shot learning

Designs, implements, and analyzes methods, prompts, and training or adaptation procedures that enable models to perform tasks from very small numbers of labeled examples or none; this includes building few-shot and zero-shot prompting workflows, exemplar selection and conditioning strategies, few-shot training/adaptation pipelines, and evaluation protocols that measure few-shot generalization and performance.

few-shotlearning

Recent Skill Trend

Momentum and market value over time
Trending
Score
No comparison yet
-0.02
Oct 01, 2026Oct 01, 2026
Career
Value
No comparison yet
$198K/year
Oct 01, 2026Oct 01, 2026

Recommended Survey Paper

Quick overview of the field
View more

Must-Read Papers

Most classic and influential ideas
View more

Improving Instruct Models for Free: A Study on Partial Adaptation

Apr 15, 2025
OI
Ozan Irsoy
🏛️ Bloomberg | NVIDIA

Instruction tuning often causes pretrained language models to forget foundational knowledge and over-specialize in conversational patterns, thereby degrading in-context learning (ICL) performance. This work identifies an intrinsic trade-off between instruction-following capability and ICL ability. To address it, we propose *partial adaptation*, a lightweight, parameter-efficient tuning paradigm: leveraging LoRA or Adapter modules, we freeze subsets of model parameters and progressively scale adaptation strength—without additional training or extra parameters. Evaluated across 12 canonical NLP few-shot tasks, our method improves average accuracy by 4.2%, while incurring only a marginal drop (−1.8%) in AlpacaEval instruction-following scores. Notably, this is the first systematic study to characterize performance trajectories across multiple model families and scales. Our approach offers a scalable, low-overhead pathway to balance instruction alignment with generalization—preserving ICL competence without compromising task-specific fidelity.

Balancing instruction tuning and pre-training knowledge retentionExploring trade-offs between instruction following and learning abilitiesMitigating performance loss in few-shot in-context learning

This study investigates the capacity of small language models to effectively use tools without relying on complex adaptation mechanisms. Focusing on Llama-3.2-3B-Instruct, the authors systematically evaluate four adaptation strategies—hypernetwork-generated LoRA weights, few-shot prompting, document-based prompting, and value-guided beam search—across four tool-use benchmarks. Experimental results demonstrate that few-shot prompting yields a 21.5% performance gain, document prompting contributes an additional 5.0%, while hypernetwork-generated LoRA weights show no significant improvement. Notably, the 3B-parameter model achieves 79.7% of GPT-5’s average performance at only one-tenth of the inference latency. These findings underscore the pivotal role of prompt engineering in enabling efficient tool use with lightweight models and offer a promising direction for resource-constrained settings.

few-shot adaptationmodel adaptationprompt engineering

Evaluating Generalization and Representation Stability in Small LMs via Prompting

Jun 15, 2025
RR
Rahul Raja
🏛️ Carnegie Mellon University | Boston University

This work systematically evaluates the generalization capability and representation robustness of small language models (SLMs) under two adaptation paradigms—few-shot prompting and supervised fine-tuning—focusing on low-resource settings, out-of-distribution (OOD) generalization, and multi-task scenarios. Methodologically, we integrate centered kernel alignment (CKA) and representational similarity analysis (RSA) for representation similarity quantification, complemented by OOD generalization benchmarks and multi-scale model comparisons. Our study is the first to characterize the knowledge internalization mechanisms of these paradigms through the lenses of representation stability and abstraction level. Results show that prompt-based learning yields more flexible representations but exhibits fragile OOD generalization; in contrast, fine-tuning achieves greater robustness yet suffers from overfitting and reduced abstraction depth. These findings provide interpretable theoretical foundations and empirically grounded guidelines for selecting adaptation strategies for SLMs in resource-constrained environments.

Analyze representation stability across adaptation strategiesAssess generalization of small LMs via prompting and fine-tuningCompare robustness in low-resource and OOD settings

This study addresses the unreliability of selecting and evaluating few-shot adaptation strategies under clinical distribution shifts by proposing the Adapter and Automator architectures. Methodologically, it defines an expanded adaptation space combined with reliability rules to automatically search for optimal strategy combinations. Furthermore, evidence-based weighted fusion and reliability screening mechanisms are introduced to achieve efficient few-shot adaptation. The approach integrates techniques from large model pre-training, few-shot learning, and AutoML. Experimental results demonstrate that the proposed method attains state-of-the-art performance on critical care datasets using only minimal patient data, providing a robust and reliable solution for transfer learning in clinical scenarios.

clinical distribution shiftsfew-shot adaptationpretrained clinical models

From Instance Training to Instruction Learning: Task Adapters Generation from Instructions

Jun 18, 2024
HL
Huanxuan Liao
🏛️ Chinese Academy of Sciences | University of Chinese Academy of Sciences | Tencent | Unisound

Instruction fine-tuning (IFT) suffers from heavy reliance on large-scale annotated examples and poor few-shot cross-task generalization. To address this, we propose an instruction-driven zero-shot adapter generation framework. Our method introduces three key innovations: (1) the first end-to-end paradigm mapping natural-language instructions directly to adapter parameters; (2) a two-stage hypernetwork training scheme that decouples instruction understanding from parameter generation; and (3) the first integration of knowledge distillation into instruction learning to align instruction-level and instance-level training signals. Evaluated on Super-Natural Instructions and P3 benchmarks, our approach matches or surpasses state-of-the-art meta-trained and hypernetwork-based models in task performance, while significantly reducing inference computational overhead. This work establishes a new paradigm for efficient, low-resource generalization of large language models.

Automate task-specific model constructionEnhance LLMs' cross-task generalizationReduce reliance on extensive task data

Latest Papers

What's happening recently
View more

This study addresses the challenge of extracting machine learning pipeline stages, which is constrained by domain diversity and where existing methods rely on manual annotation or limited classifiers. This work systematically investigates, for the first time, the potential of small language models (SLMs) to parse ML pipeline structures leveraging their inherent code comprehension capabilities without fine-tuning, employing Cochran’s Q test, McNemar’s test, and goodness-of-fit evaluations for rigorous assessment. The findings indicate that while SLMs demonstrate robust performance, they do not surpass existing classifiers; however, the core contribution lies in revealing that different classification approaches significantly influence practical insights. Despite the limitation of high inference costs, this research establishes a novel paradigm for automated ML structure parsing.

Code ClassificationMachine Learning PipelinesReverse Engineering

This study addresses the difficulty language model agents face in efficiently adapting execution frameworks to diverse tasks at test time. To this end, this work proposes "framework learning," which formulates framework revision as meta-learning over executable programs. Specifically, a proposer model is trained via reinforcement learning to iteratively refine a solver's code framework using execution feedback, thereby enabling test-time adaptation without parameter updates. By integrating large language model agents with program synthesis and automated repair techniques, this approach endows agents with the capacity to continuously generalize and improve from experience. Experimental results demonstrate significant performance gains on reasoning and multi-hop question answering tasks, validating that such test-time adaptation capabilities transfer effectively to unseen tasks.

Executable ProgramsHarness LearningMeta-Learning

This work addresses the fragmented landscape of post-training adaptation techniques, which suffer from inconsistent terminology and a lack of unified comparative or governance frameworks. To resolve this, the paper introduces the first six-dimensional taxonomy—spanning mechanism, objective, data requirements, persistence, structural scope, and model type—that systematically integrates mainstream approaches such as fine-tuning, retrieval augmentation, prompt engineering, model editing, and machine unlearning. This framework clarifies conceptual boundaries and reveals evolutionary and compositional relationships among methods. Beyond standardizing terminology, it enables standardized technical documentation, model change tracking, and AI governance analysis. The study further identifies critical challenges, including evaluation rigor, reproducibility, continual adaptation, multimodal alignment, and governance-aware workflows.

AI governancefoundation modelsmodel modification

Hot Scholars

PK

Parisa Kordjamshidi

Associate Professor, CSE, Michigan State University
Natural Language ProcessingVision & LanguageNeurosymbolic AISpatial Language Understanding
JB

Johannes Bjerva

Full Professor, Department of Computer Science, Aalborg University
Natural Language ProcessingNLP SecurityLLM SecurityComputational Linguistics
MY

Mohammad Yaqub

Researcher in Biomedical Engineering, Associate professor at MBZUAI
Artificial IntelligenceMedical Image AnalysisMachine LearningDeep learning
AJ

Alexis Joly

Research Director, Inria, Montpellier University, LIRMM
machine learningbiodiversityinformation retrievalplant identification
TC

Tianlong Chen

Assistant Professor, CS@UNC Chapel Hill; Chief AI Scientist, hireEZ
Machine LearningAI4ScienceComputer VisionSparsity