tigerchen52/awesome_role_of_small_models

a curated list of the role of small models in the LLM era

Python

113

51 commits

updated Sep 23, 2024

See the code

README

The Role of Small Models

Awesome PDF GitHub License

This work is ongoing, and we welcome any comments or suggestions.

Please feel free to reach out if you find we have overlooked any relevant papers.

What is the Role of Small Models in the LLM Era: A Survey

Lihu Chen1   Gaël Varoquaux2  

1 Imperial College London, UK    2 Soda, Inria Saclay, France   



Content List

Collaboration

SMs Enhance LLMs

Data Curation

Curating pre-training data

TitleTopicVenueCode
Data selection for language models via importance resamplingData Selection PDF Badge PDF Badge
When Less is More: Investigating Data Pruning for Pretraining LLMs at ScaleData Selection PDF Badge
CCNet: Extracting High Quality Monolingual Datasets from Web Crawl DataData Selection PDF Badge PDF Badge
QuRating: Selecting High-Quality Data for Training Language ModelsData Selection PDF Badge PDF Badge
DoReMi: Optimizing Data Mixtures Speeds Up Language Model PretrainingData Reweighting PDF Badge PDF Badge

Curating Instruction-tuning Data

TitleTopicVenueCode
MoDS: Model-oriented Data Selection for Instruction TuningData Selection PDF Badge PDF Badge
LESS: Selecting Influential Data for Targeted Instruction TuningData Selection PDF Badge PDF Badge
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction TuningData Selection PDF Badge PDF Badge

Weak-to-Strong Paradigm

Using weaker (smaller) models to align stronger (larger) models

TitleTopicVenueCode
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak SupervisionWeak-to-Strong PDF Badge PDF Badge
Weak-to-Strong Search: Align Large Language Models via Searching over Small Language ModelsWeak-to-Strong PDF Badge PDF Badge
Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of ExpertsWeak-to-Strong PDF Badge PDF Badge
Improving Weak-to-Strong Generalization with Reliability-Aware AlignmentWeak-to-Strong PDF Badge PDF Badge
Aligner: Efficient Alignment by Learning to CorrectWeak-to-Strong PDF Badge PDF Badge
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models Weak-to-Strong PDF Badge PDF Badge
Theoretical Analysis of Weak-to-Strong Generalization Weak-to-Strong PDF Badge

Efficient Inference

Ensembling different-size models to reduce inference costs

TitleTopicVenueCode
Efficient Edge Inference by Selective QueryModel Cascading PDF Badge PDF Badge
FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving PerformanceModel Cascading PDF Badge
Data Shunt: Collaboration of Small and Large Models for Lower Costs and Better PerformanceModel Cascading PDF Badge PDF Badge
AutoMix: Automatically Mixing Language ModelsModel Cascading PDF Badge PDF Badge
FrugalML: How to use ML Prediction APIs more accurately and cheaplyModel Cascading PDF Badge PDF Badge
Model Cascading: Towards Jointly Improving Efficiency and Accuracy of NLP SystemsModel Cascading PDF Badge
Routing to the Expert: Efficient Reward-guided Ensemble of Large Language ModelsModel Routing PDF Badge
Tryage: Real-time, intelligent Routing of User Prompts to Large Language ModelsModel Routing PDF Badge
OrchestraLLM: Efficient Orchestration of Language Models for Dialogue State TrackingModel Routing PDF Badge
RouteLLM: Learning to Route LLMs with Preference DataModel Routing PDF Badge PDF Badge
Fly-Swat or Cannon? Cost-Effective Language Model Choice via Meta-ModelingModel Routing PDF Badge PDF Badge
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing Model Routing PDF Badge PDF Badge
LLM-BLENDER: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion Model Routing PDF Badge PDF Badge
RouterBench: A Benchmark for Multi-LLM Routing System Model Routing PDF Badge PDF Badge
Large Language Model Routing with Benchmark Datasets Model Routing PDF Badge

Speculative Decoding

TitleTopicVenueCode
Fast Inference from Transformers via Speculative DecodingSpeculative Decoding PDF Badge PDF Badge
Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative DecodingSpeculative Decoding PDF Badge PDF Badge
Accelerating Large Language Model Decoding with Speculative SamplingSpeculative Decoding PDF Badge PDF Badge

Evaluating LLMs

Using SMs to evaluate LLM's generations

TitleTopicVenueCode
BERTScore: Evaluating Text Generation with BERTGeneral Evaluation PDF Badge PDF Badge
BARTScore: Evaluating Generated Text as Text GenerationGeneral Evaluation PDF Badge PDF Badge
Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language GenerationUncertainty PDF Badge PDF Badge
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language modelsUncertainty PDF Badge PDF Badge
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy ModelsPerformance Prediction PDF Badge PDF Badge

Domain Adaptation

Using domain-specific SMs to adjust token probability of LLMs at decoding time

TitleTopicVenueCode
CombLM: Adapting Black-Box Language Models through Small Fine-Tuned ModelsWhite-box Domain Adaptation PDF Badge
Inference-Time Policy Adapters (IPA): Tailoring Extreme-Scale LMs without Fine-tuningWhite-box Domain Adaptation PDF Badge PDF Badge
Tuning Language Models by ProxyWhite-box Domain Adaptation PDF Badge PDF Badge

Using domain-specific SMs to generate knowledge for LLMs at reasoning time

TitleTopicVenueCode
Knowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language ModelsBlack-box Domain Adaptation PDF Badge PDF Badge
BLADE: Enhancing Black-box Large Language Models with Small Domain-Specific ModelsBlack-box Domain Adaptation PDF Badge

Retrieval Augmented Generation

Using SMs to retrieve knowledge for enhancing generations:

TitleTopicVenueCode
Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksDocuments PDF Badge
KnowledGPT: Enhancing Large Language Models with Retrieval and Storage Access on Knowledge BasesKnowledge Bases PDF Badge
End-to-End Table Question Answering via Retrieval-Augmented GenerationTables PDF Badge
DocPrompting: Generating Code by Retrieving the Docs Codes PDF Badge PDF Badge
Toolformer: Language Models Can Teach Themselves to Use Tools Tools PDF Badge PDF Badge
Retrieval-Augmented Multimodal Language ModelingImages PDF Badge

Prompt-based Reasoning

Using SMs to augment prompts for LLMs

TitleTopicVenueCode
UPRISE: Universal Prompt Retrieval for Improving Zero-Shot EvaluationRetrieving Prompts PDF Badge PDF Badge
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex ReasoningDecomposing Complex Problems PDF Badge PDF Badge
Small Models are Valuable Plug-ins for Large Language ModelsGenerating Pseudo Labels PDF Badge PDF Badge
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-ThoughtGenerating Pseudo Labels PDF Badge
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation Generating Feedback PDF Badge
Small Language Models Improve Giants by Rewriting Their OutputsGenerating Feedback PDF Badge PDF Badge

Deficiency Repair

Developing SM plugins to repair deficiencies:

TitleTopicVenueCode
Small Agent Can Also Rock! Empowering Small Language Models as Hallucination DetectorHallucinations PDF Badge PDF Badge
Reconfidencing LLMs from the Grouping Loss PerspectiveHallucinations PDF Badge
Imputing Out-of-Vocabulary Embeddings with LOVE Makes LanguageModels Robust with Little CostOut-Of-Vocabulary Words PDF Badge PDF Badge

Contrasting LLMs and SMs for better generations:

TitleTopicVenueCode
Contrastive Decoding: Open-ended Text Generation as OptimizationReducing Repeated Texts PDF Badge PDF Badge
Alleviating Hallucinations of Large Language Models through Induced HallucinationsMitigating Hallucinations PDF Badge
Contrastive Decoding Improves Reasoning in Large Language ModelsAugmenting Reasoning Capabilities PDF Badge
CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction FollowingSafeguarding Privacy PDF Badge

LLMs Enhance SMs

Knowledge Distillation

Black-box Distillation:

TitleTopicVenueCode
Explanations from Large Language Models Make Small Reasoners BetterChain-Of-Thought Distillation PDF Badge
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model SizesChain-Of-Thought Distillation PDF Badge PDF Badge
Distilling Reasoning Capabilities into Smaller Language ModelsChain-Of-Thought Distillation PDF Badge PDF Badge
Teaching Small Language Models to ReasonChain-Of-Thought Distillation PDF Badge
Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step Chain-Of-Thought Distillation PDF Badge PDF Badge
Specializing Smaller Language Models towards Multi-Step ReasoningChain-Of-Thought Distillation PDF Badge
TinyLLM: Learning a Small Student from Multiple Large Language Models Chain-Of-Thought Distillation PDF Badge
Lion: Adversarial Distillation of Proprietary Large Language ModelsInstruction Following Distillation PDF Badge PDF Badge
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-TuningInstruction Following Distillation PDF Badge PDF Badge

White-box Distillation:

TitleTopicVenueCode
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighterLogits PDF Badge
ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale TransformersIntermediate Features PDF Badge PDF Badge
Less is More: Task-aware Layer-wise Distillation for Language Model CompressionIntermediate Features PDF Badge PDF Badge
MiniLLM: Knowledge Distillation of Large Language ModelsIntermediate Features PDF Badge PDF Badge
LLM-QAT: Data-Free Quantization Aware Training for Large Language ModelsIntermediate Features PDF Badge

Data Synthesis

Data Augmentation:

TitleTopicVenueCode
Improving data augmentation for low resource speech-to-text translation with diverse paraphrasingText Paraphrase PDF Badge
Paraphrasing with Large Language ModelsText Paraphrase PDF Badge
Query Rewriting for Retrieval-Augmented Large Language ModelsQuery Rewriting PDF Badge PDF Badge
LLMvsSmall Model? Large Language Model Based Text Augmentation Enhanced Personality Detection ModelSpecific Tasks PDF Badge
Data Augmentation for Intent Classification with Off-the-shelf Large Language ModelsSpecific Tasks PDF Badge PDF Badge
Weakly Supervised Data Augmentation Through Prompting for Dialogue UnderstandingSpecific Tasks PDF Badge

Training Data Generation:

TitleTopicVenueCode
Want To Reduce Labeling Cost? GPT-3 Can HelpLabel Annotation PDF Badge
Self-Guided Noise-Free Data Generation for Efficient Zero-Shot LearningLabel Annotation PDF Badge
ZeroGen: Efficient Zero-shot Learning via Dataset GenerationDataset Generation PDF Badge PDF Badge
Generating Training Data with Language Models: Towards Zero-Shot Language UnderstandingDataset Generation PDF Badge PDF Badge
Increasing Diversity While Maintaining Accuracy: Text Data Generation with Large Language Models and Human InterventionsDataset Generation PDF Badge
Synthetic Data Generation with Large Language Models for Text Classification: Potential and LimitationsDataset Generation PDF Badge
Does Synthetic Data Generation of LLMs Help Clinical Text Mining?Dataset Generation PDF Badge
Exploiting Asymmetry for Synthetic Training Data Generation: SynthIE and the Case of Information ExtractionDataset Generation PDF Badge PDF Badge
ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech DetectionDataset Generation PDF Badge PDF Badge

Competition

Computation-constrained Environment

Task-specific Environment

Interpretability-required Environment

Citation

@misc{chen2024rolesmallmodelsllm,
      title={What is the Role of Small Models in the LLM Era: A Survey}, 
      author={Lihu Chen and Gaël Varoquaux},
      year={2024},
      eprint={2409.06857},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2409.06857}, 
}

tigerchen52/awesome_role_of_small_models

a curated list of the role of small models in the LLM era

Python

113

51 commits

updated Sep 23, 2024

See the code

README

The Role of Small Models

Awesome PDF GitHub License

This work is ongoing, and we welcome any comments or suggestions.

Please feel free to reach out if you find we have overlooked any relevant papers.

What is the Role of Small Models in the LLM Era: A Survey

Lihu Chen1   Gaël Varoquaux2  

1 Imperial College London, UK    2 Soda, Inria Saclay, France   



Content List

Collaboration

SMs Enhance LLMs

Data Curation

Curating pre-training data

TitleTopicVenueCode
Data selection for language models via importance resamplingData Selection PDF Badge PDF Badge
When Less is More: Investigating Data Pruning for Pretraining LLMs at ScaleData Selection PDF Badge
CCNet: Extracting High Quality Monolingual Datasets from Web Crawl DataData Selection PDF Badge PDF Badge
QuRating: Selecting High-Quality Data for Training Language ModelsData Selection PDF Badge PDF Badge
DoReMi: Optimizing Data Mixtures Speeds Up Language Model PretrainingData Reweighting PDF Badge PDF Badge

Curating Instruction-tuning Data

TitleTopicVenueCode
MoDS: Model-oriented Data Selection for Instruction TuningData Selection PDF Badge PDF Badge
LESS: Selecting Influential Data for Targeted Instruction TuningData Selection PDF Badge PDF Badge
What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction TuningData Selection PDF Badge PDF Badge

Weak-to-Strong Paradigm

Using weaker (smaller) models to align stronger (larger) models

TitleTopicVenueCode
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak SupervisionWeak-to-Strong PDF Badge PDF Badge
Weak-to-Strong Search: Align Large Language Models via Searching over Small Language ModelsWeak-to-Strong PDF Badge PDF Badge
Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of ExpertsWeak-to-Strong PDF Badge PDF Badge
Improving Weak-to-Strong Generalization with Reliability-Aware AlignmentWeak-to-Strong PDF Badge PDF Badge
Aligner: Efficient Alignment by Learning to CorrectWeak-to-Strong PDF Badge PDF Badge
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models Weak-to-Strong PDF Badge PDF Badge
Theoretical Analysis of Weak-to-Strong Generalization Weak-to-Strong PDF Badge

Efficient Inference

Ensembling different-size models to reduce inference costs

TitleTopicVenueCode
Efficient Edge Inference by Selective QueryModel Cascading PDF Badge PDF Badge
FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving PerformanceModel Cascading PDF Badge
Data Shunt: Collaboration of Small and Large Models for Lower Costs and Better PerformanceModel Cascading PDF Badge PDF Badge
AutoMix: Automatically Mixing Language ModelsModel Cascading PDF Badge PDF Badge
FrugalML: How to use ML Prediction APIs more accurately and cheaplyModel Cascading PDF Badge PDF Badge
Model Cascading: Towards Jointly Improving Efficiency and Accuracy of NLP SystemsModel Cascading PDF Badge
Routing to the Expert: Efficient Reward-guided Ensemble of Large Language ModelsModel Routing PDF Badge
Tryage: Real-time, intelligent Routing of User Prompts to Large Language ModelsModel Routing PDF Badge
OrchestraLLM: Efficient Orchestration of Language Models for Dialogue State TrackingModel Routing PDF Badge
RouteLLM: Learning to Route LLMs with Preference DataModel Routing PDF Badge PDF Badge
Fly-Swat or Cannon? Cost-Effective Language Model Choice via Meta-ModelingModel Routing PDF Badge PDF Badge
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing Model Routing PDF Badge PDF Badge
LLM-BLENDER: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion Model Routing PDF Badge PDF Badge
RouterBench: A Benchmark for Multi-LLM Routing System Model Routing PDF Badge PDF Badge
Large Language Model Routing with Benchmark Datasets Model Routing PDF Badge

Speculative Decoding

TitleTopicVenueCode
Fast Inference from Transformers via Speculative DecodingSpeculative Decoding PDF Badge PDF Badge
Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative DecodingSpeculative Decoding PDF Badge PDF Badge
Accelerating Large Language Model Decoding with Speculative SamplingSpeculative Decoding PDF Badge PDF Badge

Evaluating LLMs

Using SMs to evaluate LLM's generations

TitleTopicVenueCode
BERTScore: Evaluating Text Generation with BERTGeneral Evaluation PDF Badge PDF Badge
BARTScore: Evaluating Generated Text as Text GenerationGeneral Evaluation PDF Badge PDF Badge
Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language GenerationUncertainty PDF Badge PDF Badge
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language modelsUncertainty PDF Badge PDF Badge
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy ModelsPerformance Prediction PDF Badge PDF Badge

Domain Adaptation

Using domain-specific SMs to adjust token probability of LLMs at decoding time

TitleTopicVenueCode
CombLM: Adapting Black-Box Language Models through Small Fine-Tuned ModelsWhite-box Domain Adaptation PDF Badge
Inference-Time Policy Adapters (IPA): Tailoring Extreme-Scale LMs without Fine-tuningWhite-box Domain Adaptation PDF Badge PDF Badge
Tuning Language Models by ProxyWhite-box Domain Adaptation PDF Badge PDF Badge

Using domain-specific SMs to generate knowledge for LLMs at reasoning time

TitleTopicVenueCode
Knowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language ModelsBlack-box Domain Adaptation PDF Badge PDF Badge
BLADE: Enhancing Black-box Large Language Models with Small Domain-Specific ModelsBlack-box Domain Adaptation PDF Badge

Retrieval Augmented Generation

Using SMs to retrieve knowledge for enhancing generations:

TitleTopicVenueCode
Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksDocuments PDF Badge
KnowledGPT: Enhancing Large Language Models with Retrieval and Storage Access on Knowledge BasesKnowledge Bases PDF Badge
End-to-End Table Question Answering via Retrieval-Augmented GenerationTables PDF Badge
DocPrompting: Generating Code by Retrieving the Docs Codes PDF Badge PDF Badge
Toolformer: Language Models Can Teach Themselves to Use Tools Tools PDF Badge PDF Badge
Retrieval-Augmented Multimodal Language ModelingImages PDF Badge

Prompt-based Reasoning

Using SMs to augment prompts for LLMs

TitleTopicVenueCode
UPRISE: Universal Prompt Retrieval for Improving Zero-Shot EvaluationRetrieving Prompts PDF Badge PDF Badge
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex ReasoningDecomposing Complex Problems PDF Badge PDF Badge
Small Models are Valuable Plug-ins for Large Language ModelsGenerating Pseudo Labels PDF Badge PDF Badge
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-ThoughtGenerating Pseudo Labels PDF Badge
CaLM: Contrasting Large and Small Language Models to Verify Grounded Generation Generating Feedback PDF Badge
Small Language Models Improve Giants by Rewriting Their OutputsGenerating Feedback PDF Badge PDF Badge

Deficiency Repair

Developing SM plugins to repair deficiencies:

TitleTopicVenueCode
Small Agent Can Also Rock! Empowering Small Language Models as Hallucination DetectorHallucinations PDF Badge PDF Badge
Reconfidencing LLMs from the Grouping Loss PerspectiveHallucinations PDF Badge
Imputing Out-of-Vocabulary Embeddings with LOVE Makes LanguageModels Robust with Little CostOut-Of-Vocabulary Words PDF Badge PDF Badge

Contrasting LLMs and SMs for better generations:

TitleTopicVenueCode
Contrastive Decoding: Open-ended Text Generation as OptimizationReducing Repeated Texts PDF Badge PDF Badge
Alleviating Hallucinations of Large Language Models through Induced HallucinationsMitigating Hallucinations PDF Badge
Contrastive Decoding Improves Reasoning in Large Language ModelsAugmenting Reasoning Capabilities PDF Badge
CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction FollowingSafeguarding Privacy PDF Badge

LLMs Enhance SMs

Knowledge Distillation

Black-box Distillation:

TitleTopicVenueCode
Explanations from Large Language Models Make Small Reasoners BetterChain-Of-Thought Distillation PDF Badge
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model SizesChain-Of-Thought Distillation PDF Badge PDF Badge
Distilling Reasoning Capabilities into Smaller Language ModelsChain-Of-Thought Distillation PDF Badge PDF Badge
Teaching Small Language Models to ReasonChain-Of-Thought Distillation PDF Badge
Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step Chain-Of-Thought Distillation PDF Badge PDF Badge
Specializing Smaller Language Models towards Multi-Step ReasoningChain-Of-Thought Distillation PDF Badge
TinyLLM: Learning a Small Student from Multiple Large Language Models Chain-Of-Thought Distillation PDF Badge
Lion: Adversarial Distillation of Proprietary Large Language ModelsInstruction Following Distillation PDF Badge PDF Badge
Selective Reflection-Tuning: Student-Selected Data Recycling for LLM Instruction-TuningInstruction Following Distillation PDF Badge PDF Badge

White-box Distillation:

TitleTopicVenueCode
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighterLogits PDF Badge
ZeroQuant: Efficient and Affordable Post-Training Quantization for Large-Scale TransformersIntermediate Features PDF Badge PDF Badge
Less is More: Task-aware Layer-wise Distillation for Language Model CompressionIntermediate Features PDF Badge PDF Badge
MiniLLM: Knowledge Distillation of Large Language ModelsIntermediate Features PDF Badge PDF Badge
LLM-QAT: Data-Free Quantization Aware Training for Large Language ModelsIntermediate Features PDF Badge

Data Synthesis

Data Augmentation:

TitleTopicVenueCode
Improving data augmentation for low resource speech-to-text translation with diverse paraphrasingText Paraphrase PDF Badge
Paraphrasing with Large Language ModelsText Paraphrase PDF Badge
Query Rewriting for Retrieval-Augmented Large Language ModelsQuery Rewriting PDF Badge PDF Badge
LLMvsSmall Model? Large Language Model Based Text Augmentation Enhanced Personality Detection ModelSpecific Tasks PDF Badge
Data Augmentation for Intent Classification with Off-the-shelf Large Language ModelsSpecific Tasks PDF Badge PDF Badge
Weakly Supervised Data Augmentation Through Prompting for Dialogue UnderstandingSpecific Tasks PDF Badge

Training Data Generation:

TitleTopicVenueCode
Want To Reduce Labeling Cost? GPT-3 Can HelpLabel Annotation PDF Badge
Self-Guided Noise-Free Data Generation for Efficient Zero-Shot LearningLabel Annotation PDF Badge
ZeroGen: Efficient Zero-shot Learning via Dataset GenerationDataset Generation PDF Badge PDF Badge
Generating Training Data with Language Models: Towards Zero-Shot Language UnderstandingDataset Generation PDF Badge PDF Badge
Increasing Diversity While Maintaining Accuracy: Text Data Generation with Large Language Models and Human InterventionsDataset Generation PDF Badge
Synthetic Data Generation with Large Language Models for Text Classification: Potential and LimitationsDataset Generation PDF Badge
Does Synthetic Data Generation of LLMs Help Clinical Text Mining?Dataset Generation PDF Badge
Exploiting Asymmetry for Synthetic Training Data Generation: SynthIE and the Case of Information ExtractionDataset Generation PDF Badge PDF Badge
ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech DetectionDataset Generation PDF Badge PDF Badge

Competition

Computation-constrained Environment

Task-specific Environment

Interpretability-required Environment

Citation

@misc{chen2024rolesmallmodelsllm,
      title={What is the Role of Small Models in the LLM Era: A Survey}, 
      author={Lihu Chen and Gaël Varoquaux},
      year={2024},
      eprint={2409.06857},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2409.06857}, 
}

Languages

Python

100.0%