-
Ensemble-Instruct: Generating Instruction-Tuning Data with a Heterogeneous Mixture of LMs
Paper • 2310.13961 • Published • 4 -
ZeroGen: Efficient Zero-shot Learning via Dataset Generation
Paper • 2202.07922 • Published • 1 -
Let's Synthesize Step by Step: Iterative Dataset Synthesis with Large Language Models by Extrapolating Errors from Small Models
Paper • 2310.13671 • Published • 17 -
Fabricator: An Open Source Toolkit for Generating Labeled Training Data with Teacher LLMs
Paper • 2309.09582 • Published • 4
Collections
Discover the best community collections!
Collections including paper arxiv:2309.09530
-
Adapting Large Language Models via Reading Comprehension
Paper • 2309.09530 • Published • 69 -
Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Paper • 2404.03715 • Published • 57 -
Ferret-UI: Grounded Mobile UI Understanding with Multimodal LLMs
Paper • 2404.05719 • Published • 56
-
Adapting Large Language Models via Reading Comprehension
Paper • 2309.09530 • Published • 69 -
TinyGSM: achieving >80% on GSM8k with small language models
Paper • 2312.09241 • Published • 33 -
How to Train Data-Efficient LLMs
Paper • 2402.09668 • Published • 33 -
Let GPT be a Math Tutor: Teaching Math Word Problem Solvers with Customized Exercise Generation
Paper • 2305.14386 • Published