tatsu-lab/alpaca
Dataset Card for Alpaca Dataset Summary Alpaca is a dataset of 52,000 instructions and demonstrations generated by OpenAI's text-davinci-003 engine. This instruction data can be used to conduct instruction-tuning for language models and make the language model follow instruction better. The authors built on the data generation pipeline from Self-Instruct framework and made the following modifications: The text-davinci-003 engine to generate the instruction data… See the full description on the dataset page: https://huggingface.co/datasets/tatsu-lab/alpaca.
mlforge datasets pull tatsu-lab/alpaca
Dataset details
About tatsu-lab/alpaca
- Homepage: https://crfm.stanford.edu/2023/03/13/alpaca.html - Repository: https://github.com/tatsu-lab/stanfordalpaca - Paper: - Leaderboard: - Point of Contact: Rohan Taori