InstructionWild

Instruction Dataset

Creating a large-scale user-based instruction dataset for natural language processing research and development

GitHub

455 stars
9 watching
41 forks
last commit: over 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
orhonovich/unnatural-instructionsA collection of automatically generated instructions for training language models.176
x2fd/lvis-instruct4vA dataset of fine-grained visual instructions generated by prompting a large language model with images from another dataset131
flagopen/flaginstructA collection of diverse instruction corpora for improving the development and tuning of Chinese Language Models173
nytud/huluA collection of linguistic datasets and benchmarks for natural language understanding tasks8
justfollowus/natural-language-processingComprehensive resource for learning natural language processing (NLP) with a structured course outline and recommended readings.834
baai-wudao/modelA repository of pre-trained language models for various tasks and domains.121
michael-wzhu/promptcblueA large-scale instruction-tuning dataset for multi-task and few-shot learning in the medical domain328
ffxsam/vue-typescript-cookbookA cookbook and resource guide for developers learning Vue.js with TypeScript273
wavelets/thinkstats2Text and supporting code for a comprehensive statistical analysis book with accompanying software8
joelcoxokc/aurelia-interfaceProvides a set of custom HTML elements and attributes to build cross-platform applications with platform-specific styles, themes, and behaviors.85
benjamintanweihao/elixir-cheatsheetsA collection of concise guides and reference materials for learning Elixir programming language and its ecosystem105
bingwen/free-programming-booksA curated list of resources for learning programming languages and software development49
mbzuai-nlp/bactrian-xA collection of multilingual language models trained on a dataset of instructions and responses in various languages.94
zhuangbiaowei/open_source_analysisAn analysis of notable software projects21
zhuiyitechnology/pretrained-modelsA collection of pre-trained language models for natural language processing tasks989