This project collects awesome resources (e.g., papers, open-source models) for large language model (LLM)
253
11 commits
updated Mar 28, 2024
This repository collects awesome projects and resources related to large language model (LLM).
StableLM
StableLM: Stability AI Language Models
Colossal-AI
Colossal-AI: Making large AI models cheaper, faster, and more accessible
ChatGLM-6B
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Moss
Moss: An open-source tool-augmented conversational language model from Fudan University GitHub: https://github.com/OpenLMLab/MOSS
LLaMA
LLaMA: Inference code for LLaMA models
Alpaca
Alpaca: The current Alpaca model is fine-tuned from a 7B LLaMA model on 52K instruction-following data generated by the techniques in the Self-Instruct paper
BELLE
BELLE: Be Everyone's Large Language model Engine GitHub: https://github.com/LianjiaTech/BELLE
Vicuna
The release repo for "Vicuna: An Open Chatbot Impressing GPT-4"
Dolly
Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
OpenAssistant
OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
LLM Zoo
LLM Zoo: democratizing ChatGPT
Chinese-LLaMA-Alpaca
Chinese LLaMA & Alpaca LLMs
BayLing
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models
PromptSource
PromptSource is a toolkit for creating, sharing and using natural language prompts. [2k][English]
T0/P3
Multitask Prompted Training Enables Zero-Shot Task Generalization[2k][English]
xP3
Crosslingual Generalization through Multitask Finetuning [Multilingual] [NMT]
Super-Natural-Instruct v2
SUPER-NATURALINSTRUCTIONS: Generalization via Declarative Instructions on 1600+ NLP Tasks [1.6k][Multilingual]
FLAN
Finetuned Language Models are Zero-Shot Learners/Flan Collection [18k][English]
CrossFit
The CrossFit Challenge 🏋️ and The NLP Few-shot Gym 💦 [English]
Self-Instruct
Self-Instruct: Aligning Language Model with Self Generated Instructions [52K] [English] [Generated]
Unnatural Instructions
Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor [240K] [English] [Generated]
Stanford Alpaca
Stanford Alpaca: An Instruction-following LLaMA Model [51.9K] [English] [Generated]
Camel
CAMEL: Communicative Agents for “Mind” Exploration of Large Scale Language Model Society [115K] [English] [Generated]
Dolly
The training data on which dolly-v2-12b is instruction tuned represents natural language instructions generated by Databricks employees [15K] [English]
GuanacoDataset
The dataset for the Guanaco model is designed to enhance the multilingual capabilities and address various linguistic tasks. [534K] [Multilingual]
Link: https://huggingface.co/datasets/JosephusCheung/GuanacoDataset
Chinese-ChatLLaMA
本项目向社区提供中文对话模型 ChatLLama 、中文基础模型 LLaMA-zh 及其训练数据。 [Multilingual]
OIG
Open Instruction Generalist (OIG) [43M] [English]
GPTeacher
GPTeacher: A collection of modular datasets generated by GPT-4, General-Instruct - Roleplay-Instruct - Code-Instruct - and Toolformer [English] [Generated]
CSL
Chinese Scientific Literature Dataset [396K] [Chinese]
GLM-130B
GLM-130B: An Open Bilingual Pre-Trained Model [Multilingual (eng, zh)]
Firefly
Firefly(流萤): 中文对话式大语言模型 [1.1M] [Chinese]
BELLE
BELLE: Be Everyone's Large Language model Engine [1.5M] [Chinese]
Chinese-Vicuna
Chinese-Vicuna: A Chinese Instruction-following LLaMA-based Model —— 一个中文低资源的llama+lora方案 [1M] [Chinese] Link: https://github.com/Facico/Chinese-Vicuna
HC3
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection [37k] [Multilingual]
Link: https://github.com/Hello-SimpleAI/chatgpt-comparison-detection
Luotuo
骆驼(Luotuo): Chinese-alpaca-lora [51.6K] [Chinese]
COIG
Chinese Open Instruction Generalist: A Preliminary Release [Chinese]
ShareGPT52K
This dataset is a collection of approximately
52,00090,000 conversations scraped via the ShareGPT API before it was shut down. These conversations include both user prompts and responses from OpenAI's ChatGPT. [90K] [Multiligual]
Stack-Exchange-Preferences
Huggingface H4 Stack Exchange Preferences Dataset [10M] [English]
Link: https://huggingface.co/datasets/HuggingFaceH4/stack-exchange-preferences
HH-RLHF
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback [169K] [English]
SHP
Stanford Human Preferences Dataset (SHP) [385K] [English]
OASST
OpenAssistant Conversations -- Democratizing Large Language Model Alignment [161K] [Multilingual]
GPT4All
GPT4All: Training an Assistant-style Chatbot with Large Scale Data Distillation from GPT-3.5-Turbo [165K] [Multilingual]
InstructWild
Instruction in the Wild: A User-based Instruction Dataset [104K] [Multilingual] [Generated]
Survey
Machine Translation
Sentiment Analysis
Multi-Lingual
1.Phoenix: Democratizing ChatGPT across Languages. Zhihong Chen, Feng Jiang, Junying Chen, Tiannan Wang, Fei Yu, Guiming Chen, Hongbo Zhang, Juhao Liang, Chen Zhang, Zhiyi Zhang, Jianquan Li, Xiang Wan, Benyou Wang, Haizhou Li
Dialogue
Summarization
Robot
Logical Reasoning
Medical AI
Commonsense
Grammatical Error Correction
Text-to-SQL
Question Answering
Keyphrase Generator
Code Intelligence
NLG
Event Extraction
Information Extraction
Data Augmentation
Keyphrase Generation
Industrial Engineering
Mathematical Word Problem
Recommendation
Chat-REC: Towards Interactive and Explainable LLMs-Augmented Recommender System. Yunfan Gao, Tao Sheng, Youlin Xiang, Yun Xiong, Haofen Wang, Jiawei Zhang
Is ChatGPT a Good Recommender? A Preliminary Study. Junling Liu, Chao Liu, Renjie Lv, Kang Zhou, Yan Zhang
Uncovering ChatGPT’s Capabilities in Recommender Systems. Sunhao Dai, Ninglu Shao, Haiyuan Zhao, Weijie Yu, Zihua Si, Chen Xu, Zhongxiang Sun, Xiao Zhang, Jun Xu
Safety
Application
AGI
Analysis, Challenge and Future Work
Comparative Analysis of CHATGPT and the evolution of language models. Oluwatosin Ogundare, Gustavo Quiros Araya
Summary of ChatGPT/GPT-4 Research and Perspective Towards the Future of Large Language Models. Yiheng Liu, Tianle Han, Siyuan Ma, Jiayue Zhang, Yuanyuan Yang, Jiaming Tian, Hao He, Antong Li, Mengshen He, Zhengliang Liu, Zihao Wu, Dajiang Zhu, Xiang Li, Ning Qiang, Dingang Shen, Tianming Liu, Bao Ge
Can we trust the evaluation on ChatGPT? Rachith Aiyappa, Jisun An, Haewoon Kwak, Yong-Yeol Ahn
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? Chaoning Zhang, Chenshuang Zhang, Sheng Zheng, Yu Qiao, Chenghao Li, Mengchun Zhang, Sumit Kumar Dam, Chu Myaet Thwal, Ye Lin Tun, Le Luang Huy, Donguk kim, Sung-Ho Bae, Lik-Hang Lee, Yang Yang, Heng Tao Shen, In So Kweon, Choong Seon Hong
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models. Junjie Ye, Xuanting Chen, Nuo Xu, Can Zu, Zekai Shao, Shichun Liu, Yuhan Cui, Zeyang Zhou, Chao Gong, Yang Shen, Jie Zhou, Siming Chen, Tao Gui, Qi Zhang, Xuanjing Huang
ChatGPT: A Meta-Analysis after 2.5 Months. Christoph Leiter, Ran Zhang, Yanran Chen, Jonas Belouadi, Daniil Larionov, Vivian Fresen, Steffen Eger
On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective. Jindong Wang, Xixu Hu, Wenxin Hou, Hao Chen, Runkai Zheng, Yidong Wang, Linyi Yang, Haojun Huang, Wei Ye, Xiubo Geng, Binxin Jiao, Yue Zhang, Xing Xie
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT. Qihuang Zhong, Liang Ding, Juhua Liu, Bo Du, Dacheng Tao
Is ChatGPT a General-Purpose Natural Language Processing Task Solver? Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, Diyi Yang
A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity. Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, Pascale Fung
A Categorical Archive of ChatGPT Failures. Ali Borji
ChatGPT and Software Testing Education: Promises & Perils. Sajed Jalil, Suzzana Rafi, Thomas D. LaToza, Kevin Moran, Wing Lam
Exploring AI Ethics of ChatGPT: A Diagnostic Analysis. Terry Yue Zhuo, Yujin Huang, Chunyang Chen, Zhenchang Xing
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection. Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, Yupeng Wu
Can ChatGPT and Bard Generate Aligned Assessment Items? A Reliability Analysis against Human Performance. Abdolvahab Khademi
Large Language Models Can Be Easily Distracted by Irrelevant Context Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed Chi, Nathanael Scharli, Denny Zhou
GPT as Knowledge Worker: A Zero-Shot Evaluation of (AI)CPA Capabilities. Jillian Bommarito, Michael J Bommarito II, Jessica Katz, Daniel Martin Katz
Exploring ChatGPT's Ability to Rank Content: A Preliminary Study on Consistency with Human Preferences. Yunjie Ji, Yan Gong, Yiping Peng, Chao Ni, Peiyan Sun, Dongyu Pan, Baochang Ma*, Xiangang Li
ChatGPT versus Traditional Question Answering for Knowledge Graphs: Current Status and Future Directions Towards Knowledge Graph Chatbots. Reham Omar, Omij Mangukiya, Panos Kalnis, Essam Mansour
Conversational Process Modelling: State of the Art, Applications, and Implications in Practice. Nataliia Klievtsova, Janik-Vasily Benzin, Timotheus Kampik, Juergen Mangler, Stefanie Rinderle-Ma
Can ChatGPT-like Generative Models Guarantee Factual Accuracy? On the Mistakes of New Generation Search Engines. Ruochen Zhao, Xingxuan Li, Yew Ken Chia, Bosheng Ding, Lidong Bing
ChatLog: Recording and Analyzing ChatGPT Across Time. Shangqing Tu, Chunyang Li, Jifan Yu, Xiaozhi Wang, Lei Hou, Juanzi Li
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination. Zihao Li
This project collects awesome resources (e.g., papers, open-source models) for large language model (LLM)
253
11 commits
updated Mar 28, 2024
This repository collects awesome projects and resources related to large language model (LLM).
StableLM
StableLM: Stability AI Language Models
Colossal-AI
Colossal-AI: Making large AI models cheaper, faster, and more accessible
ChatGLM-6B
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Moss
Moss: An open-source tool-augmented conversational language model from Fudan University GitHub: https://github.com/OpenLMLab/MOSS
LLaMA
LLaMA: Inference code for LLaMA models
Alpaca
Alpaca: The current Alpaca model is fine-tuned from a 7B LLaMA model on 52K instruction-following data generated by the techniques in the Self-Instruct paper
BELLE
BELLE: Be Everyone's Large Language model Engine GitHub: https://github.com/LianjiaTech/BELLE
Vicuna
The release repo for "Vicuna: An Open Chatbot Impressing GPT-4"
Dolly
Databricks’ Dolly, a large language model trained on the Databricks Machine Learning Platform
OpenAssistant
OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.
LLM Zoo
LLM Zoo: democratizing ChatGPT
Chinese-LLaMA-Alpaca
Chinese LLaMA & Alpaca LLMs
BayLing
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models
PromptSource
PromptSource is a toolkit for creating, sharing and using natural language prompts. [2k][English]
T0/P3
Multitask Prompted Training Enables Zero-Shot Task Generalization[2k][English]
xP3
Crosslingual Generalization through Multitask Finetuning [Multilingual] [NMT]
Super-Natural-Instruct v2
SUPER-NATURALINSTRUCTIONS: Generalization via Declarative Instructions on 1600+ NLP Tasks [1.6k][Multilingual]
FLAN
Finetuned Language Models are Zero-Shot Learners/Flan Collection [18k][English]
CrossFit
The CrossFit Challenge 🏋️ and The NLP Few-shot Gym 💦 [English]
Self-Instruct
Self-Instruct: Aligning Language Model with Self Generated Instructions [52K] [English] [Generated]
Unnatural Instructions
Unnatural Instructions: Tuning Language Models with (Almost) No Human Labor [240K] [English] [Generated]
Stanford Alpaca
Stanford Alpaca: An Instruction-following LLaMA Model [51.9K] [English] [Generated]
Camel
CAMEL: Communicative Agents for “Mind” Exploration of Large Scale Language Model Society [115K] [English] [Generated]
Dolly
The training data on which dolly-v2-12b is instruction tuned represents natural language instructions generated by Databricks employees [15K] [English]
GuanacoDataset
The dataset for the Guanaco model is designed to enhance the multilingual capabilities and address various linguistic tasks. [534K] [Multilingual]
Link: https://huggingface.co/datasets/JosephusCheung/GuanacoDataset
Chinese-ChatLLaMA
本项目向社区提供中文对话模型 ChatLLama 、中文基础模型 LLaMA-zh 及其训练数据。 [Multilingual]
OIG
Open Instruction Generalist (OIG) [43M] [English]
GPTeacher
GPTeacher: A collection of modular datasets generated by GPT-4, General-Instruct - Roleplay-Instruct - Code-Instruct - and Toolformer [English] [Generated]
CSL
Chinese Scientific Literature Dataset [396K] [Chinese]
GLM-130B
GLM-130B: An Open Bilingual Pre-Trained Model [Multilingual (eng, zh)]
Firefly
Firefly(流萤): 中文对话式大语言模型 [1.1M] [Chinese]
BELLE
BELLE: Be Everyone's Large Language model Engine [1.5M] [Chinese]
Chinese-Vicuna
Chinese-Vicuna: A Chinese Instruction-following LLaMA-based Model —— 一个中文低资源的llama+lora方案 [1M] [Chinese] Link: https://github.com/Facico/Chinese-Vicuna
HC3
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection [37k] [Multilingual]
Link: https://github.com/Hello-SimpleAI/chatgpt-comparison-detection
Luotuo
骆驼(Luotuo): Chinese-alpaca-lora [51.6K] [Chinese]
COIG
Chinese Open Instruction Generalist: A Preliminary Release [Chinese]
ShareGPT52K
This dataset is a collection of approximately
52,00090,000 conversations scraped via the ShareGPT API before it was shut down. These conversations include both user prompts and responses from OpenAI's ChatGPT. [90K] [Multiligual]
Stack-Exchange-Preferences
Huggingface H4 Stack Exchange Preferences Dataset [10M] [English]
Link: https://huggingface.co/datasets/HuggingFaceH4/stack-exchange-preferences
HH-RLHF
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback [169K] [English]
SHP
Stanford Human Preferences Dataset (SHP) [385K] [English]
OASST
OpenAssistant Conversations -- Democratizing Large Language Model Alignment [161K] [Multilingual]
GPT4All
GPT4All: Training an Assistant-style Chatbot with Large Scale Data Distillation from GPT-3.5-Turbo [165K] [Multilingual]
InstructWild
Instruction in the Wild: A User-based Instruction Dataset [104K] [Multilingual] [Generated]
Survey
Machine Translation
Sentiment Analysis
Multi-Lingual
1.Phoenix: Democratizing ChatGPT across Languages. Zhihong Chen, Feng Jiang, Junying Chen, Tiannan Wang, Fei Yu, Guiming Chen, Hongbo Zhang, Juhao Liang, Chen Zhang, Zhiyi Zhang, Jianquan Li, Xiang Wan, Benyou Wang, Haizhou Li
Dialogue
Summarization
Robot
Logical Reasoning
Medical AI
Commonsense
Grammatical Error Correction
Text-to-SQL
Question Answering
Keyphrase Generator
Code Intelligence
NLG
Event Extraction
Information Extraction
Data Augmentation
Keyphrase Generation
Industrial Engineering
Mathematical Word Problem
Recommendation
Chat-REC: Towards Interactive and Explainable LLMs-Augmented Recommender System. Yunfan Gao, Tao Sheng, Youlin Xiang, Yun Xiong, Haofen Wang, Jiawei Zhang
Is ChatGPT a Good Recommender? A Preliminary Study. Junling Liu, Chao Liu, Renjie Lv, Kang Zhou, Yan Zhang
Uncovering ChatGPT’s Capabilities in Recommender Systems. Sunhao Dai, Ninglu Shao, Haiyuan Zhao, Weijie Yu, Zihua Si, Chen Xu, Zhongxiang Sun, Xiao Zhang, Jun Xu
Safety
Application
AGI
Analysis, Challenge and Future Work
Comparative Analysis of CHATGPT and the evolution of language models. Oluwatosin Ogundare, Gustavo Quiros Araya
Summary of ChatGPT/GPT-4 Research and Perspective Towards the Future of Large Language Models. Yiheng Liu, Tianle Han, Siyuan Ma, Jiayue Zhang, Yuanyuan Yang, Jiaming Tian, Hao He, Antong Li, Mengshen He, Zhengliang Liu, Zihao Wu, Dajiang Zhu, Xiang Li, Ning Qiang, Dingang Shen, Tianming Liu, Bao Ge
Can we trust the evaluation on ChatGPT? Rachith Aiyappa, Jisun An, Haewoon Kwak, Yong-Yeol Ahn
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? Chaoning Zhang, Chenshuang Zhang, Sheng Zheng, Yu Qiao, Chenghao Li, Mengchun Zhang, Sumit Kumar Dam, Chu Myaet Thwal, Ye Lin Tun, Le Luang Huy, Donguk kim, Sung-Ho Bae, Lik-Hang Lee, Yang Yang, Heng Tao Shen, In So Kweon, Choong Seon Hong
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models. Junjie Ye, Xuanting Chen, Nuo Xu, Can Zu, Zekai Shao, Shichun Liu, Yuhan Cui, Zeyang Zhou, Chao Gong, Yang Shen, Jie Zhou, Siming Chen, Tao Gui, Qi Zhang, Xuanjing Huang
ChatGPT: A Meta-Analysis after 2.5 Months. Christoph Leiter, Ran Zhang, Yanran Chen, Jonas Belouadi, Daniil Larionov, Vivian Fresen, Steffen Eger
On the Robustness of ChatGPT: An Adversarial and Out-of-distribution Perspective. Jindong Wang, Xixu Hu, Wenxin Hou, Hao Chen, Runkai Zheng, Yidong Wang, Linyi Yang, Haojun Huang, Wei Ye, Xiubo Geng, Binxin Jiao, Yue Zhang, Xing Xie
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT. Qihuang Zhong, Liang Ding, Juhua Liu, Bo Du, Dacheng Tao
Is ChatGPT a General-Purpose Natural Language Processing Task Solver? Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, Diyi Yang
A Multitask, Multilingual, Multimodal Evaluation of ChatGPT on Reasoning, Hallucination, and Interactivity. Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, Pascale Fung
A Categorical Archive of ChatGPT Failures. Ali Borji
ChatGPT and Software Testing Education: Promises & Perils. Sajed Jalil, Suzzana Rafi, Thomas D. LaToza, Kevin Moran, Wing Lam
Exploring AI Ethics of ChatGPT: A Diagnostic Analysis. Terry Yue Zhuo, Yujin Huang, Chunyang Chen, Zhenchang Xing
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection. Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, Yupeng Wu
Can ChatGPT and Bard Generate Aligned Assessment Items? A Reliability Analysis against Human Performance. Abdolvahab Khademi
Large Language Models Can Be Easily Distracted by Irrelevant Context Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed Chi, Nathanael Scharli, Denny Zhou
GPT as Knowledge Worker: A Zero-Shot Evaluation of (AI)CPA Capabilities. Jillian Bommarito, Michael J Bommarito II, Jessica Katz, Daniel Martin Katz
Exploring ChatGPT's Ability to Rank Content: A Preliminary Study on Consistency with Human Preferences. Yunjie Ji, Yan Gong, Yiping Peng, Chao Ni, Peiyan Sun, Dongyu Pan, Baochang Ma*, Xiangang Li
ChatGPT versus Traditional Question Answering for Knowledge Graphs: Current Status and Future Directions Towards Knowledge Graph Chatbots. Reham Omar, Omij Mangukiya, Panos Kalnis, Essam Mansour
Conversational Process Modelling: State of the Art, Applications, and Implications in Practice. Nataliia Klievtsova, Janik-Vasily Benzin, Timotheus Kampik, Juergen Mangler, Stefanie Rinderle-Ma
Can ChatGPT-like Generative Models Guarantee Factual Accuracy? On the Mistakes of New Generation Search Engines. Ruochen Zhao, Xingxuan Li, Yew Ken Chia, Bosheng Ding, Lidong Bing
ChatLog: Recording and Analyzing ChatGPT Across Time. Shangqing Tu, Chunyang Li, Jifan Yu, Xiaozhi Wang, Lei Hou, Juanzi Li
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination. Zihao Li