LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

2024-03-20Code Available2· sign in to hype

Yaowei Zheng, Richong Zhang, Junhao Zhang, Yanhan Ye, Zheyan Luo, Zhangchi Feng, Yongqiang Ma

Code Available — Be the first to reproduce this paper.

Code

github.com/hiyouga/llama-factory
OfficialIn paperpytorch★ 68,911
github.com/juyongjiang/codeup
pytorch★ 127
github.com/Rcrossmeister/RLQG
pytorch★ 48
github.com/chen-gx/toolevo
none★ 13
github.com/hiyouga/hiyouga
tf★ 9
github.com/smelliecat/aaemime
none★ 1
github.com/Rcrossmeister/Knowledge-to-SQL
pytorch★ 1
github.com/BachOzean/TadE
pytorch★ 0

Abstract

Efficient fine-tuning is vital for adapting large language models (LLMs) to downstream tasks. However, it requires non-trivial efforts to implement these methods on different models. We present LlamaFactory, a unified framework that integrates a suite of cutting-edge efficient training methods. It provides a solution for flexibly customizing the fine-tuning of 100+ LLMs without the need for coding through the built-in web UI LlamaBoard. We empirically validate the efficiency and effectiveness of our framework on language modeling and text generation tasks. It has been released at https://github.com/hiyouga/LLaMA-Factory and received over 25,000 stars and 3,000 forks.

Tasks

Language Modeling Language Modelling Text Generation

LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Code

Abstract

Tasks

Reproductions