Twin-Merging. [NeurIPS2024] Twin-Merging: Dynamic Integration of Modular Expertise in Model Merging
143MIRACLE. [EMNLP2023]: MIRACLE: Towards Personalized Dialogue Generation with Latent-Space Multiple Personal Attribute Control
12lm-evaluation-harness-fast. speedup for lm-evaluation-harness; support tensor-parallel inference and data-parallel inference; support gptq, bitsandbytes, peft and exllamav2.
8Awesome-Model-Fusion. This GitHub repository summarizes most recent papers and resources related to the model fusion.
7HELM-Extended-Local. support for various name in http/local run; support for gptq/bnb/tensorparallel in local run
3CustomLLMFinetuningHandbook. an example to fine-tuning a Language Learning Model (LLM) from data preparation to deployment.
3Boys-ToolKit. multi-process openai client; vllm data parallelism
3Tools-gradio. custom Widget using gradio
2Paper-Reading-.
2axolotl. axolotl customize
1dpo. Robust recipes for to align language models with human and AI preferences
1LoRA-Pro-fix. Official code for our paper, "LoRA-Pro: Are Low-Rank Adapters Properly Optimized? "
1trl. Train transformer language models with reinforcement learning.
1PythonNotes. python notes for myself
1peft. 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
1Qwen-VL-DPO. Python
1vllm. A high-throughput and memory-efficient inference and serving engine for LLMs
1luogu-lzy. https://www.luogu.com.cn/ problems solving
1Tools-for-HuggingfaceTransformers. custom tools
1