lora
217 个项目 · ⭐ 414.1kOneTrainer is a one-stop solution for all your Diffusion training needs.
A Unified Library for Parameter-Efficient and Modular Transfer Learning
We unified the interfaces of instruction-tuning data (e.g., CoT data), multiple LLMs and parameter-efficient methods (e.g., lora, p-tuning) together for easy use. We welcome open-source enthusiasts to initiate any meaningful PR on this repo and integrate as many LLM related technologies as possible. 我们打造了方便研究人员上手和使用大模型等微调平台,我们欢迎开源爱好者发起任何有意义的pr!
基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
Code and documents of LongLoRA and LongAlpaca (ICLR 2024 Oral)
雅意大模型:为客户打造安全可靠的专属大模型,基于大规模中英文多领域指令数据训练的 LlaMA 2 & BLOOM 系列模型,由中科闻歌算法团队研发。(Repo for YaYi Chinese LLMs based on LlaMA2 & BLOOM)
ChatGPT爆火,开启了通往AGI的关键一步,本项目旨在汇总那些ChatGPT的开源平替们,包括文本大模型、多模态大模型等,为大家提供一些便利
OneDiff: An out-of-the-box acceleration library for diffusion models.
Hypernetworks that adapt LLMs for specific benchmark tasks using only textual task description as the input
Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.
Native Android messaging app using Bluetooth LE, TCP, or RNode (LoRa) over LXMF and Reticulum
从 NLP 到 LLM 的算法全栈教程,在线阅读地址:https://datawhalechina.github.io/base-llm/
Toolkit for fine-tuning, ablating and unit-testing open-source LLMs.
共 217 条 · 第 2 / 11 页