fine-tuning
612 个项目 · ⭐ 499.9kThis repository contains the source code and related resources for R-LoRA.
Fine-tune the newly released Llama-3.2 lightweight models.
📘 Taiwan-LLM Tutor: Large Language Models for Taiwanese Secondary Education
大模型学习--从模型部署到模型微调,此项目是经过训练营学习后,结合训练营项目,自我理解消化总结,以及创新型应用。可star/fork
[ACL 2026] CoCoA: Collaborative Chain-of-Agents for Parametric-Retrieved Knowledge Synergy
BoDmagh dataset is a Supervised Fine-Tuning (SFT) dataset for the Darija language
LLM fine-tuning with LoRA + NVFP4/MXFP8 on NVIDIA DGX Spark (Blackwell GB10)
Context & Guide For Reinforcement Learning with Verifiable Rewards with Large Language Models
This repository has been created to keep track of my 5 days progress.
Official Implementation of "Fine-Tuning is Fine, if Calibrated.", NeurIPS 2024
ISIC 2019 - Skin Lesion Analysis Towards Melanoma Detection
一个开箱即用、用于二分类任务的大语言微调模型框架。An out-of-the-box LLM fine-tuning framework for medical binary classification.
共 612 条 · 第 21 / 31 页