fine-tuning
612 个项目 · ⭐ 499.5kA distributed training framework for large language models powered by Lightning.
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
This repository contains code for fine-tuning the Whisper speech-to-text model.
Implementation for the different ML tasks on Kaggle platform with GPUs.
The Ultimate Developer Experience Platform for the Container Era
Train an LLM to generate cracked Manim animations for mathematical concepts.
Build a Large Language Model From Scratch
CyberBrain_Model is an advanced AI project designed for fine-tuning the model `DeepSeek-R1-Distill-Qwen-14B` specifically for cyber security tasks.
Code for TR2-D2: Tree Search Guided Trajectory-Aware Fine-Tuning for Discrete Diffusion
共 612 条 · 第 20 / 31 页