pretraining
31 个项目 · ⭐ 133.5kImplement a ChatGPT-like LLM in PyTorch from scratch, step by step
Llama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用
General technology for enabling AI capabilities w/ LLMs and MLLMs
Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family
The official implementation of MARS: Unleashing the Power of Variance Reduction for Training Large Models
Personal Project: MPP-Qwen14B & MPP-Qwen-Next(Multimodal Pipeline Parallel based on Qwen-LM). Support [video/image/multi-image] {sft/conversations}. Don't let the poverty limit your imagination! Train your own 8B/14B LLaVA-training-like MLLM on RTX3090/4090 24GB.
PyTorch code for "Unifying Vision-and-Language Tasks via Text Generation" (ICML 2021)
A Chinese Open-Domain Dialogue System
Research code for EMNLP 2020 paper "HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training"
Official codebase for "Next-Latent Prediction Transformers Learn Compact World Models"
PyTorch code for “TVLT: Textless Vision-Language Transformer” (NeurIPS 2022 Oral)
MixGen: A New Multi-Modal Data Augmentation
[ICCV 2025] Explore the Limits of Omni-modal Pretraining at Scale
[NeurIPS 2021] COCO-LM: Correcting and Contrasting Text Sequences for Language Model Pretraining
[EMNLP'21] Visual News: Benchmark and Challenges in News Image Captioning
共 31 条 · 第 1 / 2 页