pytorch
1071 个项目 · ⭐ 4272.9kEasy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.
🌈 NERpy: Implementation of Named Entity Recognition using Python. 命名实体识别工具,支持BertSoftmax、BertSpan等模型,开箱即用。
DrugHIVE: Structure-based drug design with a deep hierarchical generative model
Fine-tuning Vision Transformers on various classification datasets
NeSVoR is a package for GPU-accelerated slice-to-volume reconstruction.
Easily run text-to-video diffusion with customized video length, fps, and dimensions on 4GB video cards or on CPU.
Code for our CVPR'23 paper - "FLEX: Full-Body Grasping Without Full-Body Grasps"
[NeurIPS 2024] Official Implementation of Attention Interpolation of Text-to-Image Diffusion
Fine-tuning toolkit for Chatterbox TTS & Chatterbox TURBO models. Supports 23 languages with smart vocabulary extension. Features offline preprocessing, automatic VAD trimming, and voice cloning capabilities. Train custom TTS models with your own dataset in LJSpeech and file-based format.
A Simple Latent Diffusion Approach for Panoptic Segmentation and Mask Inpainting [ECCV 2024]
An Open-Source Package for Chinese Open-domain Conversational Chatbot (中文闲聊对话系统,一键部署微信闲聊机器人)
Official Implementation of LongLive-RAG: A general retrieval-augmented framework for long video generation.
Transformer Architectures for Generative AI
Official implementation of AAAI 2023 paper "Parameter-efficient Model Adaptation for Vision Transformers"
共 1071 条 · 第 37 / 54 页