computer-vision
507 个项目 · ⭐ 2087.8kAI-powered Claude skill for Spine 2D skeletal animation — auto-rig, animate, and preview characters
DiffSeg is an unsupervised zero-shot segmentation method using attention information from a stable-diffusion model. This repo implements the main DiffSeg algorithm and additionally includes an experimental feature to add semantic labels to the masks based on a generated caption.
Reliable, minimal and scalable library for pretraining foundation and world models
A Simple, Lightweight, and Extensible Serving Framework for X-AnyLabeling
Segment Anything Model for large-scale, vectorized road network extraction from aerial imagery. CVPRW 2024
Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis (ECCV 2024 Oral) - Official Implementation
FASHN VTON v1.5: Efficient Maskless Virtual Try-On in Pixel Space
Code and Pretrained Models for ICLR 2023 Paper "Contrastive Audio-Visual Masked Autoencoder".
Home of the AI workforce - Multi-agent system, AI agents & tools
General Multi-label Image Classification with Transformers
Diffusers-Interpret 🤗🧨🕵️♀️: Model explainability for 🤗 Diffusers. Get explanations for your generated images.
共 507 条 · 第 15 / 26 页