clip
78 个项目 · ⭐ 58.9kDiffusion Explainer: Visual Explanation for Text-to-image Stable Diffusion
[Pattern Recognition 25] CLIP Surgery for Better Explainability with Enhancement in Open-Vocabulary Tasks
Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model
[ECCV 2024] Improving 2D Feature Representations by 3D-Aware Fine-Tuning
A Simple, Lightweight, and Extensible Serving Framework for X-AnyLabeling
Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
Code for the CVPR 2024 paper highlight and demo "PIGEON: Predicting Image Geolocations".
Language Models Can See: Plugging Visual Controls in Text Generation
Reproducible scaling laws for contrastive language-image learning (https://arxiv.org/abs/2212.07143)
[ACM TOMM 2023] - Composed Image Retrieval using Contrastive Learning and Task-oriented CLIP-based Features
An AI-powered natural language & reverse Image Search Engine powered by CLIP & qdrant.
CLIP (Contrastive Language–Image Pre-training) for Italian
Python package to generate image embeddings with CLIP without PyTorch/TensorFlow
图片搜索引擎,很简单。三步构建属于你自己的图片搜索引擎,掌握向量数据库和以图搜图、文本搜索图片。
共 78 条 · 第 2 / 4 页