computer-vision
507 个项目 · ⭐ 2087.8kEva01 is NOT an assistant. She is an AI being with her own mind, feelings, and intrinsic drives. Multimodal, Modular design. Built-in voice & face recognition. Plug'n play tools. Compatible with ChatGPT, Claude, Deepseek, Gemini, Grok, and Ollama. Explore the possibilities of Human-AI Interaction.
Notebooks and Code about Generative Ai, LLMs, MLOPS, NLP , CV and Graph databases
(CVPR 2024) 🧩 TokenCompose: Text-to-Image Diffusion with Token-level Supervision
[ICLR 2023] Multimodal Analogical Reasoning over Knowledge Graphs
A minimal implementation of a denoising diffusion model in PyTorch.
NEWTON: Agentic Planning for Physically Grounded Video Generation
An unofficial implementation of ViTPose [Y. Xu et al., 2022]
Image Acceptable Alpaca (Image-Text Chat AI).
Contains code for the paper "Vision Transformers are Robust Learners" (AAAI 2022).
MinkLoc++: Lidar and Monocular Image Fusion for Place Recognition
[MICCAI 2024] Codebase for "Stable Diffusion Segmentation for Biomedical Images with Single-step Reverse Process"
Multimodal Masked Autoencoders (M3AE): A JAX/Flax Implementation
Code for our CVPR'23 paper - "FLEX: Full-Body Grasping Without Full-Body Grasps"
Iron man inspired Personal virtual assistant
共 507 条 · 第 18 / 26 页