computer-vision
507 个项目 · ⭐ 2087.8kNVIDIA AI Blueprint for video search and summarization (VSS) is a GPU-accelerated reference architecture for building video analytics agents with real-time verified alerts, visual Q&A, and automated reporting. The VSS Blueprint uses vision language models (VLMs) such as NVIDIA Cosmos, LLMs such as NVIDIA Nemotron, RAG, and NVIDIA NIMs.
:art: Semantic segmentation models, datasets and losses implemented in PyTorch.
An Embedded Computer Vision & Machine Learning Library (CPU Optimized & IoT Capable)
novel deep learning research works with PaddlePaddle
Roadmap to become a Visual-SLAM developer in 2026
MobileNetV2-YoloV3-Nano: 0.5BFlops 3MB HUAWEI P40: 6ms/img, YoloFace-500k:0.1Bflops 420KB:fire::fire::fire:
Declarative way to run AI models in React Native on device, powered by ExecuTorch.
Meta-Transformer for Unified Multimodal Learning
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
Model explainability that works seamlessly with 🤗 transformers. Explain your transformers model in just 2 lines of code.
A large-scale text-to-image prompt gallery dataset based on Stable Diffusion
共 507 条 · 第 12 / 26 页