computer-vision
507 个项目 · ⭐ 2087.8kPlantDreamer: Achieving Realistic 3D Plant Models with Diffusion-Guided Gaussian Splatting [CVPPA: ICCVW 2025]
Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator
Fine-tuned Qwen2-VL-7B for LaTeX OCR using LoRA and Unsloth on the LaTeX OCR dataset. Built augmentation pipeline (rotation, noise, contrast jitter), ran LoRA rank sweep (r=8/16/32), and evaluated across CER, Token F1, BLEU-4, and Exact Match. Deployed as a Gradio Space with live metric computation.
High-accuracy Russian ANPR system built with YOLOv8 for detection and an optimized, custom-trained PyTorch CRNN for OCR.
Learning PyTorch through the D2L book. A series of notebooks for the same
Circuitry.ai is an open-source tool that combines computer vision and large language models to detect, analyze, and explain electronic circuit diagrams. It leverages YOLOv8 for component detection and LLaMA 3 for generating intelligent textual explanations of how the circuit works.
[2025] ModalFormer: Multimodal Transformer for Low-Light Image Enhancement
VAAS is an inference-first, research-driven library for image integrity analysis. It integrates Vision Transformer Attention Mechanisms with patch-level self-consistency analysis to enable fine-grained localization and detection of visual inconsistencies across diverse image analysis tasks.
Code for paper "Semantic Diversity-aware Prototype-based Learning for Unbiased Scene Graph Generation (ECCV 2024)"
共 507 条 · 第 23 / 26 页