multi-modal
49 个项目 · ⭐ 153.5kThe TypeScript library for building AI applications.
Use late-interaction multi-modal models such as ColPali in just a few lines of code.
[CVPR 2026🔥] 🧑🎨 OmniLottie, an open-sourced multi-modal instructed vector animation generator that produces Lottie JSONs.
Multi-modal OCR pipeline optimized for ML training (text, figure, math, tables, diagrams)
Chat2Graph: Graph Native Agentic System.
A streaming multimodal database for Edge AI, and Edge Computing.
[NeurIPS 2024] MeshXL: Neural Coordinate Field for Generative 3D Foundation Models, a 3D fundamental model for mesh generation
[CVPR2020] Unsupervised Multi-Modal Image Registration via Geometry Preserving Image-to-Image Translation
[ICCV2025] Referring any person or objects given a natural language description. Code base for RexSeek and HumanRef Benchmark
Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
A large-scale multi-modal pre-trained model
152 open-source tools to run LLMs 100% locally – no cloud, no API keys, no censorship
A Python framework for multi-modal document understanding with Amazon Bedrock
Qualcomm Cloud AI SDK (Platform and Apps) enable high performance deep learning inference on Qualcomm Cloud AI platforms delivering high throughput and low latency across Computer Vision, Object Detection, Natural Language Processing and Generative AI models.
GenDB, an LLM-Powered Generative Query Engine Built for the Future
共 49 条 · 第 2 / 3 页