vision
81 个项目 · ⭐ 227.7kEnhanced ChatGPT Clone: Features Agents, MCP, Skills, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching, message search, Code Interpreter, langchain, DALL-E-3, OpenAPI Actions, Functions, Secure Multi-User Auth, Presets, open-source for self-hosting. Active
Xray, Penetrates Everything. Also the best v2ray-core. Where the magic happens. An open platform for various uses.
The Open-Source Multimodal AI Agent Stack: Connecting Cutting-Edge AI Models and Agent Infra
Caffe: a fast open framework for deep learning.
The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/
📸 A powerful, high-performance React Native Camera library.
SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal
[CVPR 2023] DepGraph: Towards Any Structural Pruning; LLMs, Vision Foundation Models, etc.
🌝 MLKit是一个强大易用的工具包。通过ML Kit您可以很轻松的实现文字识别、条码识别、图像标记、人脸检测、对象检测等功能。
Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.
Convert any web design screenshot to clean HTML/CSS code
Implementation of Bottleneck Transformer in Pytorch
Official skills for the GLM family of models.
共 81 条 · 第 1 / 5 页