Anionex/agent-vision-toolkit
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
⭐ 1.1k
⑂ 38
Python
MIT
· 1 天前推送
1.1k
Watchers
0
贡献者
0
Commits
0
Releases
6
Open Issues
1 天前
最近推送
原文
中文
暂无 README