为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, C
Equip text-only coding agents with image processing abilities like answering questions about images and understanding screenshots.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytime为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, C
agent-vision-toolkit has 946 stars on GitHub. It has been forked 34 times. agent-vision-toolkit is written mainly in Python. It has been in active development since 2026. agent-vision-toolkit is available under the MIT license. Its main topics are agent, agent-skills, claude-code, codex.
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, C
agent-vision-toolkit is an open-source project. It is released under the MIT license.
Yes. agent-vision-toolkit is free and open source — you can use, modify and self-host it.
agent-vision-toolkit is available under the MIT license.
agent-vision-toolkit is written mainly in Python.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/anionex-agent-vision-toolkit.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.