agent-vision-toolkit

agent-vision-toolkit

Anionex

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

832 Stars
30 Forks
4 Watchers
Python Language
mit License
100 SrcLog Score
Cost to Build
$1.01M
Market Value
$4.24M

Growth over time

1 data points  ·  2026-08-14 → 2026-08-14
Stars Forks Watchers
💬

How do you feel about this project?

Ask AI about agent-vision-toolkit

Question copied to clipboard

What is the Anionex/agent-vision-toolkit GitHub project? Description: "为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode". Written in Python. Explain what it does, its main use cases, key features, and who would benefit from using it.

Question is copied to clipboard — paste it after the AI opens.

How to clone agent-vision-toolkit

Clone via HTTPS

git clone https://github.com/Anionex/agent-vision-toolkit.git

Clone via SSH

[email protected]:Anionex/agent-vision-toolkit.git

Download ZIP

Download main.zip

Found an issue?

Report bugs or request features on the agent-vision-toolkit issue tracker:

Open GitHub Issues