agent-vision-toolkit — tool profile and history
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode
- Language
- Python
- License
- MIT
- Created
- Last activity
- Topics
- agent
- agent-skills
- claude-code
- codex
- computer-use
- deepseek
- dsh-plugin
- glm
- harness-engineering
- multimodal
- opencode
- text-only-llm
- vision
- vision-language-model
Install git clone https://github.com/Anionex/agent-vision-toolkit ~/.claude/skills/agent-vision-toolkit
Current measurement
Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.
Reading history
| Date | Value | Metric |
|---|---|---|
| 1 133 | GitHub stars | |
| 1 125 | GitHub stars | |
| 1 119 | GitHub stars | |
| 1 118 | GitHub stars | |
| 1 114 | GitHub stars | |
| 1 110 | GitHub stars | |
| 1 107 | GitHub stars | |
| 1 102 | GitHub stars | |
| 1 098 | GitHub stars | |
| 1 095 | GitHub stars | |
| 1 095 | GitHub stars | |
| 1 084 | GitHub stars | |
| 1 072 | GitHub stars | |
| 1 049 | GitHub stars | |
| 1 005 | GitHub stars | |
| 989 | GitHub stars | |
| 911 | GitHub stars | |
| 845 | GitHub stars | |
| 830 | GitHub stars | |
| 409 | GitHub stars | |
| 397 | GitHub stars | |
| 392 | GitHub stars | |
| 384 | GitHub stars |
Classifications
- Domains
- Type
- Agent skills
- Use cases
Neighbouring tools
These entries declare the same topics. No similarity is inferred: only the shared topics are stated.