工具说明为英文。

agent-vision-toolkit — 工具资料与历史

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screenshot OCR, frontend UI restoration, and GUI automation, with optional seamless integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode

活跃

官方来源 ↗

语言
Python
许可证
MIT
创建时间
最近活动
主题
  • agent
  • agent-skills
  • claude-code
  • codex
  • computer-use
  • deepseek
  • dsh-plugin
  • glm
  • harness-engineering
  • multimodal
  • opencode
  • text-only-llm
  • vision
  • vision-language-model

安装 git clone https://github.com/Anionex/agent-vision-toolkit ~/.claude/skills/agent-vision-toolkit

当前测量

1 125GitHub 星标+23 (+2.1 %)
实测增长

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

测量历史

90 天 · 90 次测量上限
日期数值指标
1 125GitHub 星标
1 119GitHub 星标
1 118GitHub 星标
1 114GitHub 星标
1 110GitHub 星标
1 107GitHub 星标
1 102GitHub 星标
1 098GitHub 星标
1 095GitHub 星标
1 095GitHub 星标
1 084GitHub 星标
1 072GitHub 星标
1 049GitHub 星标
1 005GitHub 星标
989GitHub 星标
911GitHub 星标
845GitHub 星标
830GitHub 星标
409GitHub 星标
397GitHub 星标
392GitHub 星标
384GitHub 星标

分类

领域
类型
智能体技能
用途

这些条目声明了相同的主题。此处不推断相似性,仅列出共同的主题。