PRODUCTIVITY/效率应用·AUTOMATION
Anionex/agent-vision-toolkit
为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-scree…
怎么装:Install: git clone https://github.com/Anionex/agent-vision-toolkit.git
02 / 现在的位置02 / Why now
首次发现FIRST SEEN
03 OCT 202603 OCT 2026
增长GROWTH
+0.2%+0.2%
1.2k → 1.2k stars · 4 个快照1.2k → 1.2k stars · 4 snapshots
状态STATUS
持续活跃Active
03 / 它能帮你做什么03 / What it helps you do
04 / 安装04 / Install
$ README 里给出的安装方式。The installation method given in the README.
$ README 里给出的安装方式。The installation method given in the README.
两种装法选一个即可 —— 上面那条给命令行用户,下面那条给写代码的用户。Pick one of the two — the first is for command-line users, the second for people writing code.
05 / 限制与风险05 / Limits and risk
风险不是警告,是可信度的一部分。以下结论只基于文档静态扫描,我们不会执行项目里的任何代码。Risk here is evidence, not an alarm. These findings come from static scanning of the docs; we never execute a project’s code.
限制LIMITS
需检查Needs review
静态扫描STATIC SCAN
01读取 API KeyReads an API keyreadme
| `VISION_API_PROTOCOL` | No | Python client/proxy protocol: `chat_completions` (default), `responses`, or `anthropic`; Anthropic mode uses
这里只做静态扫描:读 README、SKILL.md 和依赖清单,不执行代码。没有命中不代表安全。This is static scanning only: we read the README, SKILL.md and dependency list, and never execute code. No findings does not mean safe.
{
"agents": [
{
"agent": "claude-code",
"evidence": "readme: …rop-in integration for Codex, Claude Code, Pi, Oh My Pi, and OpenCode.*…"
},
{
"agent": "opencode",
"evidence": "readme: …laude Code, Pi, Oh My Pi, and OpenCode.** 🎯 An agent's vision capa…"
}
],
"apiKey": "and guide me through setting `VISION_API_KEY`, `VISION_BASE_URL`, and `VISION_MODEL`",
"docker": null,
"taxonomy": {
"scores": {
"design": 2,
"ai-tools": 2,
"dev-tools": 1,
"indie-web": 0,
"skill-agent": 9,
"productivity": 3
},
"primary": "productivity",
"secondary": "automation"
},
"localRuntime": null,
"skillMdTotal": 1,
"skillMdErrors": [],
"treeTruncated": false,
"categoryScores": {
"data": 0,
"agent": 0,
"media": 2,
"design": 8,
"browser": 0,
"devtool": 0,
"security": 0,
"marketing": 0,
"productivity": 2
},
"classification": [
{
"signal": "1 SKILL.md with name + description",
"weight": 0.95
},
{
"signal": "skills/ directory with 6 markdown files",
"weight": 0.75
},
{
"signal": "README describes a Claude / Agent Skill",
"weight": 0.7
},
{
"signal": "description mentions AI",
"weight": 0.2
}
],
"scannedSources": [
"readme",
"skills/vision-skills/SKILL.md"
],
"skillMdFetched": 1
}