From 926af005a576da245c87cc222f9f69cf089516da Mon Sep 17 00:00:00 2001 From: billowliu2 Date: Sun, 2 Aug 2026 02:32:02 +0800 Subject: [PATCH] =?UTF-8?q?docs:=20=E6=98=8E=E7=A1=AE=E9=9D=9E=E5=A4=9A?= =?UTF-8?q?=E6=A8=A1=E6=80=81=E6=A8=A1=E5=9E=8B=E4=B8=8B=20Alt+V=20?= =?UTF-8?q?=E4=BC=9A=E8=A2=AB=20CLI=20=E6=8B=A6=E6=88=AA=EF=BC=8C=E6=AD=A3?= =?UTF-8?q?=E7=A1=AE=E7=94=A8=E6=B3=95=E4=B8=BA=20@=E8=B7=AF=E5=BE=84=20?= =?UTF-8?q?=E6=88=96=20=E6=88=AA=E5=9B=BE=E5=90=8E=E6=8F=90=E9=97=AE?= MIME-Version: 1.0 Content-Type: text/plain; charset=UTF-8 Content-Transfer-Encoding: 8bit --- README.en.md | 6 +++++- README.md | 3 ++- 2 files changed, 7 insertions(+), 2 deletions(-) diff --git a/README.en.md b/README.en.md index 1d07281..1d6eeff 100644 --- a/README.en.md +++ b/README.en.md @@ -110,7 +110,7 @@ The MCP server starts automatically with the session. | --- | --- | | Multimodal model | Paste with Alt+V directly — native vision, plugin not involved | | Non-multimodal + image path | Type `@screenshot.png` or paste the path; the model calls `read_image` | -| Non-multimodal + just screenshotted/copied | Ask "analyze this screenshot"; the model calls `read_clipboard_image` to read the system clipboard | +| Non-multimodal + just screenshotted/copied | Do **not** Alt+V paste (the CLI rejects pasting on non-multimodal models with `Current model does not support image input`); just ask "analyze this screenshot" — the model calls `read_clipboard_image` to read the system clipboard | ## Activation and triggering @@ -240,6 +240,10 @@ or `xclip` (X11). ## Troubleshooting +- **Alt+V paste reports `Current model does not support image input`** → that is the + Kimi Code CLI rejecting pastes on non-multimodal models. Use `@image-path` + (`read_image`) instead, or screenshot and ask directly (`read_clipboard_image` + reads the clipboard) - **Tool returns "Vision API is not configured"** → run `node setup.mjs`, or set `VISION_API_KEY` / `VISION_API_URL` / `VISION_MODEL` - **Model list fetch fails** (404/401) → the wizard falls back to manual entry; if diff --git a/README.md b/README.md index 3e9b241..a12ca96 100644 --- a/README.md +++ b/README.md @@ -82,7 +82,7 @@ MCP 服务器会随会话自动启动。 | --- | --- | | 多模态模型 | 直接 Alt+V 粘贴图片,原生看图,插件不参与 | | 非多模态 + 有图片路径 | 输入 `@截图.png` 或直接给路径,模型自动调 `read_image` | -| 非多模态 + 刚截图/复制 | 截图后直接提问「分析这张截图」,模型自动调 `read_clipboard_image` 读系统剪贴板 | +| 非多模态 + 刚截图/复制 | 截图后**不要 Alt+V 粘贴**(CLI 会拒绝非多模态模型粘贴并报 `Current model does not support image input`),直接提问「分析这张截图」,模型自动调 `read_clipboard_image` 读系统剪贴板 | ## 激活与触发 @@ -184,6 +184,7 @@ node scripts/sync-models.mjs ## 故障排查 +- **Alt+V 粘贴报 `Current model does not support image input`** → 这是 KimiCode CLI 的拦截:非多模态模型不支持粘贴图片。改用 `@图片路径`(`read_image`)或截图后直接提问(`read_clipboard_image` 读剪贴板) - **工具返回「Vision API is not configured」** → 运行 `node setup.mjs`,或设置 `VISION_API_KEY` / `VISION_API_URL` / `VISION_MODEL` - **模型列表拉取失败**(404/401)→ 向导自动回退手动输入;若服务端不支持 `/models`,直接输入模型名即可 - **视觉验证失败** → 换一个真正支持图片输入的模型(参考服务商文档,如 `qwen-vl-max`、`glm-4v`、`gpt-4o`、`claude-3-5-sonnet` 等)