tiefeiyu/dsh-see-image ↗★ 1

dsh-see-image

see_image tool for DSH: route image files to any OpenAI-compatible vision model and relay the text description back to text-only models 适合纯文本模型需读图的场景;需配置视觉模型端点与密钥。

パッケージ
dsh-see-image
互換性
未検証
Harness ピア範囲
*
バージョン
1.0.0
ライセンス
MIT
最終更新
2026/08/14

インストール

検証済み bundle がないか、互換性チェックに失敗しています。先にリポジトリの説明を読んでください。 README 全文を読む ↗

ドキュメント

README 全文を読む ↗

Configuration

KeyDefaultDescription
baseURLhttps://api.individual.githubcopilot.comOpenAI-compatible endpoint; the plugin appends /chat/completions
modelgpt-4.1Vision model ID
apiKeyEnvVISION_API_KEYEnv var name holding the API key for non-Copilot backends; empty sends no Authorization header (keyless local endpoints like Ollama)
maxTokens1024Max output tokens (Zhipu glm-4v-flash caps at 1024; raise it for larger models)
timeoutMs90000Request timeout
maxBytes15728640Max image size (15 MB)
prompt(detailed Chinese description instruction)Default question; a question argument passed at call time takes precedence

Usage

In conversation, just say "look at this image / read this screenshot" — the model locates the file and calls the tool itself. You can also state the question explicitly:

see_image(file_path="C:\\Users\\me\\Desktop\\error.png", question="What is the full text of this error?")

question may be omitted; the default output covers: scene → verbatim text transcription → color/shape/layout details → explanation of charts/UI/errors (in Chinese).