tiefeiyu/dsh-see-image ↗★ 1

dsh-see-image

see_image tool for DSH: route image files to any OpenAI-compatible vision model and relay the text description back to text-only models 适合纯文本模型需读图的场景;需配置视觉模型端点与密钥。

패키지
dsh-see-image
호환성
미검증
Harness peer 범위
*
버전
1.0.0
라이선스
MIT
최근 업데이트
2026. 8. 14.

설치

검증된 bundle이 없거나 호환성 검사에 실패했습니다. 먼저 저장소 설명을 읽어 주세요. 전체 README 읽기 ↗

Configuration

KeyDefaultDescription
baseURLhttps://api.individual.githubcopilot.comOpenAI-compatible endpoint; the plugin appends /chat/completions
modelgpt-4.1Vision model ID
apiKeyEnvVISION_API_KEYEnv var name holding the API key for non-Copilot backends; empty sends no Authorization header (keyless local endpoints like Ollama)
maxTokens1024Max output tokens (Zhipu glm-4v-flash caps at 1024; raise it for larger models)
timeoutMs90000Request timeout
maxBytes15728640Max image size (15 MB)
prompt(detailed Chinese description instruction)Default question; a question argument passed at call time takes precedence

Usage

In conversation, just say "look at this image / read this screenshot" — the model locates the file and calls the tool itself. You can also state the question explicitly:

see_image(file_path="C:\\Users\\me\\Desktop\\error.png", question="What is the full text of this error?")

question may be omitted; the default output covers: scene → verbatim text transcription → color/shape/layout details → explanation of charts/UI/errors (in Chinese).