yauntyour/DSH-Multimodal ↗★ 1

dsh-multimodal

DSH multimodal input plugin: route file attachments (images, video, audio, text) through per-preset model chains and feed the results to the session model as prompt tokens. Adds a Multimodal settings page. 适合需让会话模型处理多种附件的用户;需在 Multimodal 设置页启用并配置各类型模型链。

패키지
dsh-multimodal
호환성
미검증
Harness peer 범위
^0.1.0-rc.6
Cordis peer 범위
^4.0.1
버전
0.3.3
라이선스
MIT
최근 업데이트
2026. 8. 15.

같은 패키지 이름의 다른 저장소

설치

$npx -p @deepseek-ai/dsh dsh plugin --profile web add github:yauntyour/DSH-Multimodal

使用

  1. 打开 设置 → Multimodal;
  2. 打开「启用 Multimodal 处理」开关;
  3. 为预设配置模型链(Provider + 模型,可多行,顺序即回退顺序;用「测试」按钮验证连通性);
  4. 回到会话:粘贴/拖入图片,或点击输入框 ➕ 附加任意文件。

示例:给图片预设配置一个视觉模型 → 发送图片时自动生成描述文本再进入文本会话模型;配置了原生多模态会话模型时,图片则原样交给会话模型。

配置(cordis.yml / settings.yaml)

plugins:
  multimodal:
    enabled: true
    presets:
      - id: image
        name: 图片
        kind: image
        patterns: []          # 空 = 匹配该类型所有文件;如 [*.png, *.jpg]
        prompt: 请详细描述这张图片…
        maxBytes: 10485760    # 覆盖大小上限(字节,默认 10MB)
        onError: note         # 默认 note(替换为失败说明);或 pass-through
        models:
          - provider: deepseek-official
            model: deepseek-chat
          - provider: openai
            model: gpt-4o