yauntyour/DSH-Multimodal ↗★ 1

dsh-multimodal

DSH multimodal input plugin: route file attachments (images, video, audio, text) through per-preset model chains and feed the results to the session model as prompt tokens. Adds a Multimodal settings page. 适合需让会话模型处理多种附件的用户;需在 Multimodal 设置页启用并配置各类型模型链。

Package
dsh-multimodal
Compatibility
Unverified
Harness peer range
^0.1.0-rc.6
Cordis peer range
^4.0.1
Version
0.3.3
License
MIT
Last updated
Aug 15, 2026

Other repositories with this package name

Install

$npx -p @deepseek-ai/dsh dsh plugin --profile web add github:yauntyour/DSH-Multimodal

使用

  1. 打开 设置 → Multimodal;
  2. 打开「启用 Multimodal 处理」开关;
  3. 为预设配置模型链(Provider + 模型,可多行,顺序即回退顺序;用「测试」按钮验证连通性);
  4. 回到会话:粘贴/拖入图片,或点击输入框 ➕ 附加任意文件。

示例:给图片预设配置一个视觉模型 → 发送图片时自动生成描述文本再进入文本会话模型;配置了原生多模态会话模型时,图片则原样交给会话模型。

配置(cordis.yml / settings.yaml)

plugins:
  multimodal:
    enabled: true
    presets:
      - id: image
        name: 图片
        kind: image
        patterns: []          # 空 = 匹配该类型所有文件;如 [*.png, *.jpg]
        prompt: 请详细描述这张图片…
        maxBytes: 10485760    # 覆盖大小上限(字节,默认 10MB)
        onError: note         # 默认 note(替换为失败说明);或 pass-through
        models:
          - provider: deepseek-official
            model: deepseek-chat
          - provider: openai
            model: gpt-4o