dsh-multimodal
DSH multimodal input plugin: route file attachments (images, video, audio, text) through per-preset model chains and feed the results to the session model as prompt tokens. Adds a Multimodal settings page. 适合需让会话模型处理多种附件的用户;需在 Multimodal 设置页启用并配置各类型模型链。
Other repositories with this package name
Install
$
npx -p @deepseek-ai/dsh dsh plugin --profile web add github:yauntyour/DSH-MultimodalREADME
Read the full README ↗使用
- 打开 设置 → Multimodal;
- 打开「启用 Multimodal 处理」开关;
- 为预设配置模型链(Provider + 模型,可多行,顺序即回退顺序;用「测试」按钮验证连通性);
- 回到会话:粘贴/拖入图片,或点击输入框 ➕ 附加任意文件。
示例:给图片预设配置一个视觉模型 → 发送图片时自动生成描述文本再进入文本会话模型;配置了原生多模态会话模型时,图片则原样交给会话模型。
配置(cordis.yml / settings.yaml)
plugins:
multimodal:
enabled: true
presets:
- id: image
name: 图片
kind: image
patterns: [] # 空 = 匹配该类型所有文件;如 [*.png, *.jpg]
prompt: 请详细描述这张图片…
maxBytes: 10485760 # 覆盖大小上限(字节,默认 10MB)
onError: note # 默认 note(替换为失败说明);或 pass-through
models:
- provider: deepseek-official
model: deepseek-chat
- provider: openai
model: gpt-4o