DSH 识图插件
Adds image understanding to DeepSeek Harness (DSH): wire up an OpenAI-compatible external vision model, and a text-only main model (e.g. deepseek) can call the describe image tool to hand an image to it and get a plain-text description — understanding screenshots, photos, charts, OCR, UIs, and more.
이것은 DeepSeek Harness(DSH) 플러그인입니다. 이 사이트는 GitHub README, 설치 정보, 유지보수 상태, 공개 보안 시그널을 모아 보여줍니다.
업스트림에서 중국어 README를 제공하지 않아 저장소 원본 내용을 표시합니다.
dsh-vision-plugin
中文 | English
Adds image understanding to DeepSeek Harness (DSH): wire up an OpenAI-compatible external vision model, and a text-only main model (e.g. deepseek) can call the describe_image tool to hand an image to it and get a plain-text description — understanding screenshots, photos, charts, OCR, UIs, and more.
Design: images go only to the secondary model (external vision model); the main model always deals with text.
Features
| 🖼️ Image understanding | The main model calls describe_image and gets a plain-text description |
| ⚙️ GUI configuration | Fill in URL / API key / model on a settings page — no config files to edit |
| 🔒 Secure credentials | API key stored in the credential store, never echoed |
| 📦 Install once, keep working | Auto-loads at DSH startup, survives restarts |
Install
One-liner (recommended)
Windows (PowerShell)
irm https://raw.githubusercontent.com/woyeshishen/dsh-vision-plugin/main/scripts/install.ps1 | iex
macOS / Linux
bash <(curl -fsSL https://raw.githubusercontent.com/woyeshishen/dsh-vision-plugin/main/scripts/install.sh)
dsh plugin command
From npm
dsh plugin --profile web add @woyeshishen/dsh-vision-plugin
From GitHub
dsh plugin --profile web add github:woyeshishen/dsh-vision-plugin
After install, the plugin auto-mounts into the profile; restart DSH (or hot-reload) to activate.
Usage
Step 1: Configure the external vision model
Open Settings → Multimodal Vision:
| Field | Description |
|---|---|
| URL (Base URL) | OpenAI-compatible endpoint, e.g. https://api.example.com/v1 |
| API key | Secret for the external model (stored encrypted, never echoed) |
| Model | Click "Load models" to fetch and pick from the endpoint |
Click Save.
Step 2: Ask the main model to look at an image
In a conversation, say:
Take a look at
D:\path\to\image.pngand describe what's in it.
The main model calls describe_image, sends the image to the external vision model, and continues reasoning from the returned description.
Tool
describe_image
| Parameter | Required | Description |
|---|---|---|
path | ✅ | Image file path; supports png / jpg / jpeg / webp / gif |
prompt | ❌ | Specific question about the image; defaults to "describe the image in detail" |
Requirements
- DeepSeek Harness
- An OpenAI-compatible (
/chat/completions), image-capable external vision model
License
보안 및 설치 증거
이 점수는 공개 저장소 메타데이터와 이 사이트에 등록된 설치 증거에만 기반하며, 코드 보안 감사와 다릅니다.
공개 플러그인 카탈로그에서 왔으며, 공개 GitHub 저장소로 연결됩니다.
저장소가 Apache-2.0 라이선스를 선언했습니다.
최근 180일 내 코드 업데이트가 있습니다.
재현 가능한 정확한 설치 메타데이터가 아직 등록되지 않았습니다. 저장소 설명에 따라 직접 확인하세요.
검사한 패키지 메타데이터에 설치 라이프사이클 스크립트가 선언되지 않았습니다.