yujiaoliang/dsh-plugin-amdgpu-inspect
Structured AMDGPU code-object and ISA inspection for DeepSeek Harness.
이것은 DeepSeek Harness(DSH) 플러그인입니다. 이 사이트는 GitHub README, 설치 정보, 유지보수 상태, 공개 보안 시그널을 모아 보여줍니다.
업스트림에서 중국어 README를 제공하지 않아 저장소 원본 내용을 표시합니다.
AMDGPU Inspect
Structured AMDGPU code-object and ISA inspection for DeepSeek Harness.
The plugin turns an existing HSACO or AMDGPU ELF file into bounded JSON that an agent can query. It reports compiler artifacts without compiling source, executing a kernel, predicting performance, or prescribing an optimization.
AMDGPU ELF -> llvm-readelf + llvm-objdump -> normalized object model -> DSH tools
Tools
amdgpu_object_inspect
List kernels and report public code-object metadata plus instruction-category counts:
{ "path": "/tmp/vector-add.hsaco" }
amdgpu_isa_query
Query one kernel's normalized instruction stream without placing the complete disassembly in model context:
{
"path": "/tmp/vector-add.hsaco",
"kernel": "vector_add",
"opcodes": ["scratch_*", "s_waitcnt"],
"cursor": 0,
"limit": 40
}
The result contains addresses, opcodes, operands, categories, and a nextCursor when more matches remain.
Install
Requirements:
- Linux
- DeepSeek Harness developer preview
- ROCm LLVM tools in
/opt/rocm/llvm/binor onPATH
Install the bundle from GitHub:
dsh plugin --profile web add github:yujiaoliang/dsh-plugin-amdgpu-inspect
dsh --profile web
The package ships plain ESM JavaScript and needs no install-time build script.
Kernel X-Ray Skill
skills/kernel-xray/SKILL.md is the opinionated layer: it tells an agent how to compile before/after artifacts, query this plugin, form hypotheses, and require a workload measurement before claiming a speedup. The plugin itself remains a low-level facts API.
Design
llvm-readelf --notessupplies public AMDGPU metadata.llvm-objdump --disassemble --demanglesupplies ISA.- A one-artifact cache is keyed by absolute path, size, and modification time.
- ISA queries are capped at 200 records and support deterministic cursors.
- Child processes run without a shell and honor tool-call cancellation.
- Inputs are limited to regular files no larger than 256 MiB.
Security boundary
This version invokes LLVM and reads the artifact directly on the DSH host. It does not inherit a deployment's virtual filesystem or subprocess provider, so treat access as host-local rather than sandboxed. It never loads the artifact into an HSA runtime or executes GPU code. LLVM diagnostics can expose local paths.
Limitations
- Parsing follows the text emitted by current ROCm LLVM tools; fixture coverage cannot guarantee every historical or future output spelling.
- Kernel arguments, symbols, relocations, control-flow graphs, and binary encodings are not exposed yet.
- The plugin reports static facts only. Resource counts and instruction counts do not establish performance.
Development
Tests use synthetic public-format metadata and disassembly; no GPU or ROCm installation is required:
npm test
License
MIT. ROCm and LLVM are external tools and are not redistributed.
보안 및 설치 증거
이 점수는 공개 저장소 메타데이터와 이 사이트에 등록된 설치 증거에만 기반하며, 코드 보안 감사와 다릅니다.
공개 플러그인 카탈로그에서 왔으며, 공개 GitHub 저장소로 연결됩니다.
저장소가 MIT 라이선스를 선언했습니다.
최근 180일 내 코드 업데이트가 있습니다.
재현 가능한 정확한 설치 메타데이터가 아직 등록되지 않았습니다. 저장소 설명에 따라 직접 확인하세요.
검사한 패키지 메타데이터에 설치 라이프사이클 스크립트가 선언되지 않았습니다.