편집자 노트

yujiaoliang/dsh-plugin-amdgpu-inspect

Structured AMDGPU code-object and ISA inspection for DeepSeek Harness.

이것은 DeepSeek Harness(DSH) 플러그인입니다. 이 사이트는 GitHub README, 설치 정보, 유지보수 상태, 공개 보안 시그널을 모아 보여줍니다.

업스트림에서 중국어 README를 제공하지 않아 저장소 원본 내용을 표시합니다.

AMDGPU Inspect

Structured AMDGPU code-object and ISA inspection for DeepSeek Harness.

The plugin turns an existing HSACO or AMDGPU ELF file into bounded JSON that an agent can query. It reports compiler artifacts without compiling source, executing a kernel, predicting performance, or prescribing an optimization.

AMDGPU ELF -> llvm-readelf + llvm-objdump -> normalized object model -> DSH tools

Tools

amdgpu_object_inspect

List kernels and report public code-object metadata plus instruction-category counts:

{ "path": "/tmp/vector-add.hsaco" }

amdgpu_isa_query

Query one kernel's normalized instruction stream without placing the complete disassembly in model context:

{
  "path": "/tmp/vector-add.hsaco",
  "kernel": "vector_add",
  "opcodes": ["scratch_*", "s_waitcnt"],
  "cursor": 0,
  "limit": 40
}

The result contains addresses, opcodes, operands, categories, and a nextCursor when more matches remain.

Install

Requirements:

  • Linux
  • DeepSeek Harness developer preview
  • ROCm LLVM tools in /opt/rocm/llvm/bin or on PATH

Install the bundle from GitHub:

dsh plugin --profile web add github:yujiaoliang/dsh-plugin-amdgpu-inspect
dsh --profile web

The package ships plain ESM JavaScript and needs no install-time build script.

Kernel X-Ray Skill

skills/kernel-xray/SKILL.md is the opinionated layer: it tells an agent how to compile before/after artifacts, query this plugin, form hypotheses, and require a workload measurement before claiming a speedup. The plugin itself remains a low-level facts API.

Design

  • llvm-readelf --notes supplies public AMDGPU metadata.
  • llvm-objdump --disassemble --demangle supplies ISA.
  • A one-artifact cache is keyed by absolute path, size, and modification time.
  • ISA queries are capped at 200 records and support deterministic cursors.
  • Child processes run without a shell and honor tool-call cancellation.
  • Inputs are limited to regular files no larger than 256 MiB.

Security boundary

This version invokes LLVM and reads the artifact directly on the DSH host. It does not inherit a deployment's virtual filesystem or subprocess provider, so treat access as host-local rather than sandboxed. It never loads the artifact into an HSA runtime or executes GPU code. LLVM diagnostics can expose local paths.

Limitations

  • Parsing follows the text emitted by current ROCm LLVM tools; fixture coverage cannot guarantee every historical or future output spelling.
  • Kernel arguments, symbols, relocations, control-flow graphs, and binary encodings are not exposed yet.
  • The plugin reports static facts only. Resource counts and instruction counts do not establish performance.

Development

Tests use synthetic public-format metadata and disassembly; no GPU or ROCm installation is required:

npm test

License

MIT. ROCm and LLVM are external tools and are not redistributed.

REPOSITORY SIGNALS

보안 및 설치 증거

이 점수는 공개 저장소 메타데이터와 이 사이트에 등록된 설치 증거에만 기반하며, 코드 보안 감사와 다릅니다.

출처 추적 가능

공개 플러그인 카탈로그에서 왔으며, 공개 GitHub 저장소로 연결됩니다.

라이선스

저장소가 MIT 라이선스를 선언했습니다.

유지보수 활동

최근 180일 내 코드 업데이트가 있습니다.

설치 증거

재현 가능한 정확한 설치 메타데이터가 아직 등록되지 않았습니다. 저장소 설명에 따라 직접 확인하세요.

설치 라이프사이클 스크립트

검사한 패키지 메타데이터에 설치 라이프사이클 스크립트가 선언되지 않았습니다.