DeepSeek Harness Plugin Hub

发布与管理完整 Harness Profiles,发现适合你的插件。

探索

插件目录环境预设文档中心动态

社区

发布插件联系我们报告问题

相关链接

Plugin Hub GitHubDeepSeek Harness 官方项目系统状态隐私说明
© 2026 DeepSeek Harness Plugin HubPowered byPaxTech

独立、非官方社区项目,与 DeepSeek 官方无隶属、授权或背书关系。

Llm Sampling — DeepSeek Harness 插件(DSH Plugin)
DeepSeek Harness Plugin Hub
ProfilesPlugins分类动态文档登录管理 Profiles
ProfilesPlugins分类动态文档登录
← Plugins
L

dsh-llm-sampling

Llm Sampling

DeepSeek Harness 的精确模型采样策略包

插件会安装到这里;不确定时保持 web。

npx -y @deepseek-ai/dsh plugin --profile web add github:kuma-loong/dsh-llm-sampling#5da04d9b3401678a188d8b318fdf43d7de1b8932
README兼容性版本

兼容性与来源证明

Llm Sampling 以 dsh-llm-sampling 发布,当前版本为 0.1.0-rc.1。Plugin Hub 会校验它的 manifest,并保存精确安装来源,便于复现安装结果。

DSH 兼容范围
*
运行环境
any
发布来源
github
Registry 更新时间
2026/8/21

版本

0.1.0-rc.1prerelease
2026/8/21

相关插件

正在加载相关插件…

最新版
0.1.0-rc.1
DSH
*
HMR
重启进程
Tree shaking
未声明可安全裁剪
解包体积
未提供
文件数
未提供
Surface
any
许可证
MIT
发布源
github
GitHub
★ 0
周下载
0
最近提交
2026/8/21
查看源码 ↗
README Badge

点击下方 Badge 复制 Markdown,粘贴到 README 即可。

这是你的 Plugin?认领权益 · 优先安全扫描

验证 package.json 声明的 GitHub 仓库,即可管理这个公开页面。认领后,Hub 会优先安排当前版本的安全扫描,并在通过后公开展示结果。

认领这个 Plugin →
报告问题

相关插件

继续浏览 models-usage 分类下经过校验的插件。

Usage@linxin666/dsh-usagedsh Web GUI 的使用统计插件:检测各提供商余额和编码计划配额,并提供实时令牌使用记录,以及当前提供商的专属宠物气泡Whale Widgetdsh-whale-widgetDSH Web 界面右下角的 DeepSeek 余额小鲸鱼挂件:余额/今日已用/峰谷定价、自定义泡泡点击序列(文本/余额/今日/峰谷/图片/随机语句与并列加权选择)、逐行样式与字体、悬浮快捷编辑、音效与每轮消耗、自定义角色/动图/音效、吸附与翻转自定义Usage Stats@ychris12138/dsh-usage-statsdsh Web GUI 的令牌使用热力图、提供商余额和订阅配额Codex Connectdsh-codex-connect用于 DeepSeek Harness 的 ChatGPT OAuth 和 Codex 模型。

README

dsh-llm-sampling

English | 中文

Installable DeepSeek Harness bundle that enforces sampling policy for exact provider/model routes through the agent/request waterfall. The Harness core owns the provider-neutral request fields and durable request header; adapters own wire translation. This plugin owns deployment policy only.

The plugin is dormant until llm-sampling.providers names a route. A configured model's default profile replaces all sampling values on every request. When reasoningEffort is explicitly off, off overlays that complete profile. Unconfigured routes pass through unchanged.

Compatibility

The plugin requires a DeepSeek Harness build whose LlmCallConfig and GenerateOptions include topP, topK, minP, presencePenalty, and repetitionPenalty. Until that core change reaches an npm release, install the plugin only with a matching Harness source checkout.

Adapters must map the configured fields. @deepseek-ai/dsh-llm-pi-ai supports the extended fields for OpenAI Chat Completions and rejects them for other protocols.

Install

Pin the reviewed commit when installing from GitHub:

dsh plugin --profile web add github:kuma-loong/dsh-llm-sampling#<commit>

Git installs run this package's prepare script. pnpm 10 and later require an explicit build allowance in the profile's pnpm-workspace.yaml:

allowBuilds:
  dsh-llm-sampling@https://codeload.github.com/kuma-loong/dsh-llm-sampling/tar.gz/<commit>: true

Copy the exact key printed by pnpm, then re-run the dsh plugin add command. Grant this permission only after reviewing the pinned source because prepare executes on the host during installation.

Configure

Add an llm-sampling section to $DSH_HOME/settings.yaml:

llm-sampling:
  providers:
    sparse-vllm:
      models:
        Qwen3.8-27B:
          default:
            temperature: 1
            topP: 0.95
            topK: 20
            minP: 0
            presencePenalty: 0
            repetitionPenalty: 1
          off:
            temperature: 0.7
            topP: 0.8
            presencePenalty: 1.5

Supported fields are temperature, topP, topK, minP, presencePenalty, and repetitionPenalty. Profiles are policy, not caller defaults: configured values win over earlier agent/request proposals. A later request policy may deliberately replace them through the normal waterfall order.

Model Experience

Exact-model sampling policy

What the model sees

No prompt text or tool schema is added. The model receives the configured sampling values in its provider request, and the effective values are recorded in the session's request/header before dispatch.

Token effect

The plugin adds no tokens. Sampling changes generation distribution and may change output length.

KV Cache effect

No prompt prefix changes. Providers may include sampling controls in request-cache identity, so a policy or reasoning-mode change can affect provider-side reuse even with identical input tokens.

Known Limitations and Deferred Work

  • The off profile is selected only for an explicit reasoningEffort: off; an omitted effort preserves the default profile because provider-owned implicit reasoning state is not guessed.
  • Extended fields require adapter support; the plugin cannot determine wire compatibility before dispatch.