Lydian815/anchored-pro

Lydian815★ 0JavaScriptLast synced: 2026-08-16

Open on GitHub

README excerpt

anchored-pro — Anchored Pro (opencode-go) dsh-anchored-standard 的 rc.6 + opencode-go(deepseek-v4-pro / pro-max)迁移版 : Pro 满血执行 —— exact RL persona ("You are a helpful software engineer assistant." 一字不改)× 官方 Minimal 真实工具对( bash + str replace editor )× 首轮 reasoningEffort=max 。 与 anchored-flash 的区别 anchored-flash anchored-pro(本预设) 目标模型 deepseek-v4-flash deepseek-v4-pro / pro-max persona w7(neutral + 分类 + 回顾/收敛/反跑题/深度思考锚) exact RL spec 句 ( You are a helpful software engineer assistant. ,零附加) 依据 dsh-router-standard P11/P24:flash 最优 w7 minimal 快照 = "the exact RL prompt and schemas";附加/改写 persona 即落入训练分布间隙(见下) 工具面 / 晋升 / 压缩纪元 Minimal 对 → 发现工具 → 按需解锁 bash + str replace editor → 发现工具 → 按需解锁 (压缩后目录只增不减) 为什么 anchored-flash「不行」:它对 pro 模型也会注入 w7 深度思考/回顾锚, 把 Pro 的轨迹拉离其 RL 分布。anchored-pro 默认只给 Pro 注入精确 spec 句, 其余全部交给低工具占比首轮与 Minimal 式引导。 为什么 persona 必须是精确 spec 句(重要) 社区与论文证据一致指向同一个结论: V4 Pro 对首行 persona 文本高度过拟合, 只有 "You are a helpful software engineer assistant." 一字不改才是「满血」条件 。 1. 本地研究 ( dsh-router-standard paper,本机 dsh-routing-suite 内有全文): - DSH 官方 minimal preset 的快照测试自称发送的是 "the exact RL prompt and schemas" —— 这一句话 + 低工具占比首轮就是后训练条件本身; - A1/A2 双吸引子理论 :spec(plan-first、集体语域 "We need"、read-first) 与 r…

View full README on GitHub →
Agentsagent-presetdeepseek-harnessdshdsh-pluginopencode

Category