看到有人发了deepseek v4 的强度控制提示词,不过我早该想到的,毕竟GPT也是juice值
不过为啥不做成参数直接传入呢?
看到有人发了deepseek v4 的强度控制提示词,不过我早该想到的,毕竟GPT也是juice值
不过为啥不做成参数直接传入呢?
佬能分享一下不。我自己感觉是如果加感叹号和语气词,都会有明显区别。
模型权重又不能变,传个参数内部又要怎么控制呢?
什么叫强度控制,你是说reasoning effort还是什么
大佬能分享一下提示词不 想看看咋搞的
# Reasoning effort levels. In thinking mode, the prompt for the selected level is
# prepended at the very beginning of the conversation. `low` is the default and
# adds nothing.
REASONING_EFFORT_PROMPTS: Dict[str, str] = {
"low": "",
"high": (
"Reasoning Effort: Absolute maximum with no shortcuts permitted.\n"
"You MUST be very thorough in your thinking and comprehensively decompose the problem to resolve the root cause, rigorously stress-testing your logic against all potential paths, edge cases, and adversarial scenarios.\n"
"Explicitly write out your entire deliberation process, documenting every intermediate step, considered alternative, and rejected hypothesis to ensure absolutely no assumption is left unchecked.\n\n"
),
"max": (
"Reasoning Effort: Beyond maximum — exhaustive, relentless, and uncompromising.\n"
"You MUST reason with the utmost depth and rigor, leaving absolutely nothing to chance: exhaustively decompose the problem into its most fundamental components, trace every causal chain to its root, and resolve the underlying cause rather than any surface symptom.\n"
"Do not stop reasoning until you have independently verified the solution from multiple angles and are certain that no assumption remains unchecked and no error remains undiscovered.\n\n"
),
}
就是不知道为什么用英语,英语效果更好吗?正式版似乎思维链更倾向使用中文的
震惊,我还真不知道原来是用提示词。
实测发现你自己写强度提示词也有一定的影响。(感觉以后在一个对话中可以通过提示词修改强度避免丢失缓存?)
不一定的,有些模型是传入推理强度信号,按理也算prompt中的文本,只不过可能不是自然语言。
模型在训练期间会使用这个推理强度信号,并跟据模型思维链长度做出奖惩,然后模型出厂就能对强度信号做出反应了
因为训练的时候就是用特定的词去训练,这些词对模型采样轨迹影响非常大
第一次听说,我以为是入参上下文窗口大小控制呢
正式版之前有看到过一篇帖子说,不仅仅是提示词,也会调整思考预算
感谢佬,这个prompt写得很好读啊。我放到我电脑里试试。
不用自己输入的,这个是选择强度的时候服务端自动注入的
还是reasoning effort决定上限吧应该是
不对吧,这个内容是哪里看到的, 有url吗?想完整的看一下
huggingface上面开源的文件里
佬,我想问一下这个是在哪里看到的,预览版我记得有看过一篇文章,说曾经的high是没有注入提示词的,max才有注入
woc,绝了,还真是,真是朴实无华的设计。。。。