我的知识库
Search
搜索
暗色模式
亮色模式
探索
llm
此标签下有8条笔记。
2026年6月14日
SFT, RL, and On-Policy Distillation — 分布视角下的 Post-Training
llm
post-training
sft
rl
on-policy-distillation
2026年6月14日
_index
paper
narrative-theory
survey
llm
story-generation
narratology
2026年6月14日
分布视角深度分析:SFT, RL, OPD
llm
distributional-lens
kl-divergence
on-policy
credit-assignment
deep-analysis
2026年6月14日
方法详细拆解:SFT / RL / OPD
llm
sft
rl
opd
opsd
post-training-pipeline
method-comparison
2026年6月14日
SFT, RL, and On-Policy Distillation Through a Distributional Lens
llm
post-training
distributional-lens
sft
rl
opd
on-policy
blog-review
2026年6月14日
cross-cutting-comparison
comparison
creative-writing
llm
cross-cutting
seven-papers
2026年6月14日
extraction
extraction
narrative-survey
narratology
llm
2026年6月08日
Where Do Deep-Research Agents Go Wrong? — 学术深度解读
paper
agent
deep-research
error-localization
llm
evaluation