feat: gate writer sections on a % teach: block (SP025)
Rewrite references/pedagogy.md around recite-vs-teach: reader model,
coverage density (sections_in is cite permission, not a to-do list),
intuition-before-formula three-beat, and paper-jump filling.
Every notes/sections/sec-*.tex must now answer gap / takeaway / jump /
omit above the first \section. lint.py checks presence only (it cannot
judge honesty); files opening with "% generated by" are exempt.
- scripts/lint.py: SP025 + is_writer_section / teach_gaps helpers
- tests/no-teach-block/: fixture with takeaway only, wired into test.sh
- examples/*: all 14 writer sections get real teach blocks
- SKILL.md, references/{agents,antipatterns,checklist}.md, DESIGN.md,
assets/notes-template.tex: route writers and consistency agent
through pedagogy.md
This commit is contained in:
@@ -1,3 +1,8 @@
|
||||
% teach:
|
||||
% gap: 知道模型要输出预测,但没想过损失算不算前向路径上的一站
|
||||
% takeaway: 这份摘录只问一件事——前向到哪里为止
|
||||
% jump: none
|
||||
% omit: 摘录之外的全文结构与实验
|
||||
\section{这篇论文在问什么}
|
||||
要把输入变成预测,最简单的机制是什么?损失要不要走在前向主路上?
|
||||
这是摘录 fixture,不假装读完全文。
|
||||
|
||||
@@ -1,3 +1,8 @@
|
||||
% teach:
|
||||
% gap: 还没有一句能离开 PDF 复述的核心主张
|
||||
% takeaway: 一次前向是 $x\to f_\theta\to\hat y$;损失在预测之后单独比较
|
||||
% jump: none
|
||||
% omit: 重复摘要的贡献条目
|
||||
\section{主张与贡献}
|
||||
\splabel{C1}
|
||||
一次前向是 $x \to f_\theta \to \hat y$。损失 $L(\hat y,y)$ 在预测之后单独计算,不是主路上的一站。
|
||||
|
||||
@@ -1,2 +1,7 @@
|
||||
% teach:
|
||||
% gap: 不知道 $f_\theta$、$\hat y$ 在本讲义里各指什么
|
||||
% takeaway: 预测器就是 $\hat y=f_\theta(x)$,其余符号查附录,不在正文重列
|
||||
% jump: none
|
||||
% omit: 论文式的记号巡游
|
||||
\section{预备:定义、假设、符号}
|
||||
预测器定义为 $\hat y = f_\theta(x)$。符号见附录。
|
||||
|
||||
@@ -1,3 +1,8 @@
|
||||
% teach:
|
||||
% gap: 会算内积,但不知道维数一涨为什么会把后面的非线性弄坏
|
||||
% takeaway: 除 $\sqrt{d}$ 是把方差按回 1 的改写,不是新算子
|
||||
% jump: 摘录直接写下 $1/\sqrt{d}$,没说维数涨会让点积方差跟着涨
|
||||
% omit: 数据集、超参、硬件
|
||||
\section{一次前向与损失}
|
||||
内积的方差会随维数 $d$ 涨。为了不让后续非线性饱和,要把点积除掉 $\sqrt{d}$。
|
||||
|
||||
|
||||
@@ -1,2 +1,7 @@
|
||||
% teach:
|
||||
% gap: 学完这一条机制,不知道同样的流程怎么套到一篇完整论文
|
||||
% takeaway: 先用 ledger 钉住 claim,再按路由决定出不出图
|
||||
% jump: none
|
||||
% omit: 摘录没覆盖的实验与消融
|
||||
\section{总结与延伸}
|
||||
摘录只保留一条机制:一次前向加侧路损失。更长的论文用 ledger 把 claim 钉住,再按路由出图。
|
||||
|
||||
Reference in New Issue
Block a user