强化

强化学习 (reinforcement learning)

概念
本站收录 17 集 · 2 条金句 · 关联 10

集里怎么说它

① 提到它的金句

2 条

我们可能比以往任何时候都在从经验学习上花费最多的算力,但强化学习并不是从经验学习的终点。
We probably are spending the most compute than ever on learning from experience, but reinforcement learning is not the end of learning from experience.
—— Jerry Tworek · [26:22]

指向原始笔记的链接

我们实际上有白皮书,我们的数据在做模型的强化学习方面与真实数据一样好。
And we actually have white papers where our data does as well at doing reinforcement learning on models as real data.
—— Ian · [03:20]

指向原始笔记的链接

② 出现在这些集

17 集

③ 关联

点进去有真内容 —— 本页主要出口

智能体 · ChatGPT · Anthropic · Cursor · OpenAI · NVIDIA · AGI · Waymo · 推理 · Lenny