Cameron R. Wolfe (Deep Learning Focus) · RSS 摘要
信源:Cameron R. Wolfe (Deep Learning Focus) · https://cameronrwolfe.substack.com/feed
- LLM 强化学习:完全指南 — Tracing the evolution of RL from first principles to the frontier of modern AI research...
- Midtraining 笔记 — 通过 midtraining 与持续预训练构建更优的专用 LLM
- 智能体世界模型 — Creating better language agents by teaching them to model their environment...
- Agentic RL: Frameworks and Best Practices — How LLMs are trained to handle long horizon tasks in complex environments...
- Agent Evaluation: A Detailed Guide — Best practices and common patterns for effectively evaluating AI agents...