Cameron R. Wolfe (Deep Learning Focus) · RSS 摘要
信源:Cameron R. Wolfe (Deep Learning Focus) · https://cameronrwolfe.substack.com/feed
- Agentic World Models — Creating better language agents by teaching them to model their environment...
- Agentic RL:框架与最佳实践 — How LLMs are trained to handle long horizon tasks in complex environments...
- Agent 评估:详细指南 — Best practices and common patterns for effectively evaluating AI agents...
- LLM 的 RL 缩放定律 — How scaling laws have evolved from pretraining to reinforcement learning...
- LLM 基准的解剖 — Common patterns used to create the most effective LLM evaluation datasets...