Cameron R. Wolfe (Deep Learning Focus) · RSS 摘要
信源:Cameron R. Wolfe (Deep Learning Focus) · https://cameronrwolfe.substack.com/feed
- Agentic World Models — Creating better language agents by teaching them to model their environment...
- Agentic RL:框架与最佳实践 — How LLMs are trained to handle long horizon tasks in complex environments...
- Agent 评估:详尽指南 — Best practices and common patterns for effectively evaluating AI agents...
- LLM 的 RL Scaling Laws — How scaling laws have evolved from pretraining to reinforcement learning...
- LLM Benchmark 的构成解析 — Common patterns used to create the most effective LLM evaluation datasets...