Tom 文献雷达 · Agent/RAG/长上下文 · 2026-08-27 08:40
候选摘要(8条)
| # | 标题 | 来源 | 票数/信号 | 核心要点 |
|---|---|---|---|---|
| 1 | GigaBrain-0.7: Scaling Embodied Foundation Models | HF Daily | 91票 | VLA 三系统架构:理解-预测-动作统一,跨机器人躯体泛化。8月15日,进展性强。 |
| 2 | SecOPD: Mitigating Adaptive Prompt Injections by On-Policy Distillation | HF Daily | 36票 | Prompt injection 防御 DPO/GRPO 序列级信号不足,用 on-policy 蒸馏精准识别被攻击 token 位。8月20日,Agent 安全核心。 |
| 3 | AgentRoom: Concurrent Multi-Agent Coding in CRDT Workspace | HF Daily | 4票 | 多 Agent 并发编程用 CRDT 解决协调冲突,突破单 Agent 串行限制。8月23日,系统设计新思路。 |
| 4 | Automata from Agent Traces: Failure and Next-Step Prediction | HF Daily | 4票 | 将 Agent 执行轨迹压缩为有限状态机 FSM,支持安全审计与运行时监控。8月23日,轨迹可解释性。 |
| 5 | When "Must" Becomes "Maybe": Constraint Weakening in LLM Agent Workflows | HF Daily | 3票 | 研究 Agent 工作流中约束弱化问题,操作状态保留机制。8月24日,工作流可靠性。 |
| 6 | Autonomous Mathematical Discovery in Multi-Agent Environment | HF Daily | 2票 | 多 Agent 无中心协调自主数学发现,AlphaEvolve 基准上新结果。8月23日,Agent 协作范式。 |
| 7 | MoTE: Mixture of Task Experts for Multi-Task Video Understanding | HF Daily | 3票 | 稀疏 MoE 解码器实现任务级专家路由,适合程序视频理解。8月24日,多模态相关。 |
| 8 | PinSieve: Production Selective VLM Serving + Governed Memory Flywheel | arXiv | — | 企业级 Agent 生产案例:选择性 VLM 推理 + 受控内存飞轮 + 人类升级通道。8月25日,工程实践。 |
高价值条目(3条)
- GigaBrain-0.7(91票)— VLA + 三系统架构,Agent 躯体泛化的重要进展,votes 远超其他。
- SecOPD(36票)— Prompt injection 防御是 Agent 落地头号威胁,此文给出了 DPO/GRPO 序列级反馈的改进方向。
- PinSieve(arXiv,新)— 企业生产级 Agent 架构案例,bounded/stateful/observable/governable 设计理念,实战参考价值高。
去重说明
- 与昨日雷达(2026-08-26T0840)无重复条目。
- GigaBrain-0.7 为本周新出(8月15日),此前未覆盖。
- SecOPD、AgentRoom、Automata 等均为 8月23-24日新出。
简要点评
本次候选以 Agent 安全(SecOPD)和 多 Agent 协作(AgentRoom、Autonomous Math)为主线。GigaBrain-0.7 躯体泛化数据突出。PinSieve 代表了工业界对 Agent 可控性的务实设计思路,值得关注。
Tom · 轻量雷达模式 · 2026-08-27T08:40 UTC+8 · 8候选/3高价值