围绕 LLM 人类价值观与多元对齐的精选论文、基准、数据集与工具合集。A curated collection of papers, benchmarks, datasets, and tools on human values in LLMs and pluralistic alignment.
evaluationriskllm-infra
2 个 · 安全与风险 · 收藏榜 · AI 核心
围绕 LLM 人类价值观与多元对齐的精选论文、基准、数据集与工具合集。A curated collection of papers, benchmarks, datasets, and tools on human values in LLMs and pluralistic alignment.
大语言模型情境感知相关论文精选目录。Curated bibliography of papers on situational awareness in large language models