NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese Extreme Long Video

  • 类型:arxiv
  • 标识:2608.13210
  • 链接:https://arxiv.org/abs/2608.13210
  • 主分类:evaluation
  • 形态:benchmark
  • TLDR:Long-form video understanding encompasses tasks that go beyond retrieving isolated events, including tracking an evolving narrative and interpreting social meaning that may remain implicit. However, existing benchmarks rarely evaluate these capabilities jointly, particularly in high-context, non-English media. To address this gap, we introduce NARU, a benchmark designed to evaluate Narrative evolution and Reasoning on cultural Understanding in Japanese long-form video. NARU consists of 1,481 questions grounded in 155 videos totaling 146.8 hours, spanning four narrative and five cultural dimens
  • 副分类:multimodal
  • 待LLM分类:否
  • 标题中文:NARU:面向日语超长视频中叙事演化与文化细微理解 benchmark
  • TLDR中文:长视频理解任务超越了孤立事件的检索,需要追踪不断演化的叙事并解读可能隐含的社会含义。然而,现有 benchmark 很少联合评估这些能力,尤其是在高语境、非英语媒体中。为弥补这一空白,我们提出 NARU,一个用于评估日语长视频中叙事演化与文化理解推理的 benchmark。NARU 包含 1,481 个问题,源自 155 个总时长 146.8 小时的视频,涵盖四个叙事维度和五个文化维度
  • 来源文件
  • /inbox/tom/_candidates/2026-08-21-agent-rag-longcontext-substack-candidates.json
  • /inbox/tom/_candidates/2026-08-21-agent-rag-longcontext-candidates.json