StudentBench: AI and human tutoring yield equivalent GRE learning gains

  • 类型:arxiv
  • 标识:2609.28470
  • 链接:https://arxiv.org/abs/2609.28470
  • 主分类:evaluation
  • 形态:method
  • TLDR:Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection with over 175,000 student-AI messages to study whether large language models (LLMs) produce learning gains equivalent to human tutoring. Using StudentBench, we measured learning gains on Quantitative and Verbal GRE questions across 2,383 human participants receiving AI tutoring, human tutoring, or n
  • 待LLM分类:否
  • 来源文件
  • /inbox/tom/_candidates/2026-09-24-agent-rag-longcontext-candidates.json