SOTA Papers
LatestTrendingDigestSearch
Newsletter
  1. Home
  2. Sydney Levine

Sydney Levine

1 article on SOTA Papers

Computer Science

CogGym Measures the Gap Between Models and Human Judgments

A unified evaluation of 50 language models across 258 cognitive experiments finds improving model--human fit, but a substantial gap from human reliability remains.

cs.AI
Sep 22, 20263 min2609.21259

Co-authors

Lance YingJinzhou WuYingshan Susan WangShivam AaryaLuca M. Schulze BuschoffHarry ChenKatherine M. CollinsAndrea de VardaShuhao FuSean Dae HoulihanAkshay K. JagadishGuangyuan JiangSamuel KiegelandTetsu KurumisawaRongzhi LiuRyan LiuNingshan MaKathryn McGregorYounes StrittmatterPolina TsvilodubJacob Hoover ViglySarah WuEnjie XuYiling YunKelsey AllenTyler Brooke-WilsonBrian ChristianEvelina FedorenkoMichael C. FrankMichael FrankeTao GaoSamuel J. GershmanRobert D. HawkinsJennifer HuJulian Jara-EttingerMax Kleiman-WeinerTal LinzenHongjing LuTimothy O'DonnellDesmond C. OngSteven T. PiantadosiRebecca SaxeEric SchulzTianmin ShuFelix A. SosaIlia SucholutskyTan Zhi-XuanTomer UllmanFei XuIlker YildirimJian-Qiao ZhuThomas L. GriffithsTobias GerstenbergKevin SmithJoshua B. Tenenbaum
SOTA Papers

Curated editorial synthesis of meaningful arXiv papers. No slop, no overclaiming.

Fields

  • Computer Science
  • Physics
  • Mathematics
  • Economics
  • Electrical Engineering

Navigation

  • Latest
  • Trending
  • Weekly Digest
  • Search
  • Contact

Legal

  • Terms of Service
  • Privacy Policy

© 2026 SOTA Papers. All rights reserved.

Data sourced from arXiv.org