pith. sign in

Delgado-Chaves, Matthew J

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

years

2026 4

clear filters

representative citing papers

Can AI Agents Synthesize Scientific Conclusions?

cs.AI · 2026-06-09 · unverdicted · novelty 7.0

A new benchmark and clean-room harness show frontier AI agents reach only 0.337 factual F1 when synthesizing conclusions from scientific evidence.

citing papers explorer

Showing 1 of 1 citing paper after filters.

  • Can AI Agents Synthesize Scientific Conclusions? cs.AI · 2026-06-09 · unverdicted · none · ref 33

    A new benchmark and clean-room harness show frontier AI agents reach only 0.337 factual F1 when synthesizing conclusions from scientific evidence.