pith. machine review for the scientific record. sign in

Natural questions: a benchmark for question answering research.Transactions of the Association for Computational Linguistics, 7:453–466

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

citation-role summary

dataset 1

citation-polarity summary

years

2026 4 2025 1

verdicts

UNVERDICTED 5

roles

dataset 1

polarities

use dataset 1

representative citing papers

Group-in-Group Policy Optimization for LLM Agent Training

cs.LG · 2025-05-16 · unverdicted · novelty 7.0

GiGPO adds a hierarchical grouping mechanism to group-based RL so that LLM agents receive both global trajectory and local step-level credit signals, yielding >12% gains on ALFWorld and >9% on WebShop over GRPO while keeping the same rollout and memory footprint.

citing papers explorer

Showing 5 of 5 citing papers.