Pith. sign in

Paper Citation Record · LEDGER

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

As of 10 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 42 inbound Pith citation observations for arXiv:2502.01100.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01100 v2

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:38:53.153716Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:11:51.671603Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:30:01.604649Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 26c1dbd9-342f-4ce9-a719-e869c9bafec1 · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.118674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.118674Z digest=sha256:74f457ac4ce3c03c4a50e738499379099b73434831f740ae9f62bef1f9b12402

Observation 1b772cbe-ddc3-4986-ba6e-285c8ea542de · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.123177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.123177Z digest=sha256:bd72b9b4ffa3dc516e22e4542cd3bbc5a196136c57fea2a97a39c71f524429ec

Observation b57ef200-9268-406f-92c3-06fd89124ec0 · outbound

This paper cites reasoning.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning reasoning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.481424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.127760Z digest=sha256:48cf9fb96b132949bf3bb1a167c458138a8de3b2e463cc601bb816401edce7cd

Observation b40429fc-af56-4806-8f07-1f62dd7c23cd · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T16:38:53.424968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.145282Z digest=sha256:07ce36e1732619a4bb76ef3283e8334ec24fca90037feec133ab2ed3d5db2bff

Observation a3a62524-938e-488f-9071-d383fdbac573 · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T16:38:53.408892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.149451Z digest=sha256:fe0fe9b43f500dbb7c3c68f12e6a558760f4c1d51450163f75290365918d6acf

Observation 62d41093-cc09-4a54-9c2e-3b7655dafba1 · outbound

This paper cites reasoning.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning reasoning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.393570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.153716Z digest=sha256:8c87309b2327f25937c9a8e209f4215eefbcfebacd1592973c8d44bb79b31355

Observation ce0d0ed2-4128-4537-a79c-a0970579a69a · outbound

This paper cites LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.064778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.064778Z digest=sha256:119430bc2d32aa087c5d0a49297cbcb663b2322ee384bf9021216bf3e8d3794d

Observation f94c5ea7-2a75-4b0f-ade2-527698736ea4 · outbound

This paper cites OpenAI o1 System Card.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning OpenAI o1 System Card

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.068978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.068978Z digest=sha256:afbc74320a55878662844b7dfe4e869048b679d0f2bdcce52fe27e8a1ae106f9

Observation 8ff28762-2050-44da-be81-e4695af44caf · outbound

This paper cites org/CorpusID:269330143.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:269330143

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.609329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.076931Z digest=sha256:8cca3e6be8a26eb2a27220337be10672cade816f644a2c1c6ff83731044b3428

Observation 5b183e58-4555-429c-b2d9-63ad359e7f06 · outbound

This paper cites org/CorpusID:245219217.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:245219217

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.579239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.086492Z digest=sha256:c6d5f31aa10ee934ba3592b1ba3469c6ee0b83671a2a596aeadc8f6476ad03ed

Observation 8fc2bf7a-f19f-4a1c-91dc-4e5e4ad7b9fa · outbound

This paper cites org/CorpusID:273233577.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:273233577

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.564681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.090982Z digest=sha256:ba63520cfada816b8cfa2199dc3b516cc72b4b7809cf726e7eab258fc492225b

Observation ad3bcf02-4584-4205-b745-f7acaa993fb1 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.099825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.099825Z digest=sha256:c868c91149e7eb8ee212ce93da30fec17fb26643e69881d4ffa1c432bce0fe34

Observation 024f6f3b-0801-4900-8ca6-cf3aee437532 · outbound

This paper cites org/CorpusID:246411621.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:246411621

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.535632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.104758Z digest=sha256:3e08859fe23e22e3c795c27f537686843bcef07b6c6d00e48c3dd14de234a0cb

Observation e88a2499-5c1f-4e82-82ba-8932c8579b56 · outbound

This paper cites Do Large Language Models Understand Logic or Just Mimick Context?.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Do Large Language Models Understand Logic or Just Mimick Context?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.109204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.109204Z digest=sha256:ca7f8d0fea4e2d9a4cc04e887d548df7fb9c0ab139ddb54ea279dfa48deada09

Observation 10bd76bd-2f6c-42c8-9f89-a83662b5c49d · outbound

This paper cites Bob cannot be in Houses 1, 4, or 5, so he must be in House 3.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Bob cannot be in Houses 1, 4, or 5, so he must be in House 3

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.520017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.113809Z digest=sha256:ab0e8a9e6f69c33ed5aa2344dcce3f0b8c58db511e1c9ef017b74c0610627803

Observation b2821de8-5e40-4420-a795-af32d99c4b48 · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-09T16:38:53.467336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.132338Z digest=sha256:ac924eca9871cfa99af24f8b460ca40d10e0c8ad97f6acd35ca603afd1db0504

Observation 9c5e17f0-e23f-4136-8b39-92f63eb3f313 · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-09T16:38:53.452360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.136805Z digest=sha256:0dcf8981846a1599a65e00a29626bf390b363905bf9a66e3ea7586b3e396922d

Observation ec099d6b-d825-4abc-8bf9-64ce025a3b83 · outbound

This paper cites an unresolved cited work.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-09T16:38:53.438938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.141051Z digest=sha256:7cc59a48f826ea194fcedd1f8f89c9f9d76183f256573ad6e14253d0204dae65

Observation 692edaeb-1f77-4679-a1b5-c8e0d8e2101a · outbound

This paper cites org/CorpusID:33813226.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:33813226

Reference 1984

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.652257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.046197Z digest=sha256:6159a6db9e4a69a56df7389222248feb974fcdfbd10fd27232f035227c2e73fa

Observation f2246a4f-3768-48d2-abed-c36232900f64 · outbound

This paper cites org/CorpusID:36951414.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:36951414

Reference 1993

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.594272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.081923Z digest=sha256:ec04485c98c2cd3976f4e075037fbe27d7f1509759b3e6d2ffbd0a3092ab2201

Observation 656f2903-5fa6-44bc-a2a8-236f8f45312e · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.060289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.060289Z digest=sha256:64bf0b6367de07035e016cfacaaee28fd6eb6a80eb8b2f83dd373fea66455bc5

Observation bf9b47b1-d2fd-47c0-b079-beda7f8e8ad4 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-09T16:38:53.051105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:38:53.051105Z digest=sha256:0b8fbdaf42a998c2bcb14cbdda8b2ccbc03754c16127df08a01614afc9304905

Observation f945cdf5-5b84-419d-9933-b7de4b9232f1 · outbound

This paper cites org/CorpusID:125304065.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:125304065

Reference 2009

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.549987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.095555Z digest=sha256:c96d391cd0e09145449493092e5300a4ab500b0588e0a77203530137acd797a7

Observation fc71e90a-497f-4fae-a867-daf25a124a3c · outbound

This paper cites org/CorpusID:211126663.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:211126663

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.665751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.041724Z digest=sha256:90239c28dcbc192ccd9c433082af59b81b6ba767033afb8791fc179efd0e0f90

Observation f33164bf-3771-45df-b8e1-cfb88cc52b4b · outbound

This paper cites org/CorpusID:247951931.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:247951931

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.678731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.036628Z digest=sha256:04a0c68674c9b12273e160ef76d2a9c8bd9083bce2b520b92eb4a653eb17ff03

Observation 5e7f503f-a5a3-4de1-ba8f-5d6783a0ef93 · outbound

This paper cites org/CorpusID:258967391.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:258967391

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.637963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.056173Z digest=sha256:847e8655d12714815717d0819a4b58c228705fda402e8fde6df284fd01f4a47a

Observation fdb46d2b-eff9-4dd5-9ba1-c5bde73af9fa · outbound

This paper cites org/CorpusID:273234137.

ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning org/CorpusID:273234137

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T16:38:53.623644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T16:38:53.073187Z digest=sha256:db53040395fec35a85960b0084543016790fc43910acd707728868435802669b

Pith citing papers

Observation f2fa3866-e3b0-435a-9775-a9b5422fefab · inbound

CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction cites this paper.

CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T13:11:51.671603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T13:11:51.671603Z digest=sha256:c8a16a66fcb8bb97300846ceda6f70c55e8bb8b2fe4490e3d414502f107eeaa6

Observation dc506c94-06cc-4389-8fe7-c0f6d25fa85b · inbound

Qwen3 Technical Report cites this paper.

Qwen3 Technical Report ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:28.564170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T06:35:27.813995Z digest=sha256:afe908a599a90a0f5a37f5c8c38c823f49318fa884881d13d7fea9de19c85702

Observation 4cca57b7-d88f-4151-bac9-e85fb5ba7eff · inbound

Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles cites this paper.

Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:07.219581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:07.219581Z digest=sha256:a5974284605e6166dd4dd5c30c4bd493df6852601e83dca8b9bc1a76f860f31d

Observation 00c3bd1a-c9f2-4983-9911-f4b4947d08cd · inbound

AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning cites this paper.

AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:24:22.051843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T14:24:21.930934Z digest=sha256:875f4419459dee3d9eb3856b695090a1162b1d6a34df2740b4856063479161ab

Observation a0974290-344d-4535-8144-9be15832edb5 · inbound

LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning cites this paper.

LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:36:35.746562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:36:35.746562Z digest=sha256:2e8a20162977fdf997e617bc03569b0dda3ca0f193218c3152be44726e9f1e82

Observation 0585ec8d-bfb5-4b8b-8c30-03a8a3682e69 · inbound

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs cites this paper.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.311682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.311682Z digest=sha256:6d0b2bb497300ab3c690658e42ad4eb8ad31f65166017d59ab8666684c3301ba

Observation 16742d04-cdeb-40b6-a297-36ef8ac30758 · inbound

MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention cites this paper.

MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:28:16.403060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T09:28:16.189617Z digest=sha256:37ba0e2ee8a5d822d17abb8b82a332ffac4d1844e9c77fbdb76e36fa1ce7fd88

Observation 351480f6-b042-44dc-a133-69877045730e · inbound

Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code cites this paper.

Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:58.771660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:43:58.771660Z digest=sha256:5bda5575154770004efbcf377ef46c06ee0f19ce95dd1fd2a4908749493167d3

Observation cb1fb8de-e48b-4175-b802-3def664383a9 · inbound

Reasoning Strategies in Large Language Models: Can They Follow, Prefer, and Optimize? cites this paper.

Reasoning Strategies in Large Language Models: Can They Follow, Prefer, and Optimize? ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:12:11.173121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:12:11.173121Z digest=sha256:fd301f231e99768c67b58c8ea33ce6463d3e1b992569d13621a2a4505c96e607

Observation 0c02e523-1b76-405a-a274-961ce3ec2b8b · inbound

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning cites this paper.

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:04.502280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:04.502280Z digest=sha256:7b899edf3b1f1b4979746f4f1efbcab2c8bddc6cc4acdb2f01ffee043b40a42c

Observation b506f12c-77cb-4296-a311-36c8da7e5cfb · inbound

Kimi K2: Open Agentic Intelligence cites this paper.

Kimi K2: Open Agentic Intelligence ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:49:28.172386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:49:27.926646Z digest=sha256:16818f893a47ef67b0c489567dabfba9e32a5025e35dbb0a546e8f3cc27a8b8f

Observation fe838800-329f-4278-95cf-3fcfd1e7784c · inbound

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling cites this paper.

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:34:23.971855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T22:33:09.674822Z digest=sha256:edc692420b5b312144bcee01c83ca5d99659fb24bef1781f45796944c9a7c3e9

Observation 7555d3b4-507f-4833-aa9b-f9b872122a48 · inbound

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications cites this paper.

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:28.861283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:26:28.861283Z digest=sha256:3cdaaa1d3e7d39fe644b03f3bd0410f4c79af06fec17fa3069e108041ff7be52

Observation f56fb2ca-c03e-4f25-9f31-00d769556327 · inbound

Investigating Advanced Reasoning of Large Language Models via Black-Box Environment Interaction cites this paper.

Investigating Advanced Reasoning of Large Language Models via Black-Box Environment Interaction ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T21:36:51.933331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T21:35:59.043898Z digest=sha256:d76835623292ad033b91b1bf125fcdc53187a527bc07049fcee595426eaa42f8

Observation 45e2190a-1617-4608-9c38-ee4eadfe2afe · inbound

Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size? cites this paper.

Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size? ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T11:47:20.035612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:47:20.035612Z digest=sha256:c34571109610c190e910996584a08c053e5a3cbd8fcec4c0294ab138825bf1d9

Observation 29969c6a-139c-4e09-a049-55ed9bb828c8 · inbound

Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL cites this paper.

Rethinking Reasoning Quality in Large Language Models through Enhanced Chain-of-Thought via RL ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T04:42:06.276239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:42:06.276239Z digest=sha256:a72e3d616e45070a86d88e80126181a56e24f134d61fa39d61de683ef1696d62

Observation cfdf86bc-2a68-4487-ac7f-fb7db5e14421 · inbound

Qwen3-Omni Technical Report cites this paper.

Qwen3-Omni Technical Report ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:37.831065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T00:20:37.406351Z digest=sha256:e6e9d952130e2f66836d2b7e82b16d37938f457dcd9d63cb5a7f7bbae476f4f3

Observation 51e388bf-3621-4467-b3d9-70da87fa89b4 · inbound

Estimating the Empowerment of Language Model Agents cites this paper.

Estimating the Empowerment of Language Model Agents ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T14:53:16.962827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:53:16.962827Z digest=sha256:3ab77aa00d6c4b46410057e8b410a3170397325a9a74b499eaaed4e7500cdc42

Observation 46bed2ee-9e03-4c97-b0cb-b55b9dadf9c4 · inbound

Soft Adaptive Policy Optimization cites this paper.

Soft Adaptive Policy Optimization ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:14:32.869131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T07:14:32.815641Z digest=sha256:b985a29ddb22ea597d532f7a38ffd33f34aafdffd8cd65a7e43aac7817cec052

Observation 58b14f13-3769-46f3-977f-02472a26027e · inbound

The Format Tax cites this paper.

The Format Tax ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T12:48:03.396039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T12:48:03.396039Z digest=sha256:e397fc1e25873e5919c21559d6a98070101c0d4d4c5bcf86caa10b78fc0d90ff

Observation 0fc03970-34a2-430d-9b45-e42b0253cc76 · inbound

SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions cites this paper.

SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructions ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:11:01.555102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:16:29.695213Z digest=sha256:747bdcf2625613ec0004d1f545111510a2768c017dc80753bf36de2a8e5f6af2

Observation b8d7f44a-eddc-4270-a920-d83b60b68027 · inbound

GroupDPO: Memory efficient Group-wise Direct Preference Optimization cites this paper.

GroupDPO: Memory efficient Group-wise Direct Preference Optimization ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:48.704714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T09:43:18.432084Z digest=sha256:cee2cc5a570d1b7b45c2ecfb7e764558ca039bc8e6879143a04e9c543d145438

Observation 2b5c0ec3-fe0c-4ae1-9d75-7af1606d9d56 · inbound

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts cites this paper.

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:56:11.325941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T05:52:28.822723Z digest=sha256:1a2152f14cb96f3e70f2a1649dee70e195ce0c4c7d65f1e3288ae3945b2cacd2

Observation 9998c1be-9a72-4b5b-8b7e-e07ae61b078e · inbound

Where Reasoning Breaks: Logic-Aware Path Selection by Controlling Logical Connectives in LLMs Reasoning Chains cites this paper.

Where Reasoning Breaks: Logic-Aware Path Selection by Controlling Logical Connectives in LLMs Reasoning Chains ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:14:46.684750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T00:12:16.581972Z digest=sha256:556e7287513776591919d61d3983b1d74c82d3133c9655c0f94fe01210cfd5de

Observation 9940588e-535f-4cc8-8aa6-316859953f91 · inbound

HyperLens: Quantifying Cognitive Effort in LLMs with Fine-grained Confidence Trajectory cites this paper.

HyperLens: Quantifying Cognitive Effort in LLMs with Fine-grained Confidence Trajectory ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:09.666429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T11:38:49.630171Z digest=sha256:b152d8fabb0df09af385fc3bd6d6620e4e9cc090d1e3a5e5d515e19a322d4648

Observation 1dd36255-27a4-4f64-8660-3cedf79632f6 · inbound

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs cites this paper.

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:40:53.495548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T03:40:04.692279Z digest=sha256:64fed2b29f3f06f8f8ae499a35116532ca6ffac7f288a27863106c33f3b80d15

Observation 6e7477b4-022d-4733-a25e-290750169ede · inbound

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs cites this paper.

Benchmarking EngGPT2-16B-A3B against Comparable Italian and International Open-source LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:19:52.885776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T08:14:55.858466Z digest=sha256:63ce9b591339f60402ba3373d3fa378d39587bdfc59e708e005d382f6095e65e

Observation 85ba0fa4-21de-4f20-9e04-e10063649e24 · inbound

MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs cites this paper.

MathConstraint: Automated Generation of Verified Combinatorial Reasoning Instances for LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:01:14.991074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T02:01:04.619417Z digest=sha256:4ab3125c5df8fa33bb844e7b8a8ada58011c1f59ac057faf2c65341921927439

Observation bd0c5757-878d-44da-a928-3fe485e5b560 · inbound

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs cites this paper.

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T02:46:18.842827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T02:44:33.143247Z digest=sha256:7d0b092ad08e69df42720d3a746853d9a2805462ead3845c8dfba3227e8368ff

Observation 87d85edb-bfbe-4555-92ae-924f7189c71d · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:02:42.348193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:caefc480e8ba02a89fd8b0fcdc1de9507c69bf8a9e56a819c2630745ff7782ec

Observation e38bfd7f-9f8e-460c-a5a2-f59fc1bf8212 · inbound

LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening cites this paper.

LLMEval-Logic: A Solver-Verified Chinese Benchmark for Logical Reasoning of LLMs with Adversarial Hardening ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:48:04.283244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T05:47:55.637898Z digest=sha256:bec87005ecf72e54c221431197c688b80ea560e460478a083c8c9336000117fa

Observation ed522808-29fc-4509-9c87-1ce094e1c85b · inbound

CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning cites this paper.

CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:53:23.843244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T11:48:18.855264Z digest=sha256:6403ecc5234d4389d17f72a59e899c9543e0a2c25e0289843680e4831743c568

Observation df6a1f46-96f6-4128-9a52-1a7f4c46de57 · inbound

CA-BED: Conversation-Aware Bayesian Experimental Design cites this paper.

CA-BED: Conversation-Aware Bayesian Experimental Design ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:26:14.632534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T17:03:01.806138Z digest=sha256:a8807e5cdc8c20d0a41124ecd028d078be339c9480304726a2e621234b6ad7f0

Observation 07e1fba3-c6e4-437e-9904-31111a72d9dd · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.872587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:3a52c3e2b4cf07e7242786bf5b858aae8d1d5b7b3de1960a289cbd2152de12b3

Observation 809e8657-9acb-4e1d-91a6-116639170ca7 · inbound

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs cites this paper.

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:41.893244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:04:05.807414Z digest=sha256:d676a1466713193e639f97b311b4e3308cfc04e01b4dee613909db486d42c2ba

Observation 7a99b32d-c596-40ef-ba5d-da3d86673af3 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.647677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:9a7ebabb44a635d7900a1f01e6dc96d25dc862e8e1f99e9d5b7e8f78b447aae6

Observation ea370ea5-625a-484e-84c7-1235846176af · inbound

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR cites this paper.

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:30:01.606119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T22:47:09.330723Z digest=sha256:994d245b86a615ce1a66120c672142103accac5a5c06a4ca6277c6a6f9ba5953

Observation 2da49ecf-37a6-460f-bca8-045dc6a0868f · inbound

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR cites this paper.

Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.466364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T09:55:19.804589Z digest=sha256:42a3b9e2c48674388704b253201056adcf6515e31dacafaffbafd7ca8e362bbd

Observation 15d5d1b2-0339-4a4e-b693-2e59d32cd680 · inbound

NebulaExp-8B: An Empirical Post-Training Pipeline via Full-Scale Ablation Research cites this paper.

NebulaExp-8B: An Empirical Post-Training Pipeline via Full-Scale Ablation Research ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:51.132305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T05:00:23.510590Z digest=sha256:6455e9289071e91cd48f48db22bcead0684a33518867b2471c71fe357fdd7b22

Observation 54c656c9-f609-458d-9dc2-5c95f6ba1637 · inbound

DOPD: Dual On-policy Distillation cites this paper.

DOPD: Dual On-policy Distillation ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-30T05:54:18.636426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T05:51:18.037199Z digest=sha256:247b1836865201bb00c4be757e15421fcc2882534304de81a086dac45bc27f1b

Observation cad6590c-cf1c-48fa-a5a2-3e51e0b2dc1d · inbound

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning cites this paper.

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T04:30:41.776241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:30:41.776241Z digest=sha256:19a8926d0a630bcdec6cf6bbbcc71549bab9343dc744c44abb9ea5d08d90849f

Observation 6159d5fd-f597-4de4-8f06-d6989927b676 · inbound

Reasoning Core: Designing Broad Procedural Data for Completion-Supervised Reasoning Training cites this paper.

Reasoning Core: Designing Broad Procedural Data for Completion-Supervised Reasoning Training ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T04:19:48.346566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:19:48.346566Z digest=sha256:f1c2ffe7767044ad607a9dc05c036228ea178726191e0661e959a08ee04f748a