Pith. sign in

Paper Citation Record · LEDGER

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction

As of 14 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 1 inbound Pith citation observation for arXiv:2501.14210.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14210 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:17:54.184997Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:28:05.353433Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:28:09.181666Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact3
  • verified fuzzy1
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c30c49d6-0afd-4754-b24a-c07db3e3ea51 · outbound

This paper cites online" 'onlinestring :=.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:53.998475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:53.998475Z digest=sha256:8d6426ba9ef1ff0112f33a7e2921d5abf96542e8c7131e29aca0df02fcc3cd5a

Observation be048d3a-de2d-4bdd-908b-414e390611c3 · outbound

This paper cites write newline.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.005148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.005148Z digest=sha256:bcf11f75d5b2d887e0f4a68fe6a355cf64db8b17466557d179dceac1f2dcdfb9

Observation 8f521cca-703d-417b-88cf-938d9f19fdc2 · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Flamingo: a Visual Language Model for Few-Shot Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.010963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.010963Z digest=sha256:b8e687cbc8ed5677ed778e2a844d13e329de5a1177ddf93cf41a99fd66e12bec

Observation 415383ee-f8bd-4f39-a131-a33511196f41 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.999405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.017581Z digest=sha256:95f093562e54f51e75518d65a4b9fe0c9a316f5dadb6947c5abce48f6e8ae366

Observation a591148b-b116-4658-b351-a3f73c806e23 · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Lawrence Zitnick, and Devi Parikh

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:17:54.980738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.023073Z digest=sha256:9377ff400e54533faf5b4a8f944c0f51638764ceec45c267a56263465d0b8445

Observation a8720c57-b617-45a1-83aa-51197d250392 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.028474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.028474Z digest=sha256:c02f4b89261476357062d03944d3a4189336f2b0360e2d434da1884bdb4d5700

Observation b4ec5e97-8b24-4e94-838b-2075b0fcb267 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.034110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.034110Z digest=sha256:8b9ffd3a2aaecb6061b270408986007287874611ab86524a20dde09995ff4dbe

Observation 7d850b0a-f59c-4fd4-93b1-e9e8deb2f5b6 · outbound

This paper cites There is a Time and Place for Reasoning Beyond the Image.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction There is a Time and Place for Reasoning Beyond the Image

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-10T15:17:54.753588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.039737Z digest=sha256:78d8c2064f2c00cb68898500b7689799f7d2ccbbb3f9c852931e5f43604f3597

Observation e413db83-12a3-4a93-b623-a79f492e4b8f · outbound

This paper cites Visual Programming: Compositional visual reasoning without training.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Visual Programming: Compositional visual reasoning without training

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.047245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.047245Z digest=sha256:a3c891f1cd28a2228d8f3aa48204e0521fb6a505b4a82adb617502a2b9163959

Observation a58f4e2e-52fc-47f9-9a51-449e5086b210 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.960224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.052657Z digest=sha256:6069e13184abefb29abfbb570dde033503315fbe256c8fcbe69079ee78f90c0a

Observation 3075c3f1-4f19-43be-86b4-e01d148f4e0b · outbound

This paper cites GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction GQA: A New Dataset for Real-World Visual Reasoning and Compositional Question Answering

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.057898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.057898Z digest=sha256:c9a0cf8fbd37c7ec96cc002fc5d356fa1e63d5e93f7f4b91f05b11258f44eba8

Observation bc1285c3-afe4-4339-9141-31ee8abbcc89 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.064584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.064584Z digest=sha256:a4e98fabd7a7d7bc65f67b29df261f2952fbc715195308ba0e38a7decd3c2fce

Observation ffc4e600-fa53-49d7-b657-c77abf6f5688 · outbound

This paper cites mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction mPLUG: Effective and Efficient Vision-Language Learning by Cross-modal Skip-connections

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.070163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.070163Z digest=sha256:e1f866cc48ba35c1ee149f027a72a4fb6fbeb3f4ecde9735f376b724be4f0b05

Observation 6a46d37a-21c2-467d-a717-bd5d0429173c · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.075828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.075828Z digest=sha256:812a8f91f43319870a33de255f9d39c514fe0fb5f53abec7e9eeaf39f39be13c

Observation 041a1335-ab7a-457a-b926-b57ed55e3e8f · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.927841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.081116Z digest=sha256:cc548d93074ac8e821aaec10672d037c092cf1088a3a4c3f2091bc4bcc11c19b

Observation 8131ef9e-9e7b-484e-9a97-ce3a5853bb34 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.086437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.086437Z digest=sha256:95e6a3aadd84aa383e5a51a3c4d47a4b8841bab9b43254a539460a5814f36a87

Observation 7368adf6-dc27-41b3-81dd-a28e72ba085b · outbound

This paper cites Visual Instruction Tuning.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Visual Instruction Tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.092735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.092735Z digest=sha256:f1d3240a67ca4aec243be5041741d549d7f7169b90a29b6df0034b4e95d63f16

Observation c97c7927-af26-4996-820a-17effd689d97 · outbound

This paper cites Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.098007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.098007Z digest=sha256:36f76f79fa867c5dfb21d2cc62f4b80a3df12224e54df15da4d6e4fec856638e

Observation 8e0d922b-041a-4fe1-8fb8-661d08c6512c · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.103443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.103443Z digest=sha256:a21fde5b347e0cb9082b50b2865dbfec812ebd09d0708533f617bb19e07ab77f

Observation 2bd9155d-b0d4-4810-bdde-591376904d75 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.109647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.109647Z digest=sha256:4666d44aa1ded0a7dce55558d9e07199ce4e287ef2f57963947b8a4110111ce1

Observation 49cd37f3-c636-4843-ab84-da7cd7b80e61 · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.115384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.115384Z digest=sha256:a3966a47c367775bd722a9cea109b5307f9e920deda19f381dfea1b7f96774f1

Observation 2b945af6-cfb0-4da5-9dc7-866af2caeacf · outbound

This paper cites Neural Programmer: Inducing Latent Programs with Gradient Descent.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Neural Programmer: Inducing Latent Programs with Gradient Descent

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.121462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.121462Z digest=sha256:8ba863a130901dc338ebbf0d7089ec99450158104f8c21149f56ad75e1fe348e

Observation 66470030-ceab-49c7-94f1-0ae0b8b6586b · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.868327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.127358Z digest=sha256:1d633dbcf6673dca62be6333396aa9c8a96fdc63a5d27b7a0f8d9bdae259d4fa

Observation a614c4b6-fc66-4afc-a26c-f431a739970a · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T15:17:54.565298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.133474Z digest=sha256:5bb00ba15929bebd5819afe759a50009e5033c5a3d49cbd730f705ba6edd0bfe

Observation 3b036b9d-8995-4c5f-aa73-7c10aed68586 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Learning Transferable Visual Models From Natural Language Supervision

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.139176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.139176Z digest=sha256:a79a055ee12e076ee0dafbb2f574fe436f08303279f44bdd55d52ebe1d06e77e

Observation a3d101e3-9fe9-4e9e-a91c-1aa577e07799 · outbound

This paper cites QR-CLIP: Introducing Explicit Open-World Knowledge for Location and Time Reasoning.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction QR-CLIP: Introducing Explicit Open-World Knowledge for Location and Time Reasoning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-10T15:17:54.322802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.144968Z digest=sha256:2f09b6dbd35245d66940d8fd585df2456aa6b46ecb99db4a785a558818a23f59

Observation eaa4216d-2878-4dc7-8db3-f7999eac5a03 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Gemini: A Family of Highly Capable Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.152319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.152319Z digest=sha256:62b896f999d6de0653abf0b91fc1f29286e1dcf72058a1134a4c1f2f8c83b431

Observation 6dfbf2cf-37b3-4eb3-a890-6b41037110d6 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction LLaMA: Open and Efficient Foundation Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.159378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.159378Z digest=sha256:de3b984bb07bbc1cc61cfa35dd0a43a5456fa6ef424fca6ec0acbc4b098a1552

Observation efaca795-5054-4363-affe-1f35efc0f06e · outbound

This paper cites Visual Entailment: A Novel Task for Fine-Grained Image Understanding.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Visual Entailment: A Novel Task for Fine-Grained Image Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.166536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.166536Z digest=sha256:c6ca09973239445a927c79cb594bb0c37fec938fe8a66d5867d8fdae7cf92622

Observation 653e19d5-2833-4234-8574-a83fbc427f1a · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.849438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.172532Z digest=sha256:b91354580ab1b5cff706f3a749e520260efe40ffe117ff43fe1d27b7c38d1da7

Observation f4d5fe1b-f6e8-4d3f-9a97-8d1d4e472ccd · outbound

This paper cites an unresolved cited work.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:17:54.830976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-10T15:17:54.178374Z digest=sha256:8750ed879f0deb8b932fbde92ff1552811c65af4626c5bdc0f928dc794fd5dd9

Observation a538b10d-ce38-44e3-a634-9e0672db419d · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:17:54.184997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:17:54.184997Z digest=sha256:14869d520f93ae74b5cead3799720350682fd8c2107a013c7cf56de780a0f6bc

Pith citing papers

Observation 82156af2-590a-4ab0-a8f7-d80e7a48338d · inbound

GETReason: Enhancing Image Context Extraction through Hierarchical Multi-Agent Reasoning cites this paper.

GETReason: Enhancing Image Context Extraction through Hierarchical Multi-Agent Reasoning PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:28:09.274276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T13:28:05.353433Z digest=sha256:d795db2b8d2dd8a18b33445eb161977661298e4d6f7d861f0476351f174ecd1f