Pith. sign in

Paper Citation Record · LEDGER

FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2402.18667.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.18667 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:36:35.362434Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T08:49:53.756683Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f1a84af6-924c-45d2-b54e-95e4d69ae204 · inbound

Verifiable Format Control for Large Language Model Generations cites this paper.

Verifiable Format Control for Large Language Model Generations FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T22:36:35.362434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:36:35.362434Z digest=sha256:42122ac2287b088736276d2755c235870d72b0649990dffbc87c1359f0078363

Observation 449781f1-aa04-4d85-9d85-8fefacf44433 · inbound

Enabling Autoregressive Models to Fill In Masked Tokens cites this paper.

Enabling Autoregressive Models to Fill In Masked Tokens FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T17:10:39.848115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:10:39.848115Z digest=sha256:cae921e88ec94800b3183bc6cba210e6eac60202cbb61e07b448f8a0488bbb61

Observation aeca3003-8e41-4678-af24-52cb62a1a04b · inbound

AraTable: Benchmarking LLMs' Reasoning and Understanding of Arabic Tabular Data cites this paper.

AraTable: Benchmarking LLMs' Reasoning and Understanding of Arabic Tabular Data FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T14:37:37.794476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:37:37.794476Z digest=sha256:b47dc8d4207c0805075f0a51d46815f660e325bc7c81f88106941d22ff5ad950

Observation 59480ad6-dce1-4844-ae79-3494d9013396 · inbound

SAGE: A Service Agent Graph-guided Evaluation Benchmark cites this paper.

SAGE: A Service Agent Graph-guided Evaluation Benchmark FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:00.959008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:41:23.956104Z digest=sha256:dab2667d66c618c9a11df2241371a72d850448e7d3f645a0323c68802f6c02ea

Observation 95d2e094-72f2-4069-8987-ecead4b90d54 · inbound

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents cites this paper.

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:49:53.758545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T08:45:56.550821Z digest=sha256:0c21d15c7f192d6233a4abfbd6bab981b49b96be133bcde2872f18b1a6234ad6

Observation 66ceaa30-b033-437f-88bb-41be7ef9e143 · inbound

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning cites this paper.

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T21:31:51.745650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:31:51.745650Z digest=sha256:49551e1996ad8ce5650a6ac2242f158dbf4a877a3deb09ca5b9b345bcfcfdef1

Observation 3826f26e-4bbb-42c4-ae95-9c57ee8835f6 · inbound

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning cites this paper.

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning FOFO: A Benchmark to Evaluate LLMs' Format-Following Capability

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T00:49:50.819098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:49:50.819098Z digest=sha256:75da77373455bac8fa90550a019d737da2cc13e004854d2424a51afce79c9429