Pith. sign in

Paper Citation Record · LEDGER

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2606.22968.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.22968 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T06:26:59.984686Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:50:52.423769Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T21:50:53.152351Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f5c072c1-75e7-4512-bfbd-b2215cd3feb2 · outbound

This paper cites In: Proceedings of the ACM SIGOPS 31st Symposium on Operating Systems Princi- ples.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: Proceedings of the ACM SIGOPS 31st Symposium on Operating Systems Princi- ples

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:c799e67e64ddda31941045390e7d3dcc2541b50ef525c80f6fe6f449d29ab8c9

Observation 3c903882-7d1c-40c9-a29e-71ccea2f3404 · outbound

This paper cites PALM: A Efficient Performance Simulator for Tiled Accelerators with Large-scale Model Training.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference PALM: A Efficient Performance Simulator for Tiled Accelerators with Large-scale Model Training

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:39:49.229562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:aff75b586d014bbe1d02c60a68902f455971b833e1c23855883c08e548d1319e

Observation 602676aa-5bc4-4c26-8bcc-5f0d44667e6a · outbound

This paper cites WaferLLM: Large Language Model Inference at Wafer Scale.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference WaferLLM: Large Language Model Inference at Wafer Scale

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:39:49.221490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:766d0398b1afd39590313ec2ce01bfabdc8af9f4b8fc3281c3a8f9fb6422c85a

Observation b3d1d655-9327-42bd-80dc-dfed180bc5e1 · outbound

This paper cites IEEE Circuits and Systems Magazine24(1), 52–81 (2024).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference IEEE Circuits and Systems Magazine24(1), 52–81 (2024)

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:c3b595e29b3e96c0a6a042f52adbe4c188a461849e9739e4bec2fbdeec9fc2de

Observation d6613e2b-d9cd-49a9-b647-bb700e1331a1 · outbound

This paper cites Advances in neural information processing systems32(2019).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Advances in neural information processing systems32(2019)

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:912ea9183a1515bc13e799c173b0aa03311f2dc826de0266770d544e0fc965d4

Observation 0f2b4e95-6100-406f-8ccc-1b177cafbf21 · outbound

This paper cites In: International Conference on Machine Learning.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: International Conference on Machine Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:1eb23419e52de2c73f0de120d14c1def1b4f24d27183632d0139a76ea6cb6359

Observation 5caec8c2-7734-47b9-8182-331c074a4dd4 · outbound

This paper cites IEEE Transactions on Circuits and Systems I: Regular Papers (2025).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference IEEE Transactions on Circuits and Systems I: Regular Papers (2025)

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:b0508bea50fb02970ba08d7f3830af396f6686f32189174fd39b84b4fd03d0d3

Observation 7fa9a676-78f2-480b-9f1c-49e941086941 · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:02c4c39bfad7fb6e7c54790834245ac5d6587bdb7fbfd5a377a096d1fbbb10aa

Observation 3d7f6b93-684b-4256-bb98-452874fd9810 · outbound

This paper cites In: International Conference on Machine Learning.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: International Conference on Machine Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:d4ebc51ae1bc2e92e761e130a120faf7a3dbd7df1f4f143852b59af30a8e3901

Observation eafb6db0-1498-427f-87fb-f7cded1c0970 · outbound

This paper cites In: 2022 IEEE Hot Chips 34 Symposium (HCS).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2022 IEEE Hot Chips 34 Symposium (HCS)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:121a98d3975fea3b6918d603ec1a288de94ccb0e7179e2c04f33d5d937f75175

Observation dff5c0af-e544-4e65-942f-3a8b60470d05 · outbound

This paper cites In: Proceedings of the 31st ACM International Conference on Architectural Support for Programming Languages and Operating Systems, Volume 2.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: Proceedings of the 31st ACM International Conference on Architectural Support for Programming Languages and Operating Systems, Volume 2

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:abfd16be1ad324685ada8c563c6e54341101bc91eebf786f4df770fbf0b74206

Observation 596d0626-2518-4580-89a5-88a6927d6dea · outbound

This paper cites In: MICRO- 54: 54th Annual IEEE/ACM International Symposium on Microarchitecture.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: MICRO- 54: 54th Annual IEEE/ACM International Symposium on Microarchitecture

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:ce92838539be7e0837c1700d72e284255d03a541810098c4f5dab6d379b5f579

Observation 29156ffe-abd8-4ed3-baac-635d78816ba5 · outbound

This paper cites Available:https://huggingface.co/ meta-llama/Meta-Llama-3-70B.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Available:https://huggingface.co/ meta-llama/Meta-Llama-3-70B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:5cfbd494f76ab5073012e19b18940634967a36f103e3d4613caa085993a6b695

Observation 14bd7cb1-7db7-4994-b834-9d26d93a1def · outbound

This paper cites Available:https://huggingface.co/ meta-llama/Llama-3.1-405B.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Available:https://huggingface.co/ meta-llama/Llama-3.1-405B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:ee9250f2eec08f010709c8d5415662161032635ebe3065878cc97d4dfd487d4e

Observation fed372b1-01c3-46c5-8485-d5c14ac56b01 · outbound

This paper cites Available:https:// huggingface.co/mistralai/Mistral-Large-Instruct-2407.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Available:https:// huggingface.co/mistralai/Mistral-Large-Instruct-2407

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:6003e1306606cf7edd073d4133807a1eba95864b665b602f95bc62f61da13e16

Observation 93d7c482-393c-4c02-a56f-dc743b43ac55 · outbound

This paper cites In: 2025 IEEE International Solid-State Circuits Conference (ISSCC).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2025 IEEE International Solid-State Circuits Conference (ISSCC)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:34cda36d323f37d0e9adca9af1759ae3a4e9a207666308d4fbe00b799c8f3b94

Observation beba5489-f8ef-4b14-a692-d0bcb33bc3d3 · outbound

This paper cites Available:https: //www.nvidia.com/en-us/data-center/dgx-vera-rubin-nvl72/.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Available:https: //www.nvidia.com/en-us/data-center/dgx-vera-rubin-nvl72/

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:67510306f64c036533144b7f3585c85cc7865d3161065e6b122357a2800867f6

Observation 0631fbe1-3e0d-4ebe-b4f9-63aaa3889ac9 · outbound

This paper cites In: 2021 58th ACM/IEEE Design Au- tomation Conference (DAC).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2021 58th ACM/IEEE Design Au- tomation Conference (DAC)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:a2edc8768711093b0ca43e7c3d0cacc548dc5b70eab6cb5c5c881c6767196ef0

Observation 401496bc-e95e-4904-aa2b-663a4cd8f246 · outbound

This paper cites Available:https://huggingface.co/ Qwen/Qwen3-235B-A22B.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Available:https://huggingface.co/ Qwen/Qwen3-235B-A22B

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:c7aa1b729d3c2e9058499858de94d68398520b4aca9fc7c252635ef8d2c98cc5

Observation 33695bc3-cce5-458c-92b8-dd4c5f9b254b · outbound

This paper cites FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:39:49.224227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:13f3d74e4a9dbc53e545385a6d9a7630011cf21fce4f9fba93d42ebae1624b9c

Observation dc50a940-a137-42c5-b63e-a4fa810d8c0c · outbound

This paper cites In: 2025 IEEE 75th Electronic Components and Technology Conference (ECTC).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2025 IEEE 75th Electronic Components and Technology Conference (ECTC)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:773dd8bf0dae99edb4e986b0cce4a6ae29b1189e50fb2249ff8116c899ac9d17

Observation f620b95b-fd81-483b-ad44-51c53e4b910b · outbound

This paper cites In: 2022 IEEE Hot Chips 34 Symposium (HCS).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2022 IEEE Hot Chips 34 Symposium (HCS)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:c1dc00ca398ff33d37c117c381f190ee8a2f765300ae05ce68471b4310614769

Observation 772081c9-eda1-4864-aef6-a541546c88ce · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:7d6b1a7fa3d77184c505cabb5ea2fe2aaffcfc30f515aaf3b061db0edc28dee1

Observation e648af3f-22d2-44a6-9ae5-02770a8f47c7 · outbound

This paper cites In: 2024 57th IEEE/ACM International Symposium on Microarchitecture (MICRO).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2024 57th IEEE/ACM International Symposium on Microarchitecture (MICRO)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:f266ace75d8a629a8abe39d6c5157e231cd9b6f7c116c573e9d82b42b50ded1d

Observation 5d3f9e29-bcf8-4592-8f29-effd411b7a41 · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Archi- tecture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Archi- tecture (HPCA)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:e8f01f459df795e30dca819f2d07455a6665e4bd1e2e59d90f0eaca8c612852d

Observation 73ec4013-42fc-4d03-8370-d3288ca9598e · outbound

This paper cites In: 2026 31st Asia and South Pacific Design Automation Conference (ASP-DAC).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 31st Asia and South Pacific Design Automation Conference (ASP-DAC)

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:20e4e7d92c9ffdfa3cdda3b97e4ec193e062e44394b22b97b6ee4c48b2639820

Observation 16a0cd12-b6e8-4f1a-a423-5a1ff96e7e2c · outbound

This paper cites In: 2026 31st Asia and South Pacific Design Automation Conference (ASP-DAC).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 31st Asia and South Pacific Design Automation Conference (ASP-DAC)

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:32ee33f2a166e2de4d85035ce480458c966fe1cd428004366fd22fca2adad3fb

Observation 4b73cca2-eda6-4b8b-812d-6f790403cf6a · outbound

This paper cites IEEE Transactions on Circuits and Systems II: Express Briefs (2025).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference IEEE Transactions on Circuits and Systems II: Express Briefs (2025)

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:1caa95ad96e251449d877424517802502c8afb2c0bbe135d59dfa6f1712461cd

Observation 3957f69e-8ab1-479c-8c89-fbc4191dad97 · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Architec- ture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Architec- ture (HPCA)

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:ffbacd4510793a2b6849161957a27e7a802d36419aa6f56e909753be45855d12

Observation 319349cc-9f56-496a-a0e0-3080260f8e01 · outbound

This paper cites In: Proceedings of the 58th IEEE/ACM International Sympo- sium on Microarchitecture®.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: Proceedings of the 58th IEEE/ACM International Sympo- sium on Microarchitecture®

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:288d86a166ddd925beba0492ab698637a5e808df4c99b0bba411212a958208f8

Observation 10889604-f3bc-4691-9620-bc8bce6011b8 · outbound

This paper cites IEEE Transactions on Computers (2025).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference IEEE Transactions on Computers (2025)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:8a73ec3223e7cb9cb9d89bf4ebe87f421adcbee26155139c24bc1c77cd07df1a

Observation 7bb9adf0-b040-4b11-9a78-7a4907146791 · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Archi- tecture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Archi- tecture (HPCA)

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:0e8b9a12f55e32dd10a7b7b87118241b804c3411e04411d3b963f8fd5c463fd7

Observation ee4ada80-d1c6-4bd5-8520-86dda3b0e0eb · outbound

This paper cites Integrated Circuits and Systems (2024).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Integrated Circuits and Systems (2024)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:e853371195a6544e455e1f560fe905ec6e3e9fc5cec7148bc0c7c56257f08fe4

Observation 76e6d9c9-f885-4dfc-93d7-ba34183d1bcb · outbound

This paper cites In: International Symposium on Advanced Parallel Processing Technologies.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: International Symposium on Advanced Parallel Processing Technologies

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:a12755f0d97b2ba8cc8c213faee61e1713dece861ddb5b5762cfee7d63c68488

Observation 33767e4d-b606-4112-850a-b75b61cfe5c3 · outbound

This paper cites 0: Modeling hierarchical networks and disaggregated systems for large-model training at scale.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference 0: Modeling hierarchical networks and disaggregated systems for large-model training at scale

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:ecc229b3f529a7a003fe30266697f0535cd89e5bfa36829b75dd1cc6b3bab263

Observation 7f0888e5-64e8-47f4-8fb9-37572d4da2ae · outbound

This paper cites In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2026 IEEE International Symposium on High Performance Computer Architecture (HPCA)

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:d93137607f2b35c9eb73e5a272abf08fa576ec764fdf8f177fdcbdf118509d96

Observation 914ddad4-90e7-4f50-9be4-922d31ddf8c5 · outbound

This paper cites In: Proceedings of the 52nd Annual International Symposium on Computer Architecture.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: Proceedings of the 52nd Annual International Symposium on Computer Architecture

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:6a4db58f55d8b2921fcb638271efb9753f3509be9cc68796425cb1b57b2384d1

Observation 1dbbde5a-4c66-40ae-b34d-002223acfc42 · outbound

This paper cites In: Proceedings of the 52nd Annual International Symposium on Computer Architecture.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: Proceedings of the 52nd Annual International Symposium on Computer Architecture

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:8453df95512e6fcb8e5518a1b0ea3911888fa3eee5297f26231a9ec31beb6c4b

Observation 90eb2d0a-6f13-4e59-b25b-7b6fef49c927 · outbound

This paper cites In: 2024ACM/IEEE51stAnnualInternationalSymposiumonComputerArchitecture (ISCA).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference In: 2024ACM/IEEE51stAnnualInternationalSymposiumonComputerArchitecture (ISCA)

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:371910dc0478d63729c2236bac032c077f66621683a7436906dacdcfc56417fe

Observation d90c3a09-fe21-4c5e-ae92-52682c9368e7 · outbound

This paper cites Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations.

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T12:39:49.226631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:00f06bb3d9599c788fc43f8fe56283b85e5e67aa082dfefbd271a814b18bc017

Observation 01d4d931-9d62-48fd-bdf1-357c2976cd3f · outbound

This paper cites IEEE Journal on Emerging and Selected Topics in Circuits and Systems (2025).

MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference IEEE Journal on Emerging and Selected Topics in Circuits and Systems (2025)

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-26T06:26:59.984686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T06:26:59.984686Z digest=sha256:745033021436039456532a86d7d09b594f004dd3ddec6d8549f82a2e3b547b52

Pith citing papers

Observation bf743749-c279-458f-b368-5f66ec08a8f6 · inbound

Fovea: Physical-Implication-Aware Wafer-Scale DSE with Decision-Domain-Guided Cross-Fidelity Refinement cites this paper.

Fovea: Physical-Implication-Aware Wafer-Scale DSE with Decision-Domain-Guided Cross-Fidelity Refinement MOCAP: Wafer-Scale-Chip-Oriented Memory-Orchestrated Chunked Pipelining Framework for Prefill-Only LLM Inference

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:50:53.247085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T21:50:52.423769Z digest=sha256:26bc5aa21eb66b9ab4854e0ef4b904181accf77d935c2d5f1e72451ed2069994