Pith. sign in

Paper Citation Record · LEDGER

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators

As of 17 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.22169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.22169 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:18:02.043090Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bbecee42-bc13-46f1-92d9-93486ec9e986 · outbound

This paper cites Antman: Dynamic scaling on gpu clusters for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Antman: Dynamic scaling on gpu clusters for deep learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:57.168913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:57.168913Z digest=sha256:1579a69953b66048b688d9fc2d3927fb70a0b70cba5afebdfe1a196219e587f6

Observation 1185e513-39cc-4cef-9cdb-7eaecf1325c5 · outbound

This paper cites Whale: Efficient giant model training over heterogeneous gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Whale: Efficient giant model training over heterogeneous gpus,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:57.209280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:57.209280Z digest=sha256:dee67186e07002fd7ce0838147ebc035cdb40a71e74112475ac92b6ebbd51c20

Observation c7e8f312-35ea-4203-b8ea-bf9f81a6929d · outbound

This paper cites Mpmoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mpmoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:09.420985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.265068Z digest=sha256:4525bdb3e2a6b5fdcadb961bd84347f1211e212cfccf044de039a2ee045b69ad

Observation 76aae5cd-6491-4bda-8d2e-1532eca578ce · outbound

This paper cites Redundancy-free high-performance dynamic GNN training with hier- archical pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Redundancy-free high-performance dynamic GNN training with hier- archical pipeline parallelism,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:09.191149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.366815Z digest=sha256:8cad38129359a383a385b0f0a5b581017675d0203127f4aa6031389e6096f026

Observation 11ae06a4-6f9f-4b71-b3a3-30d625fc11a2 · outbound

This paper cites Chimera: An analytical optimizing framework for effective compute-intensive operators fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Chimera: An analytical optimizing framework for effective compute-intensive operators fusion,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.976798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.415899Z digest=sha256:de7bae9d3c66743c1e223be044da53677f31fc656c9fa959e80a9cc5b983231d

Observation 40d23f88-78ec-45c4-a560-a1071ee9986e · outbound

This paper cites Dnnfusion: accelerating deep neural networks execution with advanced operator fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Dnnfusion: accelerating deep neural networks execution with advanced operator fusion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.820890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.516353Z digest=sha256:34602450df5fd18e02f1d9d5a991d8707b6807fe09b49e634691932463d7a807

Observation e259b4ca-baa6-4ae6-8110-70c333860490 · outbound

This paper cites Astitch: enabling a new multi- dimensional optimization space for memory-intensive ML training and inference on modern SIMT architectures,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Astitch: enabling a new multi- dimensional optimization space for memory-intensive ML training and inference on modern SIMT architectures,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.657288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.599031Z digest=sha256:b6160468d5539247ec48f337d80f7a125206758e34f927e874089cde346e28e8

Observation 89b07b5c-2f64-422d-8bca-4e012988ba5c · outbound

This paper cites Nvidia cublas.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cublas

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.480013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.686091Z digest=sha256:6cafa868b4d2451940b2c350b57c271ea9cf797431e4eb1d9d4117792bcc1480

Observation a7463e97-cdba-46a5-8226-b6c0b1eb8fbb · outbound

This paper cites Nvidia cudnn.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cudnn

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.320639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.794966Z digest=sha256:a45326c2cdd548b28edbd6ba43903461d288d4f47468c14bbbad3d9799245ccc

Observation 854fd778-21d7-419f-9b60-aab00d04dcfb · outbound

This paper cites Nvidia cutlass.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cutlass

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.146636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.833125Z digest=sha256:23b6c0e081b22f05f1353159fd404e5765da81af811381c853d2ad54ff3a3109

Observation ede060d9-800d-4524-8bb7-0330a19fc813 · outbound

This paper cites Nvidia tensorrt.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia tensorrt

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.966648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:57.951859Z digest=sha256:b21e990192d5ee8e52cd3260f312f5c640b999f9b12b8cf0e752d13bfee6325f

Observation 42a04fc8-f71d-47df-9713-a4516b44d7fa · outbound

This paper cites Mpipemoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mpipemoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.873222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.032881Z digest=sha256:f63813f2c8f620d0ae08f9c2008d64214d63f96a834732bda2ce9d1b9420615d

Observation 3d5b6d4f-535c-4e2b-b2b8-d076fe0ae1d1 · outbound

This paper cites Ansor: Generating high-performance tensor programs for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Ansor: Generating high-performance tensor programs for deep learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.739342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.151591Z digest=sha256:a1f612c4e2cc55bf596520b63985d1942829c8fc86716f71453abd568964ee6f

Observation 91b5adab-a47a-4d82-8ab8-b8ac17f1de39 · outbound

This paper cites Learning to optimize tensor programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Learning to optimize tensor programs,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.636632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.210833Z digest=sha256:4fadbe89745f88ed2e82d2f4ce3b27f492e8faf4b53d9218c1b88590090c62a2

Observation db245a2a-f1fb-427a-92c9-9fe5077cb374 · outbound

This paper cites Tiramisu: A polyhedral compiler for expressing fast and portable code,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tiramisu: A polyhedral compiler for expressing fast and portable code,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.509586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.354573Z digest=sha256:b9f7de673887781e9240baa0fe4d835f88df720c2ee4ff503ba1e816e05cb422

Observation 14913f6e-2781-4798-af21-dc23ad56e7cf · outbound

This paper cites Astra: Exploiting predictability to optimize deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Astra: Exploiting predictability to optimize deep learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.418515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.407201Z digest=sha256:22f53c533c0da8a4f75405a972f4a8b984d81cb73d1088d8be4a8ff0ddeef281

Observation c94434b6-7232-405f-baa9-59cba381aaac · outbound

This paper cites TASO: optimizing deep learning computation with automatic genera- tion of graph substitutions,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators TASO: optimizing deep learning computation with automatic genera- tion of graph substitutions,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.317022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.533271Z digest=sha256:bf48088955e9061c02ce35fa316e257664397b4d9f403e4b85aa45bb76c8960c

Observation f5de12cf-1279-4361-b77d-755584ea987a · outbound

This paper cites Scalable kernel fusion for memory-bound GPU applications,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Scalable kernel fusion for memory-bound GPU applications,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.197903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.593356Z digest=sha256:75680f253d18f5c2a7c5814d82162ce59988d20c34e0e97d958d6952a921a0c6

Observation d9eef3ca-6525-4084-b69c-59a27eeebd11 · outbound

This paper cites AKG: automatic kernel generation for neural processing units using polyhedral transformations,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators AKG: automatic kernel generation for neural processing units using polyhedral transformations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.097573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.722719Z digest=sha256:9d74a5519d0bc37c299e1df19b890aa304b9e264e64c46fa6213278edfddd7f3

Observation 38be9a61-80d0-4911-a323-514c7c633aba · outbound

This paper cites Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:58.793850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:58.793850Z digest=sha256:868ef054b6485d7d23db2c97934894ff995d0845a4f362eb30274fe06ecda2f1

Observation 457af5e1-b481-43ce-89ec-6851188d1565 · outbound

This paper cites Tensorflow xla.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensorflow xla

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.981353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.865681Z digest=sha256:8d872417f354e47551315ab00ac47460009a0b7efd59c4b001fd375ad6112011

Observation 7afa9ddf-2509-4e84-b5ad-b541e46d58f9 · outbound

This paper cites Tvm: An automated end-to-end optimizing compiler for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tvm: An automated end-to-end optimizing compiler for deep learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.886879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.918888Z digest=sha256:f1cbae75354c8181dcbf5180e12e72d09680584251efbf4ec5d51cad4436e869

Observation 5424f76d-c5aa-42d5-b1b3-97d91c422677 · outbound

This paper cites Attention is all you need,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Attention is all you need,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.779195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:58.968018Z digest=sha256:0a081e3780e26fabead11626ee99f5321d9928fdf8342195287db25ecdbfe173

Observation 84e453b0-3c1d-4654-aae4-9982a97cf419 · outbound

This paper cites Xgboost: A scalable tree boosting system,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Xgboost: A scalable tree boosting system,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.047371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.047371Z digest=sha256:2e712e770c909ffca557f297572602da982d0fe8e8a81e23bf331bd676eb793e

Observation 9e7916d3-35bf-449d-b27a-866054634b0a · outbound

This paper cites Bolt: Bridg- ing the gap between auto-tuners and hardware-native performance,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bolt: Bridg- ing the gap between auto-tuners and hardware-native performance,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.663284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.129464Z digest=sha256:6becb52b046faf91280cfd052a949cf9e28146aafc3fb6da097740325b5b9141

Observation 050e6687-95bb-44a7-9320-6e67ea74da55 · outbound

This paper cites Pytorch: An imperative style, high- performance deep learning library,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Pytorch: An imperative style, high- performance deep learning library,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.542281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.215230Z digest=sha256:b1fc6de0f642b31901c0e6e2229e803cab74b367c218973a4dc087400269fc75

Observation c85a5a73-1353-4154-a185-c9d99004c0dc · outbound

This paper cites Bytetransformer: A high-performance transformer boosted for variable-length inputs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bytetransformer: A high-performance transformer boosted for variable-length inputs,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.399045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.305175Z digest=sha256:089d38715cd5a89466a2e2a865f364034cb30e9cf3c48c7bce766e0cf600f8bc

Observation 52519be3-2c40-4494-9246-47f7201e5034 · outbound

This paper cites Raptor-t: A fused and memory-efficient sparse transformer for long and variable-length sequences.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Raptor-t: A fused and memory-efficient sparse transformer for long and variable-length sequences

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.244303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.403654Z digest=sha256:f67d07ac39c30e42e068462a26fbc3d9ec7f5822e0c62ecef9426893d8314e6d

Observation a8893b13-17ce-4513-b995-54b4b8c4b4f1 · outbound

This paper cites FusionStitching: Boosting Memory Intensive Computations for Deep Learning Workloads.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators FusionStitching: Boosting Memory Intensive Computations for Deep Learning Workloads

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.486067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.486067Z digest=sha256:6046e877506869ccf817c26f04af192380705bd23f8de2f13a195d5108766ace

Observation 80b96ac1-79d9-4651-aad5-108bd532935a · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.091515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.583398Z digest=sha256:11b3d609c2d48d1f367b1fb4f4e17376dca0152ed8483d4d9e74c7380b54d522

Observation 9101e053-b75b-409a-a3e4-5492ca3005d7 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.630141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.630141Z digest=sha256:63adeeda23803cb33bdcc03c126dea01c25063fc4696456e6dfaab2be9022f89

Observation d6c552ae-5e74-433f-915f-bd4584307700 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.732866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.732866Z digest=sha256:96101d0eee03b1544ebc9bb3019f57de3cf30ad929df11dd68f47a67d6d5a325

Observation 7fb5eb6b-639c-4275-b1a8-fa9b663f7e1c · outbound

This paper cites Hidet: Task-mapping programming paradigm for deep learning tensor programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Hidet: Task-mapping programming paradigm for deep learning tensor programs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.913727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.806635Z digest=sha256:aa1a4fe0cc7eb05151cfa25129f1b2bebbfc8d322706ab165ab75d05e5457765

Observation 130c99b9-848a-486b-85f1-cd4af975f668 · outbound

This paper cites Halide: a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Halide: a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.733116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:17:59.979562Z digest=sha256:5400f73a560cad05c1a4d6880eace4c043c9850ab19d06a3743b515e078dbe26

Observation dddcb304-b24e-4fa5-9785-ddf91d1a97cb · outbound

This paper cites Flextensor: An automatic schedule exploration and optimization framework for tensor computation on heterogeneous system,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Flextensor: An automatic schedule exploration and optimization framework for tensor computation on heterogeneous system,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.542898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.054589Z digest=sha256:4db97a1bb30b905d802db4396d6016e3563a1620fa46faaf106c5e26356436eb

Observation ec7a60cc-8591-444a-9d5a-12055042f1e7 · outbound

This paper cites AMOS: enabling automatic mapping for tensor computations on spatial accelerators with hardware abstraction,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators AMOS: enabling automatic mapping for tensor computations on spatial accelerators with hardware abstraction,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.391723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.135867Z digest=sha256:3e21643445ec43d71bd192610aa26324cf19c48a96244ccbb63c84d3244d2253

Observation 30278ef8-a73e-4540-81af-cc31728684d1 · outbound

This paper cites Atomic dataflow based graph-level workload orchestration for scalable DNN accelerators,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Atomic dataflow based graph-level workload orchestration for scalable DNN accelerators,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.211211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.213362Z digest=sha256:716be1a68043a6b2b8242f734acb6edd3ebbe19e0d6abf0f4b9fce172095d3b6

Observation 27a2a5c7-9b64-443a-89ee-aabfae0d6714 · outbound

This paper cites ROLLER: fast and efficient tensor compilation for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators ROLLER: fast and efficient tensor compilation for deep learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.031888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.316360Z digest=sha256:3ddec2fadd1a7cc3f335fc9be158a93e16c529361c64932121654f06af79b3e2

Observation 2fdbc928-81b9-4ac9-9368-a2d7a8301d09 · outbound

This paper cites Triton: an intermediate language and compiler for tiled neural network computations,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Triton: an intermediate language and compiler for tiled neural network computations,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.848685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.357498Z digest=sha256:78543742212e38e496d29dc17fe6e3c1a6734c999d9e0113ad7ff7f0e9184ad0

Observation 08894df8-1172-4a27-b829-9d08d8d64032 · outbound

This paper cites Nvidia, parallel thread execution isa.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia, parallel thread execution isa

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.710568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.444478Z digest=sha256:a93e0a4b60c3385f2a537dd894a9700484ce38dbc4bebed406d41885c0b4504a

Observation 62cbf713-64c8-4336-8933-2d8a24b8989a · outbound

This paper cites Tensorflow: A system for large-scale machine learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensorflow: A system for large-scale machine learning,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.524401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.502303Z digest=sha256:809dadba1d49abed089dcc50458fa096692937fd15240f8691f2fe5205aba06c

Observation c63e8000-3b1c-4edf-a171-006af5dcddbb · outbound

This paper cites Available: https://onnx.ai/.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Available: https://onnx.ai/

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.413135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.572160Z digest=sha256:99077c83273455e85b6f0eed92e2a116757b4e87432d7c2a2c26591965f47bab

Observation 39f3acaf-587b-43bb-8c80-425fa8f42d4d · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:00.661573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:00.661573Z digest=sha256:5cd11432f4e3da601c03813a33aba439a049f16616fac30ded3e894bcde0ee7a

Observation efe2bd7b-2c42-4c2c-9ac6-138960b9a649 · outbound

This paper cites An image is worth 16x16 words: Trans- formers for image recognition at scale,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators An image is worth 16x16 words: Trans- formers for image recognition at scale,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.281331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.750686Z digest=sha256:14096267bff079d59040c9611176aa542c6ababdd76d4a46922575fbbc34a54f

Observation 79551829-ecc3-4148-b775-011159bdbf85 · outbound

This paper cites Mlp-mixer: An all-mlp architecture for vision,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mlp-mixer: An all-mlp architecture for vision,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.147790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:00.840062Z digest=sha256:4c23beb73422a403d21129549638eb510c08c2eede3dc76a97575ce637fb8933

Observation 7c05d49d-2121-4510-ae70-e5ae4570359f · outbound

This paper cites Relay: A High-Level Compiler for Deep Learning.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Relay: A High-Level Compiler for Deep Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:00.935149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:00.935149Z digest=sha256:807283e65342e06ad6e7226cf0beeead8b3913752bfb82e6a17f424546668c9d

Observation 0e018dfd-7f5f-40ca-bc6f-70790c7bd1ce · outbound

This paper cites Intel oneapi deep neural network library.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Intel oneapi deep neural network library

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.968816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.019537Z digest=sha256:f723f3055e0dfa19255c540a190750616c3ee9f129217a9bb069329318b19084

Observation 77abbbf7-6fe3-4324-be58-7417b2784e14 · outbound

This paper cites Intel oneapi math kernel library.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Intel oneapi math kernel library

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.782986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.121244Z digest=sha256:743f9155e40c95041e44bd7b696567346410476c159b76ab18684d24b4eea2a8

Observation 39b837f5-2f97-48a1-b8fd-e28022d2d695 · outbound

This paper cites Rammer: Enabling holistic deep learning compiler optimizations with rtasks,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Rammer: Enabling holistic deep learning compiler optimizations with rtasks,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.657856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.198588Z digest=sha256:ea9e76b15a1265dec84cb1f2ebccc38c2ddda1f243d9e3d66f4ec7376d718175

Observation 972a4a4e-d287-4515-8b4d-97565b441020 · outbound

This paper cites Accelerating deep learning inference with cross-layer data reuse on gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Accelerating deep learning inference with cross-layer data reuse on gpus,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.481663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.301835Z digest=sha256:4b173d2e62fa70b42754a3ccc0b9c66853db9bdf68c69108e1bd6acb580db6e5

Observation 27e8c3fb-6275-463b-aad3-ba5f8a7e2c85 · outbound

This paper cites Efficient GPU spatial-temporal multitasking,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Efficient GPU spatial-temporal multitasking,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.325638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.414094Z digest=sha256:c6cc55c98b349cc58e49d17da3218fa1798348b46d698acc448d075eee75d55c

Observation 80618246-eca4-47aa-8edc-703189e895f5 · outbound

This paper cites On optimizing machine learning workloads via kernel fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators On optimizing machine learning workloads via kernel fusion,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.186137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.537926Z digest=sha256:d596a5ca55ae063f1028f620fefecb48c56f266b5ecd0cf84b4036d76081a7a0

Observation fcdd5e74-8860-4a8e-a15c-3728e6ee456e · outbound

This paper cites Learning to optimize halide with tree search and random programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Learning to optimize halide with tree search and random programs,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.032273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.643595Z digest=sha256:107cb2489fc08c4428662c296920e7b97c1411fb6bfc40a317448dae1d912232

Observation 3acc49c2-e698-4f4d-a533-d67a3ee08957 · outbound

This paper cites Nimble: Efficiently compiling dynamic neural networks for model inference,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nimble: Efficiently compiling dynamic neural networks for model inference,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.874308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.732170Z digest=sha256:2531d12f5769ad429605a7602b5cab28abd44808817f4867456a7078848c3b83

Observation 40723419-b400-4c46-8dad-ebf01ffe234b · outbound

This paper cites DISC: A dynamic shape compiler for machine learning workloads,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators DISC: A dynamic shape compiler for machine learning workloads,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.743768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.793959Z digest=sha256:453f4ddf0116bfe99924a4cd9a30e3f786c8b510ea399bd7b43cdbf97f90e357

Observation f6faa016-543a-48fc-a467-ed5720274a01 · outbound

This paper cites Automatic generation of high-performance quantized machine learning kernels,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Automatic generation of high-performance quantized machine learning kernels,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.607381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.864329Z digest=sha256:802ce7f49c521f61942f7a4859d6a96ea58aa387921556c9d8cac08854cf3422

Observation 5827d460-ad2e-4dfb-8d17-8b7fb5245746 · outbound

This paper cites A code generator for high-performance tensor contractions on gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators A code generator for high-performance tensor contractions on gpus,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.442784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:01.936505Z digest=sha256:80395e14aa0cbd0604ee686c69a10352d3bec5fc1430a1ef22cb4373032a8154

Observation 671f0ec3-4e56-4dc6-9149-51edd5237477 · outbound

This paper cites Neoflow: A flexible framework for enabling efficient compilation for high performance DNN training,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Neoflow: A flexible framework for enabling efficient compilation for high performance DNN training,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.297777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T22:18:02.043090Z digest=sha256:da1fe350170d68332537ef19729fe4f0ae9db11af0a82313ff6a984ab6997e3b

Pith citing papers

No inbound Pith citation observations are available.