Pith. sign in

Paper Citation Record · LEDGER

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators

As of 13 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.22169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.22169 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:18:02.043090Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bbecee42-bc13-46f1-92d9-93486ec9e986 · outbound

This paper cites Antman: Dynamic scaling on gpu clusters for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Antman: Dynamic scaling on gpu clusters for deep learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:57.168913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:57.168913Z digest=sha256:1579a69953b66048b688d9fc2d3927fb70a0b70cba5afebdfe1a196219e587f6

Observation 1185e513-39cc-4cef-9cdb-7eaecf1325c5 · outbound

This paper cites Whale: Efficient giant model training over heterogeneous gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Whale: Efficient giant model training over heterogeneous gpus,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:57.209280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:57.209280Z digest=sha256:dee67186e07002fd7ce0838147ebc035cdb40a71e74112475ac92b6ebbd51c20

Observation c7e8f312-35ea-4203-b8ea-bf9f81a6929d · outbound

This paper cites Mpmoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mpmoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:09.420985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.265068Z digest=sha256:96f78cbc4dfbb51c902e1f6395e3a4ae335db881644562a0f8c937a405840b2b

Observation 76aae5cd-6491-4bda-8d2e-1532eca578ce · outbound

This paper cites Redundancy-free high-performance dynamic GNN training with hier- archical pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Redundancy-free high-performance dynamic GNN training with hier- archical pipeline parallelism,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:09.191149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.366815Z digest=sha256:b066334796ef6d01856f1256f8e59702c42b5458019ea466ed90985d1ac92987

Observation 11ae06a4-6f9f-4b71-b3a3-30d625fc11a2 · outbound

This paper cites Chimera: An analytical optimizing framework for effective compute-intensive operators fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Chimera: An analytical optimizing framework for effective compute-intensive operators fusion,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.976798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.415899Z digest=sha256:f202ad5a252a45cd96d02846fb8c9dc18c8111d634ac123df8a034983ec83b02

Observation 40d23f88-78ec-45c4-a560-a1071ee9986e · outbound

This paper cites Dnnfusion: accelerating deep neural networks execution with advanced operator fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Dnnfusion: accelerating deep neural networks execution with advanced operator fusion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.820890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.516353Z digest=sha256:7b40efecb4d7833415a1ef85b40d72d00418b9a94a4a7228a7aa6466febd2ec2

Observation e259b4ca-baa6-4ae6-8110-70c333860490 · outbound

This paper cites Astitch: enabling a new multi- dimensional optimization space for memory-intensive ML training and inference on modern SIMT architectures,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Astitch: enabling a new multi- dimensional optimization space for memory-intensive ML training and inference on modern SIMT architectures,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.657288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.599031Z digest=sha256:91204625c8f4fa175e18a9aae15af294c0430bd4467cda98cf513f559eac296c

Observation 89b07b5c-2f64-422d-8bca-4e012988ba5c · outbound

This paper cites Nvidia cublas.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cublas

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.480013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.686091Z digest=sha256:ea86690fa0250520545a6bad27aac7a5647ba240094582bc3a901e769c7feb42

Observation a7463e97-cdba-46a5-8226-b6c0b1eb8fbb · outbound

This paper cites Nvidia cudnn.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cudnn

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.320639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.794966Z digest=sha256:3cf4d88efcd910ff0010d9fb9bcecbc408e3a7a54e35c9ade7a096044278a73c

Observation 854fd778-21d7-419f-9b60-aab00d04dcfb · outbound

This paper cites Nvidia cutlass.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia cutlass

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:08.146636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.833125Z digest=sha256:cf50d2b57254c648213bf3fde075680f776bb4a8e4d4028180081e2fec5297d3

Observation ede060d9-800d-4524-8bb7-0330a19fc813 · outbound

This paper cites Nvidia tensorrt.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia tensorrt

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.966648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:57.951859Z digest=sha256:fd338c1c85be54aef0b46d2ce442f23a46dd6ef158b236578acbd37a8a074a46

Observation 42a04fc8-f71d-47df-9713-a4516b44d7fa · outbound

This paper cites Mpipemoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mpipemoe: Memory efficient moe for pre-trained models with adaptive pipeline parallelism,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.873222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.032881Z digest=sha256:dd8be7e290cd98744e4d0c705c031b81d5d66f61f04a35470b3651ea9aaa1c80

Observation 3d5b6d4f-535c-4e2b-b2b8-d076fe0ae1d1 · outbound

This paper cites Ansor: Generating high-performance tensor programs for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Ansor: Generating high-performance tensor programs for deep learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.739342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.151591Z digest=sha256:e3b45d1b7347b81840eec36d1c2931087d4f6c27aa4f456e11adcc4008509f08

Observation 91b5adab-a47a-4d82-8ab8-b8ac17f1de39 · outbound

This paper cites Learning to optimize tensor programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Learning to optimize tensor programs,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.636632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.210833Z digest=sha256:ab379ee8bc4cb5442d28f4eefff6086ffa96a461559e15a3a518e16dc2998c13

Observation db245a2a-f1fb-427a-92c9-9fe5077cb374 · outbound

This paper cites Tiramisu: A polyhedral compiler for expressing fast and portable code,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tiramisu: A polyhedral compiler for expressing fast and portable code,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.509586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.354573Z digest=sha256:b232cadecbffb04594e91f4f868408869288a00d3de0de48143377ad6385285a

Observation 14913f6e-2781-4798-af21-dc23ad56e7cf · outbound

This paper cites Astra: Exploiting predictability to optimize deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Astra: Exploiting predictability to optimize deep learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.418515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.407201Z digest=sha256:712cbc8a3999e209a87867b100990789a069adf3e4fa9c7862608e3697cc98eb

Observation c94434b6-7232-405f-baa9-59cba381aaac · outbound

This paper cites TASO: optimizing deep learning computation with automatic genera- tion of graph substitutions,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators TASO: optimizing deep learning computation with automatic genera- tion of graph substitutions,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.317022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.533271Z digest=sha256:52b02d0eb64c33e270b210eef0612bc714a27cbd0b9d4d76de69ab3b4f37e8d9

Observation f5de12cf-1279-4361-b77d-755584ea987a · outbound

This paper cites Scalable kernel fusion for memory-bound GPU applications,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Scalable kernel fusion for memory-bound GPU applications,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.197903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.593356Z digest=sha256:be0f30ac099d57f50220ec45b82fa1ccd5d0649bdd0c5497413f8d75f6b7b5c3

Observation d9eef3ca-6525-4084-b69c-59a27eeebd11 · outbound

This paper cites AKG: automatic kernel generation for neural processing units using polyhedral transformations,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators AKG: automatic kernel generation for neural processing units using polyhedral transformations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:07.097573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.722719Z digest=sha256:f4ff5ce148e1e6a8d976c47a02492a40495c01c10890721dca6a659ae19e4f8b

Observation 38be9a61-80d0-4911-a323-514c7c633aba · outbound

This paper cites Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensor Comprehensions: Framework-Agnostic High-Performance Machine Learning Abstractions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:58.793850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:58.793850Z digest=sha256:241349ca1416d0bdadc2e90ad31ba50c05c35bc7f25da7aa5e1dec7e3cc1767e

Observation 457af5e1-b481-43ce-89ec-6851188d1565 · outbound

This paper cites Tensorflow xla.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensorflow xla

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.981353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.865681Z digest=sha256:8d108f724859a748ae715872b1c66f206ae7e56a4008abfc9ec889b96191ea58

Observation 7afa9ddf-2509-4e84-b5ad-b541e46d58f9 · outbound

This paper cites Tvm: An automated end-to-end optimizing compiler for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tvm: An automated end-to-end optimizing compiler for deep learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.886879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.918888Z digest=sha256:eeff517907e76669ee8301bf5678a169ae6ffdd799bbb002836b2b1c52a6d1ad

Observation 5424f76d-c5aa-42d5-b1b3-97d91c422677 · outbound

This paper cites Attention is all you need,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Attention is all you need,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.779195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:58.968018Z digest=sha256:004ed123fda594ddbbb1f8044eed56c2802cfb76d15e038380ebf1b70bd8c1bb

Observation 84e453b0-3c1d-4654-aae4-9982a97cf419 · outbound

This paper cites Xgboost: A scalable tree boosting system,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Xgboost: A scalable tree boosting system,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.047371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.047371Z digest=sha256:2e712e770c909ffca557f297572602da982d0fe8e8a81e23bf331bd676eb793e

Observation 9e7916d3-35bf-449d-b27a-866054634b0a · outbound

This paper cites Bolt: Bridg- ing the gap between auto-tuners and hardware-native performance,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bolt: Bridg- ing the gap between auto-tuners and hardware-native performance,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.663284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.129464Z digest=sha256:3c7f3592c0280917569dfd47654fa114ca98ab745d936875575b337f99c277e1

Observation 050e6687-95bb-44a7-9320-6e67ea74da55 · outbound

This paper cites Pytorch: An imperative style, high- performance deep learning library,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Pytorch: An imperative style, high- performance deep learning library,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.542281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.215230Z digest=sha256:005d3c4dc480325eee808f792c85facab08234190843c704d8968b63aa430aa1

Observation c85a5a73-1353-4154-a185-c9d99004c0dc · outbound

This paper cites Bytetransformer: A high-performance transformer boosted for variable-length inputs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bytetransformer: A high-performance transformer boosted for variable-length inputs,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.399045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.305175Z digest=sha256:5ad3a3e677c033b1814b287aa37345d6d3163565251bca1f1d30aad08ad135cd

Observation 52519be3-2c40-4494-9246-47f7201e5034 · outbound

This paper cites Raptor-t: A fused and memory-efficient sparse transformer for long and variable-length sequences.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Raptor-t: A fused and memory-efficient sparse transformer for long and variable-length sequences

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.244303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.403654Z digest=sha256:b9f65d776c6531364bf56c99f3be6df16bf26cd54ebe17e351494781c8e6687a

Observation a8893b13-17ce-4513-b995-54b4b8c4b4f1 · outbound

This paper cites FusionStitching: Boosting Memory Intensive Computations for Deep Learning Workloads.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators FusionStitching: Boosting Memory Intensive Computations for Deep Learning Workloads

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.486067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.486067Z digest=sha256:7575678dace60e7e8bacdf5ece6e9d08415e8dcc74d44dd7814bcf1ec0b79f35

Observation 80b96ac1-79d9-4651-aad5-108bd532935a · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Flashattention: Fast and memory-efficient exact attention with io-awareness,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:06.091515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.583398Z digest=sha256:5c98ad61c3b07a28eb9c023d362af5b975024b6f2543b6e013687737fa6b6cb3

Observation 9101e053-b75b-409a-a3e4-5492ca3005d7 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.630141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.630141Z digest=sha256:63adeeda23803cb33bdcc03c126dea01c25063fc4696456e6dfaab2be9022f89

Observation d6c552ae-5e74-433f-915f-bd4584307700 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:17:59.732866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:17:59.732866Z digest=sha256:96101d0eee03b1544ebc9bb3019f57de3cf30ad929df11dd68f47a67d6d5a325

Observation 7fb5eb6b-639c-4275-b1a8-fa9b663f7e1c · outbound

This paper cites Hidet: Task-mapping programming paradigm for deep learning tensor programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Hidet: Task-mapping programming paradigm for deep learning tensor programs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.913727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.806635Z digest=sha256:e4be6d25b4463fcf55e8b48c8032131b127496db0430e33d1a43bc4b5b07d6f1

Observation 130c99b9-848a-486b-85f1-cd4af975f668 · outbound

This paper cites Halide: a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Halide: a language and compiler for optimizing parallelism, locality, and recomputation in image processing pipelines,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.733116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:17:59.979562Z digest=sha256:a3c8b24ef4aa8a70914e45031e5a92dd3f23594dcdb59db088daeaebdbb41820

Observation dddcb304-b24e-4fa5-9785-ddf91d1a97cb · outbound

This paper cites Flextensor: An automatic schedule exploration and optimization framework for tensor computation on heterogeneous system,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Flextensor: An automatic schedule exploration and optimization framework for tensor computation on heterogeneous system,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.542898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.054589Z digest=sha256:31e334bc49fd9bcd2e34178417bc46dfd7ac6f01842294b9777eb0d535e6038a

Observation ec7a60cc-8591-444a-9d5a-12055042f1e7 · outbound

This paper cites AMOS: enabling automatic mapping for tensor computations on spatial accelerators with hardware abstraction,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators AMOS: enabling automatic mapping for tensor computations on spatial accelerators with hardware abstraction,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.391723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.135867Z digest=sha256:b0bd764190e26601b06ad45a6f9c976efb8c1dd9a224a6502124abcca82f19d0

Observation 30278ef8-a73e-4540-81af-cc31728684d1 · outbound

This paper cites Atomic dataflow based graph-level workload orchestration for scalable DNN accelerators,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Atomic dataflow based graph-level workload orchestration for scalable DNN accelerators,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.211211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.213362Z digest=sha256:60a7b5d379e518b85594a7dbec6ab1408f156b27e44b915a052939a5f73defe8

Observation 27a2a5c7-9b64-443a-89ee-aabfae0d6714 · outbound

This paper cites ROLLER: fast and efficient tensor compilation for deep learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators ROLLER: fast and efficient tensor compilation for deep learning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:05.031888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.316360Z digest=sha256:dbf9c6d7a78685c7fff4303f3b1dc71ef117f729379e2a829a7df3eae6324237

Observation 2fdbc928-81b9-4ac9-9368-a2d7a8301d09 · outbound

This paper cites Triton: an intermediate language and compiler for tiled neural network computations,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Triton: an intermediate language and compiler for tiled neural network computations,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.848685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.357498Z digest=sha256:6b5a745f1393f035389821d2a1f93e9abb30f7f543fffc1b9120695f1b393ed8

Observation 08894df8-1172-4a27-b829-9d08d8d64032 · outbound

This paper cites Nvidia, parallel thread execution isa.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nvidia, parallel thread execution isa

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.710568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.444478Z digest=sha256:5f3d85d7a0e21ed9d7aba6a5510e38a9db9dba5c1f3c6df148531954a66b83e7

Observation 62cbf713-64c8-4336-8933-2d8a24b8989a · outbound

This paper cites Tensorflow: A system for large-scale machine learning,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Tensorflow: A system for large-scale machine learning,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.524401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.502303Z digest=sha256:397ec8975784c3456c62da52b08e431840e5a781ee607e2aa80d01541d9785f9

Observation c63e8000-3b1c-4edf-a171-006af5dcddbb · outbound

This paper cites Available: https://onnx.ai/.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Available: https://onnx.ai/

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.413135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.572160Z digest=sha256:935fca1616925d5ab158ec7293de40a3559a491e475d15836b3b19b3e841339c

Observation 39f3acaf-587b-43bb-8c80-425fa8f42d4d · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:00.661573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:00.661573Z digest=sha256:5cd11432f4e3da601c03813a33aba439a049f16616fac30ded3e894bcde0ee7a

Observation efe2bd7b-2c42-4c2c-9ac6-138960b9a649 · outbound

This paper cites An image is worth 16x16 words: Trans- formers for image recognition at scale,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators An image is worth 16x16 words: Trans- formers for image recognition at scale,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.281331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.750686Z digest=sha256:5111da67c1b38b7df9373f674dcdf663c62557656ef472f55c633c45fb13114e

Observation 79551829-ecc3-4148-b775-011159bdbf85 · outbound

This paper cites Mlp-mixer: An all-mlp architecture for vision,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Mlp-mixer: An all-mlp architecture for vision,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:04.147790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:00.840062Z digest=sha256:3cd1f7d860b9b12273391ceb140c45f641e42f76efa895a2ba258b8987b39897

Observation 7c05d49d-2121-4510-ae70-e5ae4570359f · outbound

This paper cites Relay: A High-Level Compiler for Deep Learning.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Relay: A High-Level Compiler for Deep Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:00.935149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:00.935149Z digest=sha256:7575ada18d964f5bc165a796d570332a8f3e65873cc33558f4257a9ac582f09e

Observation 0e018dfd-7f5f-40ca-bc6f-70790c7bd1ce · outbound

This paper cites Intel oneapi deep neural network library.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Intel oneapi deep neural network library

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.968816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.019537Z digest=sha256:5cc137a87894cc8726b1229c7f4dcfb1ecdb24ecd6a209260b40f6084ab7a46e

Observation 77abbbf7-6fe3-4324-be58-7417b2784e14 · outbound

This paper cites Intel oneapi math kernel library.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Intel oneapi math kernel library

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.782986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.121244Z digest=sha256:63391e6a8fe64c61ae117ca818555c09976781be9ff51171545b6c7aea644797

Observation 39b837f5-2f97-48a1-b8fd-e28022d2d695 · outbound

This paper cites Rammer: Enabling holistic deep learning compiler optimizations with rtasks,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Rammer: Enabling holistic deep learning compiler optimizations with rtasks,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.657856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.198588Z digest=sha256:2ff065f3788b7750732d0a75c96f7dfbf2a06cd12b01a9445842f5094bf9376f

Observation 972a4a4e-d287-4515-8b4d-97565b441020 · outbound

This paper cites Accelerating deep learning inference with cross-layer data reuse on gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Accelerating deep learning inference with cross-layer data reuse on gpus,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.481663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.301835Z digest=sha256:155992732b31350ec948b944514111fddb670c7ad59d081a85f27925f11d8e41

Observation 27e8c3fb-6275-463b-aad3-ba5f8a7e2c85 · outbound

This paper cites Efficient GPU spatial-temporal multitasking,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Efficient GPU spatial-temporal multitasking,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.325638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.414094Z digest=sha256:ef396edd6f40b3aa51f5b768248e1ca1bca9ed67dc404f4e61f5afedf4ee8cfa

Observation 80618246-eca4-47aa-8edc-703189e895f5 · outbound

This paper cites On optimizing machine learning workloads via kernel fusion,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators On optimizing machine learning workloads via kernel fusion,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.186137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.537926Z digest=sha256:c5d6a9cb38db350e17f0d3cb9712d55506a979c28d56f107cc2c057e2e6e56b9

Observation fcdd5e74-8860-4a8e-a15c-3728e6ee456e · outbound

This paper cites Learning to optimize halide with tree search and random programs,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Learning to optimize halide with tree search and random programs,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:03.032273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.643595Z digest=sha256:aafe105585c07db43fae3888f353bb5d5e178f35981c39dcd3d407db0e972898

Observation 3acc49c2-e698-4f4d-a533-d67a3ee08957 · outbound

This paper cites Nimble: Efficiently compiling dynamic neural networks for model inference,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Nimble: Efficiently compiling dynamic neural networks for model inference,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.874308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.732170Z digest=sha256:aab383ebf6fbb81c45857bec78cbb4984d47e92423f14af4a4fd6a9cae7a65bf

Observation 40723419-b400-4c46-8dad-ebf01ffe234b · outbound

This paper cites DISC: A dynamic shape compiler for machine learning workloads,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators DISC: A dynamic shape compiler for machine learning workloads,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.743768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.793959Z digest=sha256:022a1aa70c0280cd0df732166a9b3ee2f2d4b80037605eb81aebc58d093e91dc

Observation f6faa016-543a-48fc-a467-ed5720274a01 · outbound

This paper cites Automatic generation of high-performance quantized machine learning kernels,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Automatic generation of high-performance quantized machine learning kernels,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.607381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.864329Z digest=sha256:502b0755bb1ca54793be394fb5f989337a7bd9edb9463010d27cb9c0d9480009

Observation 5827d460-ad2e-4dfb-8d17-8b7fb5245746 · outbound

This paper cites A code generator for high-performance tensor contractions on gpus,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators A code generator for high-performance tensor contractions on gpus,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.442784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:01.936505Z digest=sha256:ff519ae4fda691a7a26fc6214b6c13191b166003d2d18429a52b5baa47f6cbb2

Observation 671f0ec3-4e56-4dc6-9149-51edd5237477 · outbound

This paper cites Neoflow: A flexible framework for enabling efficient compilation for high performance DNN training,.

MCFuser: High-Performance and Rapid Fusion of Memory-Bound Compute-Intensive Operators Neoflow: A flexible framework for enabling efficient compilation for high performance DNN training,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:18:02.297777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T22:18:02.043090Z digest=sha256:a4304fd3a08789e20195ee1b1af810f55efbbe0e78ebc80024bec52dfe9c39a6

Pith citing papers

No inbound Pith citation observations are available.