Pith. sign in

Paper Citation Record · LEDGER

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage

As of 10 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.06472.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06472 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:33.427109Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2e46f6e-6bdf-45b9-938b-f02312d76d6c · outbound

This paper cites In19th USENIX Conference on File and Storage Technologies (FAST 21), pages 387–401, 2021.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage In19th USENIX Conference on File and Storage Technologies (FAST 21), pages 387–401, 2021

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:38.224953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.204819Z digest=sha256:dfdc4c1c60015150cc543c6791a750c0404e3b76ae6cbdc66812332ada1e1fa4

Observation 8b9ac2a1-8192-4455-a34a-31ae374be6b4 · outbound

This paper cites Efficient combination of rematerialization and offloading for training dnns.Advances in Neural Information Processing Systems, 34:23844–23857, 2021.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Efficient combination of rematerialization and offloading for training dnns.Advances in Neural Information Processing Systems, 34:23844–23857, 2021

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:38.076966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.210913Z digest=sha256:14cd05c6ab8aecac30f0e5758616d18231cb6e9f08f6a967227a72a193bc4dc6

Observation 75a2931f-a17e-40b1-a1bb-47b4253cdbf5 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.219076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.219076Z digest=sha256:05781c9cc025a9532e64250570a926f1930b4c7efa153e066e7c1a23dbd2abbe

Observation 4010036c-c454-4239-ae2b-b2339ce45d89 · outbound

This paper cites The Llama 3 Herd of Models.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.223943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.223943Z digest=sha256:3c7e21b18e1f7dada190e0627d580bd0541691314998121116356c8754dbd6a2

Observation 157e3f5e-29db-458b-8529-bd6344290d2c · outbound

This paper cites Exxact.https://www.exxactcorp.com/, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Exxact.https://www.exxactcorp.com/, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.885091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.229730Z digest=sha256:bb8b3a0731a69b0cf1e65bdb198616e59880bf957f656d9afc6e44307170535c

Observation 79269ad2-b570-4a8c-a116-40c88095f66c · outbound

This paper cites T5 11b.https://huggingface.co/google-t5/t5-11b, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage T5 11b.https://huggingface.co/google-t5/t5-11b, 2025

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.668737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.234022Z digest=sha256:fd8f864cf0bcac200953ad3dd1357058fbfe80b95f6940a541061baa0f8d7c71

Observation 0220b629-eb06-4ca6-af42-eb96375c0285 · outbound

This paper cites nvidia.com/blog/gpudirect-storage/.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage nvidia.com/blog/gpudirect-storage/

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.498608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.244791Z digest=sha256:503b4f91d7971753b3136d0af09f2724129493f8031335848372b6e45e751f81

Observation fa229a1c-7736-48aa-901c-63497343213c · outbound

This paper cites Swapadvisor: Pushingdeeplearningbeyondthegpu memory limit via smart swapping.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Swapadvisor: Pushingdeeplearningbeyondthegpu memory limit via smart swapping

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.315978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.249225Z digest=sha256:694fd0615677cf960ead4241d1d113614b9529dece660bab9046b6ea1daf4b41

Observation 565f4070-ef28-41fb-a56a-b5144f07d578 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism.Advances in neural information processing systems, 32, 2019.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Gpipe: Efficient training of giant neural networks using pipeline parallelism.Advances in neural information processing systems, 32, 2019

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.255025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.255025Z digest=sha256:d0f8be0715b02849a4789490a36216ea9ac6dce052dd155c92760d4933640909

Observation 90eeb0ea-75ae-43c7-9042-18b66e18393f · outbound

This paper cites Ibm granite.https://huggingface.co/ibm-granite, 2025.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Ibm granite.https://huggingface.co/ibm-granite, 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:37.132856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.262423Z digest=sha256:0a53ea567e67fcfa7c770f368f4d67bde2dc99550a3be37b48561661a4b76495

Observation 05d469d2-ce6e-4e5c-8cd2-08bfb9f65c09 · outbound

This paper cites Checkmate: Breaking the memory wall with optimal tensor rematerialization.Proceedings of Machine Learning and Systems, 2:497–511, 2020.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Checkmate: Breaking the memory wall with optimal tensor rematerialization.Proceedings of Machine Learning and Systems, 2:497–511, 2020

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.954625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.271033Z digest=sha256:0ff458601f17f815c0529dcae1439e8b81c378baff35fecd56a107e9fbf3abf6

Observation 66b4a927-a40b-4dd6-84cc-2fc2347ddce8 · outbound

This paper cites Smart-infinity: Fastlargelanguagemodeltrainingusingnear-storageprocessingonarealsystem.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Smart-infinity: Fastlargelanguagemodeltrainingusingnear-storageprocessingonarealsystem

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.734315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.281042Z digest=sha256:91834f897fb3b63913fdb973fbcbd07bbcf4a87d22cfa25fc1b9bdd066a2ecb6

Observation 62e93b52-97ee-480b-9a82-f474a889d2ba · outbound

This paper cites Deepum: Tensormigrationandprefetchinginunified memory.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Deepum: Tensormigrationandprefetchinginunified memory

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.517578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.291031Z digest=sha256:c3f525d17fb9c59d14c30e7bc3214e181636ebcb77996aa39fe69ac39e18f2c2

Observation 9e4d01fd-f963-4c16-aac4-afdd38782399 · outbound

This paper cites Scaling Laws for Neural Language Models.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Scaling Laws for Neural Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.301738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.301738Z digest=sha256:d901ff7df4e031a30e1c826eba550a3e06a89f3e44fd04a26f70b92fe2be4b27

Observation 1ee83164-88a9-44c8-8c30-d5f8387bc992 · outbound

This paper cites Beyond the memory wall: A case for memory-centric hpc system for deep learning.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Beyond the memory wall: A case for memory-centric hpc system for deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.317589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.312924Z digest=sha256:a11e336aaac9e69c9e0b7d77b58e6cd368a352cbfdeb58b7e13a1f8975b91145

Observation 25d743ac-8fb4-4732-845d-9763d4d96a9a · outbound

This paper cites TFLMS: Large Model Support in TensorFlow by Graph Rewriting.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage TFLMS: Large Model Support in TensorFlow by Graph Rewriting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.323175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.323175Z digest=sha256:0ee3412d9b28aa359d0d004c2fa6207d886397613ed77432cc5015a9eab2dc94

Observation 6c2374d8-4fe3-4c22-b633-27febef3d874 · outbound

This paper cites TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.332267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.332267Z digest=sha256:ae785bbeca5f5c86d904f52d6c64d8dc6e1603a2721cfd8309dd866620543d82

Observation 9347a12c-dc38-4126-87ec-492546f5dec0 · outbound

This paper cites LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.342788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.342788Z digest=sha256:85c8cf006b3d76f2817aa7b5f6fdce0856e3849dc972867bc167576be007502e

Observation 93c3f85c-3e80-4487-aace-506d2cf81670 · outbound

This paper cites An Empirical Model of Large-Batch Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage An Empirical Model of Large-Batch Training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.354308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.354308Z digest=sha256:c39439ec8c5ff624f7d71e279a2f809c924ae1421b5e0c2fd8832eba1650d1c6

Observation 6b5ea083-c2da-4c5a-a012-ecf0710c1997 · outbound

This paper cites Mixed Precision Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Mixed Precision Training

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.365297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.365297Z digest=sha256:b3795c9e0bfbd59a69f2ce539c06126824be50c3d9269cbfc08270de0a7b748f

Observation c7725067-0716-481a-ba8e-9b15d1ccdbdd · outbound

This paper cites Pipedream: Generalized pipeline parallelism for dnn training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Pipedream: Generalized pipeline parallelism for dnn training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.377846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.377846Z digest=sha256:e856aa3946828876610347c1fea55b33a609f11315900f2e5077d50784915e4b

Observation 4244f784-58d3-4f9b-b7ce-5b136bfbc231 · outbound

This paper cites Angel-PTM: A Scalable and Economical Large-scale Pre-training System in Tencent.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Angel-PTM: A Scalable and Economical Large-scale Pre-training System in Tencent

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.388782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.388782Z digest=sha256:6058b06f196f4132f5a7b83a6f9ec8aa59b922c0d30b9cff42b0626b23f3f376

Observation e21e5ca9-0316-495e-81fd-55617bd9174e · outbound

This paper cites Automatic differentiation in pytorch.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Automatic differentiation in pytorch

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.401389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.401389Z digest=sha256:9f5662f2bd7ff602f364d762d1697604d47235e38856a81564e825b2eb212b15

Observation 0f628802-5c38-4ba1-8595-23609fd97d9c · outbound

This paper cites Scalable diffusion models with transformers.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Scalable diffusion models with transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.412497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.412497Z digest=sha256:5762eea667233f4eb0ef0fe5d27c6a77ab69e26473bec303934f274018aaa88a

Observation 14a0829f-bdb2-48d5-916a-ca38b86e1ebd · outbound

This paper cites Training Large Neural Networks with Constant Memory using a New Execution Algorithm.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Training Large Neural Networks with Constant Memory using a New Execution Algorithm

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.422749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.422749Z digest=sha256:4607fac072d0c036ebe4915bffb06e857ef6aef74d743009e9c7dd2b6134afe3

Observation 5047daf3-6af8-4ed3-8baa-742822822aaf · outbound

This paper cites Gpu-initiated on-demand high-throughput storage access in the bam system architecture.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Gpu-initiated on-demand high-throughput storage access in the bam system architecture

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:36.064077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.431974Z digest=sha256:aa0a6c1cd93456d91623e0ceebfea720588048b4797f16ad39996b4272626844

Observation 635b8491-36e0-499a-ae3f-3c6646cb1c40 · outbound

This paper cites Language models are unsupervised multitask learners.OpenAI blog, 1(8):9, 2019.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Language models are unsupervised multitask learners.OpenAI blog, 1(8):9, 2019

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.441563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.441563Z digest=sha256:7186f0df78fe3dcec4c1f61dc5d659ca90aad462c06fa4e5bf55078335bbd45d

Observation 01ad0130-5142-4528-bb29-3d84f7a75259 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:32.485821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:32.485821Z digest=sha256:ea3f6ffa74766b78a4694dfa06c74297195460a9df855453b43319bc549cc094

Observation ea7d54cc-26f5-4bec-8c70-07fc2c4a603c · outbound

This paper cites Zero- infinity: Breaking the gpu memory wall for extreme scale deep learning.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Zero- infinity: Breaking the gpu memory wall for extreme scale deep learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.776160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.546527Z digest=sha256:46a761675a44ec912ad01a4cf39485e37313fe2fce518c61040106330aa2973a

Observation 60312ea8-13a6-49bd-acc7-eda91f132159 · outbound

This paper cites {Zero-offload}: Democratizing{billion-scale}model training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage {Zero-offload}: Democratizing{billion-scale}model training

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.580069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.628054Z digest=sha256:e7b97175fa811170641eb2e09927c6e1e044e56ca197e95aacd1d1e5a8c58ee9

Observation 612eff07-8439-4277-b73d-1afec967f1e8 · outbound

This paper cites vdnn: Virtualizeddeepneuralnetworksforscalable,memory-efficientneuralnetworkdesign.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage vdnn: Virtualizeddeepneuralnetworksforscalable,memory-efficientneuralnetworkdesign

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.359528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.723638Z digest=sha256:22238bdd13ecfae6727d22f40cec81c38cd72c1616aa52dd00fcf5d5167851de

Observation fe8d81db-435b-4143-9666-f0460fa68928 · outbound

This paper cites Building ai agents for autonomous clouds: Challenges and design principles.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Building ai agents for autonomous clouds: Challenges and design principles

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:35.090008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.812774Z digest=sha256:a14c9827d9c109e7e8d5f2711abda046d789a557f08570fb22fc5098efd9ab9b

Observation 4899a902-1259-4f04-a872-4a13a2e6d2b3 · outbound

This paper cites Stronghold: fast and affordable billion-scale deep learning model training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Stronghold: fast and affordable billion-scale deep learning model training

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.872110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.885983Z digest=sha256:632dce1b44805c38cab16fa968cc74c905f0481738eeeec169134ca0a5740da2

Observation 87b62046-85c8-4127-ae0d-7ee6f187fbb5 · outbound

This paper cites Superneurons: Dynamic gpu memory management for training deep neural networks.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Superneurons: Dynamic gpu memory management for training deep neural networks

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.654946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:32.986663Z digest=sha256:f33377d258a45b600c8a8a7758a950bffd5a284600c466566c5e8a8bf53b47f7

Observation 60397c44-59f8-443c-abe6-5f9a2c279258 · outbound

This paper cites Netllm: Adapting large language models for networking.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Netllm: Adapting large language models for networking

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.439392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:33.078030Z digest=sha256:2bb647a051a8f90591a010da5fd60faeeda5d018da2a70bb924e224699a81291

Observation 108aa304-7f91-4023-931c-365f0d00c4d5 · outbound

This paper cites SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:33.164055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:33.164055Z digest=sha256:4af76681d7b8ad4dcbb62132d7afa5e8b26ea7105ebed0bf430820d0efa0973d

Observation 57779a89-1c9f-4080-9835-0f54f9743c4d · outbound

This paper cites Acceleratingthetrainingoflargelanguagemodelsusingefficientactivation rematerializationandoptimalhybridparallelism.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Acceleratingthetrainingoflargelanguagemodelsusingefficientactivation rematerializationandoptimalhybridparallelism

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.182335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:33.266662Z digest=sha256:1bc68af797622d195fd6cd81a191f681736e2031a42e29d3deb8729a9ac4be12

Observation ca18bf44-aead-4cb4-bb78-9fb324554960 · outbound

This paper cites Zng: Architecting gpu multi-processors with new flash for scalable data analysis.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Zng: Architecting gpu multi-processors with new flash for scalable data analysis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:34.008635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:33.346043Z digest=sha256:bb831f127f30002602b0fa372c5a0909b35c39e8afd0f7e77285392ecedb990e

Observation f5f669cb-a2c0-4bdd-b073-8e9a4d7e7030 · outbound

This paper cites Flashgpu: Placing new flash next to gpu cores.

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage Flashgpu: Placing new flash next to gpu cores

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T06:00:33.797888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T06:00:33.427109Z digest=sha256:b1040086011c089a73079efdf5150e0e6b50aa1f2f505cb3069fe24b57869dd5

Pith citing papers

No inbound Pith citation observations are available.