Pith. sign in

Paper Citation Record · LEDGER

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization

As of 17 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2501.06663.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06663 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:02:34.906180Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a528b484-22b8-41de-99d1-1cf3a3a99f5f · outbound

This paper cites When machine learning meets privacy: A survey and outlook,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization When machine learning meets privacy: A survey and outlook,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.900765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.621196Z digest=sha256:b9035eb5eaf3dd92700c2fbefdcd1f733dbec212e76c739bf506dbecd2a0fa90

Observation 9af989ce-bcda-4c66-86e7-990abe64973a · outbound

This paper cites Sok: Security and privacy in machine learning,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Sok: Security and privacy in machine learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.884571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.626111Z digest=sha256:952e4619983cbdf57bbb3fe94d25af39e7a5d95000dcf93d2ff84c58e7651f27

Observation 88acd6aa-ff43-4df6-b3f2-c448727d7d2b · outbound

This paper cites DeepReach: a deep learning approach to high-dimensional reachability,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization DeepReach: a deep learning approach to high-dimensional reachability,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.867906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.632210Z digest=sha256:cbf93ca8aceaba12c445c0e8f81dca7ca3a78d4a53ef697b174a049b67a0d9ab

Observation c8afabc7-26b1-419a-a5ec-690cedf1334b · outbound

This paper cites A neural network approach applied to multi-agent optimal control,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization A neural network approach applied to multi-agent optimal control,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.852092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.637703Z digest=sha256:3d2912a801c481c66976d797de7400fd750d0934ae2013ac28e97bc9ad666467

Observation b745b666-7b76-4d58-a007-ecfe8cdfb8d5 · outbound

This paper cites Learning certified control using contraction metric,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Learning certified control using contraction metric,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.837069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.642709Z digest=sha256:7d64a473d1e810ea68d4f5738b08c19c5327647e6d7b3c79ba322fa060a70c44

Observation 3c3ef8e6-32d0-447f-a720-48282023e1fb · outbound

This paper cites Learning in the wild: When, how, and what to learn for on-device dataset adaptation,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Learning in the wild: When, how, and what to learn for on-device dataset adaptation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.820269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.648811Z digest=sha256:329ce7b032f0a1fc3c285753e089957f524d460544e5defdc93b90b8186e2768

Observation 1b9b1210-99c6-4a40-8720-67f6e87e5d94 · outbound

This paper cites EF-train: Enable efficient on-device CNN training on FPGA through data reshaping for online adaptation or personalization,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization EF-train: Enable efficient on-device CNN training on FPGA through data reshaping for online adaptation or personalization,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.800945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.655158Z digest=sha256:85ec10bf57bfb2c38fd7d9cd31f6a3486d56ca80f2e22df2ce9e375e8c8c5d1e

Observation c9790ef0-9554-4736-9790-3be9fa7cb909 · outbound

This paper cites BOOST: block minifloat-based on-device CNN training accelerator with transfer learning,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization BOOST: block minifloat-based on-device CNN training accelerator with transfer learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.784812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.660004Z digest=sha256:e01d154e83d63089be49991b6185b9ba5d0bd45169569a778a9263e1c9937429

Observation 222ec652-2c4b-438a-af1a-6de3c6671924 · outbound

This paper cites ETA: an efficient training accelerator for DNNs based on hardware-algorithm co-optimization,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization ETA: an efficient training accelerator for DNNs based on hardware-algorithm co-optimization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.768424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.665402Z digest=sha256:05afc86dcd0688e8d0e3f0e1566ce3a5b02a78b12335b21414ae00c6371cf8fc

Observation 560807a6-ef4c-4e23-906f-82521fb2e969 · outbound

This paper cites Training deep neural networks in low-precision with high accuracy using FPGAs,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Training deep neural networks in low-precision with high accuracy using FPGAs,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.751757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.670140Z digest=sha256:df603050aa655fd79cd5149155bc7ad1bc0b4cb606c327a0f1bb533b159d6f7f

Observation 1c86f4fb-efb4-441e-ba67-0f320166a3c4 · outbound

This paper cites FAST: DNN training under variable precision block floating point with stochastic rounding,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization FAST: DNN training under variable precision block floating point with stochastic rounding,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.733370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.675239Z digest=sha256:6b1a19700e542325936ff1da9009bf9e273048b319dafa3e8261e1db5e9b4dc3

Observation f26ca76e-8598-4be8-ae52-2ef688a9c662 · outbound

This paper cites Attention is all you need,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Attention is all you need,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.680780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.680780Z digest=sha256:cb96bc1eb60a41d81cbb2c53c2e0bc3f138506a1ad3a6385c79c9989f35917fe

Observation 1aad8ddd-8844-4a66-9b52-e00249f87d52 · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.705501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.685425Z digest=sha256:eee78a7699b3a560cd237699b7c3c1f0c5f17edb8955b60629b6d4b53faf6da7

Observation 29feb44a-97a5-4208-83e0-b98acbb960f2 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.689435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.689435Z digest=sha256:6a9f5818516e6414e68f52754ebdfc19cdb44d16c7d6c1d55435921e4134b72f

Observation 433a0d07-9a39-4714-a2fb-66c1eaa8693d · outbound

This paper cites Segment anything,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Segment anything,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.689465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.694119Z digest=sha256:863f98ec22341d619ba136bed278688a491616dd1af677907fb7740e7855ec3a

Observation 9eaed75d-ecba-40f5-834b-8c01c918fbe5 · outbound

This paper cites PINNsFormer: A transformer- based framework for physics-informed neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization PINNsFormer: A transformer- based framework for physics-informed neural networks,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.673202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.698271Z digest=sha256:6e41c3436d145288c0c6ff314ba446f63ea4fc098d232034afa311140b110da3

Observation ae96de93-2f5c-4ad9-a2de-f2b488434dc5 · outbound

This paper cites Federated learning in asr: Not as easy as you think,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Federated learning in asr: Not as easy as you think,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.655480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.702459Z digest=sha256:6c36675b045da0f0cc106fd511cf9d86795c662b13804fd1699d5db2e0e7319b

Observation b7f34a39-681f-4b87-acea-0fce80f52970 · outbound

This paper cites Federated Domain Adaptation for ASR with Full Self-Supervision.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Federated Domain Adaptation for ASR with Full Self-Supervision

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.706338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.706338Z digest=sha256:34b4a45f373a04a0cd4cdfbe0e1b21c315d4c1bb1a39de5e67fdd844a1b242f7

Observation 457e28ed-38f0-4f8e-bad8-703c93d922f1 · outbound

This paper cites BEBERT: efficient and robust binary ensemble BERT,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization BEBERT: efficient and robust binary ensemble BERT,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.638419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.711245Z digest=sha256:a6e3e528eabe967cffcf555fad58f1471bf97ef165f3099b59049fac795e8f06

Observation 201f0beb-952d-435f-8d1b-134d2d2ed07d · outbound

This paper cites BETA: Binarized Energy-Efficient Transformer Accelerator at the Edge.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization BETA: Binarized Energy-Efficient Transformer Accelerator at the Edge

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:02:35.025041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.716955Z digest=sha256:ea42d924829e7f0c3882d4c2552d325d042fc60ca5eff8a65e51e9ceca62a5be

Observation 53268a6c-2bfa-4786-b202-e8dd5ebf3b44 · outbound

This paper cites Training deep neural networks with 8-bit floating point numbers,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Training deep neural networks with 8-bit floating point numbers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.620937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.721805Z digest=sha256:2cac13160988f2cdd7ba04000d6aa5ad91790d415077f14207520fcd3699ad72

Observation 5b8278ff-8f14-4b08-bb71-37e1c3672f58 · outbound

This paper cites Quantization-Aware and Tensor-Compressed Training of Transformers for Natural Language Understanding.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Quantization-Aware and Tensor-Compressed Training of Transformers for Natural Language Understanding

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:02:35.002354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.726485Z digest=sha256:08262b8ec38ef48b0d5400575ae48cb13fe708d7dc98f0797f04fa7f27062168

Observation d92fa232-383f-4b18-a693-fe165f333a5a · outbound

This paper cites LoRETTA: low-rank eco- nomic tensor-train adaptation for ultra-low-parameter fine-tuning of large language models,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization LoRETTA: low-rank eco- nomic tensor-train adaptation for ultra-low-parameter fine-tuning of large language models,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.605182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.731719Z digest=sha256:2e97dd4554bcbf698a7ccb392bdb73c76ac86a1fb9e40be10d7e09822ffa6a4d

Observation d98655cd-cae4-4b81-87a1-43f52b441f17 · outbound

This paper cites AdaZeta: adaptive zeroth-order tensor-train adaption for memory- efficient large language models fine-tuning,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization AdaZeta: adaptive zeroth-order tensor-train adaption for memory- efficient large language models fine-tuning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.588268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.736335Z digest=sha256:31515cb40f3d1affa0d215b287a703ed2ba836f9395f41377487fd0d46146af7

Observation ff30906e-4784-4ce0-9e18-36450e02dc32 · outbound

This paper cites An algorithm–hardware co-optimized framework for accelerating N: M sparse transformers,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization An algorithm–hardware co-optimized framework for accelerating N: M sparse transformers,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.569929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.740905Z digest=sha256:da7833b0a17ffbec9bfa7385a185d595264d6f359acd7f836760e8e31ff22cb3

Observation 45151565-c3d7-45dd-b07e-892a0867effa · outbound

This paper cites Efficient N: M sparse DNN training using algorithm, architecture, and dataflow co-design,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Efficient N: M sparse DNN training using algorithm, architecture, and dataflow co-design,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.554111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.746477Z digest=sha256:63633d0bde5a66034036fc629409ab0a0c0e9214c0df3cb123fc3ce01330b85a

Observation 7eece556-4698-4c72-b56a-e5a6867cab5d · outbound

This paper cites MINILM: Deep self-attention distillation for task-agnostic compression of pre- trained transformers,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization MINILM: Deep self-attention distillation for task-agnostic compression of pre- trained transformers,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.537384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.751042Z digest=sha256:af3e99004fbb1c51b00b4a1214cd6595fa880e95854de2a980c2cccde3fbb681

Observation 62622bc6-62d4-4b91-a896-b92258791984 · outbound

This paper cites Wanda++: Pruning large language models via regional gradients,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Wanda++: Pruning large language models via regional gradients,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.520222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.755945Z digest=sha256:5ff1e352a0a20e913a9885bf7bf5533bf18fba84c03c886b5abf05093e472fe2

Observation 0dcc4082-d9b2-4a2f-a74d-81c3f96b9d0b · outbound

This paper cites Compressing dma engine: Leveraging activation sparsity for training deep neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Compressing dma engine: Leveraging activation sparsity for training deep neural networks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.502722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.761118Z digest=sha256:5abe5c12f83fffb14b43f6ee3b7e556ada6f613f87aade67a9333ed3dc61ee8a

Observation 11406047-74c2-4dec-8000-1a81249c299c · outbound

This paper cites Quantized neural networks: Training neural networks with low pre- cision weights and activations,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Quantized neural networks: Training neural networks with low pre- cision weights and activations,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.765940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.765940Z digest=sha256:b5326ee3063c9df96b40a6cce3500e6d02ce48f7f98d84156eb7d73746b67da8

Observation cb9ac2de-89c6-42cb-b596-2e677cfb0945 · outbound

This paper cites Deep learning with limited numerical precision,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Deep learning with limited numerical precision,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.475381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.770934Z digest=sha256:eb06c39e7d74cc6925ef34d1eb4f0f4c985c242fe25ad5fe5c983806a9b9b287

Observation d7362667-e6c8-4998-b561-0d56ff1dec9e · outbound

This paper cites Ultra- low precision 4-bit training of deep neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Ultra- low precision 4-bit training of deep neural networks,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.458527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.775666Z digest=sha256:22f4eb8a0ad7554cbb9f37dd29de29357b7a16b35377774df82666a421e55a95

Observation 771ee7b0-59ad-40c6-99e9-bf8f11d8e090 · outbound

This paper cites Tensor decompositions and applications,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensor decompositions and applications,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.780461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.780461Z digest=sha256:85ea712df535cb3cc7c5de2b5ef126b3e05f80307ecae506b6ef378e89cd5f87

Observation 296d507f-1d22-48c6-987a-8c7719f6e57d · outbound

This paper cites Speeding-up convolutional neural networks using fine-tuned cp- decomposition,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Speeding-up convolutional neural networks using fine-tuned cp- decomposition,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.430776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.784569Z digest=sha256:cdebc4979b6943cb345d5e2e6b3ac7a15fb6e915daffe8d12cf509822428cda8

Observation 9602d4be-65ce-4962-bb6e-9ecfb9e2f3ef · outbound

This paper cites Compression of Deep Convolutional Neural Networks for Fast and Low Power Mobile Applications.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Compression of Deep Convolutional Neural Networks for Fast and Low Power Mobile Applications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.789038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.789038Z digest=sha256:d260808943045c660bed89a1ee86b9f92fc37e05ec15665a888528bd16dcfbee

Observation 75f57b58-d873-4cca-bea8-9d6b4f68abd7 · outbound

This paper cites Fast video facial expression recognition by deeply tensor-compressed LSTM neural network on mobile device,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Fast video facial expression recognition by deeply tensor-compressed LSTM neural network on mobile device,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.413299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.793571Z digest=sha256:8d584aed75a380d0d1ab8f7368170d8272eebd91874a5c6cecb850b487144da3

Observation ec34acf0-f02b-4dab-aa08-215738260a02 · outbound

This paper cites Tensorizing neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensorizing neural networks,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.797827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.797827Z digest=sha256:196f924472affa15d5b00d0619cd4110e4c05f929a99689c655bd7e69aac32c5

Observation 7ae86d44-fdf6-4f37-be79-681521f0a82e · outbound

This paper cites Compression and Interpretability of Deep Neural Networks via Tucker Tensor Layer: From First Principles to Tensor Valued Back-Propagation.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Compression and Interpretability of Deep Neural Networks via Tucker Tensor Layer: From First Principles to Tensor Valued Back-Propagation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.801878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.801878Z digest=sha256:3a4fadc5b3b7bf532ec798663f0b99c94eb0273964fd2461836ee427bc3182d4

Observation fd986210-88d2-4823-84ad-cd7a985e039b · outbound

This paper cites Tensorized embedding layers,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensorized embedding layers,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.384401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.806898Z digest=sha256:efef77ee938c22c1e1aab41d1594afb8a55ec56700d8ba5443220a3a78efc5fe

Observation 1533a14d-c31f-4feb-8d1e-457bf8b41757 · outbound

This paper cites A tensorized transformer for language modeling,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization A tensorized transformer for language modeling,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.369241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.811317Z digest=sha256:ab1e357c9742e7404498da873b0b8e8d63c5bf6a6ef7b28e8401f936e190ac53

Observation eb7b8697-aa6d-4699-a0bd-2b0b95f473b9 · outbound

This paper cites TIE: energy- efficient tensor train-based inference engine for deep neural network,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization TIE: energy- efficient tensor train-based inference engine for deep neural network,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.353610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.815771Z digest=sha256:72fb6618128311ba2dd68754c3c0caea58783917d75b7ea0324a28e1178c8753

Observation f5569ca4-54fa-404b-89f4-90bf425e5a61 · outbound

This paper cites ETTE: Efficient tensor-train-based computing engine for deep neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization ETTE: Efficient tensor-train-based computing engine for deep neural networks,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.337043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.820252Z digest=sha256:a2a1c9ab4b2e5c792272a07b072c252d33caeb8c8da6e32bfab51c9da3f3c841

Observation 10e84211-9999-45b6-97d1-9f890d0e4298 · outbound

This paper cites TT-CIM: Tensor train decomposition for neural network in rram-based compute-in-memory systems,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization TT-CIM: Tensor train decomposition for neural network in rram-based compute-in-memory systems,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.320717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.824858Z digest=sha256:25b55969ffb9cc09d073ec7db839b09ac213ce7815a9ae600d3e230d685fbdfd

Observation 74892902-0616-4bf2-8e3b-06ea6f009722 · outbound

This paper cites 15.4 a 5.99-to-691.1 TOPS/W tensor-train in- memory-computing processor using bit-level-sparsity-based optimiza- tion and variable-precision quantization,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization 15.4 a 5.99-to-691.1 TOPS/W tensor-train in- memory-computing processor using bit-level-sparsity-based optimiza- tion and variable-precision quantization,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.304332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.829680Z digest=sha256:e17457f8c235fb2043b313c590b23b9cdcb199a3936b1273b8c57ab3cfbbcd08

Observation e3fff906-6096-47df-ad95-f3b2bf22f713 · outbound

This paper cites Tensor-Compressed Back-Propagation-Free Training for (Physics-Informed) Neural Networks.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensor-Compressed Back-Propagation-Free Training for (Physics-Informed) Neural Networks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.835075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.835075Z digest=sha256:d9a9b92a187b09d457beeabb555f2636616fc57f5f1019dc64ece299b564b0da

Observation e476048e-7405-4180-88bb-dad5d61f9f07 · outbound

This paper cites Tensorized optical multimodal fusion network,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensorized optical multimodal fusion network,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.288931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.839943Z digest=sha256:fe18d9de1594dd90f2022b306ea2cb670f40ed3c71ccbb50154e1b388046a62e

Observation 760367ca-702b-4081-98ce-9fadeb891b15 · outbound

This paper cites Real-time fj/mac pde solvers via tensorized, back- propagation-free optical pinn training,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Real-time fj/mac pde solvers via tensorized, back- propagation-free optical pinn training,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.271957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.845661Z digest=sha256:7c14be70e6444757765cb794163d8f0d5a5d51268f1173a74da1ee09186eb31d

Observation 4977035e-3445-4fcd-8817-9bdf888cdefb · outbound

This paper cites The ATIS spoken language systems pilot corpus,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization The ATIS spoken language systems pilot corpus,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.255993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.850857Z digest=sha256:97f6be820876db794c06cd46f786adbc351e644a47d715bb563c2cb3320354fe

Observation 787095f8-228b-43dd-ab3b-e935d0306018 · outbound

This paper cites Tensor-train decomposition,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensor-train decomposition,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:02:34.855807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:02:34.855807Z digest=sha256:baf05ba1b4b2817ac525df018b7d7b3d6f360b0ce7b3f789dfc0aeaa556dc851

Observation 84beec9b-31e1-4cc9-857f-1e28901aae08 · outbound

This paper cites Tensor-train recurrent neural net- works for video classification,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Tensor-train recurrent neural net- works for video classification,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.230465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.860510Z digest=sha256:5d0f40f1d221373bc8e635a18bf495e30e38c3d333eac16ee42a87545c887ee6

Observation f6012b34-4f7d-441d-a7cd-af7b8c441173 · outbound

This paper cites Learning compact recurrent neural networks with block-term tensor decompo- sition,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Learning compact recurrent neural networks with block-term tensor decompo- sition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.215107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.865236Z digest=sha256:0b91ea83fa48194e39fe76d39d2400261864c4d85bdc85722b276a90e6cafa1e

Observation 0808fd52-35a3-4acd-be49-6e3a8b198d0c · outbound

This paper cites Towards compact neural networks via end-to-end training: A Bayesian tensor approach with automatic rank determination,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Towards compact neural networks via end-to-end training: A Bayesian tensor approach with automatic rank determination,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.199924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.870400Z digest=sha256:e11f6138baf2b4c308d0f98098b48f83cd2e8584e36c8f7f8f5cb83a617525de

Observation 3cb058de-3320-459c-86b2-e51f08f14a82 · outbound

This paper cites Bayesian tensorized neural networks with automatic rank selection,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Bayesian tensorized neural networks with automatic rank selection,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.184999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.875116Z digest=sha256:ac65595ce5908fc9812d378b689c0ec7d34831fff3b49fecf9e41165a4e2eb33

Observation 8b321dcc-e70b-469d-8449-ac5f6d2f3161 · outbound

This paper cites Compressing 3D CNNs based on tensor train decomposition,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Compressing 3D CNNs based on tensor train decomposition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.170445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.879762Z digest=sha256:ea2d480e19bd3a911965ac9a7d04b89a59c18de17c8999aadef64459e5551c85

Observation c2f7b994-8d20-4ea8-827a-0ec5374e1923 · outbound

This paper cites Nonlinear tensor train format for deep neural network compression,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Nonlinear tensor train format for deep neural network compression,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.154340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.884097Z digest=sha256:ce08b5cb0ad70b6d9302b9bf955885f6a0a2a7bedcdd43aa628babbd5ddbed0b

Observation a07ed29c-6697-4acd-8162-6332b808ec5c · outbound

This paper cites CoMERA: Computing-and memory-efficient training via rank-adaptive tensor optimization,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization CoMERA: Computing-and memory-efficient training via rank-adaptive tensor optimization,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.136889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.888393Z digest=sha256:976cb7499dcb05644d2949877046680fad46f25cbda0a1c9081c4e90ad9f30e1

Observation d0fd9345-174b-4ce1-b0da-e0f801bac2b1 · outbound

This paper cites An FPGA-based processor for training convolutional neural networks,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization An FPGA-based processor for training convolutional neural networks,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.121361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.892435Z digest=sha256:b98ca2306c1de426025b63d48718fa465d2e8d10264f8c9b153a455acf04c389

Observation 1505b83c-b241-4ee2-ba12-8d8c38345706 · outbound

This paper cites Automatic compiler based FPGA accelerator for CNN training,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization Automatic compiler based FPGA accelerator for CNN training,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.105553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.896674Z digest=sha256:7b83383751b13b62ea532a11ec2a136b12a1ec6167b59646ee9171fdf6774217

Observation c563c2d9-fcea-4869-b147-b062f291e149 · outbound

This paper cites FPGA-based low-batch training accelerator for modern CNNs featuring high bandwidth memory,.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization FPGA-based low-batch training accelerator for modern CNNs featuring high bandwidth memory,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.089897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.901239Z digest=sha256:43385860349fcce6c375d4b273498d17868ac98a4d1815979f20cc248b93bfd9

Observation 1ed1fc16-e507-47e4-9e6e-f0eafa3719ba · outbound

This paper cites for leadership in research and development on circuits and processes for the evolution of microprocessors.

Ultra Memory-Efficient On-FPGA Training of Transformers via Tensor-Compressed Optimization for leadership in research and development on circuits and processes for the evolution of microprocessors

Reference 1983

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:02:35.075105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:02:34.906180Z digest=sha256:5761d8c55635cde22f14a8648aa9e20a3167d3722b5e1c30e64b3572e6892b4c

Pith citing papers

No inbound Pith citation observations are available.