Pith. sign in

Paper Citation Record · LEDGER

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling

As of 10 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 0 inbound Pith citation observations for arXiv:2505.17909.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17909 v1

Coverage vector

measured 85 of 85 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:23.210297Z

measured 85 of 85 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

85 of 85 outbound references displayed

  • verified exact17
  • verified fuzzy9
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66e5f206-e71b-461c-81a9-e02c56b8cf40 · outbound

This paper cites Dual Lottery Ticket Hypothesis.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Dual Lottery Ticket Hypothesis

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:27.957516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:11.919090Z digest=sha256:30b0ca01310a68e55ab54711a6e5257e2bda0e7ccffdb20c9b3a6c33f45a8578

Observation 5193e021-c86b-4be2-8e41-7df1064721d2 · outbound

This paper cites Deep Rewiring: Training very sparse deep networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Deep Rewiring: Training very sparse deep networks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.035710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.035710Z digest=sha256:b5b0e5d7ae44929195ce7fa121740dc32901e37dd28dbd0d3dfc3bcca7aecf74

Observation 09bb27e9-9ff0-450b-b125-008ed0bc1832 · outbound

This paper cites an unresolved cited work.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Unresolved cited work

Reference 3

Resolution
verified exact
doi, observed 2026-08-07T14:44:23.514478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:12.138159Z digest=sha256:a11dc5a4841a93ba327337c76d9b2df1496a62a91b41a4f3751a7ed4bf839d05

Observation 75416058-9ad1-46d7-a048-191eec4ce2c3 · outbound

This paper cites Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning Better.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning Better

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.236606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.236606Z digest=sha256:73b8dc85a6b45253b4072ce58093f07beb71526859ee5182010d9fe1b9da9d57

Observation e2e254e4-b893-42af-85eb-46e08bb5e66d · outbound

This paper cites PIQA: Reasoning about Physical Commonsense in Natural Language.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling PIQA: Reasoning about Physical Commonsense in Natural Language

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.329681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.329681Z digest=sha256:d1f36515afc9944ed33dfb9ea33cf0ccd4bafbb4365982b4be2b5bfcd3cfc43c

Observation 98e0d66b-7034-4165-84f6-fdfc12fe95fd · outbound

This paper cites Bagging predictors.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Bagging predictors

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.426680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.426680Z digest=sha256:c533ac32f8fd6d5f89845dac68af648b1468df9516375a58a0d651ff76fc1c98

Observation eb3149ad-85d8-473e-af6e-42eda8ffd709 · outbound

This paper cites Sparsity Made Easy – Introducing the Cerebras PyTorch Sparsity Library - Cerebras , 2024.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Sparsity Made Easy – Introducing the Cerebras PyTorch Sparsity Library - Cerebras , 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.542409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.542409Z digest=sha256:3365705a5dcb1713b67b221b951249ce318036f1b0fbc23473302e0cf1431b90

Observation 43462b01-4665-4a7a-8025-1233e03edb06 · outbound

This paper cites BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.608543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.608543Z digest=sha256:98feaa9ac1902ea9e9d1569320de41e8c3bcb2ca50326302650db8a8c5e22e5c

Observation 6227da71-17c8-497e-92ae-51d6ef28cd4a · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.707447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.707447Z digest=sha256:6352b5e18d6ccd4dd669f660f27afb36aea5d230ec13d9d8339eb710adf0cb0f

Observation 347dcfd5-2bc6-4ada-994e-c4e3e02bf2c2 · outbound

This paper cites Truly Sparse Neural Networks at Scale.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Truly Sparse Neural Networks at Scale

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:27.740776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:12.769776Z digest=sha256:658b816e008cca47cec6476e398773ae0aa3eb7d43a09a4dcbdf4d53bd292c7c

Observation 9e422ca2-d494-4951-85de-e46cbbe87926 · outbound

This paper cites ImageNet: A large-scale hierarchical image database.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling ImageNet: A large-scale hierarchical image database

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.870321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.870321Z digest=sha256:e1f46b87f53c1d2ec1c79692dbee4eda1a063bca949c9debc170848a901bbaea

Observation 61f28856-7849-4afc-b622-6ef12bb32ca3 · outbound

This paper cites Sparse Networks from Scratch: Faster Training without Losing Performance.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Sparse Networks from Scratch: Faster Training without Losing Performance

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:12.949488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:12.949488Z digest=sha256:ef57c80580525d1baf790b1b27370e6d9e15f80e053a70cb4587a29ecc2bc86b

Observation 69f99ec9-8bbd-4c3e-8c1b-16e8cab6e461 · outbound

This paper cites Dietterich.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Dietterich

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.057748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.057748Z digest=sha256:cabb8096b959158c632c6f92bb3c76c353f5c745c00e8ca1287417f5cf3814b9

Observation 732ef82f-9c5f-4133-9b06-b1c498aca163 · outbound

This paper cites Rigging the Lottery: Making All Tickets Winners.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Rigging the Lottery: Making All Tickets Winners

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.153977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.153977Z digest=sha256:2112827fccaad5fa6e432ebc97ff45d7f0d1fb4e76ad40e6406699eab51c4706

Observation 2c4f52e5-71e6-465e-a7bf-3064e9da51b5 · outbound

This paper cites Gradient Flow in Sparse Neural Networks and How Lottery Tickets Win.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Gradient Flow in Sparse Neural Networks and How Lottery Tickets Win

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.271860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.271860Z digest=sha256:d79daa840ed78a8c2337017bd913808029b7e744f4975541bdcb23dd40dfb8a2

Observation cbcecc6d-9459-41bf-8d81-96545c2aba95 · outbound

This paper cites Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.467858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.467858Z digest=sha256:3ccc47c9eb8771371be5ff3427f34922ed34c985c841138f359aa774918663d7

Observation a5ca5e0a-a4d4-4818-b051-271d96a6202d · outbound

This paper cites Deep Ensembles: A Loss Landscape Perspective.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Deep Ensembles: A Loss Landscape Perspective

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.632238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.632238Z digest=sha256:e8f8b2e9c2a631173f0055bedf13d1f8770b794cdf5f2eb7a2204b6da80e310c

Observation bd4e18b1-ff3b-450d-99f3-c9fae7dac3cd · outbound

This paper cites The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:13.869088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:13.869088Z digest=sha256:ef60d537373b0fe03b88f9e4f4557a3163c610f4b08435d0bb4aed0086b90c39

Observation bbcf884e-369e-42cc-912d-342783d1c755 · outbound

This paper cites A Decision-Theoretic Generalization of On-Line Learning and an Application to Boosting.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling A Decision-Theoretic Generalization of On-Line Learning and an Application to Boosting

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.079669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:14.079669Z digest=sha256:88752e211ac9a64b0f22162c927b43078996b72598a77df475eadbd258832f9d

Observation 7894540d-135f-4d5a-94ee-2b26184e4965 · outbound

This paper cites A Survey on Ensemble Learning for Data Stream Classification.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling A Survey on Ensemble Learning for Data Stream Classification

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.301408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:14.301408Z digest=sha256:354604b75bdd431eccbe442fdf10be378ccd540703d75749abc89b92c403af4f

Observation d249c34a-a4fd-4e4e-a578-5aff4c29bf8f · outbound

This paper cites The State of Sparse Training in Deep Reinforcement Learning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling The State of Sparse Training in Deep Reinforcement Learning

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:27.385114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:14.421338Z digest=sha256:624d2720945bb020af8dc4edc03fa79d95041f2cdf29d97261933df25515ee32

Observation e0e5c620-8dcf-4eda-9a30-1a7cf1bebb03 · outbound

This paper cites Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:44:27.194418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:14.586279Z digest=sha256:3e8b1f1663544d5fcd44ac20163d5a3eeac63db5c48a9ba4520c28fbace23683

Observation 4dee6fdf-9bc8-46d0-8036-f76255f032fd · outbound

This paper cites On Calibration of Modern Neural Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling On Calibration of Modern Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.769213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:14.769213Z digest=sha256:92c950b098b2990e249c2b9067c26ebe638fac8ef6537708871bf72ac84494f0

Observation 42a12602-f844-407b-8860-81bbffcb3db0 · outbound

This paper cites Learning both Weights and Connections for Efficient Neural Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Learning both Weights and Connections for Efficient Neural Networks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:14.934609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:14.934609Z digest=sha256:7a412c89686389b1b98c6ca60c4a859abad8ce28cbc88bfdb3c061639dd63c90

Observation f66e33af-f58c-45a1-8041-0852bbf8ab92 · outbound

This paper cites Neural Network Ensembles.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Neural Network Ensembles

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:15.133429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:15.133429Z digest=sha256:7f637c841df25668784d4dfd686d75ac9d45e8d0e5c470d1f1cda6ad1660bbbf

Observation 1cd3eb80-6b6d-419b-88ae-a847d889d8b4 · outbound

This paper cites The Elements of Statistical Learning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling The Elements of Statistical Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:15.312715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:15.312715Z digest=sha256:1447ebc61ec1377ef8bd665df16ecf786c785e027842d2a958aac72532cde387

Observation 9c6085fc-c066-4d71-b85a-31f4fb9a9c93 · outbound

This paper cites Training independent subnetworks for robust prediction.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Training independent subnetworks for robust prediction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:15.519215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:15.519215Z digest=sha256:50b63dfc51381c604966b8f343f04216a90ff95ecc5310dd13aa3547bc4dc427

Observation 2784e133-3518-43e1-be50-b71a84b7b959 · outbound

This paper cites Deep Residual Learning for Image Recognition.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Deep Residual Learning for Image Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:15.623137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:15.623137Z digest=sha256:69663e4674a992a3a6a3ae451dde7e60e0b92942dfcd7a14a4c05fe2f5c25dc1

Observation 5773cf56-cb35-4333-baf9-842db34ad5f8 · outbound

This paper cites Benchmarking Neural Network Robustness to Common Corruptions and Perturbations.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Benchmarking Neural Network Robustness to Common Corruptions and Perturbations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:15.789693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:15.789693Z digest=sha256:c89cc1748d422ece5e4285d3bc724d7ca5fa1dfeec4b2f8f347d6ff5bb152070

Observation bbbe08d4-a64e-4431-95d0-b789630a9671 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Measuring Massive Multitask Language Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:16.005473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:16.005473Z digest=sha256:85eca7837e1288c95762845cb054f0badf91cae00fa9399b5f94d50fc195207b

Observation 440b9ba4-f062-466e-a86f-901ea0e1ac1c · outbound

This paper cites Distilling the Knowledge in a Neural Network.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Distilling the Knowledge in a Neural Network

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:16.198598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:16.198598Z digest=sha256:2d5ae7bb5f2ad50412ff6e9b2b641fa3e3e91807094a7759c1c2fe0ebc2f3253

Observation b70f59c0-d8e2-4fcb-b744-7eba9775c5bd · outbound

This paper cites Jacobs, Michael I.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Jacobs, Michael I

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:16.342941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:16.342941Z digest=sha256:f80e3d45c315937fdd02b0b23ba09f65034fdc3c97641d11bdaeea40993c1367

Observation dd4669ac-bfaa-477c-bfba-5cbc2f465524 · outbound

This paper cites Joint Training of Deep Ensembles Fails Due to Learner Collusion.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Joint Training of Deep Ensembles Fails Due to Learner Collusion

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:26.948514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:16.441859Z digest=sha256:cd548d7824ea0b7ab50adc5cf04da89bdf48f9ae0cf36f77c06d0edc536c28db

Observation f48f0d8e-226f-4f08-aa19-125d09ad7859 · outbound

This paper cites Mercer, Lalit R.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Mercer, Lalit R

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:30.585191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:16.593230Z digest=sha256:28493f03ff10fc6b977a39d3325a1ad715f853e1a1165395268cb72a0d5201b8

Observation 3127ba84-dcf5-4e3c-8d5b-c803230311a7 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Adam: A Method for Stochastic Optimization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:16.753152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:16.753152Z digest=sha256:d672e46689431b15616415f587e8c78d282bdb7211edef06f02bd37e6dc2f61e

Observation 7cc5f1af-09cb-42ef-b42e-b9ff53c80b6d · outbound

This paper cites Learning Multiple Layers of Features from Tiny Images.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Learning Multiple Layers of Features from Tiny Images

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:30.448192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:16.927446Z digest=sha256:2b01553df5cd1bf299386096338142a3e8960b4c748dd5c11c7c3728952fa269

Observation 4812df5f-3a6a-4199-80cd-860853c6e09a · outbound

This paper cites Kuncheva and Christopher J.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Kuncheva and Christopher J

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:17.144681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:17.144681Z digest=sha256:63c4ddb437cc3224a4db0296c26d884aba51b729cf269eaee718506ff40f91af

Observation 82ad0db2-4f3d-4b22-b234-b70a37b25ae9 · outbound

This paper cites Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:17.305484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:17.305484Z digest=sha256:385a494f24b4e6cdc17598347e038da82d2a601131831c50cb53f875451c630d

Observation 18acd9d6-b85a-477c-9591-78e6ada56074 · outbound

This paper cites an unresolved cited work.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:30.228382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:17.527410Z digest=sha256:3767e4e869e323fb7f92eaae3d5920e6251605c1a19ba19aeedfc79259de7446

Observation f78425ac-28b6-4bfb-bb44-05e5bd1f1859 · outbound

This paper cites Optimal Brain Damage.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Optimal Brain Damage

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:30.005151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:17.708917Z digest=sha256:6a64f5ece8ed031ac72d6b1dfa5e8e80faccbd8c5058296b561f75efd0b6e4fc

Observation 4f89ff15-487a-4116-a9b3-3584301ab336 · outbound

This paper cites Network Fission Ensembles for Low-Cost Self-Ensembles.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Network Fission Ensembles for Low-Cost Self-Ensembles

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:26.730470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:17.858613Z digest=sha256:562ed5603991eb2cbe232cf8f3944077be39eaa0c7bca25818a79975f2addec1

Observation edcfffd4-7652-4500-a4df-1d99501792d5 · outbound

This paper cites SNIP: Single-shot Network Pruning based on Connection Sensitivity.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling SNIP: Single-shot Network Pruning based on Connection Sensitivity

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:18.076914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:18.076914Z digest=sha256:023e5ce513b54413b17294598f570381a445e74c79eeb29ec2552bc3c828bf70

Observation 0a0d4d28-85d6-4100-b22f-bb196873f680 · outbound

This paper cites Why M Heads are Better than One: Training a Diverse Ensemble of Deep Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Why M Heads are Better than One: Training a Diverse Ensemble of Deep Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:18.243783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:18.243783Z digest=sha256:13590a2d70403e320b6f7078d70334981b7e8c64b8028494291f44a2cf4c2b62

Observation 6ce60a9a-7ba9-4126-bdf4-772657ddd940 · outbound

This paper cites Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Mix-LN: Unleashing the Power of Deeper Layers by Combining Pre-LN and Post-LN

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:18.461312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:18.461312Z digest=sha256:e91221ea0b4bb8a64285e09dbf101a0a14dd05b08e444f2779650564defc8400

Observation 719d0591-8ec3-4721-8e02-4c0ffccc76ad · outbound

This paper cites Sparse evolutionary deep learning with over one million artificial neurons on commodity hardware.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Sparse evolutionary deep learning with over one million artificial neurons on commodity hardware

Reference 45

Resolution
verified exact
doi, observed 2026-08-07T14:44:23.355126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:18.616644Z digest=sha256:db48ff5fa2599b803bd5c63cf69c99190270566a3694e79bfc70baede114e133

Observation 82285b37-8be4-4244-a883-9675835242b6 · outbound

This paper cites Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:18.767822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:18.767822Z digest=sha256:29975a02d80c090f46f37ade1f25e0917aef4e8e2c65433251520ee45a7d3713

Observation 9685a34b-eabb-4197-afc9-2b1f044cc95f · outbound

This paper cites Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Deep Ensembling with No Overhead for either Training or Testing: The All-Round Blessings of Dynamic Sparsity

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:18.984031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:18.984031Z digest=sha256:dd64ffb56325b596ba9f1557221f11a9a85910bdbb19330863222fc027649860

Observation 1fa92b7c-f112-4c2e-b429-afe41b7667f2 · outbound

This paper cites The Unreasonable Effectiveness of Random Pruning: Return of the Most Naive Baseline for Sparse Training.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling The Unreasonable Effectiveness of Random Pruning: Return of the Most Naive Baseline for Sparse Training

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:19.129246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:19.129246Z digest=sha256:7a0c1e1917bdf5deea6ecc8939cda16d5bac1d420afa68ab3d1f5bd9b518be67

Observation e4ce0165-5ab5-462f-b501-107392af1db9 · outbound

This paper cites Popular Ensemble Methods: An Empirical Study.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Popular Ensemble Methods: An Empirical Study

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:19.344085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:19.344085Z digest=sha256:2066d8e0a291d5fb18c0dbb0398f8fc71ce9909e5dabdb523728cbd5980049c7

Observation 8cb67157-9dc8-4e13-a181-cd6a2d42a304 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:19.547695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:19.547695Z digest=sha256:2bf8c8061b1ee5c7d2670afaade3e3fe2f6e5b28576665c71cb6cd76c971d42b

Observation cd903a2c-20b0-4ab3-9c75-284ac6f5eb59 · outbound

This paper cites Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Scalable Training of Artificial Neural Networks with Adaptive Sparse Connectivity inspired by Network Science

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:26.358926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:19.713861Z digest=sha256:aa3913192ade2f9ab21b46a56e0670a2cafd712df02a6625b02bb5bb3d2a00f1

Observation 1895e05d-c12e-4813-8d2b-1f52d7a9badb · outbound

This paper cites Skeletonization: A Technique for Trimming the Fat from a Network via Relevance Assessment.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Skeletonization: A Technique for Trimming the Fat from a Network via Relevance Assessment

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:29.769200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:19.933039Z digest=sha256:784b42c8d8ade3e17156a4ca0f52516aa47709f71d067dcd1d6e353308d50339

Observation 330b8e86-c430-4f92-90ff-734bb8247042 · outbound

This paper cites Obtaining Well Calibrated Probabilities Using Bayesian Binning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Obtaining Well Calibrated Probabilities Using Bayesian Binning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:20.071092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:20.071092Z digest=sha256:972792e415b1ac27ff3eebb1bf76fa894d2dcaef7397f132bf006519bed1d6e8

Observation e994418e-410d-4d81-b697-afa0f2d65c6b · outbound

This paper cites DeepSparse Inference Engine , 2021.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling DeepSparse Inference Engine , 2021

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:29.514942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:20.175525Z digest=sha256:4f347266217122e5856155587959094e89f9f3e6adeeaab053052c671099d8e4

Observation 1e3ca92d-eb0d-4d54-a43e-d547ff981e2a · outbound

This paper cites Fantastic Weights and How to Find Them: Where to Prune in Dynamic Sparse Training.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Fantastic Weights and How to Find Them: Where to Prune in Dynamic Sparse Training

Reference 55

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:44:26.091130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:20.325435Z digest=sha256:8e67574d543946f96d3eabdcf8281c462d57f9fe88f7a86c4046d808e7d73b60

Observation 8428e4be-13b1-422d-a74d-97c567c7e979 · outbound

This paper cites Sparser, Better, Deeper, Stronger: Improving Sparse Training with Exact Orthogonal Initialization.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Sparser, Better, Deeper, Stronger: Improving Sparse Training with Exact Orthogonal Initialization

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:20.425205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:20.425205Z digest=sha256:9b9cacae68870bf16551971d7300c9c7b8c3b7abe13538ec65071f3488785c96

Observation dedd5263-4b20-4db2-a735-4dbd747ce3d0 · outbound

This paper cites ResNet50 v1.5 for PyTorch , 2024.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling ResNet50 v1.5 for PyTorch , 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:29.252939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:20.561166Z digest=sha256:5f49325049518118176fb361436a60f10a20fe7ad63eeb0e27a52e0c47301362

Observation 597bad49-07e9-4373-9c4d-c407433fb53c · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:29.008177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:20.688213Z digest=sha256:9b6841a3f647cc540849b2ad447a8bc058d080977fa798883ae107d361cd5407

Observation af4eac31-2d55-4676-85d0-3698fcd72da4 · outbound

This paper cites A Stochastic Approximation Method.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling A Stochastic Approximation Method

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:28.779551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:20.851505Z digest=sha256:885d3330d438b05fc33e23e61230178559019a201112808eb93af1e61c7cabe0

Observation 3d6e95a8-339c-4426-bc5a-d7d834dce895 · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling WinoGrande: An Adversarial Winograd Schema Challenge at Scale

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:21.016796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:21.016796Z digest=sha256:d0317abfab5793ae532616f3d0a4dcb6df66f2d27d9229265e4b73e31a6d5f21

Observation 9561ad24-354f-4081-a9a5-d15494df3ebf · outbound

This paper cites Towards Memory-Efficient Training for Extremely Large Output Spaces -- Learning with 500k Labels on a Single Commodity GPU.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Towards Memory-Efficient Training for Extremely Large Output Spaces -- Learning with 500k Labels on a Single Commodity GPU

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:25.765932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.148630Z digest=sha256:ef7c2a5c4207ff861e898b1f3608b9c638026e426e8f668d31658c56150f92a6

Observation 193ae81e-bd38-44a9-87b4-7858a2e693b8 · outbound

This paper cites an unresolved cited work.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:28.574747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.235606Z digest=sha256:1056f2ba0971ce47345e8cb2ab9e9cc834f3d7016f59ab042c5b00d8b18796eb

Observation aeb0dcd4-9a5e-4e31-ab05-4bde6923adfc · outbound

This paper cites Dynamic Sparse Training for Deep Reinforcement Learning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Dynamic Sparse Training for Deep Reinforcement Learning

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:25.537900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.282780Z digest=sha256:0fc0e8b43a35a984f4bb35ca1e41673e6d1a72c244efdec3a1f5a19201379473

Observation 082ecf70-70a8-4bef-9b68-6c286d644f3d · outbound

This paper cites RLx2: Training a Sparse Deep Reinforcement Learning Model from Scratch.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling RLx2: Training a Sparse Deep Reinforcement Learning Model from Scratch

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:25.215575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.344571Z digest=sha256:d31b744b8e11e670f0bb0436ccd1629f82d2caf668108b6c213f2530467a7042

Observation a06df54a-89cd-4602-842e-7f92ec4b5a76 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling LLaMA: Open and Efficient Foundation Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:21.435141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:21.435141Z digest=sha256:7f595f52aa54af44c0f9899ed472dd7eb21b5f326fddd2a29cabb2a709aec9fc

Observation ecf70624-3590-4484-9d8e-9ebc63156483 · outbound

This paper cites Varrette, H.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Varrette, H

Reference 66

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T14:44:24.896840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.512080Z digest=sha256:b54fb492888be67142a07920c4e2d5aa867acb4d76104e852d8b5ed8402c8170

Observation 5ecbf571-2046-4e99-b746-fac13ed7595b · outbound

This paper cites Attention Is All You Need.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Attention Is All You Need

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:21.615136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:21.615136Z digest=sha256:44ec81f043dca04193f1ce4a7c83875f2c60577ea8c90b1b4dda5af0a3af8cf2

Observation 609ad148-c3bc-42f5-9ba5-dd3c5f57f4aa · outbound

This paper cites Picking Winning Tickets Before Training by Preserving Gradient Flow.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Picking Winning Tickets Before Training by Preserving Gradient Flow

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:21.715735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:21.715735Z digest=sha256:9828fe2028baf89cca741f4f5b411e9bc8d913d360963b92c4b240d0fd986b87

Observation 7b458f2c-16c2-49af-baee-fd168ae858dc · outbound

This paper cites Learning Robust Global Representations by Penalizing Local Predictive Power.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Learning Robust Global Representations by Penalizing Local Predictive Power

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:44:28.359765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:21.820170Z digest=sha256:8e71744edbe97d5ca365427c6abcc617b5040a50b0e3aff98df5236a99f92adb

Observation 254e1cb9-0974-45c5-9563-18fcb67add63 · outbound

This paper cites BatchEnsemble: An Alternative Approach to Efficient Ensemble and Lifelong Learning.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling BatchEnsemble: An Alternative Approach to Efficient Ensemble and Lifelong Learning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:21.931636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:21.931636Z digest=sha256:f744d1ecff1677775c3a66f4cfb06674de2fb5f55f09b65dd494a0cbb8f6aff3

Observation 7a0f6a7e-ca27-44cb-bf6b-4861c2f88367 · outbound

This paper cites Nerva: a Truly Sparse Implementation of Neural Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Nerva: a Truly Sparse Implementation of Neural Networks

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:24.632843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.019881Z digest=sha256:e9c1d378ff5ba8757db45440032a6bdc6044ba1a0c8065c0badde2827dc7a8a9

Observation 80af6930-0542-4929-9201-1fa2278a87c8 · outbound

This paper cites Prune and Tune Ensembles: Low-Cost Ensemble Learning With Sparse Independent Subnetworks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Prune and Tune Ensembles: Low-Cost Ensemble Learning With Sparse Independent Subnetworks

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:24.414672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.092097Z digest=sha256:256bfa3585c8c1c4e531952a0be1e3a8cd1864e62d431293c90b4278015e46bd

Observation a24298ac-d6d7-44a8-8833-586e4296502e · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:22.165258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:22.165258Z digest=sha256:6fee877926c1d64f19fbbac21c3d05c9951a24d534bfd1041ad8cb854cf48cb8

Observation 7ee3a99d-23ef-40d2-b382-84c99634a2f6 · outbound

This paper cites an unresolved cited work.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:44:28.200572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.237158Z digest=sha256:a632c49cdf4b33573cf737c39a559aa53bd6fc02a12bd2f00ff0579ee0be36b4

Observation a93aa535-4cb6-49d9-a8d2-eb4c852f792c · outbound

This paper cites Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Dynamic Sparse Training versus Dense Training: The Unexpected Winner in Image Corruption Robustness

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:24.210999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.342479Z digest=sha256:d317df0f0aec5380bc37bec2fbe13dfef8eeef6bea2aa72a61b06d5745736d98

Observation 00c1b99f-a601-44c6-ae5a-b96b42b7662f · outbound

This paper cites Continual Learning with Dynamic Sparse Training: Exploring Algorithms for Effective Model Updates.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Continual Learning with Dynamic Sparse Training: Exploring Algorithms for Effective Model Updates

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:23.992254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.442515Z digest=sha256:0d6a602aaaf440b41a5586fed18b471d863cd20fcb49798e0302ed6cbc9563eb

Observation 83154707-d389-4e84-a03c-0e9a35a6b810 · outbound

This paper cites Outlier Weighed Layerwise Sparsity (OWL): A Missing Secret Sauce for Pruning LLMs to High Sparsity.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Outlier Weighed Layerwise Sparsity (OWL): A Missing Secret Sauce for Pruning LLMs to High Sparsity

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:22.533903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:22.533903Z digest=sha256:a2880417bcbbde4fa16551f4df9143465ad3753e1fe7b894ac6de45b98b5bec7

Observation 965697d7-2da6-4398-9ee5-3364f3fcd26c · outbound

This paper cites Drawing Early-Bird Tickets: Towards More Efficient Training of Deep Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Drawing Early-Bird Tickets: Towards More Efficient Training of Deep Networks

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:22.610368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:22.610368Z digest=sha256:875cae18fbbae9c1bfadeb6263fa0b4f86e41cbb2d0b0ed60086c8dabf92406f

Observation 66b6f0d6-7a53-4bbd-aff4-509c9269ea76 · outbound

This paper cites MEST: Accurate and Fast Memory-Economic Sparse Training Framework on the Edge.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling MEST: Accurate and Fast Memory-Economic Sparse Training Framework on the Edge

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:44:23.793230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:22.701532Z digest=sha256:e011a7904e381e84fa59f3b6676e76fb6dbb67776f4f7e82a372b2a54a540688

Observation ceba48ee-4491-4cff-9677-d2c3e4ec64ca · outbound

This paper cites Wide Residual Networks.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Wide Residual Networks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:22.864692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:22.864692Z digest=sha256:4d3e6c06173bff1f6951ef39c9933491ee1f57338dcdd3f69829f276d6cba7aa

Observation 3be92767-aab7-481d-9441-5efc1227d076 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling HellaSwag: Can a Machine Really Finish Your Sentence?

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:23.014536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:23.014536Z digest=sha256:9873894a0e3dfcc84f39f41fd2318b2f58d6d7655a60d306b24496caa1ce1d13

Observation 7b2f1288-742e-4992-bcec-2ad267819673 · outbound

This paper cites Brain-inspired sparse training enables Transformers and LLMs to perform as fully connected.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Brain-inspired sparse training enables Transformers and LLMs to perform as fully connected

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:23.115529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:23.115529Z digest=sha256:04f37ec304fa735068b6039280d400f68e14a46a89fa0b7be4b0fe777dbc7606

Observation 4f0d66af-7995-4241-a8f3-cbd11731efc1 · outbound

This paper cites GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:23.169462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:23.169462Z digest=sha256:b656169cb8e95ebbcbf8da3518782a3f08d0aa3e5bd9b1ce745663b2d72ca1b3

Observation fcc68dc2-e072-4c77-b691-d4632248eb78 · outbound

This paper cites Robust Lottery Tickets for Pre-trained Language Models.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Robust Lottery Tickets for Pre-trained Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:44:23.206606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:44:23.206606Z digest=sha256:64e3cb9de4c4727d0d0bc943369797e5da2a2bf67e1ccd4a5d4ce33e3631a449

Observation 64500586-55ab-44f4-9a0c-51c43c7515b2 · outbound

This paper cites Ensemble Methods: Foundations and Algorithms.

NeuroTrails: Training with Dynamic Sparse Heads as the Key to Effective Ensembling Ensemble Methods: Foundations and Algorithms

Reference 85

Resolution
verified exact
doi, observed 2026-08-07T14:44:23.274690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:44:23.210297Z digest=sha256:6f79da0010d22b7b0ab9d86ff492327c8a81ebe882e03a5122e546de18b7bcdb

Pith citing papers

No inbound Pith citation observations are available.