Pith. sign in

Paper Citation Record · LEDGER

Fast and Simplex: 2-Simplicial Attention in Triton

As of 19 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 4 inbound Pith citation observations for arXiv:2507.02754.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02754 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:31:48.003056Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:45:12.482381Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T18:45:58.257917Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4d71038d-1818-4c1e-9606-bde8cfe3d333 · outbound

This paper cites GPT-4 Technical Report.

Fast and Simplex: 2-Simplicial Attention in Triton GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.279796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.279796Z digest=sha256:a6de7a9d0be218a5ce0474975708ffddf933ed122dfd27c04a294e80eb990e6f

Observation 26f31ce2-e56f-4824-a351-c7007b99f718 · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

Fast and Simplex: 2-Simplicial Attention in Triton GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.332972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.332972Z digest=sha256:dcca3b417878ae9cba41bff5e564bb85c4e257de33390d6a3e3ab48e0c8c24e3

Observation 3487f141-73db-4bc4-953c-691206c276e1 · outbound

This paper cites Program Synthesis with Large Language Models.

Fast and Simplex: 2-Simplicial Attention in Triton Program Synthesis with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.413537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.413537Z digest=sha256:65e7fdfee26194439f3c5571e09a2b0c8b25e7e24088be245d20fbea53f86ba1

Observation 1a55b8dd-209d-48a1-836f-067391f31903 · outbound

This paper cites Explaining neural scaling laws.

Fast and Simplex: 2-Simplicial Attention in Triton Explaining neural scaling laws

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.742807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.525998Z digest=sha256:29d6fdd929c6c419f4bc69c3c36d1b9b67a54edbc4c0222fa963964389f28b59

Observation cace0b4e-ac7b-4e81-bbdd-77198cf53445 · outbound

This paper cites Systematic generalization with edge transformers.

Fast and Simplex: 2-Simplicial Attention in Triton Systematic generalization with edge transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.730842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.581513Z digest=sha256:a45ba1a17a3be70931050d80b8f99d04b44276d641b99cd7a14c5a39a1215277

Observation 7f6d9595-27b7-498f-bc6f-105d6625a25d · outbound

This paper cites Loss-to-Loss Prediction: Scaling Laws for All Datasets.

Fast and Simplex: 2-Simplicial Attention in Triton Loss-to-Loss Prediction: Scaling Laws for All Datasets

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.679967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.679967Z digest=sha256:72c8e7b0695a7a6ed01e396d6c958457b9f5994f07d9fe3a5c81b49ac9549262

Observation 1361d9f1-1229-4ba0-8034-28e3ec8c3367 · outbound

This paper cites Language models are few-shot learners.

Fast and Simplex: 2-Simplicial Attention in Triton Language models are few-shot learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.764549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.764549Z digest=sha256:2cb83f05913ee9bfc9f70c15d7b9f140b185e3285fbc8cfa2df4c775536e0fa9

Observation 0595b5d2-7b5f-46f4-a0f7-703bf88c1c54 · outbound

This paper cites Logic and the $2$-Simplicial Transformer.

Fast and Simplex: 2-Simplicial Attention in Triton Logic and the $2$-Simplicial Transformer

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.863926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.863926Z digest=sha256:424d991f3ef7e57c47854542f6a57b7e652e9c99d967175156347fd080c924ac

Observation 8932e247-1acf-4c51-8330-12fedff90499 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fast and Simplex: 2-Simplicial Attention in Triton Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.868300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.868300Z digest=sha256:5cdf733160445f2173d3737c638ca62da31f668218112b8276bb25c884ae75d9

Observation c52182ce-c3be-4338-9491-525706a39185 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness.

Fast and Simplex: 2-Simplicial Attention in Triton Flashattention: Fast and memory-efficient exact attention with io-awareness

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.872465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.872465Z digest=sha256:2d8d6a0ba33e9229e1902b84538f207c458798c6dfa77bb1e3a7b38964a028ba

Observation 93d1f437-cf8d-4e63-8be3-8c3f65402352 · outbound

This paper cites Universal Transformers.

Fast and Simplex: 2-Simplicial Attention in Triton Universal Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.876295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.876295Z digest=sha256:106b6f726900845e9ec2fa96a360f7637adb1b4729f0601f520c2a9f0d9deae4

Observation 406ebcc6-3588-47c9-a304-295ba6f9595d · outbound

This paper cites Observation on scaling laws, May 2025.

Fast and Simplex: 2-Simplicial Attention in Triton Observation on scaling laws, May 2025

Reference 12

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:31:48.443815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.880210Z digest=sha256:212367d0e1e3dfa94624748c75de27d841b04ef0372f47460e4bafc2a4f7d725

Observation 4002c051-2c80-480e-a493-7b6ee471d278 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Fast and Simplex: 2-Simplicial Attention in Triton Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.884864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.884864Z digest=sha256:16adb377a578f95dee993c0ee5811a90ca4cb382ad46004f1b7729fde354ef4a

Observation 4c514223-5d96-4024-becc-8bfcd87ce561 · outbound

This paper cites Array programming with numpy.

Fast and Simplex: 2-Simplicial Attention in Triton Array programming with numpy

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.703801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.888566Z digest=sha256:faa8c10fc122ceb5adef674afbc79d399801ba87ab536b411c6ec61be7a0054b

Observation 4fdb58ff-8252-4600-8029-4acf51ec3f0c · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Fast and Simplex: 2-Simplicial Attention in Triton Measuring Massive Multitask Language Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.892175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.892175Z digest=sha256:b979d8e3ea6fdd8cd03480543b53366564c8aa56748c208dcb95dbd3cbf69d0e

Observation 376662f9-ce1a-4eff-a1b6-398b00adfed2 · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

Fast and Simplex: 2-Simplicial Attention in Triton Deep Learning Scaling is Predictable, Empirically

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.896182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.896182Z digest=sha256:caa6f9747f312e2169a90cf047569d059997097ea042b21a7d7329cf1d6c0bf2

Observation 13036418-3af6-4cc3-afd8-77753e7ca1f7 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Fast and Simplex: 2-Simplicial Attention in Triton Training Compute-Optimal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.900090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.900090Z digest=sha256:0b71b3697ac8ad1b58efafe052b7574d78ae6e78e70c70c53b59c9b82accdb61

Observation aa2cb402-81a2-49e2-a6c6-a22fbf0238a5 · outbound

This paper cites Gpipe: Efficient training of giant neural networks using pipeline parallelism.

Fast and Simplex: 2-Simplicial Attention in Triton Gpipe: Efficient training of giant neural networks using pipeline parallelism

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.903756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.903756Z digest=sha256:cede0d59a0a8e755ff0f4ab0dfa656c8761fd93fb73bc28be73791694b63c95a

Observation f74d41ae-9f75-40f5-8469-8ffdf781a5d9 · outbound

This paper cites Hierarchical mixtures of experts and the em algorithm.

Fast and Simplex: 2-Simplicial Attention in Triton Hierarchical mixtures of experts and the em algorithm

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.907332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.907332Z digest=sha256:59a6cd28f33fd80bf8e38b4d45564ff2eee3349c578f477279d9272add7db6c2

Observation d1e28537-9b05-40cc-ab23-a18704ba6fde · outbound

This paper cites Highly accurate protein structure prediction with alphafold.

Fast and Simplex: 2-Simplicial Attention in Triton Highly accurate protein structure prediction with alphafold

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.910669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.910669Z digest=sha256:b26397dbf33f79532f9d430868edc2d98079943ff7f8c1c664f6ab5abac4e755

Observation 1a5b0c91-caab-4467-bb19-e891a840abca · outbound

This paper cites Scaling Laws for Neural Language Models.

Fast and Simplex: 2-Simplicial Attention in Triton Scaling Laws for Neural Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.913958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.913958Z digest=sha256:4b15be6def1a2d0dfab43663db7f8a509ffcb5b0bf1c89625198601139f625dc

Observation cbaf4031-3885-4590-b9fd-02197c6f2046 · outbound

This paper cites Transformers are rnns: fast autoregressive transformers with linear attention.

Fast and Simplex: 2-Simplicial Attention in Triton Transformers are rnns: fast autoregressive transformers with linear attention

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.667478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.917978Z digest=sha256:9b9ed26756b64b8b3810d94a25889f3877966b95991b7d229cc97a433d2e0964

Observation 69a3d88a-2181-42f3-9a96-0f1573a004b0 · outbound

This paper cites Strassen attention: Unlocking compositional abilities in transformers based on a new lower bound method.

Fast and Simplex: 2-Simplicial Attention in Triton Strassen attention: Unlocking compositional abilities in transformers based on a new lower bound method

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.921753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.921753Z digest=sha256:0997a7624cfb8628936d2f4b1a51ba5d91a1b3f3db6755731164cd2c2905fc0d

Observation 1f0d2026-bec1-4d54-baa4-cba9a41b96f9 · outbound

This paper cites Decoupled Weight Decay Regularization.

Fast and Simplex: 2-Simplicial Attention in Triton Decoupled Weight Decay Regularization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.925509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.925509Z digest=sha256:114954e0bbc465759b61e0f7eaa63198319fa10fc87ef75f8e09f5b0cabec897

Observation dbab16ea-a0fa-4f35-93fc-82d5fe2344cf · outbound

This paper cites Devanur, Gregory R.

Fast and Simplex: 2-Simplicial Attention in Triton Devanur, Gregory R

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.929662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.929662Z digest=sha256:42c13b4a5fc28bf0431e731eff2ebd1d4718c5e4998df8a76ec72487aa956811

Observation 05c211e6-ce1f-4e3a-a243-75e3904e0f3c · outbound

This paper cites Image transformer.

Fast and Simplex: 2-Simplicial Attention in Triton Image transformer

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.655195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.933422Z digest=sha256:3659977134ac02311473a4c72178ced41ecec6469e0bd58e7b2f99172d74da45

Observation 612c1e31-617f-40b2-8acf-a82378020f05 · outbound

This paper cites Efficient content-based sparse attention with routing transformers.

Fast and Simplex: 2-Simplicial Attention in Triton Efficient content-based sparse attention with routing transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.936863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.936863Z digest=sha256:bbd7b9f98270251a7352d7d5df6659f4be0d9fcd4dc4dea7e8a2790d61b48c2b

Observation 3b78d1ba-6456-44db-820f-2801cdf63c95 · outbound

This paper cites N-Grammer: Augmenting Transformers with latent n-grams.

Fast and Simplex: 2-Simplicial Attention in Triton N-Grammer: Augmenting Transformers with latent n-grams

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.940113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.940113Z digest=sha256:ab60bb56d26d4026ce20304b9912236dc31b4d58791db78483db6711253a7f25

Observation dd73ca29-adb4-4334-bf34-4ebf9ef17a3c · outbound

This paper cites Representational strengths and limitations of transformers.

Fast and Simplex: 2-Simplicial Attention in Triton Representational strengths and limitations of transformers

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.635934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.943571Z digest=sha256:f64199769ac46204409ecdabf59ae6c485a149c9e4ec8c4535c81a018190a460

Observation 7cb73c37-19bc-4b79-865f-0299409f9023 · outbound

This paper cites Reasoning with Latent Thoughts: On the Power of Looped Transformers.

Fast and Simplex: 2-Simplicial Attention in Triton Reasoning with Latent Thoughts: On the Power of Looped Transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.946894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.946894Z digest=sha256:20bf2260833829cf3a4d21286c1a32f9ded58f41c8069d736dc7fc58c43cc777

Observation 46102a7f-c4b2-430e-a1bb-0552fd654313 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Fast and Simplex: 2-Simplicial Attention in Triton Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.950381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.950381Z digest=sha256:6bd9da8259ffa58ab7d58eda5ba5eb9418a1645fe7572580c2b73f08e1db35af

Observation df7e133b-f805-4a35-81f7-064cee55d01d · outbound

This paper cites Scaling Laws for Linear Complexity Language Models.

Fast and Simplex: 2-Simplicial Attention in Triton Scaling Laws for Linear Complexity Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.953763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.953763Z digest=sha256:15e4b58ff2999acf557cc5c6408a1da87548fada0c507b39ce800f643ac6e21b

Observation bdf97131-425f-473d-9128-8e70286c6a36 · outbound

This paper cites Searching for efficient transformers for language modeling.

Fast and Simplex: 2-Simplicial Attention in Triton Searching for efficient transformers for language modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.624188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.957258Z digest=sha256:6a639517d9c0e05bc9c9741e3783ca211ae311829d0d4a4756406b8edfa8eabe

Observation 6801fb92-50a7-4266-800d-26187a2e3a49 · outbound

This paper cites Beyond neural scaling laws: beating power law scaling via data pruning.

Fast and Simplex: 2-Simplicial Attention in Triton Beyond neural scaling laws: beating power law scaling via data pruning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.961505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.961505Z digest=sha256:73be8a2037d1726428e76fb6aa8f0951204c27bbe5d82749951eade4c54bdf28

Observation b0fad808-806b-42f6-b332-7e33fd351338 · outbound

This paper cites Introduction to linear algebra.

Fast and Simplex: 2-Simplicial Attention in Triton Introduction to linear algebra

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.604410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.965176Z digest=sha256:13d96a965655a58beae0a92cfeb8b7e195d07f8aeb32c7d7780c76e8d7989a52

Observation 13fe905a-2e17-408a-a2f3-3951992d2336 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

Fast and Simplex: 2-Simplicial Attention in Triton Roformer: Enhanced transformer with rotary position embedding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.968446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.968446Z digest=sha256:3f6d7ebe1c89d921f8b9f264ebcd1759abc25b6007ee1e54e6cbc0f9e630d45b

Observation b5793db7-c1be-46d9-b7f8-113e2647b778 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Fast and Simplex: 2-Simplicial Attention in Triton Gemini: A Family of Highly Capable Multimodal Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.972385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.972385Z digest=sha256:a057187589a5608f141464275b8fddf9b5b212383b9cade050e0ececa939efc5

Observation cf94e19e-fde5-4398-a263-95f5696d1932 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Fast and Simplex: 2-Simplicial Attention in Triton LLaMA: Open and Efficient Foundation Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.975851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.975851Z digest=sha256:dc9271365456c589768588ea6e0f15b050aa359eb2289b4c09d1f1daa2e667b1

Observation 9bd4d8bf-6a49-442d-af78-ccc248a73793 · outbound

This paper cites On the uniform convergence of relative frequencies of events to their probabilities.

Fast and Simplex: 2-Simplicial Attention in Triton On the uniform convergence of relative frequencies of events to their probabilities

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.585422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.979234Z digest=sha256:345b792d759de6cbeeb23f18f6fe78ff598488b0dc791a6afc0012ecbe2a5426

Observation 194abe7a-bc53-4dae-a01f-f76f132f7c73 · outbound

This paper cites Attention is all you need.

Fast and Simplex: 2-Simplicial Attention in Triton Attention is all you need

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.982421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.982421Z digest=sha256:3854ea1467b3cd7bac7e8923d35b3b1e2626e57db7d37922eb7e3dce4d0b3f51

Observation 0e702b2e-fec5-4033-b4d5-6438bb9c529d · outbound

This paper cites Dcn v2: Improved deep & cross network and practical lessons for web-scale learning to rank systems.

Fast and Simplex: 2-Simplicial Attention in Triton Dcn v2: Improved deep & cross network and practical lessons for web-scale learning to rank systems

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:31:48.565542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-06T20:31:47.985661Z digest=sha256:b8e9243021851ada3d9dce8428e9e538d413b4e4d8bc92b72758964709c040c1

Observation 7c70767d-2472-41d1-b22e-34022cde12bc · outbound

This paper cites Mmlu-pro: A more robust and challenging multi-task language understanding benchmark.

Fast and Simplex: 2-Simplicial Attention in Triton Mmlu-pro: A more robust and challenging multi-task language understanding benchmark

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.989049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.989049Z digest=sha256:9cba3531df241931c19d4de5fe49c33b1a4e99d96e42336dc42b149e3c13619b

Observation 7baf9805-b726-47c2-987a-4119f5be4ac3 · outbound

This paper cites Looped Transformers are Better at Learning Learning Algorithms.

Fast and Simplex: 2-Simplicial Attention in Triton Looped Transformers are Better at Learning Learning Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.992493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.992493Z digest=sha256:28c5e9a80d7443f9c8d71253c6312982788a2f016b50e603a23ae1fe2bd4944d

Observation c5627278-022b-4e60-a142-af34a25a09b2 · outbound

This paper cites Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention.

Fast and Simplex: 2-Simplicial Attention in Triton Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.995786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.995786Z digest=sha256:c6c32e413c027c4930a5838af95c018d075266332d5cec74f7dee8380494b854

Observation 156c8414-88fe-4294-8961-e7f04c55fb01 · outbound

This paper cites Big bird: Transformers for longer sequences.

Fast and Simplex: 2-Simplicial Attention in Triton Big bird: Transformers for longer sequences

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:47.999795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:47.999795Z digest=sha256:b962825242f8d427f504f08e6dbd1584a37942082615ea38c4147c1876b20481

Observation d01a39a2-e8f4-4d62-ae49-70fe1aa6f144 · outbound

This paper cites write newline.

Fast and Simplex: 2-Simplicial Attention in Triton write newline

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:48.003056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:48.003056Z digest=sha256:b0e1cd038c81aa5e1e0ff8288700215175f06c4aeaa3879dd8277ae5f15c601f

Pith citing papers

Observation efe69367-f1cc-4101-9158-10098582d9df · inbound

TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments cites this paper.

TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments Fast and Simplex: 2-Simplicial Attention in Triton

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.773147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T03:55:15.475044Z digest=sha256:9f63ba4f206987689933207bae3dc67b2927894522700e195f0b8f8410912844

Observation b79e26a7-55a5-43c7-8f36-65502312a56e · inbound

TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments cites this paper.

TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments Fast and Simplex: 2-Simplicial Attention in Triton

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:09:44.705534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T05:08:56.760052Z digest=sha256:ed5183ab2c2a969e1004b6ab314bcdef5e31b012b57afa3bab4050f232670247

Observation d4ed6c44-3b7e-48f2-a18e-c2f7402a604c · inbound

Higher-Order Fourier Neural Operator: Explicit Mode Mixer for Nonlinear PDEs cites this paper.

Higher-Order Fourier Neural Operator: Explicit Mode Mixer for Nonlinear PDEs Fast and Simplex: 2-Simplicial Attention in Triton

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T18:45:58.259504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-29T01:49:08.216819Z digest=sha256:9d741089c7684efeaba36f4789d7232e7a3fa599d468ab3d0480a20137ce0f88

Observation 245dc3a0-35ab-45f4-9473-e9397413bd63 · inbound

Linearized 2-Simplicial Attention cites this paper.

Linearized 2-Simplicial Attention Fast and Simplex: 2-Simplicial Attention in Triton

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:45:12.482381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:45:12.482381Z digest=sha256:1dfd4085f14649dc2cc0db227709659d27dcf90525ad37fe03a445194721a233