Pith. sign in

Paper Citation Record · LEDGER

ModRWKV: Transformer Multimodality in Linear Time

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2505.14505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14505 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:35:36.773332Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b546717e-ca45-4bcd-aad1-ab69ef9020d6 · outbound

This paper cites GPT-4 Technical Report.

ModRWKV: Transformer Multimodality in Linear Time GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.126173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.126173Z digest=sha256:3c42af423769ec7434012009ff12eaaecf9b1f25c1210a38af3324616222230f

Observation 66fbd247-4f4a-4426-9aa7-f660f41a03fd · outbound

This paper cites GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation.

ModRWKV: Transformer Multimodality in Linear Time GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.216986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.216986Z digest=sha256:05efdd634631872e9de7f24a3b6a5c852fcbc5682a07fcb0320e4be23b7cbefc

Observation aaf3b6e6-43e1-4be0-b3cf-23b4959b8af4 · outbound

This paper cites AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline.

ModRWKV: Transformer Multimodality in Linear Time AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.287638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.287638Z digest=sha256:9de3c9cc6201aad637e077687e0e20cdaf7273bc83dbae5f19a7ccb5cd37d9f0

Observation a56d8b10-f90f-478d-af02-3109d9e3aeda · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.359274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.359274Z digest=sha256:1eb7f02af98bc4ef253ea6eecc8dd84379427d1664f74d9a40b3cbf41b3d9b05

Observation 68a292bd-7419-452b-bf2c-153bb6cd0355 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

ModRWKV: Transformer Multimodality in Linear Time FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.436672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.436672Z digest=sha256:9e459d6923d37a37deaddd8b6780b9ba7ba7b3c7be6caf6ad9b85a532bbe30ca

Observation 026308a5-89d4-4120-a288-666b60c7ada6 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

ModRWKV: Transformer Multimodality in Linear Time Moshi: a speech-text foundation model for real-time dialogue

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.547184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.547184Z digest=sha256:c8d35a010f17072825af36ff34453e1261f8bf88601b575185af364e8bc7e876

Observation 180b0dc2-a501-411a-aaf2-b27ae8112244 · outbound

This paper cites LLaMA-Omni: Seamless Speech Interaction with Large Language Models.

ModRWKV: Transformer Multimodality in Linear Time LLaMA-Omni: Seamless Speech Interaction with Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.644120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.644120Z digest=sha256:84ba5d6f3ead14e6e87dd36c415f8e948efedb5aaf671150034b73d62d8a0083

Observation 86587d0c-b4a8-469a-8ce9-f18d85ca108f · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.910028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.682440Z digest=sha256:17a25f583587d565b683576072aef2f619c5f4178c72042da0e5ecbdbb16203c

Observation 95732f85-aeac-46e7-89c7-635d81b40959 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

ModRWKV: Transformer Multimodality in Linear Time Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.824926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.824926Z digest=sha256:be6d7736ba0026ffc4cc8f62e00416b0bc4cd56fb70de0d1fdfc419577b31d23

Observation e5727c08-58ed-4622-97b9-84fe7113640f · outbound

This paper cites Hudson and Christopher D.

ModRWKV: Transformer Multimodality in Linear Time Hudson and Christopher D

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.813500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.901043Z digest=sha256:cf0108b8f464554e3a5b522161c0a900a71fc7f740280f1964ad3d3ba9f5fb89

Observation b988c5e1-457a-475b-9319-7ed9de739c1e · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.688259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.952880Z digest=sha256:3fbc94f27dc13bf7b50999841f5b7e1e7b42190e0197d193ea9e73500d75c761

Observation 68e29713-22c3-4a41-ae97-812550c7d3bb · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

ModRWKV: Transformer Multimodality in Linear Time Improved Baselines with Visual Instruction Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.001638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.001638Z digest=sha256:dfd3ce7e4abc28bb596d9d4359abfb8d6c7502068a60d2f0ffc6f5c425e8510c

Observation 66f5a789-2b62-4282-a8dc-d9ed01b62343 · outbound

This paper cites Visual Instruction Tuning.

ModRWKV: Transformer Multimodality in Linear Time Visual Instruction Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.068537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.068537Z digest=sha256:d7a89b096aae92c9052c8846f84935d0859cb07d638c22bc8cc84c1b144934a6

Observation ff220dcf-3e84-4379-b56a-e4d24a4a7a77 · outbound

This paper cites Timer: Generative Pre-trained Transformers Are Large Time Series Models.

ModRWKV: Transformer Multimodality in Linear Time Timer: Generative Pre-trained Transformers Are Large Time Series Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.102882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.102882Z digest=sha256:740db3d383f9a02237f8ea1d82aaccf38765f3ff767bb5f1a7e56bee3144c5b0

Observation 45b8986b-d5e7-411d-aefe-7caad2c70782 · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player?.

ModRWKV: Transformer Multimodality in Linear Time MMBench: Is Your Multi-modal Model an All-around Player?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.218235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.218235Z digest=sha256:2bed46e41da37c2b195fd7d67fa84801ba7c8101c312253849f2dba43fb546ab

Observation dad9cf08-4adc-4f1c-bbdd-c7d0208045e7 · outbound

This paper cites Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.

ModRWKV: Transformer Multimodality in Linear Time Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.273568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.273568Z digest=sha256:11d0ad5ab3009fed6d922b3a02f763917c6077766c7f855006d7849e2ccfdb57

Observation e8b66aed-7666-4ed9-9aa0-81844d3afe77 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

ModRWKV: Transformer Multimodality in Linear Time An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.359733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.359733Z digest=sha256:bd1c47ca456a4d01e8187c12133ae241f5977ee42c6f1992600babde8a11e573

Observation 5acbf1af-5252-4cb4-b22b-84b7c0dd8907 · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.407445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.407445Z digest=sha256:eec9a30a641fca91bb80ad0915fcf81f570ab2dca0f05af79ad6117d6b031b14

Observation c385da79-6aec-4424-a7ce-7bd55bfc5b77 · outbound

This paper cites RWKV-7 "Goose" with Expressive Dynamic State Evolution.

ModRWKV: Transformer Multimodality in Linear Time RWKV-7 "Goose" with Expressive Dynamic State Evolution

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.454816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.454816Z digest=sha256:b227ce86da7ba7b9128271ad4bf2780397cd7e6d519bdd941ba2100f8416f090

Observation fcd07736-4447-44c9-a2d1-1e776e13e048 · outbound

This paper cites VL-Mamba: Exploring State Space Models for Multimodal Learning.

ModRWKV: Transformer Multimodality in Linear Time VL-Mamba: Exploring State Space Models for Multimodal Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.516406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.516406Z digest=sha256:5e2a55492f6bed8eb59cc1e3394bbeb440be594923081a76851dda904e2497e0

Observation 435ba784-5698-40cc-83aa-244e67889b97 · outbound

This paper cites TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods.

ModRWKV: Transformer Multimodality in Linear Time TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.607876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.607876Z digest=sha256:b508f3b14c2eff7c4922efaacdafc6a4e75fdebd634fac9b7451ff1e94e90937

Observation a8055a95-4747-40b5-aa74-3395dde6e504 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

ModRWKV: Transformer Multimodality in Linear Time Learning Transferable Visual Models From Natural Language Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.651289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.651289Z digest=sha256:1245b0035bc47b43ff0fec41c1742e8f3a12960845763afd4e61567f7016ffeb

Observation cf7e1600-435f-49ed-9a59-da803e35217e · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

ModRWKV: Transformer Multimodality in Linear Time Robust Speech Recognition via Large-Scale Weak Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.723258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.723258Z digest=sha256:e320efd327da90b7f26338ea9337a05b9574e0d39090dc0649cb41f31e4017a7

Observation 4c1dd459-c6c3-4cf6-8903-cfb0dd1c477c · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.539494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:35:35.778808Z digest=sha256:e9b2bfc76e51f3107c18e69a2e0565dd8df01b0097f72a26f89cbcf225717a46

Observation a1b88adf-195c-48af-a715-b2060f449af2 · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.428196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T15:35:35.813739Z digest=sha256:1654e6546d94a7c5d2b33b92f771cca62e65150fc9580422a06051487ccc81d3

Observation c9d84ef3-4007-4177-a1ea-59000541265d · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ModRWKV: Transformer Multimodality in Linear Time LLaMA: Open and Efficient Foundation Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.930937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.930937Z digest=sha256:bdeeba5aa9f341002396a54f2d8416ed964e99306259686d80cc19d45a45231d

Observation cf5cd08a-3d53-4918-a8d4-7d0a897fd5dd · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

ModRWKV: Transformer Multimodality in Linear Time SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.086113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.086113Z digest=sha256:e9e2c28e2c229dbb19c2c6608e7a6ffaee3e6b3a03f29e9d84fb5319b52fd5c2

Observation b631993f-1a1e-4a52-b85e-ca91850650a3 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

ModRWKV: Transformer Multimodality in Linear Time WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.141592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.141592Z digest=sha256:145624729c7ec321d03a644ebde3aa2d02eb260223fa6f67d8961c2a4798ded1

Observation e8b704b2-7307-4d79-a53b-a6c2869b8dac · outbound

This paper cites Attention Is All You Need.

ModRWKV: Transformer Multimodality in Linear Time Attention Is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.181136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.181136Z digest=sha256:b586c071bf50772d395a5c29581d1d3150ef56b0c1bcfd9553ab4d23f7c0ee4b

Observation dd249c0d-5472-4044-b85e-3abf2164512f · outbound

This paper cites Gated Linear Attention Transformers with Hardware-Efficient Training.

ModRWKV: Transformer Multimodality in Linear Time Gated Linear Attention Transformers with Hardware-Efficient Training

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.276021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.276021Z digest=sha256:b7a1598acb485629677dbb0690442ce082c8234df8a0146b36392e5a990ccc05

Observation 18f01554-4540-4ce8-be3f-b9fbe424a5b2 · outbound

This paper cites Parallelizing Linear Transformers with the Delta Rule over Sequence Length.

ModRWKV: Transformer Multimodality in Linear Time Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.402647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.402647Z digest=sha256:3dd38336a8fbb2f73b46493cdf9fa59652987a040b9a0a9dafde64f71a2b5af6

Observation 4ea73cad-5d08-4c8d-8e82-edd91e5c2160 · outbound

This paper cites StableMask: Refining Causal Masking in Decoder-only Transformer.

ModRWKV: Transformer Multimodality in Linear Time StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.506938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.506938Z digest=sha256:76de9d9b117feb1827731b94fd97c3a0f7464662f771f4c3af7ed1deef54dd47

Observation 3b34e047-9039-40f4-8926-623e91f137d6 · outbound

This paper cites MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI.

ModRWKV: Transformer Multimodality in Linear Time MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.589234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.589234Z digest=sha256:06ff42595a660acdc65b4133387ce5cecb9466d5a735e7fabaf1b2964ce943dc

Observation cd5e4df2-1903-40a5-8390-1f4227da5100 · outbound

This paper cites URL: " 'urlintro :=.

ModRWKV: Transformer Multimodality in Linear Time URL: " 'urlintro :=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.624042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.624042Z digest=sha256:e18ceabc9366a810b72e59ad0192805bfedcc932b290e72a67acc711cc2cd22d

Observation 4620b2ad-f0f8-4949-82dd-7dd0621115f8 · outbound

This paper cites write newline.

ModRWKV: Transformer Multimodality in Linear Time write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.773332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.773332Z digest=sha256:5df1ef77cd26723cb2e6e93550640bb71f8481650b30cc23e2d27255300ad2c7

Pith citing papers

No inbound Pith citation observations are available.