Pith. sign in

Paper Citation Record · LEDGER

ModRWKV: Transformer Multimodality in Linear Time

As of 18 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2505.14505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14505 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:35:36.773332Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b546717e-ca45-4bcd-aad1-ab69ef9020d6 · outbound

This paper cites GPT-4 Technical Report.

ModRWKV: Transformer Multimodality in Linear Time GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.126173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.126173Z digest=sha256:676c2b7ecced2abaf777a72f4629cce365fa654385c2731851d477961390dc5d

Observation 66fbd247-4f4a-4426-9aa7-f660f41a03fd · outbound

This paper cites GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation.

ModRWKV: Transformer Multimodality in Linear Time GIFT-Eval: A Benchmark For General Time Series Forecasting Model Evaluation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.216986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.216986Z digest=sha256:4368b4983554de27360d24518fa372ff685eefad5319b82b74aa4d55b3c7a0d1

Observation aaf3b6e6-43e1-4be0-b3cf-23b4959b8af4 · outbound

This paper cites AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline.

ModRWKV: Transformer Multimodality in Linear Time AISHELL-1: An Open-Source Mandarin Speech Corpus and A Speech Recognition Baseline

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.287638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.287638Z digest=sha256:630d8ccb42887a33406bb6673db061c87dfb5696ec2c7843d991d8b1f910b6fc

Observation a56d8b10-f90f-478d-af02-3109d9e3aeda · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.359274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.359274Z digest=sha256:151314664ea3046bc71b513be14a80ab315bcdf8818da64d970aa3dd7b574919

Observation 68a292bd-7419-452b-bf2c-153bb6cd0355 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

ModRWKV: Transformer Multimodality in Linear Time FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.436672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.436672Z digest=sha256:d2807552aa16217592c8646181b8a5867fc2b59d5c18289dc1adbd230ff03e80

Observation 026308a5-89d4-4120-a288-666b60c7ada6 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

ModRWKV: Transformer Multimodality in Linear Time Moshi: a speech-text foundation model for real-time dialogue

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.547184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.547184Z digest=sha256:d12958df9ccb315c14ec191a5d981c68d5dbd7ac3d80f80f86948e2746bd6917

Observation 180b0dc2-a501-411a-aaf2-b27ae8112244 · outbound

This paper cites LLaMA-Omni: Seamless Speech Interaction with Large Language Models.

ModRWKV: Transformer Multimodality in Linear Time LLaMA-Omni: Seamless Speech Interaction with Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.644120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.644120Z digest=sha256:3f03511ada46f8fb5c8a4db7abec4d752bc0d9ea5cb902af674dd87982bba77a

Observation 86587d0c-b4a8-469a-8ce9-f18d85ca108f · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.910028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.682440Z digest=sha256:f6d1f854b8db5fd551c27354127641e8937529f12669a545b6e1798cd5800e96

Observation 95732f85-aeac-46e7-89c7-635d81b40959 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

ModRWKV: Transformer Multimodality in Linear Time Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:34.824926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:34.824926Z digest=sha256:fd37ed34126f6c63eeaa7b0c224bc9a8917c29d6d4b7d8d57f1ffd2494c48173

Observation e5727c08-58ed-4622-97b9-84fe7113640f · outbound

This paper cites Hudson and Christopher D.

ModRWKV: Transformer Multimodality in Linear Time Hudson and Christopher D

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:35:37.813500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.901043Z digest=sha256:2328d783d9ca58116753c11c8968323a173e1fe0fbff0432cb1dbfb40575c724

Observation b988c5e1-457a-475b-9319-7ed9de739c1e · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.688259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:35:34.952880Z digest=sha256:75de260af88b38e2178cb508093ae125f88d6dd42153e5fbb5ae25f0732181c6

Observation 68e29713-22c3-4a41-ae97-812550c7d3bb · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

ModRWKV: Transformer Multimodality in Linear Time Improved Baselines with Visual Instruction Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.001638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.001638Z digest=sha256:0bfece5fb95a0c76eb9f581bd63b0c0a5525f2f37c33b5b74cc4c50481d0dec2

Observation 66f5a789-2b62-4282-a8dc-d9ed01b62343 · outbound

This paper cites Visual Instruction Tuning.

ModRWKV: Transformer Multimodality in Linear Time Visual Instruction Tuning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.068537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.068537Z digest=sha256:1d253a1e775baf54865f7da88eb1d34e3d598b7ce7da1b6911775e8b2bd09863

Observation ff220dcf-3e84-4379-b56a-e4d24a4a7a77 · outbound

This paper cites Timer: Generative Pre-trained Transformers Are Large Time Series Models.

ModRWKV: Transformer Multimodality in Linear Time Timer: Generative Pre-trained Transformers Are Large Time Series Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.102882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.102882Z digest=sha256:294d5cf07cef575dbeab5289a4f0e32dfddae4213e3c66c5fc2d8a29badf2d66

Observation 45b8986b-d5e7-411d-aefe-7caad2c70782 · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player?.

ModRWKV: Transformer Multimodality in Linear Time MMBench: Is Your Multi-modal Model an All-around Player?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.218235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.218235Z digest=sha256:0cf4514328ee6411f9869693fd3b24c493f81f93b67dadef2420dd27dbf45aff

Observation dad9cf08-4adc-4f1c-bbdd-c7d0208045e7 · outbound

This paper cites Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.

ModRWKV: Transformer Multimodality in Linear Time Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.273568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.273568Z digest=sha256:78ff4f84791747ad7352761ed7f8a49106cc4d42221e466dc350eba0244fcbc7

Observation e8b66aed-7666-4ed9-9aa0-81844d3afe77 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

ModRWKV: Transformer Multimodality in Linear Time An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.359733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.359733Z digest=sha256:67fc290ca7b11e706d140f7e70df02287b98cea2917ba25deaad4ee7e9abeecc

Observation 5acbf1af-5252-4cb4-b22b-84b7c0dd8907 · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.407445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.407445Z digest=sha256:f9cee4469918242f8d8b95605b4deb2dca884d01d7252c1417ef987ff26cb146

Observation c385da79-6aec-4424-a7ce-7bd55bfc5b77 · outbound

This paper cites RWKV-7 "Goose" with Expressive Dynamic State Evolution.

ModRWKV: Transformer Multimodality in Linear Time RWKV-7 "Goose" with Expressive Dynamic State Evolution

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.454816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.454816Z digest=sha256:93cd63f04965fb5fd8a0efc2eb4ffe44fde90b636b8e273968f1423e23219160

Observation fcd07736-4447-44c9-a2d1-1e776e13e048 · outbound

This paper cites VL-Mamba: Exploring State Space Models for Multimodal Learning.

ModRWKV: Transformer Multimodality in Linear Time VL-Mamba: Exploring State Space Models for Multimodal Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.516406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.516406Z digest=sha256:06462c1ade0b5b4b46f3300f0dac2478fded92b250a33adf75a5a86af41bce37

Observation 435ba784-5698-40cc-83aa-244e67889b97 · outbound

This paper cites TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods.

ModRWKV: Transformer Multimodality in Linear Time TFB: Towards Comprehensive and Fair Benchmarking of Time Series Forecasting Methods

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.607876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.607876Z digest=sha256:df010fa22cb2fae9def94689cb7e4d3a495f1664eabcedfa85af93a49129a0fc

Observation a8055a95-4747-40b5-aa74-3395dde6e504 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

ModRWKV: Transformer Multimodality in Linear Time Learning Transferable Visual Models From Natural Language Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.651289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.651289Z digest=sha256:8c2395f8f15f0b821184612279b2f08e1f31d67202319e046acb56fe681edf14

Observation cf7e1600-435f-49ed-9a59-da803e35217e · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

ModRWKV: Transformer Multimodality in Linear Time Robust Speech Recognition via Large-Scale Weak Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.723258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.723258Z digest=sha256:64915e658217c3e53f1bb106c71710eaee320b02e7ba198f8aa9fd879bad05b9

Observation 4c1dd459-c6c3-4cf6-8903-cfb0dd1c477c · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.539494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:35:35.778808Z digest=sha256:edd9012f2772cd76984b693ad91fd03579592ebddeca90c897a8aa9c043714eb

Observation a1b88adf-195c-48af-a715-b2060f449af2 · outbound

This paper cites an unresolved cited work.

ModRWKV: Transformer Multimodality in Linear Time Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:35:37.428196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:35:35.813739Z digest=sha256:afcf4cd72cb80bb35cad8586f79f654a2350464e45fd65a5186bc2c7f072db05

Observation c9d84ef3-4007-4177-a1ea-59000541265d · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ModRWKV: Transformer Multimodality in Linear Time LLaMA: Open and Efficient Foundation Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:35.930937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:35.930937Z digest=sha256:34db169787f9ff1da5680108740c92b720a59fed5fa3a027e422ed8506f4ea92

Observation cf5cd08a-3d53-4918-a8d4-7d0a897fd5dd · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

ModRWKV: Transformer Multimodality in Linear Time SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.086113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.086113Z digest=sha256:f37ba1b81d0cdcb4518674f35c495e43ea85f8dabee7155f9fc943c9de15b768

Observation b631993f-1a1e-4a52-b85e-ca91850650a3 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

ModRWKV: Transformer Multimodality in Linear Time WaveNet: A Generative Model for Raw Audio

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.141592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.141592Z digest=sha256:ac73071287941a5a8b2b9ad2f50d1a94af5859672eae21297478027bfecc044c

Observation e8b704b2-7307-4d79-a53b-a6c2869b8dac · outbound

This paper cites Attention Is All You Need.

ModRWKV: Transformer Multimodality in Linear Time Attention Is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.181136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.181136Z digest=sha256:d1971a4ce3301642b63894f5a8ffd76ed3bac04d35ae7537345957aa80455e57

Observation dd249c0d-5472-4044-b85e-3abf2164512f · outbound

This paper cites Gated Linear Attention Transformers with Hardware-Efficient Training.

ModRWKV: Transformer Multimodality in Linear Time Gated Linear Attention Transformers with Hardware-Efficient Training

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.276021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.276021Z digest=sha256:b5075c07e825775440925a08c1000df7258d54fca7c5f7680f79f6bc5e394c28

Observation 18f01554-4540-4ce8-be3f-b9fbe424a5b2 · outbound

This paper cites Parallelizing Linear Transformers with the Delta Rule over Sequence Length.

ModRWKV: Transformer Multimodality in Linear Time Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.402647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.402647Z digest=sha256:3284ea331d57ee5b44a36d179eb797a8fa4b5def92713e34da1ef95897afc2a7

Observation 4ea73cad-5d08-4c8d-8e82-edd91e5c2160 · outbound

This paper cites StableMask: Refining Causal Masking in Decoder-only Transformer.

ModRWKV: Transformer Multimodality in Linear Time StableMask: Refining Causal Masking in Decoder-only Transformer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.506938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.506938Z digest=sha256:aeadfef8baffaba3c6d59ec6274cbbf6ad9aa477956cb6bf11b59d764a038b90

Observation 3b34e047-9039-40f4-8926-623e91f137d6 · outbound

This paper cites MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI.

ModRWKV: Transformer Multimodality in Linear Time MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.589234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.589234Z digest=sha256:e67c568981ffe4f83a533adc6ad4a340cae9eaad6fae9f6698b2ce38539daf51

Observation cd5e4df2-1903-40a5-8390-1f4227da5100 · outbound

This paper cites URL: " 'urlintro :=.

ModRWKV: Transformer Multimodality in Linear Time URL: " 'urlintro :=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.624042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.624042Z digest=sha256:58f8130ad70973734ca9f1d4ffde0fe4a8736e381aea803d8787f638f924a939

Observation 4620b2ad-f0f8-4949-82dd-7dd0621115f8 · outbound

This paper cites write newline.

ModRWKV: Transformer Multimodality in Linear Time write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.773332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.773332Z digest=sha256:2b5a23292df3ca7179661f597b50ab98d550821927fef2a92d7a8924a8c3839a

Pith citing papers

No inbound Pith citation observations are available.