Pith. sign in

Paper Citation Record · LEDGER

Kimi Linear: An Expressive, Efficient Attention Architecture

As of 4 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 81 inbound Pith citation observations for arXiv:2510.26692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.26692 v2

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T23:49:10.555255Z

measured 181 of 181 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 81 of 81 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T06:42:11.328087Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T14:57:14.481929Z

Reference resolution

100 of 129 outbound references displayed

  • verified exact64
  • verified fuzzy22
  • unresolved5
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch7

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75a27a1d-a0e5-4c0a-8d8f-60cfd9baf605 · outbound

This paper cites gpt-oss-120b & gpt-oss-20b Model Card.

Kimi Linear: An Expressive, Efficient Attention Architecture gpt-oss-120b & gpt-oss-20b Model Card

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.690152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:16cf0cbfc0b7936c5dbe4aed44e89e7ef9dd654f1cf1a07c9190aafb13c25f92

Observation f2d765b8-51d0-48cd-a621-a28bd003202a · outbound

This paper cites CoLT5: Faster Long-Range Transformers with Conditional Computation.

Kimi Linear: An Expressive, Efficient Attention Architecture CoLT5: Faster Long-Range Transformers with Conditional Computation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:10.697142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:c04e323b14bc5c2c336fd71b727a1fb4377220e492a1433bea1fb2eac5a884f7

Observation cc70616c-5253-48ff-8a81-228a14e420d9 · outbound

This paper cites Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers.

Kimi Linear: An Expressive, Efficient Attention Architecture Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers

Reference 3

Resolution
verified exact
doi, observed 2026-05-13T23:49:10.655674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:036d7ebb71c3226856de53e9f3afb2ac4afb9a8130cd8b0ca4189819dad4d7bc

Observation 5d1858b2-c388-4cfa-830b-efa7740dcf43 · outbound

This paper cites Simple linear attention language models balance the recall-throughput tradeoff.

Kimi Linear: An Expressive, Efficient Attention Architecture Simple linear attention language models balance the recall-throughput tradeoff

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.437770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:d914a57ada3adcc5c39ea69e54789f47e3df84b71105766722c84e0f4c6d5f79

Observation 2ddd84e6-3e43-4f36-a05c-f83b5bde0427 · outbound

This paper cites Zoology: Measuring and Improving Recall in Efficient Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Zoology: Measuring and Improving Recall in Efficient Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.704970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:73e4af6813fa8c55fd6c371e98841429b52705efe5b4527106b5fe12c047f297

Observation 636a967b-4c93-4789-a3fc-863072673810 · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

Kimi Linear: An Expressive, Efficient Attention Architecture LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:38:00.425257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:b0e771d66693a4e6bb86e6443674cabbe9fe60ab01df589f7f52c18bd8f4424c

Observation 4dfe037c-6222-4d49-aa87-1d2c90222dfa · outbound

This paper cites Round and Round We Go! What makes Rotary Positional Encodings useful?.

Kimi Linear: An Expressive, Efficient Attention Architecture Round and Round We Go! What makes Rotary Positional Encodings useful?

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.335981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:e8c7a633fa459950d2df720fc1ef4154101f6b9ef589f95eea19429345cadc4b

Observation 0a88c6d3-38f2-435a-81e3-042c953e7f23 · outbound

This paper cites ATLAS: Learning to Optimally Memorize the Context at Test Time.

Kimi Linear: An Expressive, Efficient Attention Architecture ATLAS: Learning to Optimally Memorize the Context at Test Time

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.676753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:1b27a7a18354b5fb9324a4544c14caa08d097cc4d7c782545abe5cf5517ab99d

Observation 2717f60e-5f6c-49f3-b72f-aa29a46d53e9 · outbound

This paper cites Unlimiformer: Long-range transformers with unlimited length input.

Kimi Linear: An Expressive, Efficient Attention Architecture Unlimiformer: Long-range transformers with unlimited length input

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.267868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:84df9b29b405bd0b7e6d34f19e8343617ba46980f35c5e816efd13e55c4ab3ec

Observation ded99052-3d15-4837-bd9d-9b946ac9685d · outbound

This paper cites Lessons from the Trenches on Reproducible Evaluation of Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Lessons from the Trenches on Reproducible Evaluation of Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:44:50.211576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:7f725f3ac2316d0ef6826e71f18952c64690af9b9e65554c4648b3fc147cf8f4

Observation acfb3d28-48e5-4f49-ac96-bd31f034edd1 · outbound

This paper cites The WY Representation for Products of Householder Matrices.

Kimi Linear: An Expressive, Efficient Attention Architecture The WY Representation for Products of Householder Matrices

Reference 11

Resolution
malformed identifier
raw_fallback, observed 2026-05-13T23:49:11.247558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:352b25ac81079af76b7a65270d3612cbbec9401ab87e0fee2d34f7747dc8bfe9

Observation 17536ca1-0ea1-4476-ae2f-a8009bac0b87 · outbound

This paper cites Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.991207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:d10d2bb8ae6777ed8727f958680d7c5c93d6f551446be6f688685a48135660e7

Observation 5881d725-3ebe-43e6-88bf-6fa63390aced · outbound

This paper cites Long Code Arena: a Set of Benchmarks for Long-Context Code Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Long Code Arena: a Set of Benchmarks for Long-Context Code Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.998413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:000d0660a235bad49da3f6f0e40b5f9ab9083d6239474a74de1c6b6c4ce3ac35

Observation d02947d1-91e6-43d1-ae15-88cb8b793b05 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Kimi Linear: An Expressive, Efficient Attention Architecture Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.003486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:6b621da4d6dbd48fc200819396654d14ee5f13cc27cc9587d111f4d86e6489eb

Observation c5da126a-83fa-493a-b4b9-4efedce0b6df · outbound

This paper cites The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.008835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:d4706a85667b993c4def10bf4ea024121ec0dada71e885323c718d9ac94a4a86

Observation 679ced53-34f0-4da6-bb84-d5b744f83fb7 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

Kimi Linear: An Expressive, Efficient Attention Architecture Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.661609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:547294550379bf1431fad4b3e8c42fbc9e829d579ceaa8677aed77a4d4235e25

Observation 376c4780-c4d2-4c3a-9b57-17258484afa4 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

Kimi Linear: An Expressive, Efficient Attention Architecture FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.272231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:186377122d36ab59b11aecf92b6d2e0966886f9b673ad692b93b224cc84eb54b

Observation 29562b0a-dbf4-45fb-b00f-47df45236b92 · outbound

This paper cites an unresolved cited work.

Kimi Linear: An Expressive, Efficient Attention Architecture Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-13T23:49:11.275804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:6afb51ab09e1c867791b2921101a4eddff99ff877ea9dd74786d833dc49b90cc

Observation 75b71539-d593-4791-b654-56dac58c0a93 · outbound

This paper cites DeepSeek-V3 Technical Report.

Kimi Linear: An Expressive, Efficient Attention Architecture DeepSeek-V3 Technical Report

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.013613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:c67ee5fc9b0f71941d9c6b8740007ffd42a3e9a37df8dbbbd8a8daf5ef66d89a

Observation c40f6a54-1ad0-431c-b955-beb850f50ded · outbound

This paper cites LongNet: Scaling Transformers to 1,000,000,000 Tokens.

Kimi Linear: An Expressive, Efficient Attention Architecture LongNet: Scaling Transformers to 1,000,000,000 Tokens

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.019394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:7b8ba03f1ea7b09a7171d20d04a8f8ce0ae4c8711a591d68fb078deb98c17d84

Observation 7e58ddc8-49b7-4210-99fb-3d42baf23d55 · outbound

This paper cites Flex Attention: A Programming Model for Generating Optimized Attention Kernels.

Kimi Linear: An Expressive, Efficient Attention Architecture Flex Attention: A Programming Model for Generating Optimized Attention Kernels

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:27:16.810492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:26f19f12e3650c73f64a90cb9ccc5e6b2f01b325a56746cfdb29a06933c0b575

Observation 2877e21b-05ea-4b5d-99a9-8fb659bc4e14 · outbound

This paper cites Hymba: A Hybrid-head Architecture for Small Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Hymba: A Hybrid-head Architecture for Small Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.040289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:6384f8457a816ede54d1a7c2ba8c39b54ab957969a338e37680fdc31652ee5a1

Observation a8ff28fd-ee9b-4060-a851-90f199868b08 · outbound

This paper cites Jacob Dunefsky, Philippe Chlenski, and Neel Nanda.

Kimi Linear: An Expressive, Efficient Attention Architecture Jacob Dunefsky, Philippe Chlenski, and Neel Nanda

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.066499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:5f912fca74245ba70498a0318c30f6d6e81a62685824c9fb41c92058a8bae7ec

Observation 14ffcf2d-481a-441f-8541-849cc8083d27 · outbound

This paper cites Native Hybrid Attention for Efficient Sequence Modeling.

Kimi Linear: An Expressive, Efficient Attention Architecture Native Hybrid Attention for Efficient Sequence Modeling

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.073560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8b9254b6caf32d40896b51278dd3d14da98fb77e4068281c13e088a3285bd079

Observation 4eab9d81-4c41-42d1-bbb6-c14914f23626 · outbound

This paper cites Moa: Mixture of sparse attention for automatic large language model compression.arXiv preprint arXiv:2406.14909.

Kimi Linear: An Expressive, Efficient Attention Architecture Moa: Mixture of sparse attention for automatic large language model compression.arXiv preprint arXiv:2406.14909

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.083590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:264948a99e5a1f5265c271196df7aa9cfeb7e722b29cae3280a7145931b37738

Observation ed9b015b-6b73-4953-b781-6081c3887c81 · outbound

This paper cites Are We Done with MMLU?.

Kimi Linear: An Expressive, Efficient Attention Architecture Are We Done with MMLU?

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.089956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:46cb9b48d76ae1f1bee800c4986a745fc8cc28c7f2e612ad720a4ba224eae4e9

Observation e2f45a81-9831-4efe-9fca-4ac25e8fd3f7 · outbound

This paper cites Unlocking State-Tracking in Linear RNNs Through Negative Eigenvalues.

Kimi Linear: An Expressive, Efficient Attention Architecture Unlocking State-Tracking in Linear RNNs Through Negative Eigenvalues

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.331895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:1e2ba0cf5282e8e7858516cff21d90d1d8608c72ef385146fcadc9e2870606bb

Observation 8c9cb472-c460-496f-9e04-0630f0a582ec · outbound

This paper cites How ordinary elimination became Gaussian elimination.

Kimi Linear: An Expressive, Efficient Attention Architecture How ordinary elimination became Gaussian elimination

Reference 28

Resolution
verified exact
doi, observed 2026-05-13T23:49:10.650817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:efcbf9335240b538f8fe685abfb4383edeb3e8323962a7a81e88efce4da765fc

Observation eb8bed08-9d2f-4796-952b-f217bf44ba12 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Kimi Linear: An Expressive, Efficient Attention Architecture Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.095862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9b79053a9a204da02853261990f4506e650ed313253a53254d8b1e98abf928ef

Observation d08418ee-58b2-4036-93ce-ba1bea843c40 · outbound

This paper cites an unresolved cited work.

Kimi Linear: An Expressive, Efficient Attention Architecture Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-13T23:49:11.362813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:c170b340740545480b137bcd53da4efd0d0c47ce65418f011dfaeb8f3a079e28

Observation 948722bc-294c-4f53-af35-4dc9204c0ea6 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

Kimi Linear: An Expressive, Efficient Attention Architecture Efficiently Modeling Long Sequences with Structured State Spaces

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.101786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:95caf7d8c495830c1a284dc4a84b40e1d4fcf7da3d3321a4e1b1cbd61473f839

Observation 8ac8f492-1d49-43d1-917b-195797da74e1 · outbound

This paper cites When Attention Sink Emerges in Language Models: An Empirical View.

Kimi Linear: An Expressive, Efficient Attention Architecture When Attention Sink Emerges in Language Models: An Empirical View

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:41:03.919229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:32ea6198b34920404c36dcf7edeb5308def52648b1bd381500c59ccad2b4b89a

Observation 84818da9-debf-410e-a5bb-f20cdf2b403e · outbound

This paper cites Jet-nemotron: Efficient language model with post neural architecture search.

Kimi Linear: An Expressive, Efficient Attention Architecture Jet-nemotron: Efficient language model with post neural architecture search

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.111878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:79ee851037a81453a85e0faff6f5756a94549bf5aed2c43c10f77593f14766c3

Observation 05e17856-bb91-43a2-9378-96d8571aa527 · outbound

This paper cites DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning.

Kimi Linear: An Expressive, Efficient Attention Architecture DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.379086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:d48fa0239d3237be763a4af2bc5e8fdd9aa1a21de5261f83f85003413b22dae4

Observation eb9b74d4-1501-49d9-9d20-2c8883a80055 · outbound

This paper cites Log-linear attention.

Kimi Linear: An Expressive, Efficient Attention Architecture Log-linear attention

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.120285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:581f4164a59c7da2b4cc14bf81a6dc3b502578f485bce4be167f5eb9a1e8e9d8

Observation ed4b7bb6-583d-4e6d-8d38-8bab179db898 · outbound

This paper cites Star-Transformer.

Kimi Linear: An Expressive, Efficient Attention Architecture Star-Transformer

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.126833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:788b4ffed6e396c3845a1ecbcf90bd6224664bc84dd851cad11f4121dd8e0347

Observation 0e2fe11a-7e44-4f1a-b85f-67d5ffdca829 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Kimi Linear: An Expressive, Efficient Attention Architecture Measuring Massive Multitask Language Understanding

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.133460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:6690c2b834fb0d57a7d9cf2a5a323bdde4ef1d42903f35cd1c163f7191e1f938

Observation b90d53ba-c434-43a9-ac51-62f0a70ee9d5 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Training Compute-Optimal Large Language Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.138710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:2c67020ed7a1c93e327f3812e1369f8e23ad5a9e2de5a2c1cb3dd36f0d9e9eb4

Observation 34e7ecfc-c416-4ea7-8b3f-9ba281284098 · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

Kimi Linear: An Expressive, Efficient Attention Architecture RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.144494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:99986c55021cf7ea01236eb5ce51f285e024e1eb0191aea8a63c8946d04fd177

Observation e34218d9-33eb-480f-8fd9-d3226c42f60d · outbound

This paper cites Attractor memory for long-term time series forecasting: A chaos perspective.

Kimi Linear: An Expressive, Efficient Attention Architecture Attractor memory for long-term time series forecasting: A chaos perspective

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.410394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:7110e07c1fe1c26e5f462661fb16b868aa3411b574b4d67ed29a65437355abf4

Observation 047ee04a-b254-41b2-a058-aa7b20cab988 · outbound

This paper cites Thomas Jiralerspong and Trenton Bricken.

Kimi Linear: An Expressive, Efficient Attention Architecture Thomas Jiralerspong and Trenton Bricken

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.150313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ef9a70995a055149c7a838090663362aad18836a08ad46314969d238f92de8f6

Observation 1954f43f-480a-4efb-ba53-c9008c540c5a · outbound

This paper cites Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization.

Kimi Linear: An Expressive, Efficient Attention Architecture Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.154378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:6b77aa21de46cf731a63fcf120bbd2b89091eea942181435d34da545e9d7ffe3

Observation e7e53e90-78b1-431e-9d69-755916779dfd · outbound

This paper cites Transformer Quality in Linear Time.

Kimi Linear: An Expressive, Efficient Attention Architecture Transformer Quality in Linear Time

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.422283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:50f47ef8626a2f16353c6048be7c61afe109f7aa9b0501cc82cf92a3bb2bbe06

Observation 59103331-d6ed-4ffa-9dc3-c9378bdf9eef · outbound

This paper cites C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models.

Kimi Linear: An Expressive, Efficient Attention Architecture C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.426283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:13d0ac1006ab48fc13d5954c96fce4b0241ef96e7e982873cdb6c4abb4cbbd20

Observation 68561cd7-b80b-4212-972c-1f0c39b714a3 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Kimi Linear: An Expressive, Efficient Attention Architecture LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.158394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ee809ec686e74ac371569dcdaa4ee3215d95f3db422187623d6bc658d8fc7c8b

Observation f3f7ac2c-eccf-44b4-984b-62f8888dbd4a · outbound

This paper cites Repeat After Me: Transformers are Better than State Space Models at Copying.

Kimi Linear: An Expressive, Efficient Attention Architecture Repeat After Me: Transformers are Better than State Space Models at Copying

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.163140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:a5f56749ef5985460ba8d63a8b76233c1c2e5f70cc109540dceb6b39d14446da

Observation 81013637-b007-403d-aaca-b00036e19c08 · outbound

This paper cites Accumulating Householder transformations, revisited.

Kimi Linear: An Expressive, Efficient Attention Architecture Accumulating Householder transformations, revisited

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.669304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:b4c1393af1f46eeb6b6c103a73597f7574f4536681f694512db3cb3da9f14d84

Observation 3983a4af-1161-4ea9-8e75-a971fa53e833 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

Kimi Linear: An Expressive, Efficient Attention Architecture TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.168258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:1d6073c91fbe6061f7b3d68711aebcd29a711f7b0f4419bfec8ea217eb2e03a8

Observation 35c11660-bed0-4b52-81df-d6d0d621f45f · outbound

This paper cites Transformers are RNNs: Fast Autoregressive Transformers with Linear Atten- tion.

Kimi Linear: An Expressive, Efficient Attention Architecture Transformers are RNNs: Fast Autoregressive Transformers with Linear Atten- tion

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.446228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:b5f9b26996cc8ac5a9d3f8cfdc01b53e84701920696e094079a821ab909b604e

Observation c521ae1d-8a1a-4493-a0a5-8d9449e64c12 · outbound

This paper cites The impact of positional encoding on length generalization in transformers.

Kimi Linear: An Expressive, Efficient Attention Architecture The impact of positional encoding on length generalization in transformers

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.449849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:cde52e416c7f44006641f6dc3baf6418c42ebb95ddcdeec074b391982738c8df

Observation faeaa538-5459-4389-9b05-cb3338ffc3ff · outbound

This paper cites Kimi K2: Open Agentic Intelligence.

Kimi Linear: An Expressive, Efficient Attention Architecture Kimi K2: Open Agentic Intelligence

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.174664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:3a4d807eb9009e4af57dcc9ff417e3b84093c452ac5572ec9dad8401b4fe8835

Observation b2461766-1f12-4100-b088-17f01c4a9472 · outbound

This paper cites Reformer: The Efficient Transformer.

Kimi Linear: An Expressive, Efficient Attention Architecture Reformer: The Efficient Transformer

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.179622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:f6311adfec45c95119de6cb0f467deee9ca53b8562a3ba5cf14691cd53a3ad45

Observation a6e6f308-c48e-4923-8db8-e70cfcc702be · outbound

This paper cites Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation.

Kimi Linear: An Expressive, Efficient Attention Architecture Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.184162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8deac38aa4964821f87c130e65df557c051f041b4fa401045ea11379e9ceb64c

Observation f62f61dc-fe95-4e78-89a6-67d73daa04b5 · outbound

This paper cites A Survey of Post-Training Scaling in Large Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture A Survey of Post-Training Scaling in Large Language Models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.263226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8ed5ce00ae0ea468815e966ce09dc8fab13b7abb7cb859c98a6ce28b593c5323

Observation 41f191aa-887b-4a77-a4c0-0843758cf62a · outbound

This paper cites Liger: Linearizing Large Language Models to Gated Recurrent Structures.

Kimi Linear: An Expressive, Efficient Attention Architecture Liger: Linearizing Large Language Models to Gated Recurrent Structures

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.189920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:4679495abd7bf118801192804153bf8d52f5b4b3f20c4507004f7620f3a86b61

Observation 8a1e6dc6-6aab-41c5-a027-3b284471c40a · outbound

This paper cites CMMLU: Measuring massive multitask language understanding in Chinese.

Kimi Linear: An Expressive, Efficient Attention Architecture CMMLU: Measuring massive multitask language understanding in Chinese

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.292125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ec84d1e3b4c167774582e569aa306b8a14e5a1ef24253823cb417d3c7f9ec23b

Observation 4e0ce683-6424-4b9f-a2fb-34bf4320199e · outbound

This paper cites Transmamba: Flexibly switching between transformer and mamba.arXiv preprint arXiv:2503.24067, 2025b.

Kimi Linear: An Expressive, Efficient Attention Architecture Transmamba: Flexibly switching between transformer and mamba.arXiv preprint arXiv:2503.24067, 2025b

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.195212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ceeff0184ea7164224148ed717feb167ae46517d282c24697745fc411eb8308a

Observation 20d4351b-17cc-48ec-9a5c-5248a811f133 · outbound

This paper cites an unresolved cited work.

Kimi Linear: An Expressive, Efficient Attention Architecture Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-05-13T23:49:11.305237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:7251de551d05fac4f3d02e313ceb565f419071057dbcdcfb5c856427a6e60ac3

Observation 4e60121f-f687-41c5-9d8f-0ca7cd8355d7 · outbound

This paper cites Forgetting Transformer: Softmax Attention with a Forget Gate.

Kimi Linear: An Expressive, Efficient Attention Architecture Forgetting Transformer: Softmax Attention with a Forget Gate

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.201312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:435698ea818f8a4e7e15964751d47354f051bc5f67f1c6b948a6e7ce71cb588e

Observation 4f8c4a8a-ea21-43e5-9fac-0fe197c718e9 · outbound

This paper cites Longhorn: State Space Models are Amortized Online Learners.

Kimi Linear: An Expressive, Efficient Attention Architecture Longhorn: State Space Models are Amortized Online Learners

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.207689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:630118989454d8a808a17bc94512054bedc42b4ab081087316fcf7583f2804b0

Observation 63a52458-4462-4cf0-8b58-cc0cd304b900 · outbound

This paper cites Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation.

Kimi Linear: An Expressive, Efficient Attention Architecture Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.323725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:20bf4d0fa1a40566a8834f51446cf830111e2239d33ec4ee0835de3f5e81e360

Observation 6d4576bf-5aa8-450f-a4d3-344d67b44281 · outbound

This paper cites RepoQA: Evaluating Long Context Code Understanding.

Kimi Linear: An Expressive, Efficient Attention Architecture RepoQA: Evaluating Long Context Code Understanding

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.214092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:e46cc3a7bb561c2dcdd7ec7dfcd1edb9b0af24ae1e239bf7c0222393a92fd84f

Observation 51d4eab7-67fa-46b7-8794-9dbc0d13dfbf · outbound

This paper cites Muon is Scalable for LLM Training.

Kimi Linear: An Expressive, Efficient Attention Architecture Muon is Scalable for LLM Training

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:11.220159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:f8eda41a01eb3bc2d088ab880c354e764356e9b035dc06964eed3651e10b9256

Observation 2242e950-55d2-4337-8d75-0495dc20700b · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

Kimi Linear: An Expressive, Efficient Attention Architecture MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:15:46.352230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:43a1526b2fe4076d0a1adaa2b2ce1fbdeffdb6fe4ad3b7b14166c0548f0bbc81

Observation 66200db6-62e7-4e3c-913a-2c1d69659a7b · outbound

This paper cites The Illusion of State in State-Space Models.

Kimi Linear: An Expressive, Efficient Attention Architecture The Illusion of State in State-Space Models

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.232912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:874469324766dc70f7f165931c0ab6a5cdc5d4bf84652fbe6183165638627f50

Observation 8a415a66-6174-4cbf-8015-4c889de3c5a6 · outbound

This paper cites The Parallelism Tradeoff: Limitations of Log-Precision Transformers.

Kimi Linear: An Expressive, Efficient Attention Architecture The Parallelism Tradeoff: Limitations of Log-Precision Transformers

Reference 66

Resolution
malformed identifier
raw_fallback, observed 2026-05-13T23:49:11.375241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:877e3ab61f70365c82d938353349c08bfac7bde97fb73e3f6f33b1bbabb8ee6d

Observation e911b3eb-3eb6-4c01-be23-aac7d558d4b8 · outbound

This paper cites MiniMax-01: Scaling Foundation Models with Lightning Attention.

Kimi Linear: An Expressive, Efficient Attention Architecture MiniMax-01: Scaling Foundation Models with Lightning Attention

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:26:38.921226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:beeea5ec521a291674b308b057e13d20171269af53a405b2bc104c8f718728cf

Observation 77b4b1ca-7e21-4101-ab2c-95d04947c275 · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

Kimi Linear: An Expressive, Efficient Attention Architecture Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:17:00.308428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:a2e4a089c31dd3801a790d9cacbd7335746b5eb333dad6845d5eed5808d8b021

Observation 882a95b9-ebc0-4a5d-bb80-7780dfd87455 · outbound

This paper cites Metalearning with Hebbian Fast Weights.

Kimi Linear: An Expressive, Efficient Attention Architecture Metalearning with Hebbian Fast Weights

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.720077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8476debb60adcb54834fff3489b3565fe0e0a72d86a7931d240793011c6f1e5a

Observation 441b5449-afc3-4228-9ee1-2bbd26e46761 · outbound

This paper cites Metalearned Neural Memory.

Kimi Linear: An Expressive, Efficient Attention Architecture Metalearned Neural Memory

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.726400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9204a022461360d03e0e05b84c885a649f8422b0080451d88a99722132ce8502

Observation cf36d8cf-b571-431a-a695-936aba8aa585 · outbound

This paper cites Training language models to follow instructions with human feedback.

Kimi Linear: An Expressive, Efficient Attention Architecture Training language models to follow instructions with human feedback

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.406596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:d491bf58afed12b73ec83a570db52dcfccac16e01e3fe49d6bbafb4718014123

Observation e7b91b86-1c81-4291-a7cb-f4248c0551a6 · outbound

This paper cites RWKV-7 "Goose" with Expressive Dynamic State Evolution.

Kimi Linear: An Expressive, Efficient Attention Architecture RWKV-7 "Goose" with Expressive Dynamic State Evolution

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.733656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:83e7887d6829600432373e8a295b0f50bfe2ea31f705f5ef9909052076dedc01

Observation 12509223-439c-4f37-b229-c193ab976e8e · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture YaRN: Efficient Context Window Extension of Large Language Models

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.739960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:597f0d1eed03d01ab9c7fc2038fa2ba44e4d4788c5abc3a449516e2389bb1f73

Observation 4f23b4b7-3f5f-4dcb-ba2f-797b46b5f675 · outbound

This paper cites Mixture of Sparse Attention: Content-Based Learnable Sparse Attention via Expert-Choice Routing.

Kimi Linear: An Expressive, Efficient Attention Architecture Mixture of Sparse Attention: Content-Based Learnable Sparse Attention via Expert-Choice Routing

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.746854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9497b7ca33fe2a6c38fbfbf7e67c278bcdf033adfb8be508afe6b063291511a9

Observation 4653c150-b14c-41a4-96b6-bd4b683b7159 · outbound

This paper cites Reasoning with large language models, a survey.

Kimi Linear: An Expressive, Efficient Attention Architecture Reasoning with large language models, a survey

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.434014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:a9dffb9e455f6e6b768fa38ef4629f1b29f02116f68f1e314413271261769fe5

Observation c67afc72-8999-4393-a746-66bb12c46865 · outbound

This paper cites Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation.

Kimi Linear: An Expressive, Efficient Attention Architecture Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.441905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9531563c6a4576cb1d04474689d32fbc813164f3dc70d8e01045eb2e8450e014

Observation c93cd242-856c-464a-ac27-e9f2b15970f8 · outbound

This paper cites SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling.

Kimi Linear: An Expressive, Efficient Attention Architecture SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.751896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:3c3d41a7b622846346511aeebea894c5d28a09de49f1030bbf29a789c76edec3

Observation 2fc4299c-e2db-4779-aba3-9f49db45f35f · outbound

This paper cites HGRN2: Gated Linear RNNs with State Expansion.

Kimi Linear: An Expressive, Efficient Attention Architecture HGRN2: Gated Linear RNNs with State Expansion

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.756376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:3992bdaa47ee6346b999e494cd475c773306044972a211278920bd14fa8c5c16

Observation 774a9802-38be-49da-93f6-7083d7017ea8 · outbound

This paper cites an unresolved cited work.

Kimi Linear: An Expressive, Efficient Attention Architecture Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-05-13T23:49:11.259161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:130bba380feecf49de62bb2430dc05a7b8c7397575fd2bc6cbd71c79d8c798ea

Observation b16d595b-1101-4eb5-a673-a1bd80c3aa41 · outbound

This paper cites TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer.

Kimi Linear: An Expressive, Efficient Attention Architecture TransNormerLLM: A Faster and Better Large Language Model with Improved TransNormer

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.761500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8945f2b0c0bfff1de260aade72cb186b7fcf30473fca87db013c8aed317d4815

Observation 17eba7c3-97a7-43ee-93f2-0c151eac211e · outbound

This paper cites an unresolved cited work.

Kimi Linear: An Expressive, Efficient Attention Architecture Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-05-13T23:49:11.296630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:ea8cfdb0461e08716164453044cca4fa5f7590f9a98ff6fcabfe2929c9e196aa

Observation afd76c00-2e1e-4de2-801c-17f45045ef73 · outbound

This paper cites Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free.

Kimi Linear: An Expressive, Efficient Attention Architecture Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free

Reference 82

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T23:49:10.768666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:e52fc99a71cf4f9f5fe119932d1daa26704a9a11dbe76c662ce8d7cafc6389ac

Observation 84638362-7f83-420d-90dc-ac7754632ec1 · outbound

This paper cites Xiaoye Qu, Yafu Li, Zhaochen Su, Weigao Sun, Jianhao Yan, Dongrui Liu, Ganqu Cui, Daizong Liu, Shuxian Liang, Junxian He, and 1 others.

Kimi Linear: An Expressive, Efficient Attention Architecture Xiaoye Qu, Yafu Li, Zhaochen Su, Weigao Sun, Jianhao Yan, Dongrui Liu, Ganqu Cui, Daizong Liu, Shuxian Liang, Junxian He, and 1 others

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:10.775427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:944550ebac600ad8f0a1b47480b7138773986eb63e777518d37c0531f209c743

Observation 063767e8-a890-45ae-a3c1-590696afd525 · outbound

This paper cites Accessed: 2025-10-27.

Kimi Linear: An Expressive, Efficient Attention Architecture Accessed: 2025-10-27

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.327559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:2235fa96502329045df303c7d527e96e160a404f1b0fefbbc25f433aef09879b

Observation d79e880e-6948-45ee-8ddd-e8a7dac1d769 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

Kimi Linear: An Expressive, Efficient Attention Architecture Gpqa: A graduate-level google-proof q&a benchmark

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.357913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:358d79ee464960b5e8d89af8a44101916e18e4a72c4e1c927b885a1a6f991469

Observation ab3a0f09-d42c-4ec5-87a8-6a66d8e457dc · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.

Kimi Linear: An Expressive, Efficient Attention Architecture WinoGrande: An Adversarial Winograd Schema Challenge at Scale

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:20:15.117976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:8cba974bd699814dfc75dba8db988febd036920fa5942da158e02418bd4a5619

Observation 10923a6f-0871-4b8a-b57d-d71ce819ea4c · outbound

This paper cites Linear Transformers Are Secretly Fast Weight Program- mers.

Kimi Linear: An Expressive, Efficient Attention Architecture Linear Transformers Are Secretly Fast Weight Program- mers

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.370713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:f8b17623513c4be72293aade018d5e3973ad5e4fad5cc9a6d5372ce3bc0c7601

Observation 6a949274-7c9c-47dc-ae57-35946c3ed910 · outbound

This paper cites Learning Associative Inference Using Fast Weight Memory.

Kimi Linear: An Expressive, Efficient Attention Architecture Learning Associative Inference Using Fast Weight Memory

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.787850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:4158b18142647860943fd532211def97d75be6edd3991929907855677eb8a8c2

Observation cbdc2c05-5af4-4ce8-a388-0cca208df207 · outbound

This paper cites Self-Attention with Relative Position Representations.

Kimi Linear: An Expressive, Efficient Attention Architecture Self-Attention with Relative Position Representations

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.793874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9d69e2e8c947498a969fe0adc8db914ea6915b84c1deb70b8cbe901062e2c788

Observation 85c2dd83-3392-486e-9041-9f1483d50d84 · outbound

This paper cites June 2025.URL:https: //kexue.fm/archives/11033.

Kimi Linear: An Expressive, Efficient Attention Architecture June 2025.URL:https: //kexue.fm/archives/11033

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.398060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:56acd73577a4efc4e940149f1c496fe02eaeca67416795341f56de67771ecd2a

Observation 7486701c-c8a6-4028-bc28-4c0bf795ba7d · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

Kimi Linear: An Expressive, Efficient Attention Architecture Roformer: Enhanced transformer with rotary position embedding

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T23:49:11.402606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:39076bd4257287eb101195f222912ff0f077db74af6b49e63666ad510941809c

Observation 990f7d6a-f121-4932-9ec3-3363c7c2146e · outbound

This paper cites Speed Always Wins: A Survey on Efficient Architectures for Large Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Speed Always Wins: A Survey on Efficient Architectures for Large Language Models

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.800523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:9f2aa90234b483cfc9146b9a6959ee21327cc77f9a8ed2e083e4f23413f021f5

Observation 004a5311-af12-4e4f-8e52-91223d9cc6ce · outbound

This paper cites Learning to (Learn at Test Time): RNNs with Expressive Hidden States.

Kimi Linear: An Expressive, Efficient Attention Architecture Learning to (Learn at Test Time): RNNs with Expressive Hidden States

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:20:12.625336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:4f6817ae472826abc03afd6cc6a6a3eff04cdc494e536a42d3d7c246675a1ee1

Observation 917f015e-b783-4ccf-b603-14377a5dfa85 · outbound

This paper cites Efficient attention mechanisms for large language models: A survey.

Kimi Linear: An Expressive, Efficient Attention Architecture Efficient attention mechanisms for large language models: A survey

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.813030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:a5c931474c06783bec3ac6d35f97f60fa5c775282f499aecc1ddb9f6acc10a3d

Observation e72ac2af-e83b-410e-8028-42648d562ff3 · outbound

This paper cites Retentive Network: A Successor to Transformer for Large Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture Retentive Network: A Successor to Transformer for Large Language Models

Reference 95

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.818760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:0ba7f9206b78b5de46f0ddf3c2d0753354590314cacdc35191c0576a8347a59d

Observation 6e83b91d-a9ee-496d-9d4f-db0645586b0f · outbound

This paper cites You Only Cache Once: Decoder-Decoder Architectures for Language Models.

Kimi Linear: An Expressive, Efficient Attention Architecture You Only Cache Once: Decoder-Decoder Architectures for Language Models

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.824263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:21565d66abb572ff7d5e6cc3457acbb53e73530acb08f7820eb5df424ea294b2

Observation 443e8cf7-abd1-4ef7-9828-f5ba06cca594 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Kimi Linear: An Expressive, Efficient Attention Architecture Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 97

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.829649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:1b1ef2d4c2fe02c71c9709d08ca01180c85eaddef06b9577c4371fdb6d31a1e2

Observation 3ee9956e-6d00-4ef4-aa60-664224a7a13f · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Kimi Linear: An Expressive, Efficient Attention Architecture Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 98

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:49:10.836472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:58ced97ad0bdf669a8d2994aeecc73f52b8f3f63f185945c17ad0c1dedbc3cbc

Observation 04169612-8244-4ffe-a3de-011172babba4 · outbound

This paper cites MiniCPM4: Ultra-Efficient LLMs on End Devices.

Kimi Linear: An Expressive, Efficient Attention Architecture MiniCPM4: Ultra-Efficient LLMs on End Devices

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.845126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:fa2cc4d2d0e23214ef011535cbd9dc11572c16d7bf97434647c4c6b0bfda9c84

Observation c59bac0b-b5bc-4a97-b5c1-db724dec4ad3 · outbound

This paper cites Hunyuan-TurboS: Advancing large language models through mamba-transformer synergy and adaptive chain-of-thought.

Kimi Linear: An Expressive, Efficient Attention Architecture Hunyuan-TurboS: Advancing large language models through mamba-transformer synergy and adaptive chain-of-thought

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:10.852510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:81d0b301f694268397fbd9ae3ab670932bc26f6e6c8962fc35217aad8760410a

Pith citing papers

Observation 65880f4a-7385-4491-b9c4-bcd6ac9dd8cf · inbound

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression cites this paper.

Gated KalmaNet: A Fading Memory Layer Through Test-Time Ridge Regression Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:00:27.154465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T17:59:23.826110Z digest=sha256:e67e5b40de1c6fe67de7867f1ce012d8b23eaefa2dcbd58f036813a6345bb010

Observation 4c1ace1b-38cd-492d-af25-fb949571089d · inbound

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers cites this paper.

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T06:42:11.328087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:42:11.328087Z digest=sha256:7c528461cd7a647da4ef89e22c04aff025281c613921ac7defadb63e40c8e271

Observation 15471047-c78b-4ed7-9a0d-2c02d799d208 · inbound

Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models cites this paper.

Neural Attention Search Linear: Towards Adaptive Token-Level Hybrid Attention Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T04:55:41.655757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:55:41.655757Z digest=sha256:0b84a977132c1b9835d2c80579cf3502a8dfaa815a8f5b1d3703252c699b4489

Observation e4f03d44-570b-4f54-9126-b46c9d6f5e23 · inbound

SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-Norm cites this paper.

SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-Norm Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T11:14:47.973288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T11:11:51.440058Z digest=sha256:79a0defce84baeaf5c521c184fa53f3e083c8551b0f0af56f55393fe6f5a3369

Observation 955e078b-52e1-4ca1-8981-6d06963b76a5 · inbound

Learning State-Tracking from Code Using Linear RNNs cites this paper.

Learning State-Tracking from Code Using Linear RNNs Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:46:42.351249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T21:46:14.287966Z digest=sha256:5d245c554555e89028d2e4e9935c5c9c8287869628a0a08714c967ade7da88de

Observation c7fd4020-8525-4152-ae3a-41aaa0de387b · inbound

Learning State-Tracking from Code Using Linear RNNs cites this paper.

Learning State-Tracking from Code Using Linear RNNs Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T23:07:49.480563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:07:49.480563Z digest=sha256:6182aa88f86b544cbcbfdf6ce869d04d2963ca8b0b43a8add71ec0a622f49784

Observation 811f1597-f2bd-45f1-83a5-0bea202cf12c · inbound

Test-Time Training with KV Binding Is Secretly Linear Attention cites this paper.

Test-Time Training with KV Binding Is Secretly Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T19:41:32.771173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T19:40:35.854519Z digest=sha256:de184312206c702cccae12e2081be48c9953014e46b5d5bb638d5cb9ed2afb09

Observation 82209866-5495-4f73-93cd-e3b4ad7af351 · inbound

Why Are Linear RNNs More Parallelizable? cites this paper.

Why Are Linear RNNs More Parallelizable? Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T19:12:04.686139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:12:04.686139Z digest=sha256:3791d6fbe904a362e2eaeaeb620401029564ceb9313bb055a54470aee6ae3f8f

Observation 7c9cf6dc-8503-4aa4-8453-83e63a5621bf · inbound

Attention Residuals cites this paper.

Attention Residuals Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-05-21T06:39:04.478018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T06:39:04.312270Z digest=sha256:e457abc4ee7dbb2437bf774a9a6efbb5e398e88d5ff4230f11139a34f85e404f

Observation 442ad456-ac9c-451e-8061-b757216e3775 · inbound

When Perplexity Lies: Generation-Focused Distillation of Hybrid Sequence Models cites this paper.

When Perplexity Lies: Generation-Focused Distillation of Hybrid Sequence Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T17:21:21.475017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:21:21.475017Z digest=sha256:cb0d788912edf94b86584d70f7e076353b3fc035be056113519ff862c1cd1bd4

Observation d991f051-db19-4d3a-a9ff-a8fdeb2a2fb5 · inbound

LPC-SM: Local Predictive Coding and Sparse Memory for Long-Context Language Modeling cites this paper.

LPC-SM: Local Predictive Coding and Sparse Memory for Long-Context Language Modeling Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-15T11:25:31.092475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T11:22:08.937342Z digest=sha256:ff9025da036fd9b10eaab1659f4cb28ec5723743e46ab227a23426629e3cde91

Observation c0c818a3-d5a4-44eb-b81f-9f332726c129 · inbound

Attention Editing: A Versatile Framework for Cross-Architecture Attention Conversion cites this paper.

Attention Editing: A Versatile Framework for Cross-Architecture Attention Conversion Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T19:47:34.869184Z digest=sha256:11ca05099315a3a0232ad024028181e83a0544a7aae4e34f1a74141eb956ef2c

Observation 5f27620f-5b2d-4ca6-b20c-7086603c2e58 · inbound

Mem3R: Streaming 3D Reconstruction with Hybrid Memory via Test-Time Training cites this paper.

Mem3R: Streaming 3D Reconstruction with Hybrid Memory via Test-Time Training Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:32:42.456629Z digest=sha256:5aa7ab92f240ae9f4e7cd9a6f9224c67469f2bbbdc3eafd22372aa4c95599b4c

Observation de6dd072-f4d6-4212-b67c-d5e29b8340ac · inbound

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation cites this paper.

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T16:17:09.834609Z digest=sha256:6f96199b116ef244341c6b70c41a602f69a399d70c304067555d84d8f221a0e8

Observation 3464b2cd-7eba-4619-9c7b-73360bcb2e1c · inbound

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation cites this paper.

Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T16:10:54.363560Z digest=sha256:2c0175cc5f8012799946ec83b86c41739cc70909896b29c8f0fef14dba9861c3

Observation db989a8b-4eba-46ad-864b-788b1755bb8d · inbound

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning cites this paper.

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T11:17:43.769244Z digest=sha256:5a120a7fa62d33d3ca540f90c76a423bf9f8d37bc3f7ae329bb8e655bf9cbe5a

Observation 28b51f7f-3c93-4b31-945c-5b7aaf2481a3 · inbound

Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter cites this paper.

Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T10:01:41.739908Z digest=sha256:27ae1c26a31563dbf0dd9b187da7641148f79ed4b428addb67caccdd13a6c140

Observation 689b5af6-e7a3-48db-8b59-6c774e36567a · inbound

UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training cites this paper.

UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T02:20:00.625923Z digest=sha256:10f37404f6f17e58a4f887d2c2f02d0255ce49779b1781390a0d0da7d165ca0e

Observation e11f0bb1-903b-4607-9d15-8bfde70e33b8 · inbound

Preconditioned DeltaNet: Curvature-aware Sequence Modeling for Linear Recurrences cites this paper.

Preconditioned DeltaNet: Curvature-aware Sequence Modeling for Linear Recurrences Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T00:19:24.366922Z digest=sha256:aee7cec21fa824693a50f0154cdb2321fe9102fdb9ce1baed95eb7bf0d02aa36

Observation 3ced6f9d-a599-4bd0-8e82-62244d9af53f · inbound

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference cites this paper.

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T12:18:23.898779Z digest=sha256:693fea584db9909d1f6e84efcfbac2e9c51f9951b2929a7f7214f96b4e4b01a9

Observation d93bf77b-c884-4512-960f-d3027ae1ccc6 · inbound

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing cites this paper.

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T19:56:48.015363Z digest=sha256:fb183602c4eef4efa8dbf78c56cbf594737a7272f2b3a7df7c46e796ef9adcb6

Observation fd6e21aa-18a7-4f15-8b5d-bf1ada0bdf94 · inbound

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling cites this paper.

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:39:37.485602Z digest=sha256:d5b9be2a41c05d2bdc3a74670ea4bd071ddad99913da026d5ca185791c8ac3d6

Observation 1e07b722-6554-4fb6-b79f-ffaba45d8bc5 · inbound

Heterogeneous Scientific Foundation Model Collaboration cites this paper.

Heterogeneous Scientific Foundation Model Collaboration Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T08:50:05.980191Z digest=sha256:911176d7d379bcdf169f6303108f06c691a831ac99355a4de3d84447ceaec3ca

Observation 7cc04cf2-11cb-4abb-83ee-0371e383c955 · inbound

Irminsul: MLA-Native Position-Independent Caching for Agentic LLM Serving cites this paper.

Irminsul: MLA-Native Position-Independent Caching for Agentic LLM Serving Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T05:37:02.419657Z digest=sha256:0908cde3c6fe0c714262adc539981bf57052430d3212fc4773642a2f3d6d38a9

Observation 826d8c55-c425-4196-8de9-e26fa2671c05 · inbound

MDN: Parallelizing Stepwise Momentum for Delta Linear Attention cites this paper.

MDN: Parallelizing Stepwise Momentum for Delta Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-09T15:27:55.566795Z digest=sha256:874c5424bfe270c5e9b10790ae30a9086aacc052cf3f7e15b92f3c1a6b88b46a

Observation 057d2c3b-4551-4afe-b105-39c128274991 · inbound

UniPrefill: Universal Long-Context Prefill Acceleration via Block-wise Dynamic Sparsification cites this paper.

UniPrefill: Universal Long-Context Prefill Acceleration via Block-wise Dynamic Sparsification Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 41

Resolution
malformed identifier
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T10:43:01.724760Z digest=sha256:f30102c90473df682f78b8dd4fae19d33f88cc1405b885074156b86d949f0ef7

Observation 6363b622-fe19-4bf2-b594-a1f6b0f6ab99 · inbound

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization cites this paper.

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T02:17:32.226166Z digest=sha256:a4ebaf4f2591deb6db11ffdbd9795a3bedce22a5bd04a8c2865f1a982dafa48b

Observation 8bf2528d-210e-4da0-9aec-68f888328d81 · inbound

Kaczmarz Linear Attention cites this paper.

Kaczmarz Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T01:15:58.330766Z digest=sha256:26bd746065084713d5ea6872afac1a05152a6fc0548ce21e5fdf65b5b1e41316

Observation 93db9185-49d1-429c-a0fe-4987f5c124ac · inbound

Structured Recurrent Mixers for Massively Parallelized Sequence Generation cites this paper.

Structured Recurrent Mixers for Massively Parallelized Sequence Generation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-12T01:28:46.885635Z digest=sha256:0a1783152424b756d6784685fc0bb18669ea31b9267b320f7b04cf54401c0881

Observation 5d514b23-d499-4a89-9ec5-6792b7e6a869 · inbound

Structured Recurrent Mixers for Massively Parallelized Sequence Generation cites this paper.

Structured Recurrent Mixers for Massively Parallelized Sequence Generation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:29:12.887664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T23:27:20.754475Z digest=sha256:29ce456c7b74b9087a6a578a0e26d0b3f64fe40593c13cfb638c1c6bb6a7c50f

Observation 82a8f3d5-0e84-48a7-98d9-e481335efde5 · inbound

Structured Recurrent Mixers for Massively Parallelized Sequence Generation cites this paper.

Structured Recurrent Mixers for Massively Parallelized Sequence Generation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:35:07.322917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T23:31:48.469869Z digest=sha256:c2e3a5fdd3ad10b7ef398436e465e338923ebd50ea37e96d4216e434d2d3f182

Observation dfc18780-d4cb-490d-9061-9f2eeb6e69c4 · inbound

Structured Recurrent Mixers for Massively Parallelized Sequence Generation cites this paper.

Structured Recurrent Mixers for Massively Parallelized Sequence Generation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T05:19:07.559006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T05:19:07.559006Z digest=sha256:775583620ff25dd985760e93085b40c12a559f033d2a00133a40ecb28e21fa0f

Observation 02d26702-20f5-4fea-bb0a-e9da57ba7f4c · inbound

Mela: Test-Time Memory Consolidation based on Transformation Hypothesis cites this paper.

Mela: Test-Time Memory Consolidation based on Transformation Hypothesis Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:01:01.993926Z digest=sha256:5ac5392f1443227d3211cdc271226a958de81a92a2dbb07c951544ca6b27e19c

Observation 934c8fe1-081e-45b6-8bad-f41301499ccd · inbound

$\delta$-mem: Efficient Online Memory for Large Language Models cites this paper.

$\delta$-mem: Efficient Online Memory for Large Language Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:49:11.451207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T04:08:34.936172Z digest=sha256:1ce50fcbb49c4cc389e0880f91a7f426cd75a7868bc9c6f2e05ac67dea38ae4a

Observation 8ca0cdfa-c49d-4bc3-99bc-77bc73338eed · inbound

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention cites this paper.

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-14T19:09:23.258512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T19:08:18.768344Z digest=sha256:de57d31221d18635886f20f0d07e4d21da9fa5bdd484029ceabbf6f5f54ff0d4

Observation 9f0af75d-f0be-47a4-a3d6-9783eea5be53 · inbound

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer cites this paper.

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:00:03.388528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T13:56:17.912350Z digest=sha256:322ac277ff52d1f2e1b597f6ad97f24c98efd360d7a99d65cc0b02ec9b68e517

Observation 9ab313f1-05a1-4899-a713-d4bd5d357f0e · inbound

Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps cites this paper.

Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-19T20:52:46.060876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T20:51:29.740033Z digest=sha256:2d9ea024aac348769b5fbb8585486c75a92101a5dfb4b34f5c6e237e4efe457f

Observation 5a442d03-9384-45bb-87b9-5b0b8ba0ce19 · inbound

Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps cites this paper.

Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-01T14:55:47.430149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T19:17:28.300961Z digest=sha256:40295464309e4db4d94d11dcf3f8928dde120a3bd48d63fd627521ebca7916fe

Observation 9cfb3d57-dc5d-458c-b3bf-7491a8e06de5 · inbound

KVBuffer: IO-aware Serving for Linear Attention cites this paper.

KVBuffer: IO-aware Serving for Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-20T12:03:15.266238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T12:00:06.171816Z digest=sha256:a724eb0e78202c94a7d0c75a98357e142808d285f688e2fec57f2c6b07c599a6

Observation 9fe4a42a-9aa2-448a-a3e9-2f9a8264f3c0 · inbound

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond cites this paper.

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-20T07:58:07.546337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T07:57:51.032025Z digest=sha256:6156263c692f50122ac5085a28478bf11aae2d62788f9771130181f959044676

Observation 8131eb8b-a3f0-4773-a222-da8eead1d4e7 · inbound

EntmaxKV: Support-Aware Decoding for Entmax Attention cites this paper.

EntmaxKV: Support-Aware Decoding for Entmax Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-22T09:04:45.948518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T09:01:28.988942Z digest=sha256:29f27fa831201fab93a8325db58d82c54b2b3a90133cf968b389d283fa6ce48e

Observation 3457a483-f5a9-42bd-ad98-72e9216099a4 · inbound

Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention cites this paper.

Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:54:36.599302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T04:53:24.400091Z digest=sha256:286684b942f28511ade88b7b6880585fa6c2499770ddb544e2264d10199fda45

Observation a5364585-c31b-4f32-8630-7d9e9ba2b820 · inbound

HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction cites this paper.

HorizonStream: Long-Horizon Attention for Streaming 3D Reconstruction Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:35:20.918343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T04:33:48.127707Z digest=sha256:bc2f947d7e996b78e8f6d593372d50fc069bc2f6fb239baa7039076215baebb4

Observation d81d6ed5-5c7d-4098-8621-edf1d52bd63a · inbound

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer cites this paper.

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:43:13.644858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T07:39:41.553623Z digest=sha256:81282cd918903820c9cf9014e229c207f8c73b9836e404f2e8eec0a6d56ebb63

Observation 3eac1317-6324-45fb-8d4c-d9fa6af0819c · inbound

Memory by Design: Probabilistic Sequence Layers cites this paper.

Memory by Design: Probabilistic Sequence Layers Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:26:12.886984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T21:07:31.407554Z digest=sha256:490ad4c3bf4a3c73164ae5b28d042b529743ad796bb7f78c7ce75ab3a85d0ff5

Observation 1f728b27-b06d-4858-b0bb-82c388acf58d · inbound

Zamba2-VL Technical Report cites this paper.

Zamba2-VL Technical Report Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 138

Resolution
verified exact
local_arxiv, observed 2026-07-01T19:26:00.707508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:34:20.970856Z digest=sha256:531b2a5b4e8774cdf6471e51d66db9d67717f4a847067148c7f9e07be1dcf156

Observation 23b270cb-f7b0-4820-865a-39acde4c9bae · inbound

When Good Enough Is Optimal: Multiplication-Only Matrix Inversion Approximation for Quantized Gated DeltaNet cites this paper.

When Good Enough Is Optimal: Multiplication-Only Matrix Inversion Approximation for Quantized Gated DeltaNet Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-06-28T02:41:31.882425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T02:36:03.243732Z digest=sha256:20f4564edb0d846dfb1288f33a1b0b92383bcedb3bf0dfe4d1fdee1206fb075e

Observation b2c0cd7f-ed2c-41dc-8f88-72c9eda16f27 · inbound

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents cites this paper.

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:36:59.413041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:07:14.691347Z digest=sha256:ea79e0307e1cf53ddbd77a904c17be93794176e66e83efa82ff0109af7d4c00c

Observation 14d4d18f-62c0-451e-b82a-5aaccabedd55 · inbound

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing cites this paper.

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:36:59.549388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:06:04.896501Z digest=sha256:3dde34d88bdf2a9adeed8e9c4df986b75d09a9a5b024ae56cc88d169cf580587

Observation e13357e3-3a68-4815-8d38-a254333c20f8 · inbound

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory cites this paper.

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 100

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T11:46:55.258244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T03:07:52.730713Z digest=sha256:84b0068e49af414a34d0951f3c50f85ab8d482b5fb8f04ebb8ab4c16cec1de67

Observation 156b442d-5f90-425f-9378-ec32da37f074 · inbound

Gated Bidirectional Linear Attention for Generative Retrieval cites this paper.

Gated Bidirectional Linear Attention for Generative Retrieval Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:17:22.261079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T20:38:01.086990Z digest=sha256:6f663ace860868fc1548f63993ed08675693b338fe46d172110174b1fb55ba27

Observation 7f2316ad-a695-45b7-996a-6ec6dca2dc96 · inbound

End-to-End Context Compression at Scale cites this paper.

End-to-End Context Compression at Scale Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:17:31.552992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T16:36:54.699174Z digest=sha256:eec7f2768abcc0f69858cd140d6d6b8a2c64b46aa7ac4766e1dd7f276ce87452

Observation 0bf84b1f-5931-462a-8a99-ba2c7ec45df0 · inbound

Blurry Window Attention cites this paper.

Blurry Window Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:46:14.018167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T17:43:34.429061Z digest=sha256:c414b2bef31962c9467e9525e0d31305cc6da7550f253a55ca17580a536d86fc

Observation 3e276bbf-6ebc-4ead-83e0-f496a9d6499a · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T10:48:03.158827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:920461f7487d31cdc0f7fcf4ed4c981a1f9d84ad35ea3606bd5e8d3e08305cb9

Observation 1c535d7f-e6a3-4aed-9d67-a1712083fb23 · inbound

On Subquadratic Architectures: From Applications to Principles cites this paper.

On Subquadratic Architectures: From Applications to Principles Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-06-27T10:30:51.596621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T10:27:34.213676Z digest=sha256:8c00b1db046c555af131a93a644008b038ea53de8ba86793ec28e0ddeccf0f9e

Observation 9413d482-399b-42ab-9b45-5f4c896c5c29 · inbound

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention cites this paper.

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-04T14:09:53.423055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T04:26:00.066003Z digest=sha256:822ef06e162c7dd94d7aea4ba640c64ca90ff796a9b8c64fb2e743d29e12620f

Observation 9ae232e7-76ac-4164-9581-5afdf26ceaf0 · inbound

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention cites this paper.

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-30T09:44:37.439663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T09:42:52.096671Z digest=sha256:a3a3dbaf3bd9dda77ad391ef5d0463c9634671a9f2dd54eeed99abaeb36a927e

Observation cdf5ff6a-d52e-47a5-ade3-dd4a0d2711c7 · inbound

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention cites this paper.

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T11:49:31.618283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:49:31.618283Z digest=sha256:2991a47e22d79a2e4febd637c44fe36e54273eb8e36f8f6e09b54f6954c4f4aa

Observation a6805378-502e-440b-bfc4-d5429812d0fa · inbound

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback cites this paper.

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-30T09:54:35.307609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T09:46:57.030469Z digest=sha256:b7f6eb11043572af73b719860fc62d7c70d4f85815f91006399d2f76d42ec037

Observation 8c7bc288-168b-40c1-aa95-35ec9b136aad · inbound

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback cites this paper.

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T07:16:45.194991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:16:45.194991Z digest=sha256:f0d37b580ad8dc38f8a30ddbc5be83ca700a7fd744139026d9892de4352e84ab

Observation e9c4921b-d4b3-4170-9c5b-84096877ab69 · inbound

Timesteps of Mamba Align with Human Reading Times cites this paper.

Timesteps of Mamba Align with Human Reading Times Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:54:20.515930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T06:49:07.298160Z digest=sha256:128142870381b1bd948bb97310e43c50d8f652c5828ba59a857f244356c762fa

Observation 3b8599c0-e6fb-4d1b-9c1e-94ba99e40b3b · inbound

Morphing into Hybrid Attention Models cites this paper.

Morphing into Hybrid Attention Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-06-30T08:44:28.058764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T05:56:51.447893Z digest=sha256:64f11fa05337475583ea429b3f349a66d7c103a51932bfd51d051e63cf6f8bdf

Observation d7719444-cfa5-4438-8ade-213476af67cb · inbound

HYPIC: Accelerating Hybrid-Attention LLM Serving with Position-Independent Caching cites this paper.

HYPIC: Accelerating Hybrid-Attention LLM Serving with Position-Independent Caching Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-03T18:58:50.728456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-03T18:53:12.126023Z digest=sha256:453d3645233ba1d009bd28b1edc87f920da73b20b5a247d03e71f5e3ec9343bc

Observation 2f90d8db-4490-4bb7-9888-6e93d334bb60 · inbound

HYPIC: Accelerating Hybrid-Attention LLM Serving with Position-Independent Caching cites this paper.

HYPIC: Accelerating Hybrid-Attention LLM Serving with Position-Independent Caching Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T16:49:27.301911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:49:27.301911Z digest=sha256:bbb0d16cdce33cd524b4b78384e204141b44605546e39b32fa215b8e1796a481

Observation a9428ec2-acc6-48c0-8c06-c2c89d3b19fb · inbound

Evidence-State Rewards for Long-Context Reasoning cites this paper.

Evidence-State Rewards for Long-Context Reasoning Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T13:28:18.204607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-03T13:25:14.844589Z digest=sha256:e1996a42716a9bdf3f33b7ce93f17493b5cc6cee1c32ea347543c4f428f12642

Observation e03c1908-7587-4fbe-ba75-4b9a4d46c8ca · inbound

SHiPPO: Recurrent Memory with Transported Polynomial Projections cites this paper.

SHiPPO: Recurrent Memory with Transported Polynomial Projections Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T05:11:52.395478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:11:52.395478Z digest=sha256:4324aebbf1fa809dcbb17d9eca0bf1155bffc2fa2f3bdd70727aa9cb5b980502

Observation c2eb8bcd-5ac9-4afe-8186-fee909aa5f6e · inbound

The Key to Going Linear: Analysis-Driven Transformer Linearization cites this paper.

The Key to Going Linear: Analysis-Driven Transformer Linearization Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-09T01:45:50.809547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-09T01:44:40.722957Z digest=sha256:abf82734009a4a4b6e2e8e62d3b3e4f97142a2deea42e8ecbe0cf769d58cae31

Observation bc7af603-b610-4bab-b73c-5877701bfd55 · inbound

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing cites this paper.

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-10T14:57:14.483248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-10T14:48:23.587464Z digest=sha256:c97ab87cb5e7ae670d2d992dbe4bfc518027c7ea54fe8e935d21c65713f168f9

Observation 67b185ac-4f92-45ff-9857-6e3e85f7ef35 · inbound

MOSAIC: Adaptive Inter-layer Composition for Efficient Heterogeneous Vision-Language Models cites this paper.

MOSAIC: Adaptive Inter-layer Composition for Efficient Heterogeneous Vision-Language Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-13T00:53:20.749426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:53:20.749426Z digest=sha256:3d5625065bd2648c756c3a37a5d6fd0f489c0fa917443df2699db36d4dac1fe3

Observation 4cf222d0-c388-486a-a352-10976826336b · inbound

The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale cites this paper.

The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T06:39:22.418349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:39:22.418349Z digest=sha256:3aa13889187b95773cf8fd5774d556d15fe6e048df17930ac9c98e282937da8b

Observation def46d64-c570-43ff-9648-1dea801a2e65 · inbound

The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale cites this paper.

The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T02:03:24.541627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:03:24.541627Z digest=sha256:ddb7505fc52770f4086d648869184255e5010726d74e0e319ef1810c5ee1fe5c

Observation f1f013e8-402b-4073-bf80-5e11e636ebd8 · inbound

Beyond Memory Leaderboards: Evaluating Scientific Memory as Budgeted Context Restoration cites this paper.

Beyond Memory Leaderboards: Evaluating Scientific Memory as Budgeted Context Restoration Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T19:49:13.762529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:49:13.762529Z digest=sha256:261dad4138e4e71e6ddcc702e7be9917789ca6a8a4001b0533e9f13001a973d6

Observation fb3387b6-c813-4594-b2b7-01f201f96f88 · inbound

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions cites this paper.

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T09:17:47.299700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:17:47.299700Z digest=sha256:f38643321f509516129dd7604c227023270d109ec54bba8f893314e9a34889d4

Observation 432cb132-374e-4e15-b6a4-79b5b8590cae · inbound

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions cites this paper.

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T02:08:20.290764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:08:20.290764Z digest=sha256:fc310686f2664c62c93a93ad48e824db3d9aabac1eb8beaf38da72d361c4ff57

Observation 6dd4bf36-fe90-46fd-a305-59ffa98fa70e · inbound

From Scalars to Time Series: Rethinking Implicit Neural Representations for Time-Varying Volumetric Data cites this paper.

From Scalars to Time Series: Rethinking Implicit Neural Representations for Time-Varying Volumetric Data Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T08:57:43.970524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:57:43.970524Z digest=sha256:2b700bf5cb41eecef182519b9436d650590a7db7be833d5b0a609e172270c42b

Observation d3fed15d-513b-4eb8-acb4-9120230e8599 · inbound

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation cites this paper.

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T07:09:05.662964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:09:05.662964Z digest=sha256:f3670b1e954f1ae3ef95310d720d990d9488b6103385691771a7efa99e58b370

Observation 9ab79ce3-e90f-454b-96df-5f08e5183d8d · inbound

Raven: High-Recall Sequence Modeling with Sparse Memory Routing cites this paper.

Raven: High-Recall Sequence Modeling with Sparse Memory Routing Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T02:44:01.780935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:44:01.780935Z digest=sha256:0a5a46af9f0633689c8299770147a2d9b414bd5487fd9e55565106b6c0926cc8

Observation 73ef5fbe-a39a-49f4-93d1-58c426563884 · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.231897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.231897Z digest=sha256:4dd5a0b431ef9ae30ae1e75d4a5cc3291d86123c1f29610da5f534ff41df2e3f

Observation 8a0057b2-c7fc-4d64-9f00-23ae7b625bfa · inbound

Subtract or Replay? Exact Deletion from Language-Model Memory cites this paper.

Subtract or Replay? Exact Deletion from Language-Model Memory Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:27.930847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T06:08:27.930847Z digest=sha256:d28019ed34ea0b4f124d34dfede39d118511cc1165dd268f83de7b86f6bfcbfc

Observation cee43a81-79fd-4491-bbc7-17946f0ff2af · inbound

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers cites this paper.

Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-31T02:16:08.200768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:16:08.200768Z digest=sha256:a8c30a5ea8fb4bc75ffb3918e25bc0fbf12bf82bd20397230d4a8208cddce217

Observation b4d13edd-e7a3-4f69-9bd6-f04eb7f08839 · inbound

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference cites this paper.

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference Kimi Linear: An Expressive, Efficient Attention Architecture

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T05:46:28.137780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T05:46:28.137780Z digest=sha256:7f256b589123a591a31a6a2d311104340a48dee1bc28df27200db910c748535f