Pith. sign in

Paper Citation Record · LEDGER

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?

As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.10415.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10415 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:36:57.791208Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31c65fd3-40c8-4b8d-858a-08d5bda3d075 · outbound

This paper cites online" 'onlinestring :=.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.842851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.842851Z digest=sha256:e276b3ab42bbe794a9b324489ad3bb896ad43d55491713baea4e7039cae0d431

Observation b5e898a2-cd89-4029-b420-8c86610052cf · outbound

This paper cites write newline.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.910445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.910445Z digest=sha256:f3f63b71955fb4ba878c9e47bf9c4c09c03a8af1eeab54740cf6f919c552c213

Observation 90da2a9b-2ad0-4a37-81c4-fcba53e4aa8d · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.968362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.968362Z digest=sha256:cb4951e8b33d108c49022bcf632aa318c48fbe453575d6caeffe5c78de8a4646

Observation aa144632-b19d-4ee2-892e-12eaceb40209 · outbound

This paper cites GPT-4 Technical Report.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? GPT-4 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.005935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.005935Z digest=sha256:e8a2c0cbccf671d950ed5761cd50e65e8a947f78219755b9e35ae53b3b5cbee9

Observation d384cc06-f01b-4501-9b84-265165211a64 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.093486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.093486Z digest=sha256:32b968a3364629d3b6d69dd2ca826fd3ffb95b8b5bac8d820362a46fd928c504

Observation 26f63b4d-6496-40e8-924d-82df3d97b04a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T04:36:57.849183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.141937Z digest=sha256:4afe93007ff29b3421bfac449cc766ddaecc71120bd058b2ca035d29fb4485f2

Observation a4ed913e-e0ba-4396-8743-f4378946d707 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.306506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.185800Z digest=sha256:283c8e60aa551f665b1d6b9ba2283ac2349a367f35f19c7b226da18d013bcd87

Observation 19f9dbbe-8ed9-411c-bb04-0d4f57fc7fb6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.228030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.228030Z digest=sha256:28702c3ff7cdc69e382ae8220c6df0192771349236d5e790eab356556c267759

Observation fb374256-0dd1-4b4d-a3ca-c07e0744ebe3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.290223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.267735Z digest=sha256:ffe3892cdc51c6459a8e75be119ffc48476d2846eecd9d1816c06ec5b6700203

Observation d4690e8f-a6ff-4b89-af78-e6dd86db56b8 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.303597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.303597Z digest=sha256:e453983a6f0a6d689f3fa00b71df1b9520107796eb0fa9dc63bc9465de6fd9a5

Observation 7166f125-6432-4892-9b2b-cf08503f23c5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.352974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.352974Z digest=sha256:f6b292178aa27ff3248209d6d86544396f38d4f2dd05e695cc63ddc2685581cc

Observation 3d591363-7b61-4c01-b648-91b2e7daa879 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.418189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.418189Z digest=sha256:750486a5197a52fb0e927d9646ed813dcb058216800a45c75840db5c8bf99a7f

Observation 9a505477-76a1-41a3-bd06-4a82bd59794f · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.274288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.484181Z digest=sha256:02a813d8438c6abdb8cdd535e65c9902a9a9c679879aecc2910bf391ba4dead4

Observation f6ca50f9-4f86-4913-9064-40a506cc1344 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.265367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.552826Z digest=sha256:f8df1a1b9f85353ddbf67f36d4c4fea2f369e87e524658fe8b8d58532abaac2a

Observation 554cc7f2-cfbf-4f4f-a6e6-a6711c45b3ac · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.586973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.586973Z digest=sha256:13e305d125930da0a7c641cb8cb8cc9b12d68a429516c2ac91b6c5a7320dd405

Observation 851cadbf-3692-4203-9ab4-74d172aada91 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.251177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.623691Z digest=sha256:2e63e59138cfa5eb0639f1ce04328fad34fdfad998a07a6196594f4089340c40

Observation 6357a1b2-4a06-4e72-ae96-19caa511c722 · outbound

This paper cites Ku, Qian Liu, and Wenhu Chen.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Ku, Qian Liu, and Wenhu Chen

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:36:58.241963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.689422Z digest=sha256:449c4dc019e95d7558031a47fd22e05f2cb1d86b0461c143cdfb0c3dc3193438

Observation 61e8e627-bce3-4b03-b763-7090f19bff57 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.756183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.756183Z digest=sha256:8919af9817264a33eaee60a2ab9f1528dcf73b28df6b8222c87f1cb722ba5c77

Observation ee18bd61-5898-473d-b604-74072ed4d97b · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.800127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.800127Z digest=sha256:804aec3d175e4ed708c7c8c8d0ebb2db84b6caa59b6d839a8c68b4b41a25c5ae

Observation 9cbdcc6f-eed7-49c7-b3ef-b58c46889f77 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.871215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.871215Z digest=sha256:0b5323a623fee92acb93ba48c299e3367b9ef86c0d06c8106314628718af3123

Observation aec1934d-55f9-4823-b709-84fa00b75b98 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.903250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.903250Z digest=sha256:318af8b1f9435c7c8bff165b6b2135f4dee953079577eab3c9d0185044520785

Observation b3b66b49-a8e3-45ae-99c4-5c84ace87e7b · outbound

This paper cites A Survey on Benchmarks of Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? A Survey on Benchmarks of Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.960671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.960671Z digest=sha256:63e766e1e78616704d940a580c59ac827f85572dbfb89ad9c9ea51b81c909b7e

Observation 38f2f181-a075-4453-b6d4-6c11922694ef · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.221435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.035431Z digest=sha256:91bff0a879b3fb0cb0fa7efcd26eff9ace70d0c40979caa11842018538699db4

Observation 84c2d7c2-8b62-4e08-8ebc-ea82b2ad1476 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.212845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.083977Z digest=sha256:6f71055c85cb6c032cbb53aa3bcfa87aa95a620b4b9419ff6dd38b6cf7b96b74

Observation b565f4d8-0151-4476-b5da-228d5d773e7e · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.203854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.178525Z digest=sha256:ba6f17956528f4e20de903902074b484b2d4f74104749922ffc31540bb96db5e

Observation 2382cdb4-0975-4350-b76c-b6060a36dc1b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.195026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.223584Z digest=sha256:bc64376e1a8e5bfc472250a66a39afd21949e343400b37262759c243a2d078e7

Observation a3a0c202-9922-420b-9427-584464a29986 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.184983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.312487Z digest=sha256:88e8af23b19cdcaba8264cee539acff374753df7ab2587d6d88bb692ae0f156a

Observation 2a5ac443-47a1-4cc5-9b78-d33bc962200b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.175717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.355871Z digest=sha256:2533ccf17c7e40f7bcc84bd9f205dbb18b69303fb53d1086567693f648d2a499

Observation 9f40534c-5146-4bc6-bf49-8f471341bfa3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.165973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.406025Z digest=sha256:0f5c9e71000fbc64cf5f2963724fea03b1afe054d4dce199ab720c25e9fa29f6

Observation a779e404-8c5e-4802-935e-36120f0f9bea · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.156334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.493936Z digest=sha256:32f5261c027233f5a0471ed354e15dee79847bd4c225e2342e6c93f3a25fce28

Observation cfa1177c-7f38-4a46-a26e-1c46c67e3bdd · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.146711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.529318Z digest=sha256:df87d804de12728b6089c1c22c700c556c18b7e188da4cbb049bb32501e5818e

Observation 6f3ec60e-cef3-4ab6-ac1b-2abc939dcfe2 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.614358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.614358Z digest=sha256:f27eaf869e24ecd46afef38c8e1d52325365766f558e1e204d3239d7cf5a0e8d

Observation d5b9b921-0af2-411c-99ba-ff972055fb86 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.130607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.660424Z digest=sha256:e73d5b7be849e1cc08df730eb4fd5c45061bab10d18a9ba6e77f0431f8d32ace

Observation fb57719b-1494-4abc-8166-e03811368756 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.121927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.724102Z digest=sha256:5afc38c1ac57ac8d8fdfff8423ead8ecd7a774a68b39e77dbe1b0c1d30514d03

Observation 315654a7-0260-4135-a85d-a0d4189a2a95 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.113028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.745924Z digest=sha256:7f6af4dd0505a58ebab66839dea8d15a825d9c6deb57da1d48d89398f7ad9dc4

Observation 75cecedd-f28b-458e-8c73-a1db0c65008d · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.103796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.749772Z digest=sha256:c274126d871372da8cc5374237392d43975842a9d5ed826d155f66256024be6e

Observation 7e92028f-d1cf-455d-b143-4a5abba33883 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaMA: Open and Efficient Foundation Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.752458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.752458Z digest=sha256:e67167fb36c1a6192ef3061aa32db4452e2edea866b80d333549e2442ffc75bf

Observation dca4467a-b6e6-4873-a438-5a406bf2eba0 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.094986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.755306Z digest=sha256:b0b60b011f67bf64f065a27195edd0174047a6e2122e488900b6ca1525d65e77

Observation 7a6459ba-220b-4ae5-9fed-9da86d444cbf · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.758726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.758726Z digest=sha256:bdea2e9022fe211c93dba3a20705ae8221f6659969ecddab35f98eac8dec9031

Observation 1d00e8dc-26b0-4a19-b456-da1fac3709de · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.761786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.761786Z digest=sha256:e7f40305a58bed72018b043181e960b15e82b9ad2ac39b1e4b0c3c27f9d690e6

Observation a7b68da8-5e05-4361-96a0-b382231b5f3a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.764923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.764923Z digest=sha256:283fd2979008a4993b7eadb774790eda08946bcacf8bc978d99ad4a130bea768

Observation 60289f6a-c095-4d0f-8598-7e7a8d5c3156 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.768267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.768267Z digest=sha256:38a9c93bbe3acabfcb8c9069d290fa433507e1781f8e117029ce4ed80a479e91

Observation ea391491-2233-462b-99df-43c0d61d2dd6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.078067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.771470Z digest=sha256:8ceedb5c1fb3e5333811e8d298fc71b53707a082ff1e73ddee1bb7ad9379c17d

Observation ae19ed72-386f-4706-9dd8-fc22b2680059 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.774458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.774458Z digest=sha256:4799674a94601c02e5ee491cf51f1d7435604c16d62c4c0c78e6457cd0bbebac

Observation 48c77cfe-290f-4270-adb6-2b1adcf3f0d6 · outbound

This paper cites Long Context Transfer from Language to Vision.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Long Context Transfer from Language to Vision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.778449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.778449Z digest=sha256:04017e3f2c1d262e0fc7ea125c6def761e8d5524d7f27f283cf99dc3f979afe5

Observation 6e1a7f37-a7c4-4c2d-9e09-8187eea79bce · outbound

This paper cites Weinberger, and Yoav Artzi.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Weinberger, and Yoav Artzi

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.782456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.782456Z digest=sha256:4e70a259c456ea4d0698961bf37c3156fed1c8b1728158d41b29853b1070d47f

Observation ac596079-51f6-453a-9ae1-83a32905c794 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.785475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.785475Z digest=sha256:a1226e8f7e7ace6d2441e6e969e0ec2f00a31cb3b8dc9d03172b7f1ea80f57c7

Observation 7632a9aa-e197-41da-a670-197147fc38e5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.050574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.788190Z digest=sha256:f3d0496cd77d1af44d3e78f09e68df8a036f9eec69a80eb7341ecde31f3ce733

Observation a3890a0f-9ccd-4788-a921-54571984658d · outbound

This paper cites Zwaan and Gabriel A.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Zwaan and Gabriel A

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:36:57.954724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.791208Z digest=sha256:b0985402e1e45b169ccdedce1cb6fe0567b7ad8f39c0f4cd992ff007f1db0e4c

Pith citing papers

No inbound Pith citation observations are available.