Pith. sign in

Paper Citation Record · LEDGER

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences?

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.10415.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10415 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:36:57.791208Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31c65fd3-40c8-4b8d-858a-08d5bda3d075 · outbound

This paper cites online" 'onlinestring :=.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.842851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.842851Z digest=sha256:7f794f45daf9c1ddb8387943e2491c3ea1f23252ea0d9a17b4850ed57fd68550

Observation b5e898a2-cd89-4029-b420-8c86610052cf · outbound

This paper cites write newline.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.910445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.910445Z digest=sha256:b83b622304ba8391e21b21147b5a178d2080e31067146ff833e1054f63fb9bda

Observation 90da2a9b-2ad0-4a37-81c4-fcba53e4aa8d · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:55.968362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:55.968362Z digest=sha256:8c64034fc1fab3bf6ec7ba64a514109168e8820be13090ff7b89d714170c46a8

Observation aa144632-b19d-4ee2-892e-12eaceb40209 · outbound

This paper cites GPT-4 Technical Report.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? GPT-4 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.005935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.005935Z digest=sha256:19fe52ee626a9f4be15b1e341a676426f82395c021b4520c2006b2a79272f235

Observation d384cc06-f01b-4501-9b84-265165211a64 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.093486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.093486Z digest=sha256:181fa054216024bc0b06ce22e42ff97e9389eff48952ecf69e57837cb47b1209

Observation 26f63b4d-6496-40e8-924d-82df3d97b04a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T04:36:57.849183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.141937Z digest=sha256:409e3e50ab140fb789d011c6431e8820ea25fc811e3f68b721f7b6de58a5d5dc

Observation a4ed913e-e0ba-4396-8743-f4378946d707 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.306506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.185800Z digest=sha256:77f389430661459d2c2878c049a8270e42661a0dfc0c6bef6be7acb94785649e

Observation 19f9dbbe-8ed9-411c-bb04-0d4f57fc7fb6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.228030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.228030Z digest=sha256:7989390400f1f8f2213cb9ff984b7c3eed75f8a1e0734d40a269a9d5d327ca78

Observation fb374256-0dd1-4b4d-a3ca-c07e0744ebe3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.290223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.267735Z digest=sha256:a8c258179ca7a30ec0119f2f90c1e2b66bb54a89c008e18ab2b31462beddb450

Observation d4690e8f-a6ff-4b89-af78-e6dd86db56b8 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.303597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.303597Z digest=sha256:7297f757342a3ccd9240555e54a9a8d6f742c3a2fd4eca34c46204ee8468a736

Observation 7166f125-6432-4892-9b2b-cf08503f23c5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.352974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.352974Z digest=sha256:2ebe3c982bd6b7b5ec0640db9c2296e2ff2123b726288ebd7ccce7f1b3fea791

Observation 3d591363-7b61-4c01-b648-91b2e7daa879 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.418189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.418189Z digest=sha256:d774d8bfdf03cdc4449b81a70f2eedd10b0f9e7c39963e7c63c9a4af2e3d0b3a

Observation 9a505477-76a1-41a3-bd06-4a82bd59794f · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.274288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.484181Z digest=sha256:ef5b7f3007095a401fca5ea6fcefea69fd8b85db5ea294d62bf3bc1bd03d1265

Observation f6ca50f9-4f86-4913-9064-40a506cc1344 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.265367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.552826Z digest=sha256:076202e958c5dae1858c08c9e9fd8737f1ac3dd10592cf678c406bae9152b4b9

Observation 554cc7f2-cfbf-4f4f-a6e6-a6711c45b3ac · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.586973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.586973Z digest=sha256:89d153b793e74b2f7adaabb720eaaaa55cab9c7d7f493e4cfdede0a318beb0aa

Observation 851cadbf-3692-4203-9ab4-74d172aada91 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.251177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.623691Z digest=sha256:52a645973b8f3a627a387dedb88e484505b2be07f77270ce2857e1c98c2aefa4

Observation 6357a1b2-4a06-4e72-ae96-19caa511c722 · outbound

This paper cites Ku, Qian Liu, and Wenhu Chen.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Ku, Qian Liu, and Wenhu Chen

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:36:58.241963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:56.689422Z digest=sha256:68059d35c8cbf6643e4d56cb4f2e39b4f312d36e3c069bd5de9a1d40d0750a20

Observation 61e8e627-bce3-4b03-b763-7090f19bff57 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.756183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.756183Z digest=sha256:4e5511d6c7b0733530e8f4b30c93ce97feb2e4caa20f45a415857004b82c9ae1

Observation ee18bd61-5898-473d-b604-74072ed4d97b · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.800127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.800127Z digest=sha256:689fb0bd93a16c5882ad3e2354cd645462c010d3e467f8e56d0c305c84cc400f

Observation 9cbdcc6f-eed7-49c7-b3ef-b58c46889f77 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.871215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.871215Z digest=sha256:b6f5b1e698c450cdf6268582d7ea7a74ef83c8af7c18264243dba6f86e38facf

Observation aec1934d-55f9-4823-b709-84fa00b75b98 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.903250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.903250Z digest=sha256:af331c2a26ac4f278ebbe959ef970721e1ea21d509199adf3684ed7391ef714b

Observation b3b66b49-a8e3-45ae-99c4-5c84ace87e7b · outbound

This paper cites A Survey on Benchmarks of Multimodal Large Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? A Survey on Benchmarks of Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:56.960671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:56.960671Z digest=sha256:d81cbc0ee419a62220a72eca6e5b3affd0a1f6ba50d075085beda3741011b547

Observation 38f2f181-a075-4453-b6d4-6c11922694ef · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.221435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.035431Z digest=sha256:04b73ff9a73823b97c34e0fa4072695b3085711f3d3a2f76f27b8d2234cf1c6d

Observation 84c2d7c2-8b62-4e08-8ebc-ea82b2ad1476 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.212845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.083977Z digest=sha256:6a842b168bf0df3701ec93f5b666a1e67f539091e4bb50eb5befea5a5bacc531

Observation b565f4d8-0151-4476-b5da-228d5d773e7e · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.203854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.178525Z digest=sha256:9e4e0ddb7cf809fc796de2a8efbed5f5a0bae09a553ad1992ae4b364921b37f6

Observation 2382cdb4-0975-4350-b76c-b6060a36dc1b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.195026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.223584Z digest=sha256:04cc0dddc04fb4e97fcd9c1fbade69cf5d7d9edebf93460758f10b4fbaefb91c

Observation a3a0c202-9922-420b-9427-584464a29986 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.184983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.312487Z digest=sha256:5a5e32239d9bfa2629aa83dbaac730b26d5362d189e001e389e43c7aef1b40e6

Observation 2a5ac443-47a1-4cc5-9b78-d33bc962200b · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.175717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.355871Z digest=sha256:3791f56e5b3bea19ba4f67720e1c63db21df9157b18e14d46e5ef7b3ba07530f

Observation 9f40534c-5146-4bc6-bf49-8f471341bfa3 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.165973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.406025Z digest=sha256:2977d599adc1c1a35bc715b513e2565e1caadc520318aa756b01c2baa8dfdb22

Observation a779e404-8c5e-4802-935e-36120f0f9bea · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.156334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.493936Z digest=sha256:cc299ffd942969558c55492007773946e52385dc547c67504200f5e16f2b4f59

Observation cfa1177c-7f38-4a46-a26e-1c46c67e3bdd · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.146711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.529318Z digest=sha256:875106c7a21abade39a517b43ac8e415ef8ee27e1644857ad14e108049145a02

Observation 6f3ec60e-cef3-4ab6-ac1b-2abc939dcfe2 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.614358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.614358Z digest=sha256:8885b243affa235e237b07090669d2c97e5c379183a307314da745b02988d9e8

Observation d5b9b921-0af2-411c-99ba-ff972055fb86 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.130607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.660424Z digest=sha256:e7b476a2abee223a61897da61bf8577aa447b859d65651663113c65177eb4bb3

Observation fb57719b-1494-4abc-8166-e03811368756 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.121927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.724102Z digest=sha256:3d7da464ed092cb257239a91d55e60389d76bef4ecb8fc07d2b5b5a63e66bdef

Observation 315654a7-0260-4135-a85d-a0d4189a2a95 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.113028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.745924Z digest=sha256:ac7bc8ca111fcac3e83b61bfd8972d1d13b039dec8252d4af70b3b5d73dc1f96

Observation 75cecedd-f28b-458e-8c73-a1db0c65008d · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.103796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.749772Z digest=sha256:2bc0dbaf26966b6aa178d096a88f4387d3d3cc549de88522684ea71fd29fcd41

Observation 7e92028f-d1cf-455d-b143-4a5abba33883 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? LLaMA: Open and Efficient Foundation Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.752458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.752458Z digest=sha256:a55a236b72505295df039cd31ebfe4033253cb9b6cae2fe11736a9f6dfe85000

Observation dca4467a-b6e6-4873-a438-5a406bf2eba0 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.094986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.755306Z digest=sha256:7927c28245d468ff775af90e04f6c3e8ed2427185924c9dcfc7e806043e1d83f

Observation 7a6459ba-220b-4ae5-9fed-9da86d444cbf · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.758726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.758726Z digest=sha256:e958ac62acd062de94eded7cd2f42edff41891465f6aaaef66064a4a60db89f6

Observation 1d00e8dc-26b0-4a19-b456-da1fac3709de · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.761786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.761786Z digest=sha256:9482bfa449433834f655143474e44bea108ca420640a0a9135f61d38a91fbd3a

Observation a7b68da8-5e05-4361-96a0-b382231b5f3a · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.764923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.764923Z digest=sha256:f927d0105e34363f22d627b0567bc3bc99fef9d317c8421aa6afd5bcb08961df

Observation 60289f6a-c095-4d0f-8598-7e7a8d5c3156 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.768267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.768267Z digest=sha256:f1f6d279041511fcf4acc28a1eeaca6c10de13f1b99d9cf854e9323491321a11

Observation ea391491-2233-462b-99df-43c0d61d2dd6 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.078067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.771470Z digest=sha256:748b77d76a25cd350e1a90e56b1d4115cff60536c74f2a001532d6364f526534

Observation ae19ed72-386f-4706-9dd8-fc22b2680059 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.774458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.774458Z digest=sha256:062fd243ba5613a3adb2c520bfa01a1c376fd58dc102f10ddc2f33ce306762b7

Observation 48c77cfe-290f-4270-adb6-2b1adcf3f0d6 · outbound

This paper cites Long Context Transfer from Language to Vision.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Long Context Transfer from Language to Vision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.778449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.778449Z digest=sha256:e31750bf69c5ec2974e610fbbdec721d3c78680a236e2fac2aca4ea7df785b40

Observation 6e1a7f37-a7c4-4c2d-9e09-8187eea79bce · outbound

This paper cites Weinberger, and Yoav Artzi.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Weinberger, and Yoav Artzi

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.782456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.782456Z digest=sha256:5c73af0db2a114f14b10ef33d9916d71fdceb8710027bf1d471c296d5881fcf2

Observation ac596079-51f6-453a-9ae1-83a32905c794 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T04:36:57.785475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:36:57.785475Z digest=sha256:75989f13985d996de0488f50ccbe96e4c760d3aae95612811fccd5b443739c78

Observation 7632a9aa-e197-41da-a670-197147fc38e5 · outbound

This paper cites an unresolved cited work.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:36:58.050574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.788190Z digest=sha256:266f37461fe7b6f1b596f47a48ae793dc94ff069a4a17502217f3a0e04f270cf

Observation a3890a0f-9ccd-4788-a921-54571984658d · outbound

This paper cites Zwaan and Gabriel A.

Burn After Reading: Do Multimodal Large Language Models Truly Capture Order of Events in Image Sequences? Zwaan and Gabriel A

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:36:57.954724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:36:57.791208Z digest=sha256:88ce5d30cb351ec1844890cd8221bdf3f7d71364a18dfe60fb36537c5f1bfec7

Pith citing papers

No inbound Pith citation observations are available.