Pith. sign in

Paper Citation Record · LEDGER

Token Communication for Multimodal Large Language Model

As of 17 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2608.07279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07279 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:07.775436Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ecaad9c-8d8b-4506-b89d-9adb72e8e9ae · outbound

This paper cites GPT-4 Technical Report.

Token Communication for Multimodal Large Language Model GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.632006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.632006Z digest=sha256:0c57032e55aa12609f8fc6d287c04a3d0e1f02aac3129727a028f9c6d39170c5

Observation a5db22f9-d36f-4ce1-97e9-7c72983add4f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Token Communication for Multimodal Large Language Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.636889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.636889Z digest=sha256:aad1a9a8a577324603b136692081f92cad426ff1725d7fe66b0ad47c8011d26b

Observation bbf9c675-e846-45cc-b5c7-4a800a01415b · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Token Communication for Multimodal Large Language Model Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.641325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.641325Z digest=sha256:6b2530f15f5844e24c3b49b8302c202a18b5f1866cf75f08cc51e71c034771c2

Observation 7153ef3e-b8b9-4038-b605-09d8508c45c5 · outbound

This paper cites Qwen3-VL Technical Report.

Token Communication for Multimodal Large Language Model Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.645580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.645580Z digest=sha256:38caa93cab51c99bbc9803eb74dc6c8cd9e2e52a656d79ed7728eabe23d7f8b2

Observation 855e989b-70ce-4808-95ec-40eb6bb7ad47 · outbound

This paper cites Attention is all you need,.

Token Communication for Multimodal Large Language Model Attention is all you need,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.058515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.649471Z digest=sha256:63dacba452c371dff59f7d22649ee3bb35dc9a73699411fd5b000ecb73b04afd

Observation 888e5bbc-85e5-4bcb-a713-1c94cebfac09 · outbound

This paper cites State of AI: An empirical 100 trillion token study with openrouter,.

Token Communication for Multimodal Large Language Model State of AI: An empirical 100 trillion token study with openrouter,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.653190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.653190Z digest=sha256:7e6975a6b1f24aa6c33bca157e7f396faefffa5405938e7286182b371f0e9212

Observation 98920b2d-aa90-4611-9177-bb0b1f590f15 · outbound

This paper cites Token communications: A large model-driven framework for cross-modal context-aware semantic communications,.

Token Communication for Multimodal Large Language Model Token communications: A large model-driven framework for cross-modal context-aware semantic communications,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.047850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.657057Z digest=sha256:7a367c2b188469122f2f4cd861b1ba32f65596558cb50af4b45fbf339690ee39

Observation bcd11c82-346c-4ccb-818a-026a3019c624 · outbound

This paper cites Adaptive semantic token communication for Transformer-based edge inference,.

Token Communication for Multimodal Large Language Model Adaptive semantic token communication for Transformer-based edge inference,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.036162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.660459Z digest=sha256:5e8632834d4e042c040552b170cae503907426631a78ea0974b5f45b9de6710b

Observation 0faf6efa-f02f-416d-bf62-9454c1f8736d · outbound

This paper cites ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,.

Token Communication for Multimodal Large Language Model ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.024453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.663886Z digest=sha256:f7dcfbb7e2678e1f64f650c9d1ac9ff5258e0e83a57af2c31197d33534fcfeaf

Observation 53ef0cef-ddbe-48fa-87cf-f983fe34e989 · outbound

This paper cites Joint semantic-channel coding and modulation for token communications,.

Token Communication for Multimodal Large Language Model Joint semantic-channel coding and modulation for token communications,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.667285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.667285Z digest=sha256:97e2568cf299ae8d1c6ac9a9e062ea9d29c3903a36381714b331fe3e93ea0a35

Observation 5360908d-3f45-436c-9f14-52728af9f7ae · outbound

This paper cites Towards practical real-time neural video compression,.

Token Communication for Multimodal Large Language Model Towards practical real-time neural video compression,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.006774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.670497Z digest=sha256:734f05c8dfbe6e52145c0d2d307a1530d4278f1fed1f15e00c50188a2850b38d

Observation 774c5ba9-3112-4a31-bc91-973c930037de · outbound

This paper cites ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,.

Token Communication for Multimodal Large Language Model ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.996273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.674091Z digest=sha256:f91e7d87e5bcc5b10985cfb4e0c9cbc903fd76d36f6a15fa90456e49449aeebc

Observation d19045bc-e121-48b4-ade7-c00e01d47e4a · outbound

This paper cites Cache-to-cache: Direct semantic communication between large lan- guage models,.

Token Communication for Multimodal Large Language Model Cache-to-cache: Direct semantic communication between large lan- guage models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.985849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.677376Z digest=sha256:6db600f5f09872f42b2a44dd907adcb7bfbdade71532deb9683bd82a29846655

Observation 8a5e0cce-f707-4287-b5a5-4e5159a1190f · outbound

This paper cites Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication.

Token Communication for Multimodal Large Language Model Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:08.499206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.680779Z digest=sha256:5a576f74e47198ffa99e85db69b1049c67c4daf28d0c947ec1140f8dbe4ee511

Observation 45f102f9-a02a-4783-998e-39f9f8d459e8 · outbound

This paper cites Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,.

Token Communication for Multimodal Large Language Model Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.975666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.684504Z digest=sha256:32ae20633852eca61ee4c200fbdd442f43945bb6f80f9f84c5959aa5918653c1

Observation dad7d109-7b26-47bd-b32e-385d048287fb · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token Communication for Multimodal Large Language Model Learning transferable visual models from natural language supervision,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.965249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.687950Z digest=sha256:71bff602d8040b9fe7d5bf0303d2e57e05dfb315600a1761bf5e9f2d63f5f5bb

Observation dddb6459-27f1-4f03-97d1-5f2186057d3a · outbound

This paper cites Sigmoid loss for language image pre-training,.

Token Communication for Multimodal Large Language Model Sigmoid loss for language image pre-training,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.955174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.691678Z digest=sha256:4894d0d1c59e9f6c90b42fba30099820ffdaf3319ba8257ddbda2b0e1396d0ba

Observation 5f90155f-efc7-482f-89fa-8ba0e75786c4 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Token Communication for Multimodal Large Language Model SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.695904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.695904Z digest=sha256:d91cde9486f3b6d579467af3fffc42e856eeace6e6cd9ea10946615e77c21442

Observation 927cf759-a421-45b5-8124-c9557287e158 · outbound

This paper cites FiLM: Visual reasoning with a general conditioning layer,.

Token Communication for Multimodal Large Language Model FiLM: Visual reasoning with a general conditioning layer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.945436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.699701Z digest=sha256:284c62370e995762c11d81c94a78b02f58b61418c07fabaacfb4d301036e19e7

Observation c034656f-aca6-429a-94bb-8a9fd4d49eb1 · outbound

This paper cites Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,.

Token Communication for Multimodal Large Language Model Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.703005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.703005Z digest=sha256:288f00e6803859a5480d92e0c44ffd873235bd25d4fea247166b57b8e6b572ec

Observation ddf72d9f-e075-4134-a2b7-d38ba30fd2c4 · outbound

This paper cites Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,.

Token Communication for Multimodal Large Language Model Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.934886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.706786Z digest=sha256:ae5a0b267837ee0c6e28fde93d45f315b1ad896f45bf5f7c0539b8219d4e0f29

Observation d173f0e0-cbec-4829-9f1f-162432ab935f · outbound

This paper cites Semantic Packet Aggregation for Token Communication via genetic beam search,.

Token Communication for Multimodal Large Language Model Semantic Packet Aggregation for Token Communication via genetic beam search,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.923828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.710123Z digest=sha256:1d232bc09d05b8c446605088bbee874826a5c99413972d8e1e4ef62306641a34

Observation 541ad56a-b543-4f96-9afe-ed9c7da91386 · outbound

This paper cites Vector quantized se- mantic communication system,.

Token Communication for Multimodal Large Language Model Vector quantized se- mantic communication system,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.913321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.713881Z digest=sha256:49d0ec98283178f263225dacab88c577757b9deab646421baded156932a46238

Observation 9e96452b-eee0-4301-ac7f-e1f9140813ea · outbound

This paper cites TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,.

Token Communication for Multimodal Large Language Model TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.717438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.717438Z digest=sha256:fd791fd63f187c7bb161bec71c7dde6955b7ddf08a5940581a5745e7a2e02ba8

Observation 078f5ff2-1485-42be-87a6-914b35814f3a · outbound

This paper cites VILA-U: A unified foundation model integrating visual understanding and generation,.

Token Communication for Multimodal Large Language Model VILA-U: A unified foundation model integrating visual understanding and generation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.901925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.720344Z digest=sha256:26e3d79e6eb7b4edd9f0d2fa83c9994f780ea3d87e23bd5808cc3f5ece246740

Observation f7c954d0-481f-4901-bf9f-a70be26be229 · outbound

This paper cites BPG Image Format,.

Token Communication for Multimodal Large Language Model BPG Image Format,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.891074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.723210Z digest=sha256:871f2ae0025b286fd9e4c212d85f0aebb1b16b8eeeace4f9dbc3c1f0057e8153

Observation 9ef23ea2-cd1d-406b-9d6b-40c50e7c09e6 · outbound

This paper cites VVC Test Model,.

Token Communication for Multimodal Large Language Model VVC Test Model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.880908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.726261Z digest=sha256:97df6652763314f84187c71e32865fdd07f7a2c59e1416341538cd008fda0841

Observation 2badacf3-8c4f-49da-817b-18633f90c9de · outbound

This paper cites Generative latent coding for ultra-low bitrate image compression,.

Token Communication for Multimodal Large Language Model Generative latent coding for ultra-low bitrate image compression,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.870828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.729308Z digest=sha256:a5d5a79617b893ee7c3f4c5ab07f2284c2235a3ed5d9d376ed86d016ba2a3300

Observation fd8b8371-275f-4f28-877e-9c1ffcbff8b0 · outbound

This paper cites ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization.

Token Communication for Multimodal Large Language Model ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:07.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.732195Z digest=sha256:784cd5ca71128bca2e5ffc6f773a5223db0345398044531c2a7b6d3d2ea5ca8e

Observation c8e0059a-576e-43b0-8343-2033df4b29f9 · outbound

This paper cites TransTIC: Transferring Transformer-based image compression from human perception to machine perception,.

Token Communication for Multimodal Large Language Model TransTIC: Transferring Transformer-based image compression from human perception to machine perception,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.859941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.735347Z digest=sha256:b47379858899351d099c34bdd2f6fa9b85a4e66a963f04e0fc37db5c32bb9095

Observation 79ecba71-da28-4fcb-a584-28c608137512 · outbound

This paper cites High efficiency image compression for large visual-language models,.

Token Communication for Multimodal Large Language Model High efficiency image compression for large visual-language models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.848679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.738505Z digest=sha256:2f2d8295e188e19ec30c419f355e721e09b7cd19523ae2b53edea52d902a075b

Observation b5a30225-6560-4c45-b896-b19b1500cc10 · outbound

This paper cites Bridging compressed image latents and multimodal large language models,.

Token Communication for Multimodal Large Language Model Bridging compressed image latents and multimodal large language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.837576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.741570Z digest=sha256:e15b4500f5dc5247dba463463cde7266cd8e9def4314ab4a645d0f041323ffbd

Observation e59c94a8-4d6a-4ddf-865c-978c7f874692 · outbound

This paper cites When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,.

Token Communication for Multimodal Large Language Model When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.826328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.744536Z digest=sha256:2de8df7e4be63dea3be401975032d5c09345345c8aef6d259caf1c38acf5fa8b

Observation f7ebd7c4-0b21-4201-9a06-f44acca20650 · outbound

This paper cites Variational image compression with a Scale Hyperprior,.

Token Communication for Multimodal Large Language Model Variational image compression with a Scale Hyperprior,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.814243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.748016Z digest=sha256:ee20a2cac25729f45dd80d4bc68f3d3fadec543075c538ee51c982694560681e

Observation a22fd1e2-0f2c-42c1-a700-d72dd191200a · outbound

This paper cites Deep joint source- channel coding for wireless image transmission,.

Token Communication for Multimodal Large Language Model Deep joint source- channel coding for wireless image transmission,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.751553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.751553Z digest=sha256:ad277df6221ece6d453f489f930f24dc0b8f861b5bcdafcaaa8aa06359f49afd

Observation 1e87d3b4-ce2e-4e55-91a1-157cc5fff32c · outbound

This paper cites Qwen-Image Technical Report.

Token Communication for Multimodal Large Language Model Qwen-Image Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.755370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.755370Z digest=sha256:8d7f74f7da00e7b8eef93b44b9fb2bc9741455a4c2f192417a6c08766934dd67

Observation 1c680780-4ab5-4fe4-afa5-a90c939647ee · outbound

This paper cites RoFormer: En- hanced transformer with rotary position embedding,.

Token Communication for Multimodal Large Language Model RoFormer: En- hanced transformer with rotary position embedding,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.794331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.759074Z digest=sha256:fcfaba47a5cbf2eb1ac09a6618bac98af8c50f1043702d3b181991aa2216fe1a

Observation 03d46eb8-dd22-46a2-b7ad-dae4a800caeb · outbound

This paper cites LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,.

Token Communication for Multimodal Large Language Model LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.780968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.762426Z digest=sha256:29bbb89f516aebfb64ca2f8ebb93b785d86a4db9b9cf4a47f441eb658aeff823

Observation 5de70b01-9f75-4c32-a013-bc5e60a6bc2e · outbound

This paper cites MME: A comprehen- sive evaluation benchmark for multimodal large language models,.

Token Communication for Multimodal Large Language Model MME: A comprehen- sive evaluation benchmark for multimodal large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.768647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.765590Z digest=sha256:0d7f6249ea7abb22102303c1f305e228b4ad79d0bb4443dccbd23124de9ef4e0

Observation 84fce95a-d273-4994-96e6-61589768554b · outbound

This paper cites Evaluating object hallucination in large vision-language models,.

Token Communication for Multimodal Large Language Model Evaluating object hallucination in large vision-language models,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.755815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.768948Z digest=sha256:02a576bdd743beb3830f365246a71d6a837e50217e97f389167cd437c86ee18f

Observation ec0a164b-bf2c-4c03-b2e8-1e7f83cc6415 · outbound

This paper cites SEED-Bench: Benchmarking multimodal large language models,.

Token Communication for Multimodal Large Language Model SEED-Bench: Benchmarking multimodal large language models,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.743228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.772067Z digest=sha256:70d92703a11df34ad5870e1c1a0a137b2767e71c8b17cf172bf49525e644411e

Observation 00ae0b86-3e77-46be-939d-ed1154bc40a9 · outbound

This paper cites Deep visual-semantic alignments for gen- erating image descriptions,.

Token Communication for Multimodal Large Language Model Deep visual-semantic alignments for gen- erating image descriptions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.730688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T11:13:07.775436Z digest=sha256:e8285ab78ad8f02b40c29a628849efa2c18c753d765a8ab2a3d415ac7fbf7425

Pith citing papers

No inbound Pith citation observations are available.