Pith. sign in

Paper Citation Record · LEDGER

Token Communication for Multimodal Large Language Model

As of 18 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2608.07279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07279 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:07.775436Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ecaad9c-8d8b-4506-b89d-9adb72e8e9ae · outbound

This paper cites GPT-4 Technical Report.

Token Communication for Multimodal Large Language Model GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.632006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.632006Z digest=sha256:0c57032e55aa12609f8fc6d287c04a3d0e1f02aac3129727a028f9c6d39170c5

Observation a5db22f9-d36f-4ce1-97e9-7c72983add4f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Token Communication for Multimodal Large Language Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.636889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.636889Z digest=sha256:aad1a9a8a577324603b136692081f92cad426ff1725d7fe66b0ad47c8011d26b

Observation bbf9c675-e846-45cc-b5c7-4a800a01415b · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Token Communication for Multimodal Large Language Model Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.641325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.641325Z digest=sha256:6b2530f15f5844e24c3b49b8302c202a18b5f1866cf75f08cc51e71c034771c2

Observation 7153ef3e-b8b9-4038-b605-09d8508c45c5 · outbound

This paper cites Qwen3-VL Technical Report.

Token Communication for Multimodal Large Language Model Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.645580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.645580Z digest=sha256:38caa93cab51c99bbc9803eb74dc6c8cd9e2e52a656d79ed7728eabe23d7f8b2

Observation 855e989b-70ce-4808-95ec-40eb6bb7ad47 · outbound

This paper cites Attention is all you need,.

Token Communication for Multimodal Large Language Model Attention is all you need,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.058515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.649471Z digest=sha256:52ddb5d404171b84c3c81c5394fb649c3642fc227557b71a8f385d359effd4d7

Observation 888e5bbc-85e5-4bcb-a713-1c94cebfac09 · outbound

This paper cites State of AI: An empirical 100 trillion token study with openrouter,.

Token Communication for Multimodal Large Language Model State of AI: An empirical 100 trillion token study with openrouter,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.653190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.653190Z digest=sha256:7e6975a6b1f24aa6c33bca157e7f396faefffa5405938e7286182b371f0e9212

Observation 98920b2d-aa90-4611-9177-bb0b1f590f15 · outbound

This paper cites Token communications: A large model-driven framework for cross-modal context-aware semantic communications,.

Token Communication for Multimodal Large Language Model Token communications: A large model-driven framework for cross-modal context-aware semantic communications,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.047850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.657057Z digest=sha256:9f7bb302cf5eee48039f2c02c5ca236b3610440672ecc1d0b41fb2cc03494d9b

Observation bcd11c82-346c-4ccb-818a-026a3019c624 · outbound

This paper cites Adaptive semantic token communication for Transformer-based edge inference,.

Token Communication for Multimodal Large Language Model Adaptive semantic token communication for Transformer-based edge inference,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.036162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.660459Z digest=sha256:f6d19e9a8154da21281d993c0da9d43c5a65633d07be0b3c4dcd939b96495fa6

Observation 0faf6efa-f02f-416d-bf62-9454c1f8736d · outbound

This paper cites ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,.

Token Communication for Multimodal Large Language Model ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.024453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.663886Z digest=sha256:a5b3590f1424a7ccf8fcc35c5418225721c5a9a8779e4ec8fcad2a5132154c91

Observation 53ef0cef-ddbe-48fa-87cf-f983fe34e989 · outbound

This paper cites Joint semantic-channel coding and modulation for token communications,.

Token Communication for Multimodal Large Language Model Joint semantic-channel coding and modulation for token communications,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.667285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.667285Z digest=sha256:97e2568cf299ae8d1c6ac9a9e062ea9d29c3903a36381714b331fe3e93ea0a35

Observation 5360908d-3f45-436c-9f14-52728af9f7ae · outbound

This paper cites Towards practical real-time neural video compression,.

Token Communication for Multimodal Large Language Model Towards practical real-time neural video compression,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.006774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.670497Z digest=sha256:7382adfddaf083a1ee571f67d85df408baef58ed48683d93c10c85e79bd03319

Observation 774c5ba9-3112-4a31-bc91-973c930037de · outbound

This paper cites ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,.

Token Communication for Multimodal Large Language Model ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.996273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.674091Z digest=sha256:28ba8f6f2e059e8e2f5dfab2f0dcd93e46ff3a0f5f4ec79c6c86cc79b47e45e3

Observation d19045bc-e121-48b4-ade7-c00e01d47e4a · outbound

This paper cites Cache-to-cache: Direct semantic communication between large lan- guage models,.

Token Communication for Multimodal Large Language Model Cache-to-cache: Direct semantic communication between large lan- guage models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.985849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.677376Z digest=sha256:8dc195788c44d5a69df732ff10b7d7640b87cd30c0c87e47d0e50886f6b5d10e

Observation 8a5e0cce-f707-4287-b5a5-4e5159a1190f · outbound

This paper cites Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication.

Token Communication for Multimodal Large Language Model Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:08.499206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.680779Z digest=sha256:840c17402d363dc7665860697ed7b400f9607f9d01c61b78abe70d0f51ae14ab

Observation 45f102f9-a02a-4783-998e-39f9f8d459e8 · outbound

This paper cites Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,.

Token Communication for Multimodal Large Language Model Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.975666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.684504Z digest=sha256:5139bf9cacb48bc24dab8d7e337c970d30e47da30db6d025605344d844d4060f

Observation dad7d109-7b26-47bd-b32e-385d048287fb · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token Communication for Multimodal Large Language Model Learning transferable visual models from natural language supervision,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.965249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.687950Z digest=sha256:cc417546e9c9dd260b35b5867b3d28a62bd412ba358e78405f0f217603ddc0c4

Observation dddb6459-27f1-4f03-97d1-5f2186057d3a · outbound

This paper cites Sigmoid loss for language image pre-training,.

Token Communication for Multimodal Large Language Model Sigmoid loss for language image pre-training,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.955174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.691678Z digest=sha256:23ddc91238ad56dd5fa5a3a5fe7ad08b6e8c732974627b15b3917b45b33e27ae

Observation 5f90155f-efc7-482f-89fa-8ba0e75786c4 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Token Communication for Multimodal Large Language Model SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.695904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.695904Z digest=sha256:d91cde9486f3b6d579467af3fffc42e856eeace6e6cd9ea10946615e77c21442

Observation 927cf759-a421-45b5-8124-c9557287e158 · outbound

This paper cites FiLM: Visual reasoning with a general conditioning layer,.

Token Communication for Multimodal Large Language Model FiLM: Visual reasoning with a general conditioning layer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.945436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.699701Z digest=sha256:e16ad183d1c83110d48fca533e1f898cea0288cb9d5e41040cb1e2ed0a201f0c

Observation c034656f-aca6-429a-94bb-8a9fd4d49eb1 · outbound

This paper cites Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,.

Token Communication for Multimodal Large Language Model Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.703005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.703005Z digest=sha256:288f00e6803859a5480d92e0c44ffd873235bd25d4fea247166b57b8e6b572ec

Observation ddf72d9f-e075-4134-a2b7-d38ba30fd2c4 · outbound

This paper cites Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,.

Token Communication for Multimodal Large Language Model Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.934886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.706786Z digest=sha256:d89b3962d9f89255e8ad812ab8782db777895f10651aa361da10c936bc823daa

Observation d173f0e0-cbec-4829-9f1f-162432ab935f · outbound

This paper cites Semantic Packet Aggregation for Token Communication via genetic beam search,.

Token Communication for Multimodal Large Language Model Semantic Packet Aggregation for Token Communication via genetic beam search,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.923828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.710123Z digest=sha256:44bf9462b819f61c9a4c0ac9bba3e8736859c58ea8ca6fc88a3b1c72a4cd7e94

Observation 541ad56a-b543-4f96-9afe-ed9c7da91386 · outbound

This paper cites Vector quantized se- mantic communication system,.

Token Communication for Multimodal Large Language Model Vector quantized se- mantic communication system,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.913321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.713881Z digest=sha256:d783a63b2d9a02f720746a0b0fb89b014051d1b0ab8ee93dfa09920f1fb098ef

Observation 9e96452b-eee0-4301-ac7f-e1f9140813ea · outbound

This paper cites TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,.

Token Communication for Multimodal Large Language Model TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.717438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.717438Z digest=sha256:fd791fd63f187c7bb161bec71c7dde6955b7ddf08a5940581a5745e7a2e02ba8

Observation 078f5ff2-1485-42be-87a6-914b35814f3a · outbound

This paper cites VILA-U: A unified foundation model integrating visual understanding and generation,.

Token Communication for Multimodal Large Language Model VILA-U: A unified foundation model integrating visual understanding and generation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.901925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.720344Z digest=sha256:c6e4ce0f9efa68e33adba96fc02771d67a19762e51f37ddd1d45aa57f1ebe6be

Observation f7c954d0-481f-4901-bf9f-a70be26be229 · outbound

This paper cites BPG Image Format,.

Token Communication for Multimodal Large Language Model BPG Image Format,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.891074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.723210Z digest=sha256:a49b6ee04c1c900703435c461a183d5cfa78887206a1d2774d40bdba43048e8e

Observation 9ef23ea2-cd1d-406b-9d6b-40c50e7c09e6 · outbound

This paper cites VVC Test Model,.

Token Communication for Multimodal Large Language Model VVC Test Model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.880908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.726261Z digest=sha256:4da90a49086842d8e63cd206a8af6222a80d7421d13b351bfc321d1dd345e643

Observation 2badacf3-8c4f-49da-817b-18633f90c9de · outbound

This paper cites Generative latent coding for ultra-low bitrate image compression,.

Token Communication for Multimodal Large Language Model Generative latent coding for ultra-low bitrate image compression,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.870828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.729308Z digest=sha256:7368230fa6b147d8e6d69c0fed93b67a1e69ecc5cd11224b6d73520e22f4311f

Observation fd8b8371-275f-4f28-877e-9c1ffcbff8b0 · outbound

This paper cites ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization.

Token Communication for Multimodal Large Language Model ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:07.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.732195Z digest=sha256:5dac0aa24bf04d295dbecbd5c02e32534743acb93c740cc697ca0ab2cb8d7ea6

Observation c8e0059a-576e-43b0-8343-2033df4b29f9 · outbound

This paper cites TransTIC: Transferring Transformer-based image compression from human perception to machine perception,.

Token Communication for Multimodal Large Language Model TransTIC: Transferring Transformer-based image compression from human perception to machine perception,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.859941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.735347Z digest=sha256:fc937057eeb336485add829bd2c4cb4fd5b4f507ad083710dd8ee65149efd369

Observation 79ecba71-da28-4fcb-a584-28c608137512 · outbound

This paper cites High efficiency image compression for large visual-language models,.

Token Communication for Multimodal Large Language Model High efficiency image compression for large visual-language models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.848679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.738505Z digest=sha256:11da2fed635821465e6f6dc49f30fcb0c3f19a905acf4c80715cb8c001ff4baa

Observation b5a30225-6560-4c45-b896-b19b1500cc10 · outbound

This paper cites Bridging compressed image latents and multimodal large language models,.

Token Communication for Multimodal Large Language Model Bridging compressed image latents and multimodal large language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.837576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.741570Z digest=sha256:d44d0890063fadda488de28976269f7c5016090be10d90344ef73323be6250d6

Observation e59c94a8-4d6a-4ddf-865c-978c7f874692 · outbound

This paper cites When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,.

Token Communication for Multimodal Large Language Model When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.826328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.744536Z digest=sha256:9cec2431b275d3e624d77d5fcdc2118650f08ea83aad8f10f1e59d7b7557a48f

Observation f7ebd7c4-0b21-4201-9a06-f44acca20650 · outbound

This paper cites Variational image compression with a Scale Hyperprior,.

Token Communication for Multimodal Large Language Model Variational image compression with a Scale Hyperprior,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.814243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.748016Z digest=sha256:6d028cc904d7565af1c05ef1372f5b31d05508cf04fda54a94ba01c3b38dac18

Observation a22fd1e2-0f2c-42c1-a700-d72dd191200a · outbound

This paper cites Deep joint source- channel coding for wireless image transmission,.

Token Communication for Multimodal Large Language Model Deep joint source- channel coding for wireless image transmission,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.751553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.751553Z digest=sha256:ad277df6221ece6d453f489f930f24dc0b8f861b5bcdafcaaa8aa06359f49afd

Observation 1e87d3b4-ce2e-4e55-91a1-157cc5fff32c · outbound

This paper cites Qwen-Image Technical Report.

Token Communication for Multimodal Large Language Model Qwen-Image Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.755370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.755370Z digest=sha256:8d7f74f7da00e7b8eef93b44b9fb2bc9741455a4c2f192417a6c08766934dd67

Observation 1c680780-4ab5-4fe4-afa5-a90c939647ee · outbound

This paper cites RoFormer: En- hanced transformer with rotary position embedding,.

Token Communication for Multimodal Large Language Model RoFormer: En- hanced transformer with rotary position embedding,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.794331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.759074Z digest=sha256:73af222f3ed14c132a7482166e22223831fb4b7f0befb811b68b4524c4173f7f

Observation 03d46eb8-dd22-46a2-b7ad-dae4a800caeb · outbound

This paper cites LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,.

Token Communication for Multimodal Large Language Model LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.780968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.762426Z digest=sha256:22e027a7a753a46e2d98eeb583991d1732d139bc4ab6c4f0f252163d0292803a

Observation 5de70b01-9f75-4c32-a013-bc5e60a6bc2e · outbound

This paper cites MME: A comprehen- sive evaluation benchmark for multimodal large language models,.

Token Communication for Multimodal Large Language Model MME: A comprehen- sive evaluation benchmark for multimodal large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.768647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.765590Z digest=sha256:0f2ea1f666bb0d35827a64c9062abdc20a729ccd43753df02560d66ed9ffea30

Observation 84fce95a-d273-4994-96e6-61589768554b · outbound

This paper cites Evaluating object hallucination in large vision-language models,.

Token Communication for Multimodal Large Language Model Evaluating object hallucination in large vision-language models,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.755815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.768948Z digest=sha256:3c9d05a3551eb7afd952e1e6aa1d9dcb30873b2045e00100b0a6fc3fbb6fc385

Observation ec0a164b-bf2c-4c03-b2e8-1e7f83cc6415 · outbound

This paper cites SEED-Bench: Benchmarking multimodal large language models,.

Token Communication for Multimodal Large Language Model SEED-Bench: Benchmarking multimodal large language models,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.743228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.772067Z digest=sha256:a81a015d833dbc21116910c587a16b82bdceefb3f472e65f389ce9f0a063b550

Observation 00ae0b86-3e77-46be-939d-ed1154bc40a9 · outbound

This paper cites Deep visual-semantic alignments for gen- erating image descriptions,.

Token Communication for Multimodal Large Language Model Deep visual-semantic alignments for gen- erating image descriptions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.730688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T11:13:07.775436Z digest=sha256:270e7ae7d3d3a807fc087cde0ae9e1d3fe58beaabb0c75a2f92205904d01f58a

Pith citing papers

No inbound Pith citation observations are available.