Pith. sign in

Paper Citation Record · LEDGER

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.01805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.01805 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:28:31.007992Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact3
  • verified fuzzy23
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 58c4b7eb-5136-45ed-9d07-f7ef1e6eed79 · outbound

This paper cites Attention is all you need,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Attention is all you need,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.887978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.162115Z digest=sha256:e6dea1ad0a452c0c71bbe52d6f90f132e1edab0f77ed54b38d8191082e2b1e0b

Observation ed321ea6-c919-4db1-908b-82444a3d586c · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Flamingo: a visual language model for few-shot learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.678178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.244021Z digest=sha256:1dd69559f2425f4f6b0bf70b779e8f6567161e07d2855e7965ead2847720ebbe

Observation b15e45c5-ffd7-4bc5-a1bc-8c6eecc18dd3 · outbound

This paper cites BLIP-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks BLIP-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.468973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.357176Z digest=sha256:be6c3986c0e54b9636f43f68c5a5a1aa038ff1331d25e79f7977c5f89e6680c8

Observation 77a02c35-d8e1-46ea-aab5-a0461c5db666 · outbound

This paper cites Visual instruction tuning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Visual instruction tuning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.154463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.480107Z digest=sha256:08b4450970c0af2df7eeebe98ef9095bfe3c630d9069ac47cfbc28c127d87a75

Observation d7c8ea46-75e0-40fd-a5c4-67bee9205e83 · outbound

This paper cites GPT-4V(ision) system card,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks GPT-4V(ision) system card,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.901875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.566506Z digest=sha256:27c9777316ef56bf0b8cd33677ede51d9aae73e3ac14722f7566eaa09973b0e1

Observation a7de76dd-a279-4fba-bad9-47a991aed3c7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Gemini: A Family of Highly Capable Multimodal Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:27.682913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:27.682913Z digest=sha256:d8809e8e5eee297d362d8659b64d7cb5bf9dc4fe961fd70385816061297b5c02

Observation 3a6a6579-913a-4db8-8c3f-40291285f713 · outbound

This paper cites Multimodal Large Language Models: A Survey.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Multimodal Large Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:27.791343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:27.791343Z digest=sha256:8636b2342363df3829cfb3b1e83e2500f264c24857ac6fb934c44a381ca941a7

Observation 4b7ad9f2-723d-495f-8c70-bbbcceb3df7e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Learning transferable visual models from natural language supervision,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.663757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:27.907772Z digest=sha256:b2b0ca99c6a85b309bcb5b280b06fa7243560e0fcf8de40b44249b0eae251a60

Observation 1c579584-1d74-401b-bf4d-c766e6744607 · outbound

This paper cites GPipe: Efficient training of giant neural networks using pipeline parallelism,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks GPipe: Efficient training of giant neural networks using pipeline parallelism,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.382048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:28.008715Z digest=sha256:686344a650786082f92ad8d66196b73e8a2d874acff8adba727078e533859d82

Observation 2fc2a36b-fc36-4aed-bc34-11bb0d7e18c6 · outbound

This paper cites Are we ready for autonomous driving? The KITTI vision benchmark suite,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Are we ready for autonomous driving? The KITTI vision benchmark suite,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.161539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:28.112758Z digest=sha256:5ad80bddeb3ff29fbc105e9692c005b7652fbe89f0c8bb8b93134616ba1e0e2b

Observation e968f2ef-eb65-473b-a658-f297bfcc937b · outbound

This paper cites CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.979321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:28.245719Z digest=sha256:0993818617003e04b24ab602eff615881fc326a056122fba997cd72bd87b8a11

Observation 4693fc8e-b960-4e65-a0da-3112d475814f · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks On the Opportunities and Risks of Foundation Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.330922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.330922Z digest=sha256:b8c451e092a4b5712eb34adb8b73cd19d935980749b7693e0e5e9b0dacc0fd50

Observation ded83ce8-1ede-430b-8f3a-0071dfc61015 · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.696692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:28.464722Z digest=sha256:7681ed381b46a6a48904faa4009e4d3852fdf1f802756993b5f2053c679ef1fd

Observation 6dbc7e4e-478f-414e-ac8f-4d80661a55fa · outbound

This paper cites MoVA: Adapting Mixture of Vision Experts to Multimodal Context.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MoVA: Adapting Mixture of Vision Experts to Multimodal Context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.582750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.582750Z digest=sha256:96cdc612a500c74cdad5373c97d399c5c1ae8ed5476e4b07eda3ac21a54d3997

Observation fc9e6c91-c1ad-4e8a-853f-ccccfda0095e · outbound

This paper cites MoE-LLaVA: Mixture of Experts for Large Vision-Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MoE-LLaVA: Mixture of Experts for Large Vision-Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.723551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.723551Z digest=sha256:bde354534ce6b88778ce08633c0a80c34ed0c9b1d75adeb892dac1b61d0fef87

Observation 800b636e-a3e2-47da-b929-62c22976b28c · outbound

This paper cites Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.885000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.885000Z digest=sha256:538c273c532a7b56daf32f80124ba2072544a73f4f6359497e6eacdfdfbff2b9

Observation 5053dd87-4df4-425f-a960-ea5c9efa6517 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.028589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.028589Z digest=sha256:599f71be57381571af571d24b91b2314010304e239047e7bcd22ebe8ed38fa58

Observation a1c95fab-a0bd-4c02-bdee-1baf6f6db446 · outbound

This paper cites Quantization and training of neural networks for efficient integer-arithmetic-only inference,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Quantization and training of neural networks for efficient integer-arithmetic-only inference,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.498999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.118082Z digest=sha256:e2aef877c7d2566d018287e71c9bd76cf79b4ab74c9a0eb58083518cf047d8ea

Observation 7f590711-5ed2-4868-9ad3-f1c73bf8941a · outbound

This paper cites Semantic communi- cations for future Internet: Fundamentals, applications, and challenges,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Semantic communi- cations for future Internet: Fundamentals, applications, and challenges,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.259834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.175454Z digest=sha256:0c917f8cc271b6a546e8da894b9f72f40efa43ac3c35b5076eb5274df7bbdfda

Observation 7d2684f1-2d5e-4d65-bbb2-20e441305abc · outbound

This paper cites Model Context Protocol: An open standard for connecting AI assistants to the world,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Model Context Protocol: An open standard for connecting AI assistants to the world,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.048642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.274728Z digest=sha256:7a519f4a9d2a1bee6202352d9fb588e067cd549805de56ce0c8d937247ac8f6f

Observation 2a38a444-a312-4dcd-892d-59ae20333ce1 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:34.758087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.371945Z digest=sha256:2fea3199328fa9d78eab51964348a88543d2a23ed5af992f06ecfbea7e959d84

Observation cce9d416-f28e-4688-b77b-f1004b543d4d · outbound

This paper cites Variational inference: A review for statisticians,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Variational inference: A review for statisticians,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:34.538165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.453585Z digest=sha256:b237dd4a02c138d48622b3d33ebd0e551b51e2be2433af853516155ad11264d8

Observation 91906a3f-c0c1-4c59-989a-69f55e9b67da · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T05:28:34.258420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.523900Z digest=sha256:ca32abad4ab2190c734bb81af208bb42c3d511d07fa915b41d2abf8971ae1b0f

Observation 8d491f23-dcb5-4121-b3bb-c2effb726810 · outbound

This paper cites Goldsmith, Wireless Communications.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Goldsmith, Wireless Communications

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.582311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.582311Z digest=sha256:ba8bed61d1883f44e2cfb9c6959887cb652b20676114e388651708090f73ff00

Observation 0362e2d8-cdf2-47dc-a51d-fdc6fd1c5fe3 · outbound

This paper cites Correlation model for shadow fading in mobile radio systems,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Correlation model for shadow fading in mobile radio systems,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.658115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.658115Z digest=sha256:ace13136f72d53a8f345bc9624fbf0ce6fd5745915d35f7a3b00eae6fb902b65

Observation b0baa3cc-65b8-466e-955f-beb4bfd46632 · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.723707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.723707Z digest=sha256:e4984f703a8a3f8acc1b296da5713bc8758b53f20da70dc58661fb281e1d0d34

Observation 36a5f0af-7e02-4208-ac75-e46934ddb293 · outbound

This paper cites Billion-scale similarity search with GPUs,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Billion-scale similarity search with GPUs,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.784574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.784574Z digest=sha256:095438f182c37430ab5603031d7f82857bc2f0050c9b372f3cf0d7bed4e10aac

Observation 0da89616-29ca-4e87-a233-58bc37123e45 · outbound

This paper cites Design of coherence- aware channel indication and prediction for rate adaptation,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Design of coherence- aware channel indication and prediction for rate adaptation,

Reference 28

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.716166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.850481Z digest=sha256:02ca34329eaa5c799cf3d4405925cd5d0224931435a6e60c98171dd97a58efea

Observation 7ce1394c-4d72-4683-a49f-41b8d0a5dfd1 · outbound

This paper cites Bayesian Forecasting and Dynamic Models,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Bayesian Forecasting and Dynamic Models,

Reference 29

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.477624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.908023Z digest=sha256:da86d6f05b07522765a917e586f5aecfdace423a876bd79ed8a26be3bbe13a00

Observation 28e9d839-f3f2-4cc9-8531-de9553b01419 · outbound

This paper cites Exact Expressions for Kullback–Leibler Divergence for Univariate Distributions,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Exact Expressions for Kullback–Leibler Divergence for Univariate Distributions,

Reference 30

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.241572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:29.970215Z digest=sha256:0d20197811bca118c3e72fb582f2e06c86143f4187e1651708e2bf9a98decd45

Observation f56c7da4-9cd1-4f91-9c7f-8c997624180d · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.019543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.019543Z digest=sha256:0194f84a8ddbaaad6ad07ba5b22b6da12b3106a5b3bde00860333e9e508c315f

Observation a31d8df1-eb42-4528-9056-a0e2f1567566 · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Learn to explain: Multimodal reasoning via thought chains for science question answering,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.966178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.065261Z digest=sha256:997f47825ece9eb719348a9aaa1e16e09d52cc86ba47c8e89279c3bbf81234d8

Observation 95778c97-8116-42e5-b4f8-c7e2cab0d1e7 · outbound

This paper cites EdgeViT: Efficient visual modeling for edge computing,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks EdgeViT: Efficient visual modeling for edge computing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.744255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.150058Z digest=sha256:2ce45ab74751f59ceba80d33f4b5e1cddaa43c0b8b437174770fba62aef6d501

Observation 08aff055-873e-4ef2-bb1c-568401557bfb · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks DINOv2: Learning Robust Visual Features without Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.225411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.225411Z digest=sha256:074b8659972a6a748f0239fd6f579065672de2680556a03ed5c08af309be264d

Observation 9058a541-53e6-438f-a111-27bfd4bec2f3 · outbound

This paper cites Carion, F.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Carion, F

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.517529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.260805Z digest=sha256:189b75c55da77d61a65b42b29f4ea56b7f2971ad983192d783d1e520724041f9

Observation 7863eb52-77b9-423b-a111-ea6429067e00 · outbound

This paper cites ”Segment anything.” Proceedings of the IEEE/CVF international conference on computer vision.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks ”Segment anything.” Proceedings of the IEEE/CVF international conference on computer vision

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.272765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.311042Z digest=sha256:84b075d5024dd79b7c77e1919f181eb59f69ca57c907afae427a5102fc9d6fdc

Observation ce69559f-c102-46ab-b8e5-a147c4252005 · outbound

This paper cites ”Pix2struct: Screenshot parsing as pretraining for visual language understanding.” International Conference on Machine Learning.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks ”Pix2struct: Screenshot parsing as pretraining for visual language understanding.” International Conference on Machine Learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.075370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.385137Z digest=sha256:e52a12e76feceebe8c90d82e2db8d2a58e0f3e7ad51e8584619b315b3392ae3c

Observation be705dd5-c18e-4b77-9ace-73c05ce9de7e · outbound

This paper cites DePlot: One-shot visual language reasoning by plot-to-table translation,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks DePlot: One-shot visual language reasoning by plot-to-table translation,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.440513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.440513Z digest=sha256:273939a8580ff92f9e7cf31c5931d8ede9be0bb0ffd721ec6002b4fa5a28757d

Observation 2746403a-70de-4f48-8073-6af1059634fe · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-06T05:28:32.867902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.521602Z digest=sha256:671487cead51c6bd04b48d57c5eb2f4ff2cfad985844c628d11f340484997d8b

Observation 50dad1d6-a554-4ca9-a40a-e9e7f5189ced · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.630380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.630380Z digest=sha256:844d92682326490c6f74709314d006b0d6cd086471e96717fed7c2b084c54091

Observation 6da73563-e187-441f-8a2f-5f71241f4148 · outbound

This paper cites MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.709973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.709973Z digest=sha256:d72537ba2c30f1101ea53c9a2a1f6c88b9fd8ed341d99da0df5633566cd2fb8a

Observation c4d33446-d77e-48f7-a07a-b4a0f4f72c5a · outbound

This paper cites Resource manage- ment with deep reinforcement learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Resource manage- ment with deep reinforcement learning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.621476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.819463Z digest=sha256:dfac8d4f6f861f070486b32a080688b1dff028fe0a244cf1f7e2f1f0f2b0d3b8

Observation c0ffb6cc-622d-4746-a4ae-1b65ca98c86e · outbound

This paper cites Human-level control through deep reinforcement learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Human-level control through deep reinforcement learning,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.486135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:30.918159Z digest=sha256:85b075670648d9b04dca3e6b43e930a5e862d29ddf05e9efa6553b454f8e788a

Observation 46009e0b-efe7-40a3-a11a-4b523ee682f1 · outbound

This paper cites Adaptive computation time for recurrent neural networks,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Adaptive computation time for recurrent neural networks,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.290618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T05:28:31.007992Z digest=sha256:3aa4c05c4e41163201ab95d377651538299712d946a7d8bddc82aead5283c4a5

Pith citing papers

No inbound Pith citation observations are available.