Pith. sign in

Paper Citation Record · LEDGER

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System

As of 16 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2508.11885.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11885 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:29:34.062251Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1342987-a4d9-4c15-afab-6330f34a3b28 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in Neural Informa- tion Processing Systems, 35:23716–23736, 2022.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Flamingo: a visual language model for few-shot learning.Advances in Neural Informa- tion Processing Systems, 35:23716–23736, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.516943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.798435Z digest=sha256:6a449aa02688803d5e1fc61112d08004959ce33ac3d63c91af2246402d752357

Observation c669e027-0e8e-4318-a411-87b7b1f8ecef · outbound

This paper cites Deep Variational Information Bottleneck.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Deep Variational Information Bottleneck

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.804065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.804065Z digest=sha256:38b1690aec8dd731db8df5c60aff64017abaa447fda53bae5a25d530fd995e81

Observation 417d974b-8cff-4c08-86f8-c18136e04d28 · outbound

This paper cites Divprune: Diversity-based visual token pruning for large multimodal models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Divprune: Diversity-based visual token pruning for large multimodal models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.501792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.809188Z digest=sha256:0f1ca0387f4e49cb29a2f01c91b8524b5d277b9e4c222c33aa1f63ed4a71618d

Observation fa8cd2cc-21ab-4ef7-8593-b4369c6a903b · outbound

This paper cites A closer look at referring expressions for video object segmentation.Multimedia Tools and Applications, 82 (3):4419–4438, 2023.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System A closer look at referring expressions for video object segmentation.Multimedia Tools and Applications, 82 (3):4419–4438, 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.490357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.813249Z digest=sha256:4e606c5fc2b567f8275db7a4cf9268384310d927257510f4576c6ea320681137

Observation c8a16dc0-b6ce-4fb1-8c34-ec37cdc63fd4 · outbound

This paper cites Token Merging: Your ViT But Faster.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Token Merging: Your ViT But Faster

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.817206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.817206Z digest=sha256:d91ed23874f6af9716ae5c7773d6d2760223568d439f516d3f64b58eca6eda9f

Observation 89887543-2312-40b9-9386-9a6c66ff89e3 · outbound

This paper cites Matryoshka multimodal models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Matryoshka multimodal models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.479102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.932998Z digest=sha256:c344e199ef1684f76d702f2cefe6ce07eea406ae37f2d7052f05bae85e2588cf

Observation 3194bba4-5780-41b9-a2da-efe947697eff · outbound

This paper cites An im- age is worth 1/2 tokens after layer 2: Plug-and-play in- ference acceleration for large vision-language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System An im- age is worth 1/2 tokens after layer 2: Plug-and-play in- ference acceleration for large vision-language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.469859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.936978Z digest=sha256:b36ede0661e1d2ee1c0b52a414837d9f9f6d810be87135d14b6bf83657031f17

Observation 7f07d4f5-a8f3-4c80-82f1-b7de5527d8c0 · outbound

This paper cites Sequence complementor: Complementing transformers for time series forecasting with learnable sequences.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Sequence complementor: Complementing transformers for time series forecasting with learnable sequences

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.459928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.940928Z digest=sha256:2328f86fa23f96ce8f836f590ce242512ae36a0cd5333e4793557d448be5bb5f

Observation 128e75b4-12f7-4efb-8fc4-5e939be8866f · outbound

This paper cites Masked- attention mask transformer for universal image seg- mentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Masked- attention mask transformer for universal image seg- mentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.449726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.944480Z digest=sha256:4df4c30aef65f5539f312ca2bffeffdb4f2457b84fc64217547774337660f0ab

Observation 86784cca-0919-4f77-b1d6-d1a3a87d05d1 · outbound

This paper cites Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.438582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.948263Z digest=sha256:d4005b1f2744d289929e9a481870ee00f71448fe43f8902a3dd26392a7a840c4

Observation 819eefe8-fdf4-423e-bacb-9d06bef93de4 · outbound

This paper cites John Wiley & Sons, 2nd edition,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System John Wiley & Sons, 2nd edition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.427386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.951951Z digest=sha256:5a0487681c5550d3ff13752eb0110b89274b34187dc0ff6fb1786b6e10292ee0

Observation 45307e83-cfdc-4395-80f4-ed9d8a2d5e5d · outbound

This paper cites Phi-2: The surprising power of small language models.Microsoft Research Blog, 1:3, 2023.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Phi-2: The surprising power of small language models.Microsoft Research Blog, 1:3, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.416035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.955972Z digest=sha256:39f344029c96977069246df99bae8036657d9294f638a16a69254a7be353ae21

Observation 70165da1-020b-4f72-8c26-eee384ecae7e · outbound

This paper cites Segment anything.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Segment anything

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.404631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.959773Z digest=sha256:a15592222b58e44326f494024efc6ac944b62d285badd2e876e8d72a16e71e62

Observation d6dd719d-9f45-4fdc-b39d-8b796b6657d7 · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System LISA: Reasoning Segmentation via Large Language Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.963290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.963290Z digest=sha256:1e55cd6ca7ff6b0e4c18dbcdd527e66ac7eb4e29cf2c642508535ae517dcb256

Observation b5584475-201e-4788-ae9e-e6b8bd5bb732 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.967349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.967349Z digest=sha256:21f0c26ec62cca013c7ce2afaab3e046693a0050fa6bc6159e87b7d5b6cf6986

Observation 69159d6c-a40a-4b04-8776-ad1578e16ab0 · outbound

This paper cites Referring transformer: A one-step approach to multi-task visual grounding.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Referring transformer: A one-step approach to multi-task visual grounding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.392940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.971599Z digest=sha256:4587af5d4c5b4e879def6868dd4d9cccc7fd6164454723220d8e22fb7870e0df

Observation e61207bf-1808-4bdb-a967-d6947c340696 · outbound

This paper cites To- kenpacker: Efficient visual projector for multimodal llm.International Journal of Computer Vision, pages 1–19, 2025.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System To- kenpacker: Efficient visual projector for multimodal llm.International Journal of Computer Vision, pages 1–19, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.381446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.975868Z digest=sha256:4c834824d9f8a019b87d2b1120ab232b0d052c2477582362f0b86081468b0981

Observation e4bceed2-a8c3-467f-b214-0d01ecdc27b9 · outbound

This paper cites Boosting multimodal large language models with visual tokens withdrawal for rapid inference.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Boosting multimodal large language models with visual tokens withdrawal for rapid inference

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.370933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.979318Z digest=sha256:4d425462d09bf56496198104eb9ba6dd0d2bb8cbcb99cba5ff2ce63737cf7674

Observation a72d34e1-9562-447f-9e85-6886b6f6c7c3 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Swin transformer: Hierarchical vision transformer using shifted windows

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.360956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.982412Z digest=sha256:84ebd88528aa233ec93788df7107aa88ae8da0dd06b7e0a284ec8c8dc8fe7db4

Observation cb9f644e-fb0f-448d-afd3-3fa6d21b2bde · outbound

This paper cites Modeling context between objects for referring ex- pression understanding.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Modeling context between objects for referring ex- pression understanding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.350384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.985618Z digest=sha256:6ddb405508847474f7be60bc06006ab9ead1bd0f8b2ddc6d34b19edeb41c850d

Observation 18f123f2-0628-49c6-89d8-8e1ca92ee7d6 · outbound

This paper cites PerceptionGPT: Effectively Fusing Visual Perception into LLM.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PerceptionGPT: Effectively Fusing Visual Perception into LLM

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.989125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.989125Z digest=sha256:ca1975c98784f055e0cdfdf2776809834db7d1cc8a9bad074cb6d1008c1adb3a

Observation c34e7079-30e6-4004-92c1-01f0c58eccf7 · outbound

This paper cites PixelLM: Pixel Reasoning with Large Multimodal Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.993033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.993033Z digest=sha256:14bd8e4f6abad33987a14d36f17768874f59399dc4e4d1e88a0c8de5d7ee74d3

Observation d9d77504-c35e-4ef2-ab93-bd11a2850368 · outbound

This paper cites Urvos: Unified referring video object segmentation network with a large-scale benchmark.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Urvos: Unified referring video object segmentation network with a large-scale benchmark

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.340883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:33.996873Z digest=sha256:489f244683b90b1abee4085b526683961eb76b3e0977eade4451c72820c08079

Observation f4b03301-9631-425c-b52a-3bdea729dcae · outbound

This paper cites Llava-prumerge: Adaptive token re- duction for efficient large multimodal models.ICCV,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Llava-prumerge: Adaptive token re- duction for efficient large multimodal models.ICCV,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.330531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.000423Z digest=sha256:a27939f6cce4778a4ce1593e1798c6ea96720ca1dee9956db50787db3dffc1e8

Observation a5bc1ebb-aaef-4bb8-a004-3b6ea746b88e · outbound

This paper cites Contrastive grouping with transformer for referring image segmentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Contrastive grouping with transformer for referring image segmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.319946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.004669Z digest=sha256:4fa94937d01d8d04c4211277375cf96fc5bfcd04d35f321980b6ac745da018b5

Observation 07155c16-30d9-4768-9404-dcc028f130ad · outbound

This paper cites Deep learning and the information bottleneck principle.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Deep learning and the information bottleneck principle

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.307134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.008176Z digest=sha256:a314941c41acf3b7ae38d9065a9973c0a8cc696c79baa879c31afd764b8d9707

Observation bde20625-28c0-429e-9e9e-f5bd4619eb66 · outbound

This paper cites Cris: Clip-driven referring image segmentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Cris: Clip-driven referring image segmentation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.296740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.011745Z digest=sha256:bc124c7b31d48c5de016509abb01717780e1f9d8972ee93419adfe67ab9450e1

Observation 95dff278-d6b0-4fd9-a107-a561770720b9 · outbound

This paper cites LaSagnA: Language-based Segmentation Assistant for Complex Queries.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System LaSagnA: Language-based Segmentation Assistant for Complex Queries

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.015425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.015425Z digest=sha256:4066fec9addfacb04653048cd35fb36df2eb36936175454acaa0fb62b605025c

Observation a7f7e31d-8a5e-444a-9e1c-af8ca44499ff · outbound

This paper cites Instructseg: Unifying instructed visual segmentation with multi- modal large language models.ICCV 2025, 2025.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Instructseg: Unifying instructed visual segmentation with multi- modal large language models.ICCV 2025, 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.284853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.020519Z digest=sha256:86ec2b32644366641eaad5c31601cdba225c871b56d451d21abfaa6e892990da

Observation 337b2236-973c-46ee-ad98-70736df1d5e5 · outbound

This paper cites GSVA: Generalized Segmentation via Multimodal Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System GSVA: Generalized Segmentation via Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.025187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.025187Z digest=sha256:634257c4ed5e294d0b22b26ea16b2e14f1c5f52863bc172b37e4c1a4899d726c

Observation 555d31c1-f140-4e16-a401-0627bf54e0cf · outbound

This paper cites VISA: Reasoning Video Object Segmentation via Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System VISA: Reasoning Video Object Segmentation via Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.029272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.029272Z digest=sha256:f5ea0cd4eafc01bc63ad23c9bfe0ad0da9ddc8b4a977b06e682cd879d76adde2

Observation 4dac4429-0caa-4fba-a3ba-531d83d4e636 · outbound

This paper cites Visionzip: Longer is better but not necessary in vision language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Visionzip: Longer is better but not necessary in vision language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.272300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.033613Z digest=sha256:a13325eeddba214146a1e1998f1923f457db6ffc290154a20871e9b4ca2b5818

Observation 8a94a600-f29b-404d-8fb4-d65a86e297fe · outbound

This paper cites Fit and prune: Fast and training-free visual token pruning for multi-modal large language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Fit and prune: Fast and training-free visual token pruning for multi-modal large language models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.260031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.037936Z digest=sha256:79d3f8b93ea14a50552236973b7763e2a90bb476fa7c2aebf186f57251ed3f95

Observation 18f9d2a3-d968-471b-8379-df5177abe9e9 · outbound

This paper cites Modeling context in re- ferring expressions.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Modeling context in re- ferring expressions

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.247202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.041698Z digest=sha256:0ccd70774ebac15d5b4e1cba9281a3dfd7201312486b10c9a3790cda4621bb49

Observation 5382ee08-f445-477c-829b-c1a83b00a8c4 · outbound

This paper cites Sigmoid loss for language im- age pre-training.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Sigmoid loss for language im- age pre-training

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.234978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.045442Z digest=sha256:c1c7b9fc719f14a77e13ec9feada96789f14f10642ddbbac799cc3906d1c2af4

Observation b844293a-ca70-4fa6-93ae-5f2b30ab76a5 · outbound

This paper cites Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.049030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.049030Z digest=sha256:c85b73b91aaffe5fc9d31799e26dec36b30f1e05376fd145d76409e214fdc82e

Observation cb30834c-3eac-4f51-8643-68ed8c81b4de · outbound

This paper cites PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.053765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.053765Z digest=sha256:1331dbf0d23957ada055fc5a95fa9827921a7f1f55e4eb07c89176f304da7775

Observation 29e03bcb-c8e3-49f7-96ae-7cdd54b1217f · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.058509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.058509Z digest=sha256:a1a0df61019c4c8077048383416f1c94f420efe8b1d71f724fb16ad0a8cfca85

Observation 8f09e585-558a-4f4e-b564-6776761766f7 · outbound

This paper cites dog with its mouth open,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System dog with its mouth open,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.220714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T17:29:34.062251Z digest=sha256:3056c3eb3b2396cf4d5c93621be7fb8e99654d8491f110faa6e5a83354bfbeba

Pith citing papers

No inbound Pith citation observations are available.