Pith. sign in

Paper Citation Record · LEDGER

By My Eyes: Grounding Multimodal Large Language Models with Sensor Data via Visual Prompting

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2407.10385.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.10385 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:14:20.410753Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T18:48:17.204802Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 86ec7702-2def-4a6e-85f1-eb67fbd5026d · inbound

Time2Lang: Bridging Time-Series Foundation Models and Large Language Models for Health Sensing Beyond Prompting cites this paper.

Time2Lang: Bridging Time-Series Foundation Models and Large Language Models for Health Sensing Beyond Prompting By My Eyes: Grounding Multimodal Large Language Models with Sensor Data via Visual Prompting

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T12:14:20.410753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T12:14:20.410753Z digest=sha256:0fd3eee9fbe4ca22050e66853cc61f4d1d4450e1d17e3e16b5bb97dc368747f1

Observation e50ae3c4-1777-4649-9969-ebeb8cdbb092 · inbound

LENS: LLM-Enabled Narrative Synthesis for Mental Health by Aligning Multimodal Sensing with Language Models cites this paper.

LENS: LLM-Enabled Narrative Synthesis for Mental Health by Aligning Multimodal Sensing with Language Models By My Eyes: Grounding Multimodal Large Language Models with Sensor Data via Visual Prompting

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:48:17.207135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T18:46:45.455333Z digest=sha256:501e26cb2b38b16b14bf42e36d03c270fd321af71dae30f2ff41ad46620cc439