Pith. sign in

Paper Citation Record · LEDGER

GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2406.09781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.09781 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:45:17.713849Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T16:24:41.484573Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6393b5fd-af2c-483f-bb4e-852502383b96 · inbound

VisGraphVar: A Benchmark Generator for Assessing Variability in Graph Analysis Using Large Vision-Language Models cites this paper.

VisGraphVar: A Benchmark Generator for Assessing Variability in Graph Analysis Using Large Vision-Language Models GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T14:52:57.291073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:52:57.291073Z digest=sha256:1e844b9e00748903cf5d6811ffd57df019834d6522a9ba9c20f0fe24c5321fb4

Observation 87a6c221-2386-4a65-871c-cc21054f86ae · inbound

TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation cites this paper.

TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T22:54:20.647753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:54:20.647753Z digest=sha256:b198a237a28268a9c0470db9b7de0cd4fad88f22ac61fdebc226a9e7f81862da

Observation 6bbaa4dd-644f-4626-8b5b-c5de0f8e0c1a · inbound

Leveraging Large Language Models for Generating Labeled Mineral Site Record Linkage Data cites this paper.

Leveraging Large Language Models for Generating Labeled Mineral Site Record Linkage Data GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:59:18.260001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:59:18.260001Z digest=sha256:c4c9ad59ad49304a56ea45df926f2e16d83b9e6703ef29e6abcbea99e7f5f6af

Observation e98a992b-af01-4104-b6ec-704ab9ef44fc · inbound

RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering cites this paper.

RAMQA: A Unified Framework for Retrieval-Augmented Multi-Modal Question Answering GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-10T16:24:41.489971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-10T16:24:40.869195Z digest=sha256:37089b961466f883aedf8944a7bc3c5904ecb3c208d11600b61c8880bd39a8b3

Observation 459f9a35-a0bf-40e3-bc93-dc13061978b3 · inbound

Towards Scalable Human-aligned Benchmark for Text-guided Image Editing cites this paper.

Towards Scalable Human-aligned Benchmark for Text-guided Image Editing GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T04:45:17.713849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:45:17.713849Z digest=sha256:8fc7c6101a683b7925cf2665f68707ab496ac9401cabebb1c3fff5ec47cc9293