Pith. sign in

Paper Citation Record · LEDGER

HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2506.03922.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03922 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:12.998437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:29:59.531643Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cc4463f7-9f06-4b0f-add8-442fc11eea2e · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 207

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:12.998437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:12.998437Z digest=sha256:f27bbd2791824fb4871bf2badcd692539ca96e3f1243a117bf6fe1def072db43

Observation 45c223d0-eb99-4368-90ca-c3fc613922f9 · inbound

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI cites this paper.

KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speech-to-Speech Conversational AI HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:53.978211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:49:53.978211Z digest=sha256:4713466be8b2e8a84409f8e18d5e8e3d4e09f138c40ebfbc38fb63c61c7bf0a8

Observation 6b5301ed-f230-46c9-b5e6-a0b5cc23ad5f · inbound

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine cites this paper.

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T18:32:04.537458Z digest=sha256:bfa01c7f98045c1b950d4813efe3c98162b58275c196ff42b6915bd3b1d0abc1

Observation bd8cc177-28b7-40c9-a578-96508a619b44 · inbound

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond cites this paper.

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T12:02:07.027775Z digest=sha256:dc0d0209f0496aa7a04d2eb8515360ccb281ebfc1b9c875eb11a46a22243e084

Observation 45873610-c958-4a40-8b84-4db700ad89e1 · inbound

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond cites this paper.

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-04T17:29:43.764085Z digest=sha256:d2170b3b12b43b24565236d59eab63ce1bbaa762ffd313e0a82733014a89a777

Observation c3d0a8d3-b80a-4149-81f1-88b93ac0756e · inbound

BenCSSmark: Making the Social Sciences Count in LLM Research cites this paper.

BenCSSmark: Making the Social Sciences Count in LLM Research HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T17:02:15.446874Z digest=sha256:ba165e87c0405c08cc4a5d1c6bb0f1614985d8c52c59c93fecee470fbafa1bee

Observation d761669d-a7f3-482b-a0e8-b8b0b6c6f9eb · inbound

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench cites this paper.

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T15:31:25.079191Z digest=sha256:7ec6534effb8d6dfd59a878b19e7c1a13956849a9ce9e58a5a98f1929b1201ef

Observation e18acb3e-7388-4a86-a7c2-219f0e7a4b2c · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:10:32.522453Z digest=sha256:9a2f3a8af48a485516228156ff4b7d8dca0a0b03d50632302d97a5fce08eb7e9

Observation 45a9b7c3-7ce1-40dd-8ffe-f8184af67398 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T16:40:22.441025Z digest=sha256:5a89dd06a5563680a23cfadcf5744fa5b16c04f29fa2b9d0d704b09a2abfbf47

Observation d30b77b3-aa77-406f-8e46-e568ea9681e5 · inbound

NormAct: Benchmarking Embodied Agents' Proactive Compliance with Unspoken Social Norms cites this paper.

NormAct: Benchmarking Embodied Agents' Proactive Compliance with Unspoken Social Norms HSSBench: Benchmarking Humanities and Social Sciences Ability for Multimodal Large Language Models

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-08-12T02:23:50.971990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-29T04:26:43.018373Z digest=sha256:44b8f272f7c491c41dd201dd4276816fa169ffb83215bfe0aada395fe17095dc