Pith. sign in

Paper Citation Record · LEDGER

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting

As of 17 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2608.11692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11692 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:37:37.763247Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ab5cceb-4c71-4e7d-9272-8da36b3c4417 · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.690779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.690779Z digest=sha256:ff06c6b08d9dcd3d515d808a5ae4bfc4ee17d7fb0a4d610bcab9c35c6c7addd3

Observation b9bf3255-27ed-402c-94f3-4c0577ccd266 · outbound

This paper cites InFindings of the Association for Computational Linguistics: EMNLP 2023, 9318–9333.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting InFindings of the Association for Computational Linguistics: EMNLP 2023, 9318–9333

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:37:38.323587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T00:37:37.704729Z digest=sha256:3bc0cdbbef50f068c07fae4bb0c0100e9cdf9e2e39ff3a11594c1da351cdfbb2

Observation 3e8520e9-b3db-4a27-b9b2-f6f26e69c525 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.709205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.709205Z digest=sha256:c33482e7664a743daa1ee17cd27185b9a1d6faeef269f09413b7196a4f06f38d

Observation e652f8dd-5352-4e4d-aa24-659e96ccbef8 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting OpenVLA: An Open-Source Vision-Language-Action Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.713734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.713734Z digest=sha256:436cd288211fb1e4adb4cb782b6ee495dfa28841ec11b1660ab73e5cba8d7967

Observation 51bbf3c2-da75-40eb-b74e-2e4028d001e0 · outbound

This paper cites InProceedings of the IEEE/CVF international conference on computer vision, 4015–4026.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting InProceedings of the IEEE/CVF international conference on computer vision, 4015–4026

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:37:38.309905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T00:37:37.718478Z digest=sha256:ec91431b17e35c237022fdc9f485681b9f1efab9eaefe66d3fa124ef3d5a9828

Observation 3fb753b9-0fe8-4aa3-917c-c424329fd286 · outbound

This paper cites MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting MIA-DPO: Multi-Image Augmented Direct Preference Optimization For Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.723260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.723260Z digest=sha256:5f30a4f4e202a885dbff8c3946409713744e9444566cf60e5a82b1b0b0c5747f

Observation 4dc63216-ddde-45fb-aec0-a02b16d96103 · outbound

This paper cites Vaswani,A.;Shazeer,N.;Parmar,N.;Uszkoreit,J.;Jones,L.; Gomez,A.N.;Kaiser,Ł.;andPolosukhin,I.2017.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Vaswani,A.;Shazeer,N.;Parmar,N.;Uszkoreit,J.;Jones,L.; Gomez,A.N.;Kaiser,Ł.;andPolosukhin,I.2017

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:37:38.295423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T00:37:37.732044Z digest=sha256:a277b3d3e1bd13fddb5fd0ccb6431efcfc9e90828cafcab4c5029676deedc998

Observation 54dc09b7-aebb-423a-9b3d-caacdfa07516 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.740964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.740964Z digest=sha256:57dfca11f16628454e032bcca602b982d979a8795df1d259045be0837b9246b4

Observation 5ad4d5f8-8d0a-48fe-b1b1-2e3cf2886768 · outbound

This paper cites Yu, T.; Wang, Z.; Wang, C.; Huang, F.; Ma, W.; He, Z.; Cai, T.; Chen, W.; Huang, Y.; Zhao, Y.; et al.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Yu, T.; Wang, Z.; Wang, C.; Huang, F.; Ma, W.; He, Z.; Cai, T.; Chen, W.; Huang, Y.; Zhao, Y.; et al

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.745184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.745184Z digest=sha256:c0e98b37ceff1c4ea2429d0cd4d4f7415b58ea7dadfb6bf1e0e44fc07b389c30

Observation 56453918-3ec6-4d22-a7e4-772b5df82ca9 · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.749745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.749745Z digest=sha256:98f2c614ea58e0c96f3970da3b7ff6ec3cef9f97a2577cfc395e70c12b0144dd

Observation 9f0967a6-2c06-47f7-b2cf-f07f610f245b · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Multimodal Chain-of-Thought Reasoning in Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.754197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.754197Z digest=sha256:bc130df09e98e83226e99072afd44d860042a52ed040f6317866af3c136236a2

Observation 8ed7a7ac-2639-471a-bb04-ecffe359935a · outbound

This paper cites SVIT: Scaling up Visual Instruction Tuning.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting SVIT: Scaling up Visual Instruction Tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.758655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.758655Z digest=sha256:3a253c709334ab3f02edc40a3b6cb21fbf30f7dd3e1e2053429705320199be36

Observation d6769a7f-e629-4851-b91a-75ae5eeccb8e · outbound

This paper cites VEGA: Learning Interleaved Image-Text Comprehension in Vision-Language Large Models.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting VEGA: Learning Interleaved Image-Text Comprehension in Vision-Language Large Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.763247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.763247Z digest=sha256:fb02f1ea8af6ec2210942e663ec211b6b3315d3772302f9dd093ad5426d7c8a1

Observation 7687aeb8-95f7-4836-abfb-7e3ed22b3efc · outbound

This paper cites MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.736322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.736322Z digest=sha256:2f60fbda25bdfbe9019683f9822a05b5a9f0f74f3219e4a74341bcceb4268758

Observation 194b83cd-941b-4aee-aabb-d5bdec776142 · outbound

This paper cites SimCSE: Simple Contrastive Learning of Sentence Embeddings.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting SimCSE: Simple Contrastive Learning of Sentence Embeddings

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.695698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.695698Z digest=sha256:7aaa45ff90f1b159a1e38f7c7f059df50abac2cc625926072b2fb5721d837b57

Observation e09af384-f4b1-4135-bd92-9edfb0b24bf1 · outbound

This paper cites Rephrase and Respond: Let Large Language Models Ask Better Questions for Themselves.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Rephrase and Respond: Let Large Language Models Ask Better Questions for Themselves

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.686110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.686110Z digest=sha256:125642c46e166c347ebfebc3e2d7ad0d394760758dbe46be647d7075b74af540

Observation 2b587c88-39b0-4a6b-abf9-8fffb9632a08 · outbound

This paper cites Pixtral 12B.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Pixtral 12B

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.680904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.680904Z digest=sha256:b776ae077830d539efa7babf45b75777e7c936c632054d8027c3229dba49ac19

Observation 470c10ff-8a77-438f-9442-ef0896ca14b9 · outbound

This paper cites arXiv:2509.06266.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting arXiv:2509.06266

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.700455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.700455Z digest=sha256:2e78afdd54248664130bf814a116c24606ae07b703d224967532913583afd8b3

Observation e266b8a4-7ba7-4dad-9593-950f50fa6f82 · outbound

This paper cites Tian, X.; Zou, S.; Yang, Z.; and Zhang, J.

HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting Tian, X.; Zou, S.; Yang, Z.; and Zhang, J

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-16T00:37:37.727540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:37:37.727540Z digest=sha256:4d54181c564e274e9fddf4604616ffcc7e70f030552806e38ccc16088b12a1a4

Pith citing papers

No inbound Pith citation observations are available.