Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T22:27:52.861022Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2607.05859.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T22:27:52.861022Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c4c887a8-553f-4055-92b9-b475a4163984 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a86a0f61-0c8d-4e7c-8256-b81a4a0d4b11 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a8204efc-3275-425c-b072-f346397254a0 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Spice: Semantic propositional image caption evaluation, in: European conference on computer vision, Springer
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7d62fa39-4409-48df-8b4a-4e9159c111c0 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5609b157-c88b-48ac-a63b-e511b5f6377f · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c04100e3-cd8e-4ba3-a284-21d5c58c1a73 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Context-awarevision-languagemodelagentenriched with domain-specific ontology for construction site safety monitoring
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation db022c64-081e-45f2-8731-59aa3cee9036 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Enhancing vision-language model for construction safety inspection via visually grounded reasoning and reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 15e6ae3a-8d24-4bda-a6f4-97f1cf8b41fa · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Augmented reality, deep learning and vision-language query system for construction worker safety
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b97a474a-15c1-478c-a0d9-7f4144f25c6a · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a8b5b3c8-cc53-4915-913d-0bbd6b662b3b · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 50a78136-ca05-4dfb-b2c1-b1ec9f796d1c · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Tailored vision-language framework for automated hazard identification and report generation in construc- tion sites
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 79dd781b-948c-4f84-b7e2-c20924fd489e · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Canmultimodallargelanguagemodelstrulyperform multimodalin-contextlearning?,in:2025IEEE/CVFWinterConferenceonApplicationsofComputerVision(WACV),IEEE.pp.6000–6010
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1db83eaa-394d-4813-a22f-8ed71a09fd3d · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Are large pre-trained vision language models effective construction safety inspectors
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e8539243-9771-4b00-9826-177b464188bd · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Expandingperformanceboundariesof open-source multimodal models with model, data, and test-time scaling
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 41ae7e5f-ee17-45c6-a353-c484f16f95ec · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Chain of Thought Prompt Tuning in Vision Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1d220ae5-72a7-43e9-8af5-4a2abd499063 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Intelligent virtual assistants with llm-based process automation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 64dbea47-ccd8-4b3e-a774-abadb1139a23 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring LoRA: Low-Rank Adaptation of Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 34e022b9-5582-4941-837f-beb8b64786bd · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8cd8e49b-6f76-4ef8-9070-523d55ac4409 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Vru-accident: A vision-language benchmark for video question answering and dense captioning for accident scene understanding
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 39d9e36e-2391-4d2c-8153-146d5dd508ea · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Safe-llava:Aprivacy-preservingvision-languagedatasetandbenchmarkforbiometricsafety
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 33641ca8-8c68-45d1-b53b-fddf2cd063c1 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Res-bench: Benchmarking the robustness of multimodal large language models to dynamic resolution input, in: Proceedings of the AAAI Conference on Artificial Intelligence, pp
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ea9338c7-0f61-4527-99cb-c07ee5d7f74b · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Construction site fall hazard identification and automated captioning using adapted vision-language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f3830f6e-72a3-44e8-b6b8-88a0da821de7 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring AdaptVision: Efficient vision-language models via adaptive visual acquisition.arXiv preprint arXiv:2512.03794, 2025
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3761a0a6-e163-4383-8877-fd02b169a6b3 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Llava-next: Improved reasoning, ocr, and world knowledge
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 40e2f463-7018-460c-b20b-fa0004ecc42c · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Visual instruction tuning, in: Oh, A., Naumann, T., Globerson, A., Saenko, K., Hardt, M., Levine, S
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0782e0f5-ef7b-4b50-bb0d-c21cd184fb6a · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 675779f7-775f-42f0-8e63-74351b6ab11e · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring The Llama 3 Herd of Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7b142371-113d-4e6a-bd64-fcc9f6097183 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring GPT-4 Technical Report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b911e575-3131-4eb4-a6d0-3ddc7845e535 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Privacy-Aware Visual Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 81a41807-1877-427e-a3fc-02ad0209459e · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Vlm-robustbench: A comprehensive benchmark for robustness of vision-language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fa13d331-0a3d-4d9b-a0a6-27624befe93f · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Vision-language models for edge networks: A comprehensive survey
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4ef34baf-f4b8-43b6-b59e-71e9e1cfc553 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Upop: Unified and progressive pruning for compressing vision-language transformers, in: International Conference on Machine Learning, PMLR
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 01db1329-854a-4e32-a16d-f8f63d9592ac · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Real-time safety detection on construction sites using a vision-language and nlp- based model
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ef1d064f-93a1-4846-9930-8abc511a1541 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring A double thinking enabled visual language model for open-set construction site safety inspections
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5a347992-f026-482e-9bb5-576ef879123e · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Region-level vision-language model for detecting distraction behavior and mobility attributes of vulnerable road users
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a9ca898f-8578-4960-ac15-8a82a00adb27 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Visual question answering-based referring expression segmentation for construction safety analysis
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 03238cdb-052e-49f1-8880-de628907165f · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Selma: A speech-enabled language model for virtual assistant interactions
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 34c9128a-0134-4b53-9856-11c20b6f6627 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b89f6176-1db0-4b14-99ce-d82dc42d1d68 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 845780f1-5d67-4024-8896-568df97ed835 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Efficient Vision-Language Models by Summarizing Visual Tokens into Compact Registers
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 55cfce7d-25f9-42f7-ad8f-44484ae995b2 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 309f9b4b-22ac-4542-9287-bb29ee9c9b1e · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Llava-cot: Let vision language models reason step-by-step, in: Proceedings of the IEEE/CVF International Conference on Computer Vision
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ba0ffa3c-a517-4530-aa69-323d59955cd3 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Edgevideoanalytics:Asurveyonapplications,systemsandenablingtechniques
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b0423c76-7678-438b-bf77-54c87ea12c95 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Qwen2.5 Technical Report
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1c8752b1-5b71-4006-b0fb-0b9598995377 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Vision transformer-based visual language understanding of the con- struction process
Reference 45
Source-reported events for the cited work
correction dated 2024-05-21. Source: crossref record 10.1016/j.aej.2024.05.064->10.1016/j.aej.2024.05.015:correction, observed 2026-07-11T03:14:22.632143+00:00. This notice travels one citation hop only.
Observation d220c57d-6314-422a-94b9-67342edbee60 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Visionzip: Longer is better but not necessary in vision language models, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3f04773a-9683-4cc4-9277-5bd91c9cf822 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Visionthink: Smart and efficient vision language model via reinforcement learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation aec977bd-d0e7-416b-90e9-f9dca407730b · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring BERTScore: Evaluating Text Generation with BERT
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f908bd59-8cb1-41b8-b779-8b2e0a937fef · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 01c5da41-714e-464d-9706-0f192a0a4c77 · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Multimodal Chain-of-Thought Reasoning in Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 370f17ac-fcac-4490-a6de-21e2d736c31f · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Mmicl: Empowering vision-language model with multi-modal in-context learning, in: International Conference on Learning Representations, pp
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 31197953-97f3-4862-a023-8cb7add18a2a · outbound
AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring Internvl3: Exploring advanced training and test-time recipes for open-source multimodal models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
No inbound Pith citation observations are available.